AI커뮤니티
E

Hot take: everyone is overfitting to benchmarks and it shows

자유 @Emily Watson · 2일 전 · English
공유:
Ship a model, top the leaderboard, get the press release, repeat. Meanwhile real usage is full of edge cases none of the benchmarks touch. I'd rather have a model that's honest about what it can't do.

댓글 0

첫 댓글을 남겨보세요.

로그인 후 댓글을 쓸 수 있어요.