New ask Hacker News story: Ask HN: How come everyone is an LLM expert?

Ask HN: How come everyone is an LLM expert?
3 by delis-thumbs-7e | 3 comments on Hacker News.
It is a bit strange how new model is released and hour after there is commentators declaring it complete trash and embarrassment to the AI industry, or the best thing since sliced bread. Surely they have not had the opportunity to test the ins and outs of the model yet? Or do people just blindly trust benchmarks as if they were not pretty easy to manipulate, as research has shown quite a few times now? Or is it just all vibe? So how do you measure how one model is better than another?

Comments

Popular posts from this blog

New ask Hacker News story: Ask HN: Releasing code under AGPLv3, but want to block LLM reconstruction?