Ask HN: How come everyone is an LLM expert?

  • Posted 3 hours ago by delis-thumbs-7e
  • 2 points
It is a bit strange how new model is released and hour after there is commentators declaring it complete trash and embarrassment to the AI industry, or the best thing since sliced bread. Surely they have not had the opportunity to test the ins and outs of the model yet? Or do people just blindly trust benchmarks as if they were not pretty easy to manipulate, as research has shown quite a few times now? Or is it just all vibe?

So how do you measure how one model is better than another?

3 comments

    Loading..
    Loading..
    Loading..