Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

But the end user doesn't care about flops for closed models. All they care about is how much it ends up costing them.


Which is why I said that cost might be a better metric than output tokens. But even that is somewhat misleading -- because there isn't really a single price -- there are a variety of offers / deals / subsidies, including subscription plans.

But if we're actually talking benchmarking intelligence -- as evidenced by "Artificial Analysis Intelligence Index" and the rest of the reporting -- then comparing systems with different amounts of recurrence on output tokens is flawed.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: