Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.


Using more tokens is a significant problem if you pay per token?


Not if the price per token is significantly lower.

Also this arm of the discussion was about speed, not price.


> Using more tokens is a significant problem if you pay per token?

Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: