Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I'm the author - hello! Added to the post! Qwen averages 325 tok/s in processing prompts, and 34 tok/s in token generation. That isn't instant, but it's quick enough that I never really think about it.
 help



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: