Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
phazonoverload
12 days ago
|
parent
|
context
|
favorite
| on:
My local model setup on an M4 Pro Mac Mini
I'm the author - hello! Added to the post! Qwen averages 325 tok/s in processing prompts, and 34 tok/s in token generation. That isn't instant, but it's quick enough that I never really think about it.
help
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: