Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I'm watching people host models on things like LangSmith and OpenRouter for a fraction of the cost you are talking about. We have other people reporting their M4 Macs providing them with performance close to what they get with ChatGPT and Claude all locally with just a 24 GB M4 Mac. We already spend money on laptops. I can put in a ticket for an M4 Macbook Pro from IT right now.


Fantastic. I run local models too. This was specifically about APIs


Why do you think costs will go up when the competition is increasing while hardware prices per compute cycle go down? That's the part I don't get.


Going from a fixed monthly fee to a per token billing structure is not a cost reduction, and we are watching that happen in real time.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: