Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Yes, 15-30 t/sec is pretty slow for local models so I recommend running local LLM tasks overnight where (vs paid plans) there isn't a risk of chewing through your token budget from a rogue loop or sub-agent. Even if it takes hours, you're sleeping anyway so no concern. herdr + pi works great for this but there are lots of harnesses.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: