Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

We could make LLM inference 100x cheaper to run at home efficiently, but that solution might need to be updated every 1-2 years, whereas current GPU are useful for various others tasks and last longer
 help



> We could make LLM inference 100x cheaper to run at home efficiently

do you genuinely think that's going to happen?




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: