Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
ttoinou
7 days ago
|
parent
|
context
|
favorite
| on:
Kimi-K3 on HuggingFace
We could make LLM inference 100x cheaper to run at home efficiently, but that solution might need to be updated every 1-2 years, whereas current GPU are useful for various others tasks and last longer
help
nektro
6 days ago
[–]
> We could make LLM inference 100x cheaper to run at home efficiently
do you genuinely think that's going to happen?
reply
Consider applying for YC's Fall 2026 batch!
Applications
are open till July 27.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: