The short version
A cross-platform inference engine that can run small machine-learning models in browsers.
From the HackerLinks archive
A cross-platform inference engine that can run small machine-learning models in browsers.
The short version
A cross-platform inference engine that can run small machine-learning models in browsers.
Why it caught our attention
A commenter reported using it for small in-browser models.
Where it surfaced on Hacker News
“We use the ONNX runtime for small models in the browser”
seamossfet · recommendation · production use
Direct HN comment
WebLLM: high-performance in-browser LLM inference engine