No AI summary available for this article.
Why It Matters
Hey HN, Anders and Tom here.
Provenance
Discovered via Hacker News and published by GitHub.
Key Claims
Original description
Hey HN, Anders and Tom here. We're building Magnitude, an inference engine for agents that optimizes itself to run as fast as possible on your hardware. It works on Mac, Linux, and Windows on any hardware and is up to 2x faster than llama.cpp. We're both software engineers and previously built an open source browser agent to 4k+ GH stars and 100k+ downloads. We increasingly wanted to run it on local models, but found that no inference engine worked for our use case. Inference engines today all make a performance tradeoff. They are either: - Built for batched inference on datacenter hardware at...
Discovered via Hacker News
Community-ranked links and discussion from the HN front page.
Publisher: github.com
ID: 49911995 · Indexed about 1 hour ago