olla
High-performance lightweight proxy and load balancer for LLM infrastructure. Intelligent routing, automatic failover and unified model discovery across local and remote inference backends.
Not yet testedsource: GitHubhomepageGoApache-2.0commit bb0612193168
Go, Apache-2.0 licensed. The project labels itself: ai, amd, golang, intel, llama cpp, llamacpp, llm inference and llm proxy.
olla has not been verified yet.
Measured by Argusic on a fresh machine every time. Every number links to its evidence.
At a glance
Subject data from GitHub, linked at the top of this page, refreshed . Test data by Argusic (CC BY 4.0); every number links to a run page with the full log, the recording, and their sha256 hashes.
Also tested, in the same area
Every one of these was installed and run by Argusic on a clean machine. Nothing appears here that was not tested.
Run history
No recorded runs.
Topics (from GitHub)
aiamdgolangintelllama-cppllamacppllm-inferencellm-proxyllm-routerllm-routinglmstudiolocal-aimlxnvidiaollamaproxyself-hostedself-hosted-aisglangvllm
Embed the badge
Markdown for the project README. It links back here; terms on the terms page.
[](https://argusic.com/subject/olla)Questions
- Does olla run?
- olla has not been fully verified yet. No recorded run has produced a verdict yet.
- How did Argusic test olla?
- On a fresh, disposable machine, with every command recorded. 0 attempts are recorded, and the full method is on the methodology page.
- Where is the evidence for olla?
- All 0 recorded runs are on this page, each linking to its full log and terminal recording, stored with a sha256 fingerprint so it cannot be quietly altered.