Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent model selection.
Run AI models on demand with a system that optimizes memory use by loading and unloading models as needed.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeStateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent model selection.
Chameleon has 42 stars on GitHub. It has been forked 3 times. Chameleon is written mainly in Rust. It has been in active development since 2026. Chameleon is available under the MIT license. Its main topics are ai-infrastructure, generative-ai, latency-optimization, llm.
Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent model selection.
Chameleon is an open-source project. It is released under the MIT license.
Yes. Chameleon is free and open source — you can use, modify and self-host it.
Chameleon is available under the MIT license.
Chameleon is written mainly in Rust.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/megeezy-chameleon.html)