Home Projects tiny-llm
tiny-llm
Python

tiny-llm

learn LLM inference system on Apple Silicon for systems engineers: build a tiny vLLM + Qwen

by skyzh · GitHub
Stars
Forks
License
Created
Last commit
Category
Language
courselarge-language-modelllmApache-2.0Python
View on GitHub
In plain words

Follow a course to understand how to serve large language models on Apple Silicon, from start to finish.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
tiny-llm — GitHub preview card
📈 Star history
5k4k
2026-07-072026-08-31
📈 Track tiny-llm

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

learn LLM inference system on Apple Silicon for systems engineers: build a tiny vLLM + Qwen

tiny-llm has 4.5k stars on GitHub. It has been forked 369 times. tiny-llm is written mainly in Python. It has been in active development since 2025. tiny-llm is available under the Apache-2.0 license. Its main topics are course, large-language-model, llm, python.

Frequently asked questions

What is tiny-llm?

learn LLM inference system on Apple Silicon for systems engineers: build a tiny vLLM + Qwen

Is tiny-llm open source?

tiny-llm is an open-source project. It is released under the Apache-2.0 license.

Is tiny-llm free?

Yes. tiny-llm is free and open source — you can use, modify and self-host it.

What license does tiny-llm use?

tiny-llm is available under the Apache-2.0 license.

What language is tiny-llm written in?

tiny-llm is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — tiny-llm

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=skyzh-tiny-llm)](https://olud.ai/project/skyzh-tiny-llm.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.