Private on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometric lock, tamper detection. + GPU Acceleration
Chat with an AI on your Android device without needing an internet connection, ensuring your privacy.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimePrivate on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometric lock, tamper detection. + GPU Acceleration
OfflineLLM has 225 stars on GitHub. It has been forked 21 times. OfflineLLM is written mainly in Kotlin. It has been in active development since 2026. Its main topics are android, android-ai, android-ai-app, android-llm.
Private on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometric lock, tamper detection. + GPU Acceleration
OfflineLLM is an open-source project.
Yes. OfflineLLM is free and open source — you can use, modify and self-host it.
OfflineLLM is written mainly in Kotlin.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/jegly-offlinellm.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.