Home Projects nanodl
nanodl
Python

nanodl

JAX library for training sub-4B foundation models for edge

by HenryNdubuaku · GitHub
Stars
Forks
License
Created
Last commit
Category
Language
attentionattention-mechanismdeep-learningMITPython
View on GitHub
In plain words

Train smaller versions of advanced AI models efficiently for specific tasks using a JAX library.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
nanodl — GitHub preview card
📈 Star history
307306
2026-07-202026-08-31
📈 Track nanodl

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

JAX library for training sub-4B foundation models for edge

nanodl has 306 stars on GitHub. It has been forked 13 times. nanodl is written mainly in Python. It has been in active development since 2023. nanodl is available under the MIT license. Its main topics are attention, attention-mechanism, deep-learning, distributed-training.

Frequently asked questions

What is nanodl?

JAX library for training sub-4B foundation models for edge

Is nanodl open source?

nanodl is an open-source project. It is released under the MIT license.

Is nanodl free?

Yes. nanodl is free and open source — you can use, modify and self-host it.

What license does nanodl use?

nanodl is available under the MIT license.

What language is nanodl written in?

nanodl is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — nanodl

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=henryndubuaku-nanodl)](https://olud.ai/project/henryndubuaku-nanodl.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →
self-attention-cv
Implementation of various self-attention mechanisms focused on computer vision. Ongoing reposit…
1.2k · artificial-intelligence
sharesattentionattention-mechanism
performer-pytorch
An implementation of Performer, a linear attention-based transformer, in Pytorch
1.2k · artificial-intelligence
sharesattentionattention-mechanism
lambda-networks
Implementation of LambdaNetworks, a new approach to image recognition that reaches SOTA with le…
1.5k · artificial-intelligence
sharesattentionattention-mechanism
Deepdive-llama3-from-scratch
Achieve the llama3 inference step-by-step, grasp the core concepts, master the process derivati…
631 · attention
sharesattentionattention-mechanism
pytorch-original-transformer
My implementation of the original transformer model (Vaswani et al.). I've additionally include…
1.1k · attention
sharesattentionattention-mechanism
pytorch-GAT
My implementation of the original GAT paper (Veličković et al.). I've additionally included the…
2.7k · attention
sharesattentionattention-mechanism
neural_sp
End-to-end ASR/LM implementation with PyTorch
594 · asr
sharesattentionattention-mechanism
seq2seq-summarizer
Pointer-generator reinforced seq2seq summarization in PyTorch
358 · abstractive-summarization
sharesattentionattention-mechanism
Paddle
PArallel Distributed Deep LEarning: Machine Learning Framework from Industrial Practice (『飞桨』核心…
24k · deep-learning
sharesdistributed-training
HandyRL
HandyRL is a handy and simple framework based on Python and PyTorch for distributed reinforceme…
304 · deep-learning
sharesdistributed-training

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.