
Foundation Text-Generation Models Below 360M Parameters
Great candidates for fine-tuning targeting Wllama and Transformers.js for mobile devices, ordered by number of parameters.
Text Generation • Updated • 125k • 50Note License: apache-2.0 Context Length: 8k
PleIAs/Pleias-350m-Preview
Updated • 409 • 22Note License: apache-2.0 Context Length: 2k
OuteAI/Lite-Oute-1-300M
Text Generation • Updated • 322 • 7Note License: apache-2.0 Context Length: 4k
keeeeenw/MicroLlama
Text Generation • Updated • 1.58k • 48Note License: apache-2.0 Context Length: 2k
cerebras/Cerebras-GPT-256M
Text Generation • Updated • 265 • 25Note License: apache-2.0 Context Length: 2k
UUFO-Aigis/Pico-OpenLAiNN-250M
Updated • 2 • 3Note License: apache-2.0 Context Length: 2k
upstage/TinySolar-248m-4k
Text Generation • Updated • 607 • 7Note License: apache-2.0 Context Length: 4k
M4-ai/TinyMistral-248M-v3
Text Generation • Updated • 82 • 8Note License: apache-2.0 Context Length: 2k
MiniLLM/MiniPLM-llama3.1-212M
Text Generation • Updated • 167 • 4Note License: apache-2.0 Context Length: 1k
MiniLLM/MiniPLM-Qwen-200M
Text Generation • Updated • 287 • 4Note License: apache-2.0 Context Length: 1k
princeton-nlp/Sheared-Pythia-160m
Text Generation • Updated • 7 • 4Note License: apache-2.0 Context Length: 2k
JackFram/llama-160m
Text Generation • Updated • 362k • 34Note License: apache-2.0 Context Length: 2k
SmallDoge/Doge-160M
Text Generation • Updated • 49 • 5Note License: apache-2.0 Context Length: 2k
EleutherAI/pythia-160m
Text Generation • Updated • 95.8k • 31Note License: apache-2.0 Context Length: 2k
openai-community/gpt2
Text Generation • Updated • 10.9M • 2.75kNote License: mit Context Length: 1k
HuggingFaceTB/SmolLM2-135M
Text Generation • Updated • 695k • 94Note License: apache-2.0 Context Length: 8k
amd/AMD-Llama-135m
Text Generation • Updated • 7.19k • 113Note License: apache-2.0 Context Length: 2k
MiniLLM/MiniPLM-Mamba-130M
Text Generation • Updated • 33 • 3Note License: apache-2.0 Context Length: 1k
EleutherAI/gpt-neo-125m
Text Generation • Updated • 181k • 205Note License: mit Context Length: 2k
cerebras/Cerebras-GPT-111M
Text Generation • Updated • 11.7k • 76Note License: apache-2.0 Context Length: 2k
BEE-spoke-data/smol_llama-101M-GQA
Text Generation • Updated • 555 • 28Note License: apache-2.0 Context Length: 1k
UUFO-Aigis/Pico-OpenLAiNN-100M
Updated • 1 • 1Note License: apache-2.0 Context Length: 2k
Felladrin/Qwen2-96M
Text Generation • Updated • 15 • 2Note License: apache-2.0 Context Length: 8k
Felladrin/Minueza-2-96M
Text Generation • Updated • 116 • 6Note License: apache-2.0 Context Length: 4k
distilbert/distilgpt2
Text Generation • Updated • 3.38M • 531Note License: apache-2.0 Context Length: 1k
weiser/82M-0.4
Text Generation • Updated • 70Note License: apache-2.0 Context Length: 1k
BEE-spoke-data/smol_llama-81M-tied
Text Generation • Updated • 11 • 6Note License: apache-2.0 Context Length: 1k
EleutherAI/pythia-70m
Updated • 110k • 68Note License: apache-2.0 Context Length: 2k
JackFram/llama-68m
Text Generation • Updated • 485k • 27Note License: apache-2.0 Context Length: 2k
OuteAI/Lite-Oute-1-65M
Text Generation • Updated • 48 • 9Note License: apache-2.0 Context Length: 2k
SmallDoge/Doge-60M
Text Generation • Updated • 42 • 4Note License: apache-2.0 Context Length: 2k
Felladrin/Minueza-32M-Base
Text Generation • Updated • 27 • 18Note License: apache-2.0 Context Length: 2k
GerbilLab/Gerbil-A-32m
Text Generation • Updated • 13 • 2Note License: apache-2.0 Context Length: 2k
EleutherAI/pythia-31m
Text Generation • Updated • 25.2k • 5Note License: apache-2.0 Context Length: 2k
SmallDoge/Doge-20M
Text Generation • Updated • 41 • 9Note License: apache-2.0 Context Length: 2k
EleutherAI/pythia-14m
Text Generation • Updated • 208k • 23Note License: apache-2.0 Context Length: 2k