AI & ML interests
Hardware-aware AI Model Optimization
Recent Activity
View all activity
Optimized Kimi models
Mixture-of-Experts Large Language Models with Advanced Quantization
-
nota-ai/Solar-Open-100B-NotaMoEQuant-NVFP4
Text Generation • 59B • Updated • 242 • 23 -
nota-ai/Solar-Open-100B-Nota-FP8
Text Generation • 103B • Updated • 252 • 46 -
nota-ai/Solar-Open-100B-NotaMoEQuant-Int4
Text Generation • 2B • Updated • 586 • 62 -
nota-ai/Qwen3-30B-A3B-NotaMoEQuant-Int4
Text Generation • 0.6B • Updated • 27 • 8
Block-removed Knowledge-distilled SD models; https://github.com/Nota-NetsPresso/BK-SDM
Compressed NVIDIA Models (Nemotron and Cosmos)
Solar Open models officially optimized by Nota AI for Korea’s government-led Sovereign AI initiative as part of the Upstage consortium.
-
nota-ai/Solar-Open2-250B-Nota-INT4
Text Generation • 41B • Updated • 5.69k • 40 -
nota-ai/Solar-Open2-250B-Nota-NVFP4
Text Generation • 145B • Updated • 66.9k • 174 -
nota-ai/Solar-Open2-250B-Nota-INT4-GlobalPruned
Text Generation • 35B • Updated • 472 • 43 -
nota-ai/Solar-Open2-250B-Nota-NVFP4-GlobalPruned
Text Generation • 117B • Updated • 478 • 34
ERGO: LVLM trained with RL on efficiency objectives; https://github.com/nota-github/ERGO
Shortened LLMs from Depth Pruning; https://github.com/Nota-NetsPresso/shortened-llm
-
nota-ai/cpt_st-vicuna-v1.3-1.5b-ppl
Text Generation • 1B • Updated • 12 • 4 -
nota-ai/cpt_st-vicuna-v1.3-2.7b-ppl
Text Generation • 3B • Updated • 13 • 5 -
nota-ai/cpt_st-vicuna-v1.3-3.7b-ppl
Text Generation • 4B • Updated • 17 • 4 -
nota-ai/cpt_st-vicuna-v1.3-5.5b-ppl
Text Generation • 6B • Updated • 10 • 4
Compressed NVIDIA Models (Nemotron and Cosmos)
Optimized Kimi models
Solar Open models officially optimized by Nota AI for Korea’s government-led Sovereign AI initiative as part of the Upstage consortium.
-
nota-ai/Solar-Open2-250B-Nota-INT4
Text Generation • 41B • Updated • 5.69k • 40 -
nota-ai/Solar-Open2-250B-Nota-NVFP4
Text Generation • 145B • Updated • 66.9k • 174 -
nota-ai/Solar-Open2-250B-Nota-INT4-GlobalPruned
Text Generation • 35B • Updated • 472 • 43 -
nota-ai/Solar-Open2-250B-Nota-NVFP4-GlobalPruned
Text Generation • 117B • Updated • 478 • 34
Mixture-of-Experts Large Language Models with Advanced Quantization
-
nota-ai/Solar-Open-100B-NotaMoEQuant-NVFP4
Text Generation • 59B • Updated • 242 • 23 -
nota-ai/Solar-Open-100B-Nota-FP8
Text Generation • 103B • Updated • 252 • 46 -
nota-ai/Solar-Open-100B-NotaMoEQuant-Int4
Text Generation • 2B • Updated • 586 • 62 -
nota-ai/Qwen3-30B-A3B-NotaMoEQuant-Int4
Text Generation • 0.6B • Updated • 27 • 8
ERGO: LVLM trained with RL on efficiency objectives; https://github.com/nota-github/ERGO
Block-removed Knowledge-distilled SD models; https://github.com/Nota-NetsPresso/BK-SDM
Shortened LLMs from Depth Pruning; https://github.com/Nota-NetsPresso/shortened-llm
-
nota-ai/cpt_st-vicuna-v1.3-1.5b-ppl
Text Generation • 1B • Updated • 12 • 4 -
nota-ai/cpt_st-vicuna-v1.3-2.7b-ppl
Text Generation • 3B • Updated • 13 • 5 -
nota-ai/cpt_st-vicuna-v1.3-3.7b-ppl
Text Generation • 4B • Updated • 17 • 4 -
nota-ai/cpt_st-vicuna-v1.3-5.5b-ppl
Text Generation • 6B • Updated • 10 • 4