ModernBERT Bringing BERT into modernity via both architecture changes and scaling answerdotai/ModernBERT-base Fill-Mask • 0.1B • Updated Jan 15 • 776k • 953 answerdotai/ModernBERT-large Fill-Mask • 0.4B • Updated Jan 15 • 83.4k • 428 Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Paper • 2412.13663 • Published Dec 18, 2024 • 157
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Paper • 2412.13663 • Published Dec 18, 2024 • 157
Multi-Vector Retrievers answerdotai/answerai-colbert-small-v1 33.4M • Updated Nov 18, 2024 • 1.37M • 155 answerdotai/JaColBERTv2.5 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 4.97k • 22
Japanese Retrieval answerdotai/JaColBERTv2.5 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 4.97k • 22 answerdotai/JaColBERTv2.4 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 3 • 4 answerdotai/MMARCO-japanese-32-scored-triplets Viewer • Updated Jul 31, 2024 • 8.64M • 156 • 6
CLA-Experiments answerdotai/llama3-8b-instruct-CLA-3 Text Generation • Updated Jul 18, 2024 • 2 • 1 answerdotai/llama3-8b-instruct-CLA-2 Text Generation • Updated Jul 18, 2024 • 1 • 1
Quantized-FT-Orca-Math Models trained during quantization aware fine-tuning experiments using PyTorch's FSDP. answerdotai/llama-7b-orca-math-10k-full Text Generation • Updated Mar 28, 2024 • 2 answerdotai/llama-7b-orca-math-10k-bnb-qlora Updated Mar 28, 2024 answerdotai/llama-7b-orca-math-10k-bnb-qdora Updated Mar 28, 2024 answerdotai/llama-7b-orca-math-10k-bnb-llama-pro Updated Mar 28, 2024 • 1
Function Calling A list of datasets, models and papers for making LLMs better at function calling and tool usage argilla/Synth-APIGen-v0.1 Viewer • Updated Oct 10, 2024 • 49.4k • 131 • 66 Salesforce/xlam-function-calling-60k Viewer • Updated Jan 24 • 60k • 4.97k • 544 sanjay920/gemma-function-calling Viewer • Updated Feb 21, 2024 • 112k • 63 • 9 gorilla-llm/Berkeley-Function-Calling-Leaderboard Preview • Updated Feb 14 • 3.26k • 87
ModernBERT Bringing BERT into modernity via both architecture changes and scaling answerdotai/ModernBERT-base Fill-Mask • 0.1B • Updated Jan 15 • 776k • 953 answerdotai/ModernBERT-large Fill-Mask • 0.4B • Updated Jan 15 • 83.4k • 428 Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Paper • 2412.13663 • Published Dec 18, 2024 • 157
Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference Paper • 2412.13663 • Published Dec 18, 2024 • 157
CLA-Experiments answerdotai/llama3-8b-instruct-CLA-3 Text Generation • Updated Jul 18, 2024 • 2 • 1 answerdotai/llama3-8b-instruct-CLA-2 Text Generation • Updated Jul 18, 2024 • 1 • 1
Multi-Vector Retrievers answerdotai/answerai-colbert-small-v1 33.4M • Updated Nov 18, 2024 • 1.37M • 155 answerdotai/JaColBERTv2.5 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 4.97k • 22
Quantized-FT-Orca-Math Models trained during quantization aware fine-tuning experiments using PyTorch's FSDP. answerdotai/llama-7b-orca-math-10k-full Text Generation • Updated Mar 28, 2024 • 2 answerdotai/llama-7b-orca-math-10k-bnb-qlora Updated Mar 28, 2024 answerdotai/llama-7b-orca-math-10k-bnb-qdora Updated Mar 28, 2024 answerdotai/llama-7b-orca-math-10k-bnb-llama-pro Updated Mar 28, 2024 • 1
Japanese Retrieval answerdotai/JaColBERTv2.5 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 4.97k • 22 answerdotai/JaColBERTv2.4 Sentence Similarity • 0.1B • Updated Jul 31, 2024 • 3 • 4 answerdotai/MMARCO-japanese-32-scored-triplets Viewer • Updated Jul 31, 2024 • 8.64M • 156 • 6
Function Calling A list of datasets, models and papers for making LLMs better at function calling and tool usage argilla/Synth-APIGen-v0.1 Viewer • Updated Oct 10, 2024 • 49.4k • 131 • 66 Salesforce/xlam-function-calling-60k Viewer • Updated Jan 24 • 60k • 4.97k • 544 sanjay920/gemma-function-calling Viewer • Updated Feb 21, 2024 • 112k • 63 • 9 gorilla-llm/Berkeley-Function-Calling-Leaderboard Preview • Updated Feb 14 • 3.26k • 87