Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up

Collections

Discover the best community collections!

Collections including paper arxiv:2502.15814

All resources for SpeechLMs from "Slamming: Training a Speech Language Model on One GPU in a Day". We provide tokeniser, lm, and datasets

slprl/slam

Audio-to-Audio • Updated about 21 hours ago • 45 • 5
slprl/slam_scaled

Audio-to-Audio • Updated about 22 hours ago • 27 • 3
slprl/SpokenSwag

Viewer • Updated 1 day ago • 67.9k • 114 • 3
slprl/mhubert-base-25hz

Feature Extraction • Updated Jan 10 • 2.75k • 4

Slamming: Training a Speech Language Model on One GPU in a Day

Paper • 2502.15814 • Published 7 days ago • 45

about 5 hours ago

Slamming: Training a Speech Language Model on One GPU in a Day

Paper • 2502.15814 • Published 7 days ago • 45

about 8 hours ago

Slamming: Training a Speech Language Model on One GPU in a Day

Paper • 2502.15814 • Published 7 days ago • 45

FM_Training_Infra

about 11 hours ago

Slamming: Training a Speech Language Model on One GPU in a Day

Paper • 2502.15814 • Published 7 days ago • 45

about 8 hours ago

Adding NVMe SSDs to Enable and Accelerate 100B Model Fine-tuning on a Single GPU

Paper • 2403.06504 • Published Mar 11, 2024 • 53
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Paper • 2502.06703 • Published 16 days ago • 137
Slamming: Training a Speech Language Model on One GPU in a Day

Paper • 2502.15814 • Published 7 days ago • 45

about 15 hours ago

LinFusion: 1 GPU, 1 Minute, 16K Image

Paper • 2409.02097 • Published Sep 3, 2024 • 33
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion

Paper • 2409.11406 • Published Sep 17, 2024 • 26
Diffusion Models Are Real-Time Game Engines

Paper • 2408.14837 • Published Aug 27, 2024 • 123
Segment Anything with Multiple Modalities

Paper • 2408.09085 • Published Aug 17, 2024 • 22

Perception and abstraction. Each modality is tokenized and embedded into vectors for model to comprehend.

about 15 hours ago

VILA^2: VILA Augmented VILA

Paper • 2407.17453 • Published Jul 24, 2024 • 40
Octopus v4: Graph of language models

Paper • 2404.19296 • Published Apr 30, 2024 • 117
Octo-planner: On-device Language Model for Planner-Action Agents

Paper • 2406.18082 • Published Jun 26, 2024 • 48
Dolphin: Long Context as a New Modality for Energy-Efficient On-Device Language Models

Paper • 2408.15518 • Published Aug 28, 2024 • 43

interesting stuff

about 5 hours ago

Chain-of-Verification Reduces Hallucination in Large Language Models

Paper • 2309.11495 • Published Sep 20, 2023 • 38
Adapting Large Language Models via Reading Comprehension

Paper • 2309.09530 • Published Sep 18, 2023 • 77
CulturaX: A Cleaned, Enormous, and Multilingual Dataset for Large Language Models in 167 Languages

Paper • 2309.09400 • Published Sep 17, 2023 • 85
Language Modeling Is Compression

Paper • 2309.10668 • Published Sep 19, 2023 • 83

Company

TOS Privacy About Jobs

Website

Models Datasets Spaces Pricing Docs