Skip to content

๐ŸงŠ 3D Vision

๐Ÿ’ฌ ACL2025 ยท 1 paper notes

๐Ÿ“Œ Same area in other venues: ๐Ÿ“ท CVPR2026 (751) ยท ๐Ÿ”ฌ ICLR2026 (197) ยท ๐Ÿงช ICML2026 (30) ยท ๐Ÿค– AAAI2026 (79) ยท ๐Ÿง  NeurIPS2025 (116) ยท ๐Ÿ“น ICCV2025 (267)

Slamming: Training a Speech Language Model on One GPU in a Day

This paper proposes the Slam training recipe, which systematically optimizes model initialization, architectural choices, synthetic data, and preference alignment to train a speech language model on a single A5000 GPU within 24 hours, achieving performance comparable to large-scale SLMs.