Running Featured 84 QED-Nano: Teaching a Tiny Model to Prove Hard Theorems 📝 84 Who needs 1T parameters? Olympiad proofs with a 4B model
Running on CPU Upgrade Featured 3.29k The Smol Training Playbook 📚 3.29k The secrets to building world-class LLMs
SmolVLM Collection State-of-the-art compact VLMs for on-device applications: Base, Synthetic, and Instruct. Check our blog: https://huggingface.co/blog/smolvlm • 5 items • Updated May 5, 2025 • 44
Running on CPU Upgrade Agents Featured 1.45k Open ASR Leaderboard 🏆 1.45k Explore and compare ASR model performance across datasets
view article Article Prefill and Decode for Concurrent Requests - Optimizing LLM Performance tngtech • Apr 16, 2025 • 94
view article Article Efficient Request Queueing – Optimizing LLM Performance tngtech • Apr 2, 2025 • 27