U-Space: Uncovering When and Why Uncertainty Arises in Language Models Paper • 2610.09087 • Published 5 days ago • 33
ConwayResearch/Underdog-Saluki-27B-1.0 Text Generation • 27B • Updated about 5 hours ago • 59.2k • 265
DecepEval: A Benchmark for Evaluating Deception in LLM Agents Paper • 2610.07967 • Published 5 days ago • 74
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 3 days ago • 65
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 24 days ago • 228
Follow the Entities: A Corpus Map for Agentic Search Paper • 2609.37226 • Published 12 days ago • 100
Rethinking Training-Inference Mismatch in LLM Reinforcement Learning: Where It Arises and How to Correct It Paper • 2609.32444 • Published 15 days ago • 35
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 23 days ago • 138
yandex/AliceAI-Foundation-80B-A3B-Base Text Generation • 81B • Updated about 10 hours ago • 5.97k • 384