When to Switch: Reliable Action-Chunk Extension for Vision-Language-Action Models Paper • 2610.05719 • Published 2 days ago • 14
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 9 days ago • 138
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-Coder-GGUF Image-Text-to-Text • 117B • Updated 8 days ago • 479k • 326
FLUX 3 Action Collection Open weights 7B world action model: base and shared encoders, SO-101 and DROID policies. Code: https://github.com/black-forest-labs/flux-action • 3 items • Updated 5 days ago • 41
view article Article **Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization** nvidia • 14 days ago • 68
primitive-ai/DeepSeek-V4-Flash-Vision-Exp-REAP-145B Text Generation • 146B • Updated about 1 month ago • 424 • 8
NeoMME Collection Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual Encoders • 12 items • Updated 9 days ago • 33
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • Sep 3 • 147
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published Aug 31 • 69