LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents Paper • 2608.17393 • Published 9 days ago • 24
Lego-RL Collection Harness-native RL for coding agents: the trained policy and the training task index. • 6 items • Updated about 14 hours ago • 11
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 15 days ago • 30
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 15 days ago • 30
SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining Paper • 2602.10718 • Published Apr 28
XStreamVGGT: Extremely Memory-Efficient Streaming Vision Geometry Grounded Transformer with KV Cache Compression Paper • 2601.01204 • Published Jan 3
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 15 days ago • 30
Massive Activations in Hybrid Linear Attention Models Collection Official PAS, ISP, gating-ablation, and scale-study checkpoints for Massive Activations in Hybrid Linear Attention Models. • 10 items • Updated 14 days ago