Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 3 days ago • 108
Modular TTT: Rethinking Test-Time Training as Composable Modules Paper • 2608.07110 • Published 13 days ago • 8
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 10 days ago • 334
A Quantized Native Runtime for On-Device Semantic Audio Generation Paper • 2607.08526 • Published Jul 9 • 4
Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process Paper • 2607.03748 • Published Jul 4 • 41
BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery Paper • 2606.20997 • Published Jun 19 • 13
Value-Aware Stochastic KV Cache Eviction for Reasoning Models Paper • 2606.03928 • Published Jun 2 • 8
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction Paper • 2605.26115 • Published May 25 • 53