MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors Paper • 2607.12000 • Published 14 days ago • 39
Towards Autonomous and Auditable Medical Imaging Model Development Paper • 2607.10522 • Published 15 days ago • 20
WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces Paper • 2606.09426 • Published Jun 8 • 107
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining Paper • 2605.14747 • Published May 14 • 147
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence Paper • 2605.12882 • Published May 13 • 274
Reinforcing Multimodal Reasoning Against Visual Degradation Paper • 2605.09262 • Published May 10 • 7
PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments Paper • 2605.02240 • Published May 4 • 9
Soft Anisotropic Diagrams for Differentiable Image Representation Paper • 2604.21984 • Published Apr 27 • 5
RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time Paper • 2604.11626 • Published Apr 13 • 103
Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Paper • 2604.10949 • Published Apr 13 • 40
Scientific Graphics Program Synthesis via Dual Self-Consistency Reinforcement Learning Paper • 2604.06079 • Published Apr 7 • 8
Adam's Law: Textual Frequency Law on Large Language Models Paper • 2604.02176 • Published Apr 2 • 510
An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU Paper • 2603.16428 • Published Mar 17 • 51