Learning Multimodal Embeddings with Evidence-Aligned Readout Paper • 2609.33659 • Published 14 days ago • 10
DecepEval: A Benchmark for Evaluating Deception in LLM Agents Paper • 2610.07967 • Published 5 days ago • 75
Harness-Aware Distillation for Small Language Model Agents Paper • 2610.02858 • Published 9 days ago • 13
Taming VLAs under Robot Execution Errors: Self-Compensation and Stress Testing Paper • 2609.37334 • Published 12 days ago • 40
DiVeR: Decision-Critical Verifier Learning for VLA Test-Time Scaling Paper • 2610.04933 • Published 7 days ago • 15
EVISKILL: Grounding Skill Evolution in Replayable Evidence Paper • 2610.05030 • Published 7 days ago • 52
HLA-WM: Hybrid Linear Attention for Long-Horizon Video World Models Paper • 2610.05739 • Published 6 days ago • 21
RealtimeWAM: One-Step Asynchronous World Action Models Paper • 2610.06617 • Published 6 days ago • 24
LLM-as-Jev: LLMs Are Already Jev-Style Decision Models -- When and How to Fine-Tune Them Paper • 2610.02076 • Published 7 days ago • 14
WM-VLM: Probing Internal World Models for Interleaved Visual-Textual Reasoning Paper • 2609.34826 • Published 13 days ago • 12
Spatial Memory Intelligence: Endowing World Models with Understanding-Driven Long-Term Memory Paper • 2610.02521 • Published 10 days ago • 59
VeriHarness: Scaling Agentic Verification for Long-Horizon Tasks Paper • 2610.00972 • Published 10 days ago • 61
Native Action-Prior Learning from Videos for World Action Models Paper • 2610.03391 • Published 9 days ago • 88
World Action Modeling with Progressive Visual Planning Paper • 2610.02508 • Published 10 days ago • 97
Omni-Embed-Mini: Binding Modalities Without Forgetting via Dense Distillation Paper • 2610.02148 • Published 10 days ago • 23
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 10 days ago • 92