A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples Paper • 2607.29122 • Published 9 days ago • 4
WorldDiT: A Unified Diffusion Architecture for World and Action Modeling Paper • 2607.23909 • Published 13 days ago • 7
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Paper • 2607.11562 • Published 27 days ago • 77
Xiaomi-Robotics-U0: Unified Embodied Synthesis with World Foundation Model Paper • 2607.11643 • Published 27 days ago • 45
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning Paper • 2607.07508 • Published Jul 8 • 29
Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity Paper • 2607.07386 • Published Jul 8 • 13
SiamJEPA: On the Role of Siamese Student Encoders in JEPA Paper • 2607.04044 • Published Jul 4 • 1
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better Paper • 2607.04884 • Published Jul 6 • 9
Unified Audio Intelligence Without Regressing on Text Intelligence Paper • 2607.05196 • Published Jul 6 • 23
view post Post 748 Uhh did Opus 4.8 cheat on PostTrainBench??it found an API key in the PostTrainBench environment that allowed it to generate synthetic training data without using GPU hours, boosting the base model by 0.4913Source: https://posttrainbench.com/traces/run.html?id=claude_non_api_max_claude-opus-4-8_10h_run1__healthbench_Qwen_Qwen3-4B-Base_17315102#tab=trace See translation 1 reply · 👀 2 2 🔥 1 1 + Reply
Duration Aware Scheduling for ASR Serving Under Workload Drift Paper • 2603.11273 • Published Mar 11 • 3
Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models Paper • 2606.03748 • Published Jun 2 • 22
DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation Paper • 2605.30350 • Published May 28 • 13
Contrastive Distribution Matching for Amortized Sequential Monte Carlo in Discrete Diffusion Paper • 2605.23346 • Published May 22
Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini Paper • 2605.27295 • Published May 26 • 23
optimize_anything: A Universal API for Optimizing any Text Parameter Paper • 2605.19633 • Published May 19 • 6
Geometric Context Transformer for Streaming 3D Reconstruction Paper • 2604.14141 • Published Apr 15 • 38