-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
Collections
Discover the best community collections!
Collections including paper arxiv:2608.15089
-
AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Paper • 2607.28618 • Published • 303 -
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
Paper • 2607.28227 • Published • 309 -
Metis: Memory Foundation Model
Paper • 2607.26760 • Published • 272 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 507
-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Paper • 2402.04252 • Published • 31 -
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
Paper • 2402.03749 • Published • 15 -
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 24
-
DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation
Paper • 2608.13489 • Published • 98 -
VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End?
Paper • 2608.15265 • Published • 59 -
MOSS-VL Technical Report
Paper • 2608.15045 • Published • 47 -
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling
Paper • 2608.14783 • Published • 19
-
Writing in the Margins: Better Inference Pattern for Long Context Retrieval
Paper • 2408.14906 • Published • 144 -
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Paper • 2410.10819 • Published • 7 -
LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models
Paper • 2410.09342 • Published • 39 -
PDFTriage: Question Answering over Long, Structured Documents
Paper • 2309.08872 • Published • 55
-
LASER: LLM Agent with State-Space Exploration for Web Navigation
Paper • 2309.08172 • Published • 14 -
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
Paper • 2606.15007 • Published • 19 -
StateM: Reaching 95.3% Raw Accuracy, or a \$15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling
Paper • 2608.15089 • Published • 441 -
FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis
Paper • 2608.18580 • Published • 119
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 7 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 12 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation
Paper • 2608.13489 • Published • 98 -
VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End?
Paper • 2608.15265 • Published • 59 -
MOSS-VL Technical Report
Paper • 2608.15045 • Published • 47 -
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling
Paper • 2608.14783 • Published • 19
-
AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Paper • 2607.28618 • Published • 303 -
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
Paper • 2607.28227 • Published • 309 -
Metis: Memory Foundation Model
Paper • 2607.26760 • Published • 272 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 507
-
Writing in the Margins: Better Inference Pattern for Long Context Retrieval
Paper • 2408.14906 • Published • 144 -
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Paper • 2410.10819 • Published • 7 -
LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models
Paper • 2410.09342 • Published • 39 -
PDFTriage: Question Answering over Long, Structured Documents
Paper • 2309.08872 • Published • 55
-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Paper • 2402.04252 • Published • 31 -
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
Paper • 2402.03749 • Published • 15 -
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 24
-
LASER: LLM Agent with State-Space Exploration for Web Navigation
Paper • 2309.08172 • Published • 14 -
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
Paper • 2606.15007 • Published • 19 -
StateM: Reaching 95.3% Raw Accuracy, or a \$15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling
Paper • 2608.15089 • Published • 441 -
FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis
Paper • 2608.18580 • Published • 119