SparseDecoding: Decoding-Aware Pruning for Accurate and Efficient LLM Inference Paper • 2610.12327 • Published 3 days ago • 25
Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs Paper • 2501.19164 • Published Jan 31, 2025
CGGS: Consistency-Augmented Geometric Gaussian Splatting for Ego-centric 3D Scene Generation Paper • 2607.03819 • Published Jul 4 • 11
EarlyTom: Early Token Compression Completes Fast Video Understanding Paper • 2605.30010 • Published May 28 • 30
RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Paper • 2605.21195 • Published May 20 • 19
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs Paper • 2603.19217 • Published Mar 19 • 29
ConCuR: Conciseness Makes State-of-the-Art Kernel Generation Paper • 2510.07356 • Published Oct 8, 2025 • 2
OmniAgent: Audio-Guided Active Perception Agent for Omnimodal Audio-Video Understanding Paper • 2512.23646 • Published Dec 29, 2025 • 15
TARS: MinMax Token-Adaptive Preference Strategy for Hallucination Reduction in MLLMs Paper • 2507.21584 • Published Jul 29, 2025 • 11
ERC-SVD: Error-Controlled SVD for Large Language Model Compression Paper • 2505.20112 • Published Mar 16
SparseDecoding: Decoding-Aware Pruning for Accurate and Efficient LLM Inference Paper • 2610.12327 • Published 3 days ago • 25
SparseDecoding: Decoding-Aware Pruning for Accurate and Efficient LLM Inference Paper • 2610.12327 • Published 3 days ago • 25
RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Paper • 2605.21195 • Published May 20 • 19
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs Paper • 2603.19217 • Published Mar 19 • 29
MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding Paper • 2510.23479 • Published Oct 27, 2025 • 18
OmniAgent: Audio-Guided Active Perception Agent for Omnimodal Audio-Video Understanding Paper • 2512.23646 • Published Dec 29, 2025 • 15