Hogwild! Inference: Parallel LLM Generation via Concurrent Attention Paper • 2504.06261 • Published Apr 8, 2025 • 110
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs Paper • 2510.11696 • Published Oct 13, 2025 • 183
AttentionPredictor: Temporal Pattern Matters for Efficient LLM Inference Paper • 2502.04077 • Published Feb 6, 2025 • 2
An Embarrassingly Simple Approach for Wafer Feature Extraction and Defect Pattern Recognition Paper • 2303.11632 • Published Mar 21, 2023 • 1
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published 10 days ago • 94
Large Language Models are Pattern Matchers: Editing Semi-Structured and Structured Documents with ChatGPT Paper • 2409.07732 • Published Sep 12, 2024
Human Guided Exploitation of Interpretable Attention Patterns in Summarization and Topic Segmentation Paper • 2112.05364 • Published Dec 10, 2021 • 1
ThinkPatterns-21k: A Systematic Study on the Impact of Thinking Patterns in LLMs Paper • 2503.12918 • Published Mar 17, 2025
Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs Paper • 2601.08763 • Published Jan 13 • 150
Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models Paper • 2505.10554 • Published May 15, 2025 • 119
When Implausible Tokens Get Reinforced: Tail-Aware Credit Calibration for LLM Reinforcement Learning Paper • 2607.07976 • Published Jul 8
Observable Patterns Are Not Explanations: A Causal-Geometric Analysis of Latent Reasoning Models Paper • 2606.12689 • Published Jun 10 • 2
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 6 days ago • 331
ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads Paper • 2608.02703 • Published 13 days ago • 6