PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents Paper • 2609.06702 • Published 7 days ago • 20
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents Paper • 2609.06702 • Published 7 days ago • 20
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents Paper • 2609.06702 • Published 7 days ago • 20
ConMax: Confidence-Maximizing Compression for Efficient Chain-of-Thought Reasoning Paper • 2601.04973 • Published Jan 8
TreePS-RAG: Tree-based Process Supervision for Reinforcement Learning in Agentic RAG Paper • 2601.06922 • Published Jan 11 • 1
Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed Chains Paper • 2410.18415 • Published Oct 24, 2024
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers Paper • 2406.10991 • Published Jun 16, 2024 • 1
TreePS-RAG: Tree-based Process Supervision for Reinforcement Learning in Agentic RAG Paper • 2601.06922 • Published Jan 11 • 1
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward Paper • 2510.03222 • Published Oct 3, 2025 • 76
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers Paper • 2406.10991 • Published Jun 16, 2024 • 1