Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Paper • 2605.02290 • Published May 4 • 43
SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations Paper • 2606.05563 • Published Jun 4 • 56
Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization Paper • 2608.09043 • Published 3 days ago • 6
Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization Paper • 2608.09043 • Published 3 days ago • 6
Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Paper • 2508.19578 • Published Aug 27, 2025
Towards Multi-dimensional Evaluation of LLM Summarization across Domains and Languages Paper • 2506.00549 • Published May 31, 2025
Completing Missing Annotation: Multi-Agent Debate for Accurate and Scalable Relevant Assessment for IR Benchmarks Paper • 2602.06526 • Published Feb 6
Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization Paper • 2608.09043 • Published 3 days ago • 6
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback Paper • 2503.21332 • Published Mar 27, 2025 • 25
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback Paper • 2503.21332 • Published Mar 27, 2025 • 25