Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 2 days ago • 1
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 2 days ago • 1
StochBench: A Domain-Specific Benchmark for Stochastic Processes in Lean Paper • 2609.09264 • Published 4 days ago • 7
Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility Paper • 2608.04001 • Published Aug 4 • 1
Overcoming Dynamics-Blindness: Training-Free Pace-and-Path Correction for VLA Models Paper • 2605.11459 • Published May 14
Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers Paper • 2601.07036 • Published Jan 11
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 2 days ago • 1
StochBench: A Domain-Specific Benchmark for Stochastic Processes in Lean Paper • 2609.09264 • Published 4 days ago • 7
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7 • 5
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
Forte : Finding Outliers with Representation Typicality Estimation Paper • 2410.01322 • Published Oct 2, 2024 • 2
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks Paper • 2505.20047 • Published May 26, 2025 • 3
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks Paper • 2505.20047 • Published May 26, 2025 • 3
LLM-Drop Collection Resources for studies on redundancy in LLMs, including layer dropping and representation-hierarchy-based pruning. • 16 items • Updated Jul 31 • 7