Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 27 days ago • 252
Unlocking Lossless Speedups in LLMs via Discrete Diffusion Paper • 2609.04010 • Published Sep 3 • 114
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent Paper • 2502.16852 • Published Feb 24, 2025
TSAQA: Time Series Analysis Question And Answering Benchmark Paper • 2601.23204 • Published Jan 30 • 3
ALERT: Zero-shot LLM Jailbreak Detection via Internal Discrepancy Amplification Paper • 2601.03600 • Published Jan 7
Subspace Alignment for Vision-Language Model Test-time Adaptation Paper • 2601.08139 • Published Jan 13
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Paper • 2605.21468 • Published May 20 • 51