#reasoning
12 briefs tagged #reasoning.
R3-Bench tests LLM reasoning under shared budgets
The benchmark finds six models often underperform simple allocation baselines across reasoning suites.
Anthropic model makes progress on Riemann hypothesis
An unreleased Anthropic model advanced work on the long-unsolved math problem, TechCrunch reported.
OpenAI says AI solved 10 long-standing math problems
The Verge reports mathematicians are weighing how AI could change a traditionally slow-moving field.
BDH-CQ improves ARC-AGI-1 cost efficiency
A 150M-parameter model combines in-context learning with recurrent latent reasoning.
Researchers report jailbreak for encrypted reasoning traces
The paper says interchangeable encrypted trace blocks can expose proprietary model reasoning.
OpenAI gives free ChatGPT users unlimited text chats
Free and Go users will lose text-chat rate limits, while file and image messages remain capped.
Researchers probe on-policy distillation for multilingual LLMs
New papers test OPD variants for math, ASR and low-resource reasoning across languages.
New papers target weak points in on-policy distillation
Researchers propose SPOT, recoverability control and other methods to improve student-model training.
Researchers target reasoning gaps in model post-training
New arXiv papers propose denser supervision for language and multimodal reasoning models.
Researchers propose Skill Entropy for LLM reasoning
Skill^2-Bench tests how models switch across 558 skills in long-horizon tasks.
New papers probe privileged information in OPSD
Researchers test how rubrics, anchors and problem structure affect self-distilled model training.
ReflectRL trains on failed expert reasoning traces
The paper proposes using flawed expert trajectories as reflection signals in on-policy training.