LLM reasoning / chain-of-thought
From Simulation to Enaction: Post-trained language models recognize and react to their own generations
Current frontier — recent, not yet superseded in the knowledge base
0 papers critique it · 0 beat it on benchmarks
Recent methods in the same sub-problem, not yet superseded in the knowledge base — arXiv benchmark leaders, not vetted production recommendations.