Improving Small Language Model Reasoning With A* Search @ycrootaccess
Improving Small Language Model Reasoning With A* Search  @ycrootaccess
Uploaded August 2026 | Updated September 2026, 2 weeks ago
At our inaugural YCML at Startup School, YC Partner Ankit Gupta speaks with Alexander Braverman about a test-time scaling method for improving reasoning in smaller language models.

Instead of relying on a larger teacher model or an external reward model, the method uses the language model’s own self-critique as a heuristic in an A*-inspired search. It explores multiple reasoning paths, deprioritizes weaker branches, and searches for a stronger answer using the same underlying model. On mathematical reasoning benchmarks, the method improved accuracy more efficiently than other test-time approaches at comparable token and runtime budgets.

Apply to Y Combinator: ycombinator.com/apply
Work at a startup: ycombinator.com/jobs
Improving Small Language Model Reasoning With A* SearchAI, Startups, & Competition: Shaping California’s Tech FutureThe Past and Future of YC BioDavid AI: Powering the Voice Era of AILecture 16 - How to Run a User Interview (Emmett Shear)Infisical: The Open Source Security StackAleph: The AI Platform for Modern FinanceThe Right Reason and Way to Approach StrategicsHow Diode Is 10x-ing Hardware DesignThe Future of AI Molecular DiscoveryContext Engineering: Lessons Learned from Scaling CoCounselLecture 7 - How to Build Products Users Love (Kevin Hale)
YC Root Access |

Improving Small Language Model Reasoning With A* Search

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER