OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
FreeOutcome-supervised value models for mathematical reasoning planning
FreeFree tier
About OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
OVM (Outcome-supervised Value Models) is a research model for planning in mathematical reasoning tasks. It uses outcome supervision to train value functions that evaluate the correctness of intermediate reasoning steps, enabling more effective search and planning during problem solving. The model is described in the paper 'OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning' and is available as open-source software.
Key Features
Outcome-supervised value function training
Step-by-step reasoning evaluation and planning
Search-guided reasoning in mathematical problems
Open-source research implementation
Best For
Mathematical problem solving with step-level verificationPlanning and search in multi-step reasoning tasksResearch on value-based reasoning optimization