Preprint2023
WizardMath
Unknown
WizardMath enhances mathematical reasoning in Llama-2 via Reinforcement Learning from Evol-Instruct Feedback (RLEIF), outperforming many open and closed-source LLMs.
0Aug 1, 2023Reinforcement LearningReasoning
arXiv
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.