avatar
Big Data Science
@bdscience
24.07.2024 15:58
⚡️The largest collection of datasets of ~ 1 million pairs of problems and solutions for mathematical competitions

NuminaMath - datasets consisting of 1 million pairs of problems and solutions for various mathematical problems.

🔎Chain of Reasoning (CoT): 860 thousand pairs of problems and solutions created using CoT.

🛠 Tool-Integrated Reasoning (TIR): 73K synthetic solutions derived from GPT-4 with code execution feedback to break complex problems into simpler subproblems that can be solved using Python.

According to the researchers, models trained on NuminaMath achieve best-in-class performance among open-weight models and approach or beat their own models in math competition scores.
huggingface.co
NuminaMath - a AI-MO Collection
Datasets and models for training SOTA math LLMs. See our GitHub for training & inference code: https://github.com/project-numina/aimo-progress-prize
👍 1
1.6K

Обсуждение 0

Обсуждение не доступно в веб-версии. Чтобы написать комментарий, перейдите в приложение Telegram.

Обсудить в Telegram