Adaptive Math LLM Tuner
8.2
A software tool designed to fine-tune Large Language Models (LLMs) specifically for solving open-ended mathematical problems, incorporating reinforcement learning with tailored reward functions beyond simple answer-based evaluation.
250h
mvp estimate
8.2
viability grade
27
views
technology stack
Python
Difficult
inspired by
Fine-tuning LLMs for solving open-ended math problems