Institution profile

ModelBest Inc.

Industry researchasia · cn
Official website
Research library28linked papers
Opportunities0open roles
Selected work

Representative Papers

MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement

Aug 14, 2026

This study addresses the challenges of parameter memorization dependency, insufficient library knowledge, and inadequate feedback mechanisms in automatic mathematical formalization. We propose an iterative optimization framework integrating Mathlib retrieval with compiler-driven feedback. Leveraging the newly constructed FormalVerse dataset comprising 367K samples, we train an 8B model via supervised fine-tuning and reinforcement learning, introducing a novel retrieval planner and verification-guided refinement mechanism to overcome single-pass generation limitations. The resulting model achieves Pass@8 scores of 88.06% (SC) and 72.37% (CC) across six benchmarks, significantly outperforming multiple specialized 32B models. These results demonstrate substantial improvements in both syntactic correctness and semantic consistency for mathematical formalization tasks.

0 citationsRead paper
Recent publications

Latest Papers

MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement

Aug 14, 2026

This study addresses the challenges of parameter memorization dependency, insufficient library knowledge, and inadequate feedback mechanisms in automatic mathematical formalization. We propose an iterative optimization framework integrating Mathlib retrieval with compiler-driven feedback. Leveraging the newly constructed FormalVerse dataset comprising 367K samples, we train an 8B model via supervised fine-tuning and reinforcement learning, introducing a novel retrieval planner and verification-guided refinement mechanism to overcome single-pass generation limitations. The resulting model achieves Pass@8 scores of 88.06% (SC) and 72.37% (CC) across six benchmarks, significantly outperforming multiple specialized 32B models. These results demonstrate substantial improvements in both syntactic correctness and semantic consistency for mathematical formalization tasks.

0 citationsRead paper