openbmb/AgentCPM-Report-GGUF
8B • Updated • 179 • 40
Large Language Models
MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement
Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning