MiniCPM5
Collection
A SOTA 1B on-device LLM, small yet powerful. • 17 items • Updated • 41
Large Language Models
MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement
Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning