MiniCPM5-2B: The Best Sub-Agent Model Yet?
Sam Witteveen explores the MiniCPM 2.5B model, examining how scaling its data and training recipe from the previous 1B version affects its performance in function calling and agentic tasks. The analysis includes a look at new RL2 training techniques, benchmark results, and practical testing for sub-agent applications using RunPod serverless deployments.
