Nanbeige4.1-3B v4.1-3B
Open Source
Agentic AI
Coding
Nanbeige4.1-3B is an enhanced iteration of the Nanbeige4-3B-Base model, achieved through further post-training optimization with supervised fine-tuning (SFT) and reinforcement learning (RL). This model demonstrates strong reasoning capabilities, robust preference alignment, and effective agentic behaviors, making it a competitive open-source model at a small parameter scale.
Strong reasoning capabilities
Robust preference alignment
Effective agentic behaviors
PricingFree
Version4.1-3B