Trinity Large
Trinity Large is a 400-billion parameter sparse Mixture-of-Experts (MoE) model developed by Arcee AI. It utilizes 256 experts with 4 experts active per token, resulting in a sparsity ratio of 1.56%. The model was trained on 17 trillion tokens using 2,048 Nvidia B300 GPUs over a period of 33 days.
400-billion parameter sparse MoE model
256 experts with 4 experts active per token
Sparsity ratio of 1.56%
PricingFree on OpenRouter for a limited time
RegionUnited States