Params: 398B
Trinity Large Thinking is a reasoning-optimized sparse Mixture-of-Experts language model with 398B total parameters and approximately 13B active parameters per token. It features 60 layers, 48 attention heads, and 256 experts, activating 4 experts per token alongside 1 shared expert. The model combines sliding-window and full attention with a 262K-token context window and is optimized for agentic reasoning, multi-step planning, tool use, and long-context workflows, generating explicit reasoning enclosed in dedicated reasoning tags before producing its final response.
Text GenerationInstruction Following+7