inclusionAI

2026inclusionAI

LLaDA2.1 v2.1

Open Source
LLMs
ML

LLaDA2.1 is the latest iteration in the LLaDA series, focusing on accelerating text diffusion through token editing. This model aims to enhance the efficiency and performance of diffusion-based language models.

Accelerated text diffusion via token editing
Enhanced efficiency in diffusion-based models
Improved performance over previous LLaDA versions
PricingFree
Version2.1
2026inclusionAI

LLaDA2.2-flash v2.2

Open Source
diffusion
llm

LLaDA2.2-flash is an agent-oriented Mixture-of-Experts (MoE) diffusion language model designed for long-context agentic workloads. It introduces Levenshtein Editing with DELETE and INSERT control tokens to enable efficient parallel generation, error correction, and multi-turn tool use.

128K context window with Block Routing for efficient MoE expert activation
Levenshtein Editing using DELETE and INSERT control tokens for structural sequence modification
Agentic Reinforcement Learning via L-EBPO for improved error correction
PricingOpen source under Apache License 2.0
Version2.2
2026inclusionAI

Ring-2.5-1T v2.5-1T

Open Source
Coding
Agentic AI

Ring-2.5-1T is the world's first open-source trillion-parameter reasoning model based on hybrid linear attention architecture. It introduces a high-ratio linear attention mechanism that reduces memory access overhead by over 10× and increases generation throughput by more than 3× for sequences exceeding 32K tokens, making it particularly suitable for deep reasoning and long-horizon task execution.

High-ratio linear attention mechanism
Enhanced generation efficiency
Deep reasoning capabilities
PricingFree
Version2.5-1T
2026inclusionAI

Ming-flash-omni 2.0 v2.0

Open Source
Agentic AI
Audio

Ming-flash-omni 2.0 is an advanced multimodal foundation model developed by inclusionAI, leveraging a sparser Mixture-of-Experts (MoE) variant of Ling-Flash-2.0 with 100 billion total parameters, of which only 6.1 billion are active per token. This architecture enables efficient scaling and empowers unified multimodal intelligence across vision, speech, and language, representing a significant advancement toward Artificial General Intelligence (AGI).

Expert-level Multimodal Cognition: Accurately identifies plants and animals, recognizes cultural references, and delivers expert-level analysis of artifacts.
Immersive and Controllable Unified Acoustic Synthesis: Integrates speech, audio, and music within a single channel, enabling zero-shot voice cloning and nuanced attribute control.
High-Dynamic Controllable Image Generation and Manipulation: Unifies segmentation, generation, and editing, allowing for sophisticated spatiotemporal semantic decoupling.
PricingFree
Version2.0
2026inclusionAI

Ming-flash-omni 2.0 v2.0

Open Source
Audio
TTS

Ming-flash-omni 2.0 is an open-source, state-of-the-art multimodal large language model developed by inclusionAI. It leverages the Ling-2.0 architecture, a Mixture-of-Experts (MoE) framework comprising 100 billion total parameters, with 6 billion active parameters per token. This design enables efficient scaling and empowers unified multimodal intelligence across vision, speech, and language, representing a significant advancement toward Artificial General Intelligence (AGI).

Expert-level Multimodal Cognition: Accurately identifies plants, animals, cultural references, and artifacts, delivering expert-level analysis.
Immersive and Controllable Unified Acoustic Synthesis: Integrates speech, audio, and music within a single channel, enabling zero-shot voice cloning and nuanced attribute control.
High-Dynamic Controllable Image Generation and Manipulation: Unifies segmentation, generation, and editing, allowing for sophisticated spatiotemporal semantic decoupling.
PricingFree
Version2.0
2025inclusionAI

Ming-omni-tts-16.8B-A3B v1.0

Open Source
Audio

Ming-omni-tts-16.8B-A3B is a high-performance unified audio generation model developed by inclusionAI. It enables precise control over speech attributes and facilitates the synthesis of speech, environmental sounds, and music in a single channel.

Fine-grained Vocal Control: Supports precise control over speech rate, pitch, volume, emotion, and dialect through simple commands.
Intelligent Voice Design: Features 100+ premium built-in voices and supports zero-shot voice design via natural language descriptions.
Immersive Unified Generation: Jointly generates speech, ambient sound, and music in a single channel, delivering a seamless auditory experience.
Version1.0
RegionChina

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode