Back to News
Dongchao YangFebruary 4, 2026

UniAudio 2.0

Open Source
Audio
TTS

Explore UniAudio 2.0

Visit the official website to learn more and get started

### TL;DR

UniAudio 2.0 is a unified audio language model designed to handle text, speech, sound, and music. It introduces ReasoningCodec, a discrete audio codec that factorizes audio into reasoning and reconstruction tokens, enhancing both understanding and generation tasks. The model is trained on 100 billion text tokens and 60 billion audio tokens, demonstrating strong performance across various audio tasks.

Key Insights & Metrics

Pricing
Free
Cost structure
Version
2.0
Current release version
Hardware
NVIDIA A100 GPU, 64GB RAM
Compute requirements
Category
Open Source
Licensing model
Region
China
Primary region

Key Features

  • ReasoningCodec for audio factorization
  • Unified autoregressive architecture for text and audio
  • Trained on extensive text and audio datasets
  • Strong few-shot and zero-shot generalization
  • Competitive performance across speech, sound, and music tasks

Discussion

0
Upvotes
0
Downvotes
0 reviews

Sign in to leave a review

Reviews

No reviews yet. Be the first to review!

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode