Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingNewsWatching0 independent reports0

Nemotron-Labs-Audex-30B-A3B: Unified Audio Intelligence Without Regressing on Text Intelligence

First seen · 7/7/2026, 12:00 PMLatest activity · 7/7/2026, 12:00 PM

Nemotron Labs introduces Nemotron-Labs-Audex-30B-A3B, a unified audio-text mixture-of-experts model built on Nemotron-Cascade-2-30B-A3B. A single Transformer decoder handles projected audio inputs, text tokens, and quantized audio output tokens in one generation space. Training uses 157.4B audio tokens and 320.5B text tokens, followed by supervised training, text-only Cascade RL, and multi-domain on-policy distillation. The abstract reports leading results across audio understanding, speech recognition and translation, text-to-speech, audio generation, and speech-to-speech generation, with marginal or no regression in text reasoning and agentic capabilities. Checkpoints are released for research.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/7, 12:00 PMnot independentRepresentative
    Nemotron-Labs-Audex-30B-A3B: Unified Audio Intelligence Without Regressing on Text Intelligence