# Alibaba — Qwen3.5-Omni: Scaling Up, Toward Native Omni-Modal AGI

- Company: Alibaba (alibaba.com)
- Announced: 2026-03-29T20:00:00+00:00
- Category: new-model
- Subject: Qwen
- Models affected: Qwen3.5-Omni, Qwen3.5-Omni-Plus, Qwen3.5-Omni-Flash, Qwen3.5-Omni-Light, Qwen3-Omni
- Context window: 256k long-context input
- Source: https://qwen.ai/blog?id=qwen3.5-omni
- Record: https://forck.live/items/4600-qwen3-5-omni-scaling-up-toward-native-omni-modal-agi

Alibaba's Qwen announces Qwen3.5-Omni, a new fully omnimodal LLM supporting text, images, audio, and audio-visual content. It comes in three sizes: Plus, Flash, and Light, with 256k long-context input. It is natively pretrained on massive data and offers enhanced multilingual capabilities compared to Qwen3-Omni. The model is available via Offline and Realtime APIs, and features include audio-visual captioning, voice cloning, and ARIA technique for speech stability.

## Evidence

Verbatim from https://qwen.ai/blog?id=qwen3.5-omni:

> Qwen3.5-Omni is Qwen’s latest generation of fully omnimodal LLM, supporting the understanding of text, images, audio, and audio-visual content. Both the Thinker and Talker in Qwen3.5-Omni adopt the Hybrid-Attention MoE. Qwen3.5-Omni series includes Instruct versions in three sizes: Plus, Flash, and Light, with support for 256k long-context input.

---

Record: https://forck.live/items/4600-qwen3-5-omni-scaling-up-toward-native-omni-modal-agi
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
