# Alibaba — Qwen-VLA: From Understanding the World to Acting in It

- Company: Alibaba (alibaba.com)
- Announced: 2026-05-29T09:00:00+00:00
- Category: new-model
- Subject: Qwen
- Models affected: Qwen-VLA
- Source: https://qwen.ai/blog?id=qwenvla
- Record: https://forck.live/items/4608-qwen-vla-from-understanding-the-world-to-acting-in-it

Alibaba announces Qwen-VLA, a general-purpose Vision-Language-Action model built on the Qwen multimodal backbone, designed to extend visual perception, language understanding, and spatial reasoning into continuous action generation and trajectory prediction for embodied intelligence tasks such as robotic manipulation and vision-language navigation.

## Evidence

Verbatim from https://qwen.ai/blog?id=qwenvla:

> Qwen-VLA is a general-purpose Vision-Language-Action model. Built upon the Qwen multimodal backbone, it extends visual perception, language understanding, and spatial reasoning into continuous action generation and trajectory prediction.

---

Record: https://forck.live/items/4608-qwen-vla-from-understanding-the-world-to-acting-in-it
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
