Qwen
Alibaba Cloud profile, latest version intelligence, capability movement, and source-aware release history.
Profile
Qwen is Alibaba Cloud’s open and commercial model ecosystem, including coding- and agent-oriented releases across the Qwen stack.
Latest Version
Qwen-Image-3.0
Jul 2026
Confidence
94%
extracted
Context
N/A
last known / not captured
Versions
5
documented releases
Current Release Intelligence
Latest published version and metadata capture state
Release
Qwen-Image-3.0
minorQwen-Image-3.0 is the third-generation foundational image generation model in the Qwen-Image series. The core focus of this release is 'Real' (实), emphasizing authenticity across three dimensions, building upon previous versions that focused on precision, variety, completeness, beauty, and authenticity. It represents a significant advancement in rich content generation, authentic details, and deep knowledge.
Decision Fit
Best For
Repository coding and rapid iteration
Avoid If
You need non-technical, no-setup workflows
Release Diff
Core fields and capture notes for this release
Version
Context Window
Not captured in this release; showing last known value when available
Not captured
Pricing
Not captured in this release; showing last known value when available
Not captured
Current Capabilities
Qwen-Image-3.0 with previous-release movement
Previous score: 8
Previous score: 4
Version History
5 releases documented
Qwen-Image-3.0
July 21, 2026
Qwen-Image-3.0 is the third-generation foundational image generation model in the Qwen-Image series. The core focus of this release is 'Real' (实), emphasizing authenticity across three dimensions, building upon previous versions that focused on precision, variety, completeness, beauty, and authenticity. It represents a significant advancement in rich content generation, authentic details, and deep knowledge.
Qwen-AgentWorld
June 23, 2026
Qwen released Qwen-AgentWorld, a native language world model that simulates agent environments across seven domains. It features native world modeling where environment modeling is the training objective from continual pre-training onward (CPT → SFT → RL), not a post hoc adaptation. A single model simulates text-based environments including MCP, Search, Terminal, and SWE.
Qwen-Robot Suite (Qwen-RobotNav, Qwen-RobotManip, Qwen-RobotWorld)
June 16, 2026
Qwen released the Qwen-Robot Suite, a collection of three foundation models for physical world intelligence: Qwen-RobotNav for agentic navigation, Qwen-RobotManip for robotic manipulation, and Qwen-RobotWorld for embodied world modeling. These models bridge the gap between vision-language understanding and physical control for embodied intelligence. The suite enables open-ended instruction following and generalization across real-robot platforms and tasks.
Qwen3.7-Plus
June 1, 2026
Qwen3.7-Plus is a multimodal agent model that unifies vision and language into a single, versatile agent foundation. It builds on Qwen3.7's strong text backbone and delivers a comprehensive upgrade in vision-language capabilities. It is designed to function as a versatile multimodal agent.
Qwen-VLA
May 29, 2026
Qwen-VLA is a new embodied intelligence model that goes beyond understanding images and videos to enable real-world action. It builds on multimodal large language model capabilities to support embodied agents that can act in the world, not just understand it. This represents a significant expansion from perception to physical action for the Qwen family.