
RESEARCH
- E-Commerce Bench: Long-Horizon Operations, Multi-Dimensional EvaluationAgent benchmarks over the past few years have mostly followed one pattern. A goal is handed to the model, and the model tries to reach it within a bounded number of turns, whether that means finding the treasure in a maze, producing a report, or fixing a piece of code. Performance is then scored on the quality of the deliverable or on how much of the task got done, and evaluations of this kind usually come with a well-defined natural stopping point. Most long-horizon tasks… Read more: E-Commerce Bench: Long-Horizon Operations, Multi-Dimensional Evaluation
- Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous DrivingIntroduction We introduce Qwen-Drive-1.0, the first vision-language foundation model for autonomous driving that unifies 3D perception and visual question answering at the pretraining stage and further extends to motion planning, while keeping the pretrained VLM architecture entirely untouched. Built on the natively multimodal Qwen3.5-4B, it attaches two external modules. A BEV perception head serves as an explicit, inspectable 3D probe, jointly performing 3D object detection, semantic occupancy prediction, and BEV map segmentation, and a Planning Expert generates future ego trajectories through flow matching. Through staged training, we… Read more: Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving
- Qwen3.8-Max: A New Bar for Coding and CoworkToday, we are officially releasing Qwen 3.8-Max, the most capable model in the Qwen family to date. This also marks the first time we will open-source the weights of a Qwen-Max-class model — the open weights will be released next week. Built upon the architectural foundation of Qwen 3.5, Qwen 3.8-Max scales to 2.4 trillion parameters, delivering comprehensive improvements across coding, work, research, and long-horizon tasks. It can not only answer more challenging questions, but also complete complex tasks end-to-end with greater reliability, producing dependable deliverables. Coding For… Read more: Qwen3.8-Max: A New Bar for Coding and Cowork
- Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-EfficiencyIntroduction In this release we are opening the weights of Qwen3.8-Flash-Next, a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4. It plays the same role that Qwen3-Next played for Qwen3.5: the hybrid Gated DeltaNet + Gated Attention design introduced at that time has since been used across the Qwen3.5, Qwen3.6, Qwen3.7 and Qwen3.8 series. We are again releasing the architectural changes early, so that the community can examine them before the full Qwen4 model family is built on top of them. Qwen3.8-Flash-Next upgrades… Read more: Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency



