Did Codex Reset
GitHub

Qwen Team presents natively multimodal Qwen3.8-Omni-Flash

DAIR.AI

The Qwen Team presents Qwen3.8-Omni-Flash, a natively multimodal model trained for long-horizon agent tasks across text, audio, and video, including video editing and long-form audio and video translation. It uses the sparse mixture-of-experts design of Qwen3.8-Next and a context window of one million tokens. A co-training strategy keeps text performance while carrying agent skills over to audio and video tasks.

Two open-source frameworks come with the model. Qwen-MM-Plugins adds audio and video support to existing agent harnesses, and Qwen-Live-Harness handles real-time multimodal interaction with context and memory management, tool use, and sub-agent delegation. The Qwen3.8-Omni paper covers the model.