AI Briefing
Key AI developments: model releases, research standards, funding, and policy. Minor product updates and unverified rumors are omitted.
Model releases
Alibaba releases Qwen3.8-Flash-Next open multimodal MoE
Alibaba released Qwen3.8-Flash-Next, an open-weight multimodal MoE model with 125B total parameters and about 6B active per token, native roughly 262k context extendable to 1M, previewing Qwen4 architecture with low training and inference costs.
Why it matters
Another efficient open model pressures the low-cost tier where DeepSeek and Gemini Flash variants compete, which matters for teams optimizing spend per token.
Model releases
DeepSeek ships open weights for V4-Flash-Vision-Exp
DeepSeek released DeepSeek-V4-Flash-Vision-Exp, a 305B-parameter experimental multimodal mixture-of-experts model under an MIT license on Hugging Face, adding competitive vision and agent capabilities on the V4-Flash text stack.
Why it matters
A high-performing, freely downloadable multimodal open model speeds up agent and vision work outside closed APIs, at pricing comparable to the text-only build.
Model releases
Tencent open-sources Hy4 preview at 770B parameters
Tencent open-sourced Hy4 preview, a 770B-parameter MoE model with 49B active parameters and a 1M-token context window under Apache 2.0, aimed at coding, office, and research productivity.
Why it matters
Another major Chinese open release expands self-hosted, long-context options for teams that cannot rely on a single closed API vendor.
Summaries of publicly reported developments, with links to originals. Not original reporting. Minor product updates and unverified rumors are excluded.