AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
Join Now
EN

AI News

View More

Qwen3.8-Flash-Next Released, Early Spoiler of Qwen4 Architecture

Alibaba's Qwen team open-sourced Qwen3.8-Flash-Next, a multimodal MoE model previewing Qwen4. It has a highly sparse MoE design with 125B parameters, plus a 51B N-gram embedding table and 4B multi-token prediction module, delivering superior performance at very low compute cost.....

13.9k 10 minutes ago
Qwen3.8-Flash-Next Released, Early Spoiler of Qwen4 Architecture

Qwen3.8-Flash-Next Architecture: A Lightweight Multimodal Large Model Achieving Performance Breakthrough

Alibaba Qwen released multimodal MoE Qwen3.8-Flash and open-sourced next-gen Qwen3.8-Flash-Next (Qwen4 prototype). 125B total params, 51B N-gram embedding, 6B active per token; native 260K context, up to 1M. Training cost 1/9 of previous; API: input ¥1/M tokens, output ¥3/M.....

12.5k 21 minutes ago
Qwen3.8-Flash-Next Architecture: A Lightweight Multimodal Large Model Achieving Performance Breakthrough
AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map