AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

32B Inference Performance Surpasses o1-mini! Alibaba Tongyi Launches FIPO Algorithm to Make Large Models Think Deeper

Alibaba's Tongyi Lab introduces the FIPO algorithm, which overcomes traditional reinforcement learning bottlenecks in complex logical reasoning. Using the Future-KL mechanism, it accurately identifies key reasoning steps, effectively addressing model stagnation in tasks like mathematics, thereby enhancing both accuracy and efficiency.....

14.5k 1 days ago
32B Inference Performance Surpasses o1-mini! Alibaba Tongyi Launches FIPO Algorithm to Make Large Models Think Deeper

Tongyi Lab Launches FIPO Algorithm, 32B Model Inference Performance Surpasses o1-mini

Alibaba's Tongyi Lab introduces FIPO, a novel algorithm with a 'Future-KL' mechanism that addresses the 'reasoning length stagnation' in pure reinforcement learning for long-text reasoning, enhancing complex logic alignment training.....

13.5k yesterday
Tongyi Lab Launches FIPO Algorithm, 32B Model Inference Performance Surpasses o1-mini

AliTongyi Lab Launches FIPO Algorithm to Significantly Enhance Large Model Inference Capabilities

Alibaba's Qwen Pilot team introduces the FIPO algorithm, which uses a Future-KL mechanism to identify key tokens in reasoning chains, enhancing large model inference and overcoming limitations of traditional reinforcement learning methods.....

16.6k yesterday
AliTongyi Lab Launches FIPO Algorithm to Significantly Enhance Large Model Inference Capabilities
AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map