AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

Born for Complex Visual Reasoning! Microsoft Releases Phi-3.5-vision Lightweight, Multimodal Open Source Model

Microsoft has released Phi-3.5-vision, a lightweight, multimodal open source AI model designed for processing textual and visual inputs, supporting a context length of 128K. This model is suitable for resource-constrained environments and features capabilities such as image understanding, OCR, chart parsing, and multi-image summarization, showcasing excellent performance and low latency. Comprised of 4.2 billion parameters, it is trained with high-quality data to ensure performance and privacy. It includes three models: lightweight AI, expert mix, and multimodal model, all demonstrating outstanding performance in image and video processing benchmarks.

18k 14 hours ago
Born for Complex Visual Reasoning! Microsoft Releases Phi-3.5-vision Lightweight, Multimodal Open Source Model

AI Products

View More
Phi-3.5-vision

Phi-3.5-vision

An advanced multimodal model that supports image and text understanding.

AI model
10.9k

Models

View More

Phi 3.5 Vision Instruct

FriendliAI

P

Phi-3.5-vision is a lightweight and advanced open-source multimodal model that supports a 128K context length and focuses on processing high-quality, inference-rich text and visual data.

MultimodalTransformersTransformersOther
FriendliAI
370
0

Phi 3.5 Vision Instruct

microsoft

P

Phi-3.5-vision is a lightweight, cutting-edge open multimodal model supporting 128K context length, focusing on high-quality, reasoning-rich text and visual data.

MultimodalTransformersTransformersOther
microsoft
397.4k
679
AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map