AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

NVIDIA Open Sources OmniVinci All-Modal Understanding Model with Only 1/6 of the Training Data

NVIDIA released the OmniVinci all-modal understanding model, leading top models by 19.05 points in multiple benchmark tests. The model uses only 0.2 trillion training tokens, achieving six times the data efficiency of competitors. It aims to achieve unified understanding of vision, audio, and text, advancing machine multimodal cognitive capabilities.

14.9k 4 days ago
NVIDIA Open Sources OmniVinci All-Modal Understanding Model with Only 1/6 of the Training Data

NVIDIA Launches OmniVinci, a Multimodal Understanding Model That Sets a New SOTA with 19.05 Points Higher

NVIDIA released the multimodal understanding model OmniVinci, which outperformed top models by 19.05 points in benchmark tests. The model achieves excellent performance with only 1/6 of the training data. It aims to enable AI systems to simultaneously understand vision, audio, and text, simulating human multisensory perception of the world.

13.1k 2 days ago
NVIDIA Launches OmniVinci, a Multimodal Understanding Model That Sets a New SOTA with 19.05 Points Higher

Models

View More

Omnivinci

nvidia

O

OmniVinci is a large language model for full-modal understanding developed by NVIDIA. It has the capabilities of visual, text, and audio processing, as well as voice interaction, and supports multi-modal reasoning and understanding.

MultimodalTransformersTransformers
nvidia
383
71
AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map