Wire
Nvidia Launches Multimodal AI Model to Combine Vision, Speech and Language
Nvidia (NVDA) said Tuesday it has launched Nemotron 3 Nano Omni, an open multimodal AI model designed to combine vision, speech and language capabilities into a single system.The model can process text, images, audio and video together, eliminating the need for separate models, and it has more accuracy in tasks such as document intelligence, audio-video reasoning and computer-use applications, the company saidNvidia said the model delivers up to nine times higher throughput than comparable models, reducing costs and improve scalability while maintaining responsiveness.Nemotron 3 Nano Omni has been adopted by companies such as Foxconn and Palantir (PLTR), and others such as Dell Technologies (DELL) and DocuSign (DOCU) are evaluating the technology, Nvidia said.Price: $209.59, Change: $-7.02, Percent Change: -3.24%
$DELL$DOCU$NVDA$PLTR