SKYNET://COUNTDOWN SYS:MONITORING

Multimodal AI AI News & Updates

[Commercial Release] [SRC↗]

Mistral Launches Voxtral: Open-Source Speech AI Models Challenge Closed Corporate Systems

French AI startup Mistral has released Voxtral, its first open-source audio model family designed for speech transcription and understanding. The models offer multilingual capabilities, can process up to 30 minutes of au...

Risk: [+0.01% ↑] [0 days]
AGI: [+0.02% ↑] [0 days]
Analyze >>
[Commercial Release] [SRC↗]

Google Integrates Project Astra's Real-Time Multimodal AI Across Search and Developer APIs

Google announced Project Astra will power new real-time, multimodal AI experiences across Search, Gemini, and developer tools through its Live API. The technology enables low-latency voice and visual interactions, with p...

Risk: [+0.05% ↑] [0 days]
AGI: [+0.04% ↑] [0 days]
Analyze >>
[Commercial Release] [SRC↗]

Amazon Releases Nova Premier: High-Context AI Model with Mixed Benchmark Performance

Amazon has launched Nova Premier, its most capable AI model in the Nova family, which can process text, images, and videos with a context length of 1 million tokens. While it performs well on knowledge retrieval and visu...

Risk: [+0.04% ↑] [-1 days ↑]
AGI: [+0.03% ↑] [-1 days ↑]
Analyze >>

OpenAI Releases Advanced AI Reasoning Models with Enhanced Visual and Coding Capabilities

OpenAI has launched o3 and o4-mini, new AI reasoning models designed to pause and think through questions before responding, with significant improvements in math, coding, reasoning, science, and visual understanding cap...

Risk: [+0.09% ↑] [-2 days ↑]
AGI: [+0.09% ↑] [-2 days ↑]
Analyze >>

Google Plans to Combine Gemini Language Models with Veo Video Generation Capabilities

Google DeepMind CEO Demis Hassabis announced plans to eventually merge their Gemini AI models with Veo video-generating models to create more capable multimodal systems with better understanding of the physical world. Th...

Risk: [+0.05% ↑] [-1 days ↑]
AGI: [+0.04% ↑] [-1 days ↑]
Analyze >>

Meta Launches Advanced Llama 4 AI Models with Multimodal Capabilities and Trillion-Parameter Variant

Meta has released its new Llama 4 family of AI models, including Scout, Maverick, and the unreleased Behemoth, featuring multimodal capabilities and more efficient mixture-of-experts architecture. The models boast improv...

Risk: [+0.06% ↑] [-1 days ↑]
AGI: [+0.05% ↑] [-1 days ↑]
Analyze >>
[Commercial Release] [SRC↗]

Microsoft Enhances Copilot with Web Browsing, Action Capabilities, and Improved Memory

Microsoft has significantly upgraded its Copilot AI assistant with new capabilities including performing actions on websites, remembering user preferences, analyzing real-time video, and creating podcast-like content sum...

Risk: [+0.05% ↑] [-1 days ↑]
AGI: [+0.04% ↑] [-1 days ↑]
Analyze >>