AI Models news — Page 14

The latest AI model releases, capability updates, evaluations, and major advances from leading labs and research teams.

ModelsOfficialMidjourney Updates

Midjourney Seeks User Help for High-Res Image Ranking Ahead of v8.1/8.2 Updates

Midjourney is preparing major aesthetic updates for versions 8.1 and 8.2 and is asking users to help rank images at full 2K resolution for the first time. This user-driven ranking process aims to improve the model's image quality.

Why it matters: User participation in high-resolution image ranking could lead to significant improvements in the visual quality of future Midjourney releases.

ModelsOfficialAmazon Science

Customized Amazon Nova models improve molecular-property prediction in drug discovery

Amazon Science reports that a single, optimized large language model (Amazon Nova) now unifies molecular-property prediction tasks that previously required multiple models. The model can serve as a reasoning partner for medical chemists in drug discovery.

Why it matters: This advancement could accelerate drug discovery by providing a unified AI tool for molecular property prediction, reducing the need for multiple specialized models.

ModelsOfficialGoogle DeepMind

Google DeepMind Unveils Gemini 3.1 Flash TTS with Granular Audio Control

Google DeepMind has introduced Gemini 3.1 Flash TTS, a new audio model featuring granular audio tags that allow for precise control over AI-generated speech. This enables more expressive and finely directed audio generation.

Why it matters: The model offers users enhanced control over AI speech, supporting more natural and expressive audio for various applications.

ModelsOfficialMidjourney Updates

Midjourney Releases V8.1 Alpha with Enhanced Moodboards and srefs

Midjourney has released V8.1 Alpha, the latest version of its V8 model, following feedback from V8.0. The update maintains a consistent aesthetic similar to V7 and introduces enhancements to moodboards and srefs functionality.

Why it matters: The update offers users improved creative tools and consistency in AI-generated imagery.

ModelsOfficialGoogle DeepMind

Google DeepMind Unveils Gemini Robotics-ER 1.6 with Enhanced Spatial Reasoning

Google DeepMind has released Gemini Robotics-ER 1.6, an update to its embodied reasoning model that enhances spatial reasoning and multi-view understanding for autonomous robotics. The model aims to improve the interpretation of 3D environments from multiple camera angles, supporting real-world robotics tasks.

Why it matters: This advancement could improve robots' ability to navigate and manipulate objects in complex, real-world environments.

ModelsOfficialGoogle DeepMind

Google DeepMind Releases Gemma 4, Its Most Capable Open Models Yet

Google DeepMind has announced Gemma 4, which it describes as its most intelligent open models to date. The models are designed for advanced reasoning and agentic workflows, aiming to be both highly capable and accessible.

Why it matters: Gemma 4 marks a notable advancement in open model development, potentially enabling more sophisticated AI applications.

ModelsOfficialGoogle DeepMind

Google DeepMind announces Gemini 3.1 Flash Live for more natural audio AI

Google DeepMind has released Gemini 3.1 Flash Live, a new voice model designed to improve precision and reduce latency in voice interactions. The model aims to make audio AI more fluid, natural, and reliable.

Why it matters: This advancement could enhance user experience in voice-based AI applications by reducing delays and improving accuracy.

ModelsOfficialGoogle DeepMind

Google DeepMind Launches Lyria 3 Pro with Longer Tracks and Broader Integration

Google DeepMind has introduced Lyria 3 Pro, a new version of its music generation model that enables the creation of longer tracks with structural awareness. The model is also being integrated into more Google products and surfaces.

Why it matters: This update advances AI music generation by improving track length and structural coherence, and expands its availability across Google's ecosystem.

ModelsOfficialMistral AI News

Mistral AI Releases Voxtral TTS: Open-Weight Text-to-Speech Model

Mistral AI has announced Voxtral TTS, an open-weights text-to-speech model that is fast, instantly adaptable, and produces lifelike speech for voice agents. The model is designed for use in voice agent applications.

Why it matters: This release marks a significant advancement in open-weight TTS technology, enabling developers to build more natural and responsive voice agents.

ModelsOfficialMidjourney Updates

Midjourney V8.1 Now Available on Discord and Web with Improved Sharpness

Midjourney has released V8.1 on both Discord and midjourney.com, featuring improved sharpness and image quality. The enhancements are especially noticeable for style references (SREFs) and moodboards, but apply to all images.

Why it matters: This update directly improves image quality for all Midjourney users, particularly enhancing style consistency features.

ModelsOfficialMistral AI News

Mistral AI Introduces Physics AI Models for Engineering Acceleration

Mistral AI has announced a new class of AI models designed to predict the behavior of physical systems. These models aim to support engineers and hardware product development by simulating physical phenomena.

Why it matters: This development represents a notable expansion of AI applications into physical simulation, with potential implications for engineering and hardware design.

ModelsReportedTechCrunch / AI

OpenAI says GPT 5.6 is the ‘preferred model’ for Microsoft Copilot 365 amid breakup chatter

OpenAI has named GPT 5.6 as the preferred model for Microsoft Copilot 365, addressing rumors about a potential split between the two companies. OpenAI's latest models will continue to power Microsoft's productivity suite.

Why it matters: This announcement clarifies the ongoing collaboration between OpenAI and Microsoft, providing reassurance to Copilot 365 users.