Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents
Homepage items are hidden so this list continues with earlier updates. Keep scrolling to load more.
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents
A glaring loophole in digital music distribution makes it very easy to piggyback on the talent of real artists with AI generated music.
Google’s Gemini Notebook is adding grounded voice chats, lecture recording, interactive study overviews and short videos for students.
DeepSeek V4 is a family of models, and only some of them accept images
Google is rolling out Gemini 3.8 Live voice models with real-time dialogue, visual grounding and parallel reasoning across apps and APIs.
ElevenLabs has added its ElevenCreative generation tools to the ElevenLabs MCP, giving Claude, ChatGPT and Cursor access to over 50 models.
Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
Our OpenAI-compatible speech endpoint puts TTS models from Mistral, xAI, Microsoft, and more behind one request shape
GPT-Live-1 brings simultaneous listening and speaking to developers, with interruption handling, backend delegation, and 12 voices.
Our long-term research effort to build AI systems that understand the visual world and its dynamics.
DeepSeek-V4.1-Flash Release Today, we officially release the DeepSeek-V4.1-Flash model
Apple says its new watches won’t save raw audio, but features that can transcribe recent speech and summarize ambient conversations raise new questions about consent, privacy
Apple introduced Apple Reference Image to help users determine whether photos have been edited, including alterations made by AI.
Suno's new v6 AI music model is its first made with support from the record industry
At Wednesday's iPhone Duo launch event, Apple announced a handful of new Siri AI Audio Intelligence features, including Siri Recap, Live Rewind, Sound Recognition
Apple is launching a new way to prove that the picture you took isn't manipulated by AI
At the IBC conference, running Sept. 11-14 in Amsterdam, the creative
Suno has unveiled a new AI music model generation, v6, in three versions, built together with Warner Music Group, BMG, and Believe. All older models are being shut down
OpenAI is releasing two new image models with ChatGPT Images 2.5. Flare handles faster generation, Sunburst delivers more precise edits
As it grapples with a bevy of lawsuits, Suno said its new model, Suno v6, is not trained using music it used to train previous versions of the AI model.
Amazon's Prime Video is launching a new AI-powered feature that lines up an actor's mouth with "human-dubbed" audio
Overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5. The most jam packed, feel the AGI day in the history of AI.
Send a source image and an edit prompt in one request, get the edited image back, and change the editing model by editing a single field. Runnable Python and TypeScript included.
OpenAI announced ChatGPT Images 2.5 on Tuesday and is adding a new way to tell ChatGPT what you want it to make an image of: by drawing a doodle