ScrollEd wants to turn textbooks into TikTok
ScrollEd turns textbooks into a scrollable, Instagram-like feed with video, audio, and quizzes
ScrollEd turns textbooks into a scrollable, Instagram-like feed with video, audio, and quizzes
Alibaba's Qwen team has released Qwen-Image-2.1, an open-weight model that generates and edits images on powerful consumer GPUs
Runway wants to stream AI video as users prompt it, rather than make them wait for finished clips
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents
A glaring loophole in digital music distribution makes it very easy to piggyback on the talent of real artists with AI generated music.
Google’s Gemini Notebook is adding grounded voice chats, lecture recording, interactive study overviews and short videos for students.
DeepSeek V4 is a family of models, and only some of them accept images
Google is rolling out Gemini 3.8 Live voice models with real-time dialogue, visual grounding and parallel reasoning across apps and APIs.
ElevenLabs has added its ElevenCreative generation tools to the ElevenLabs MCP, giving Claude, ChatGPT and Cursor access to over 50 models.
Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
Our OpenAI-compatible speech endpoint puts TTS models from Mistral, xAI, Microsoft, and more behind one request shape
GPT-Live-1 brings simultaneous listening and speaking to developers, with interruption handling, backend delegation, and 12 voices.
Our long-term research effort to build AI systems that understand the visual world and its dynamics.
DeepSeek-V4.1-Flash Release Today, we officially release the DeepSeek-V4.1-Flash model
Apple says its new watches won’t save raw audio, but features that can transcribe recent speech and summarize ambient conversations raise new questions about consent, privacy
Apple introduced Apple Reference Image to help users determine whether photos have been edited, including alterations made by AI.
Suno's new v6 AI music model is its first made with support from the record industry
At Wednesday's iPhone Duo launch event, Apple announced a handful of new Siri AI Audio Intelligence features, including Siri Recap, Live Rewind, Sound Recognition
Apple is launching a new way to prove that the picture you took isn't manipulated by AI
At the IBC conference, running Sept. 11-14 in Amsterdam, the creative
Suno has unveiled a new AI music model generation, v6, in three versions, built together with Warner Music Group, BMG, and Believe. All older models are being shut down
OpenAI is releasing two new image models with ChatGPT Images 2.5. Flare handles faster generation, Sunburst delivers more precise edits
As it grapples with a bevy of lawsuits, Suno said its new model, Suno v6, is not trained using music it used to train previous versions of the AI model.
Amazon's Prime Video is launching a new AI-powered feature that lines up an actor's mouth with "human-dubbed" audio