ScrollEd wants to turn textbooks into TikTok
ScrollEd turns textbooks into a scrollable, Instagram-like feed with video, audio, and quizzes
Track image, video and audio tools by event and medium.
ScrollEd turns textbooks into a scrollable, Instagram-like feed with video, audio, and quizzes
Alibaba's Qwen team has released Qwen-Image-2.1, an open-weight model that generates and edits images on powerful consumer GPUs
Runway wants to stream AI video as users prompt it, rather than make them wait for finished clips
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents
A glaring loophole in digital music distribution makes it very easy to piggyback on the talent of real artists with AI generated music.
Google’s Gemini Notebook is adding grounded voice chats, lecture recording, interactive study overviews and short videos for students.
DeepSeek V4 is a family of models, and only some of them accept images
Google is rolling out Gemini 3.8 Live voice models with real-time dialogue, visual grounding and parallel reasoning across apps and APIs.
ElevenLabs has added its ElevenCreative generation tools to the ElevenLabs MCP, giving Claude, ChatGPT and Cursor access to over 50 models.
Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
Our OpenAI-compatible speech endpoint puts TTS models from Mistral, xAI, Microsoft, and more behind one request shape
GPT-Live-1 brings simultaneous listening and speaking to developers, with interruption handling, backend delegation, and 12 voices.
Our long-term research effort to build AI systems that understand the visual world and its dynamics.
DeepSeek-V4.1-Flash Release Today, we officially release the DeepSeek-V4.1-Flash model
Apple says its new watches won’t save raw audio, but features that can transcribe recent speech and summarize ambient conversations raise new questions about consent, privacy
Apple introduced Apple Reference Image to help users determine whether photos have been edited, including alterations made by AI.
Suno's new v6 AI music model is its first made with support from the record industry
At Wednesday's iPhone Duo launch event, Apple announced a handful of new Siri AI Audio Intelligence features, including Siri Recap, Live Rewind, Sound Recognition
Apple is launching a new way to prove that the picture you took isn't manipulated by AI
At the IBC conference, running Sept. 11-14 in Amsterdam, the creative
Suno has unveiled a new AI music model generation, v6, in three versions, built together with Warner Music Group, BMG, and Believe. All older models are being shut down
OpenAI is releasing two new image models with ChatGPT Images 2.5. Flare handles faster generation, Sunburst delivers more precise edits
As it grapples with a bevy of lawsuits, Suno said its new model, Suno v6, is not trained using music it used to train previous versions of the AI model.
Amazon's Prime Video is launching a new AI-powered feature that lines up an actor's mouth with "human-dubbed" audio
Overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5. The most jam packed, feel the AGI day in the history of AI.
Send a source image and an edit prompt in one request, get the edited image back, and change the editing model by editing a single field. Runnable Python and TypeScript included.
OpenAI announced ChatGPT Images 2.5 on Tuesday and is adding a new way to tell ChatGPT what you want it to make an image of: by drawing a doodle
Google's Gemini student app surfaces Immersive View, an image-based topic explorer with nested nodes spanning dinosaurs to molecules.
Title, source, and timing are the first filters that decide whether a story deserves more of your attention.
Title, source, and timing are the first filters that decide whether a story deserves more of your attention.
Search by product, company, medium or event.
36 stories available