Gemini Immersive View appears in the "Students" section
Google's Gemini student app surfaces Immersive View, an image-based topic explorer with nested nodes spanning dinosaurs to molecules.
Browse by source
Google's Gemini student app surfaces Immersive View, an image-based topic explorer with nested nodes spanning dinosaurs to molecules.
Google’s Gemini Notebook is adding grounded voice chats, lecture recording, interactive study overviews and short videos for students.
Google is rolling out Gemini 3.8 Live voice models with real-time dialogue, visual grounding and parallel reasoning across apps and APIs.
ElevenLabs has added its ElevenCreative generation tools to the ElevenLabs MCP, giving Claude, ChatGPT and Cursor access to over 50 models.
GPT-Live-1 brings simultaneous listening and speaking to developers, with interruption handling, backend delegation, and 12 voices.
ScrollEd turns textbooks into a scrollable, Instagram-like feed with video, audio, and quizzes
Our long-term research effort to build AI systems that understand the visual world and its dynamics.
DeepSeek V4 is a family of models, and only some of them accept images
Our OpenAI-compatible speech endpoint puts TTS models from Mistral, xAI, Microsoft, and more behind one request shape
Send a source image and an edit prompt in one request, get the edited image back, and change the editing model by editing a single field. Runnable Python and TypeScript included.
Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
At the IBC conference, running Sept. 11-14 in Amsterdam, the creative
A glaring loophole in digital music distribution makes it very easy to piggyback on the talent of real artists with AI generated music.
Overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5. The most jam packed, feel the AGI day in the history of AI.
DeepSeek-V4.1-Flash Release Today, we officially release the DeepSeek-V4.1-Flash model
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents
Runway wants to stream AI video as users prompt it, rather than make them wait for finished clips
Alibaba's Qwen team has released Qwen-Image-2.1, an open-weight model that generates and edits images on powerful consumer GPUs
Apple says its new watches won’t save raw audio, but features that can transcribe recent speech and summarize ambient conversations raise new questions about consent, privacy
Apple introduced Apple Reference Image to help users determine whether photos have been edited, including alterations made by AI.
As it grapples with a bevy of lawsuits, Suno said its new model, Suno v6, is not trained using music it used to train previous versions of the AI model.
Suno's new v6 AI music model is its first made with support from the record industry
At Wednesday's iPhone Duo launch event, Apple announced a handful of new Siri AI Audio Intelligence features, including Siri Recap, Live Rewind, Sound Recognition