ElevenLabs AI Voice Generation and Speech Technology ElevenLabs Adds Character-Aware Audiobook Production to ElevenCreative ElevenLabs has expanded ElevenCreative with an audiobook workflow that supports multi-voice narration, detailed Studio direction, and distribution from one platform.
Gemini by Google: Multimodal AI for Text, Code, and Media Gemini Robotics 2 Has Not Been Announced: What Google DeepMind Actually Offers Gemini Robotics 2 is not a verified Google DeepMind release. The documented lineup includes Gemini Robotics 1.5, Gemini Robotics-ER 1.6, and On-Device variants.
n8n Automation, Workflows & Automation Resources n8n Adds Qwen Cloud Node for Text, Image and Video AI Automation Workflows n8n has introduced a dedicated Qwen Cloud node, giving workflow builders access to Alibaba Cloud's Qwen model family for text, image and video tasks.
OpenAI Artificial Intelligence Models Research and Technology OpenAI’s National Science Initiative Brings Frontier AI Into Research Workflows OpenAI has outlined a national science initiative that pairs frontier AI with researchers, national laboratories, universities, and scientific infrastructure.
OpenAI Artificial Intelligence Models Research and Technology OpenAI’s Goblin Post Highlights an Emerging Risk in AI Alignment and Reliability OpenAI’s post-mortem on goblin-like model language offers a practical warning for developers and enterprises: emergent personas can be an alignment and reliability concern, not simply a curiosity.
OpenAI Artificial Intelligence Models Research and Technology OpenAI’s Evaluation Playbook Puts Harness Design at the Center of Model Testing OpenAI’s new evaluation guidance argues that model benchmarks measure more than models. Harness design, tools, budgets, and prompting can materially affect reported results.
OpenAI Artificial Intelligence Models Research and Technology OpenAI Says AI Benchmark Scores Depend on Harnesses, Budgets, and Memory Design OpenAI's evaluation guidance argues that benchmark scores are contingent on the setup used to test a model, particularly for long-horizon agents using tools and retained context.
OpenAI Artificial Intelligence Models Research and Technology OpenAI Says Two API Settings Tripled GPT-5.6 Sol's ARC-AGI-3 Score OpenAI found that its ARC-AGI-3 harness was preventing GPT-5.6 Sol from retaining useful reasoning. Enabling retained reasoning and compaction substantially improved the reported score while cutting output tokens.
Gemini by Google: Multimodal AI for Text, Code, and Media Gemini Omni Brings AI Video Relighting to Google Photos and Google Vids Google is extending Gemini Omni's prompt-based video editing into Google Photos and Google Vids, with relighting designed to improve poorly lit footage.
Gemini by Google: Multimodal AI for Text, Code, and Media Google Gemini Omni Brings Prompt-Based Video Background Changes to Vids and Photos Google has expanded Gemini Omni's video editing workflow with prompt-based background changes across Google Vids and Google Photos Video Remix.
Gemini by Google: Multimodal AI for Text, Code, and Media Google Photos Video Remix Brings Gemini Omni Video Styles to Eligible Subscribers Google Photos is rolling out Video Remix, a Gemini Omni-powered tool that applies cinematic and artistic templates to users' video clips.
Gemini by Google: Multimodal AI for Text, Code, and Media Google Brings Gemini Omni to Vids for Instruction-Driven Video Editing and Generation Google has expanded Gemini Omni in Google Vids, enabling creators to generate clips, edit existing footage through conversation and refine videos with media references.
OpenAI Artificial Intelligence Models Research and Technology OpenAI GPT-5.6 Puts Efficiency at the Center of Its New Model Family OpenAI has released the GPT-5.6 family, positioning Sol as its flagship model and adding Terra and Luna for lower-cost workloads across its platforms.
OpenAI Artificial Intelligence Models Research and Technology OpenAI Says GPT-5.6 Stack Optimizations Cut End-to-End Serving Costs by About 20% OpenAI says its GPT-5.6 efficiency work spans model design, inference infrastructure and agent orchestration, with cumulative gains across cost and throughput.
ElevenLabs AI Voice Generation and Speech Technology ElevenLabs’ Image and Video Tools Show the Path to AI-Assisted Launch Video Production ElevenLabs has documented multimodal video tools that combine generated visuals with audio capabilities. Its public materials point to Image & Video beta and Flows, rather than a separately announced Fable product.
Generative Engine Optimization GEO for AI Search Visibility Google AI Mode Citations Are Not an Above-the-Fold Game, SALT Research Finds SALT.agency's AI Mode research suggests that content position is not a reliable citation advantage. Clear headings and relevant opening sentences matter more.
OpenAI Artificial Intelligence Models Research and Technology OpenAI Plans Free Frontier-Model Access for 100,000 Academic Researchers OpenAI is planning a program to give 100,000 academic researchers year-long access to frontier models, pairing advanced model availability with default data protections.
OpenAI Artificial Intelligence Models Research and Technology OpenAI’s Academic Researcher Access Plan Could Expand Frontier AI Use in Science A reported OpenAI initiative could give a large group of academic researchers free access to advanced AI models through 2027, though public program details remain limited.
Gemini by Google: Multimodal AI for Text, Code, and Media Waymo Ojai Debuts a Three-Screen Cabin and Sixth-Generation Driver in Rider Rollout Waymo has begun introducing Ojai, a rider-focused vehicle platform with a redesigned cabin, accessibility features, app controls, and its sixth-generation Waymo Driver.
Gemini by Google: Multimodal AI for Text, Code, and Media Google Expands Gemini Spark as a Persistent AI Agent Across Its Product Ecosystem Gemini Spark is Google's early-stage persistent AI agent, designed to work across Gemini and Google services with user oversight for major actions.
Generative Engine Optimization GEO for AI Search Visibility Claude Code for SEO Shows How Governed Update Pipelines Can Join Search and Business Data Claude Code can be used to bring search, analytics, advertising, and AI-visibility data into a more governed SEO update workflow, with verification built into the process.
Generative Engine Optimization GEO for AI Search Visibility Claude Code SEO Workflow: Assessing a Reported Content Update Result A reported ranking and click improvement tied to a Claude Code workflow illustrates the appeal of targeted content optimization. The specific outcome remains uncorroborated, but the underlying approach offers useful lessons for SEO teams.
Generative Engine Optimization GEO for AI Search Visibility Rand Fishkin on Zero-Click Search and Why Web Discovery Is Becoming More Concentrated Search Engine Land's interview with Rand Fishkin examines why zero-click behavior predates generative AI and why web traffic is concentrating among fewer platforms.
Generative Engine Optimization GEO for AI Search Visibility Younger Consumers Are Leaning Toward AI Answers, but Trust Still Shapes Search Vox Media survey data points to a growing preference for AI-driven answers among younger consumers. The shift raises practical questions for publishers and brands, but it does not make traditional search obsolete.
n8n Automation, Workflows & Automation Resources n8n AI Workflow Builder Brings Natural Language Automation to Cloud Workflows n8n's AI Workflow Builder lets Cloud customers create, test, and refine automation workflows from natural-language prompts.