Pippa's Artist-Paid Model Tests Whether Ethical Incentives Can Reshape AI Video Generation
A new text-to-video startup is attempting to address the art-theft controversy by paying creators directly—but micro-payments and scraped training data complicate the pitch.
Runway launches Media Router to triage crowded generative media landscape
Runway's new routing infrastructure automatically selects the optimal image, video, or audio model based on developer priorities like cost, speed, and quality.
NVIDIA and Hugging Face integrate NeMo Automodel for distributed diffusion fine-tuning at scale
NVIDIA NeMo Automodel now works seamlessly with Hugging Face Diffusers, enabling production-grade distributed training of video and image models without checkpoint conversion.
Google Vids Adds Personal AI Avatars and Gemini Omni Integration
Google expands Vids with custom digital avatars, multi-modal video generation, and step-by-step editing, positioning the tool against specialized AI video platforms.
Odyssey raises $310M Series B at $1.45B valuation, backed by Amazon and Natural Capital
World model startup Odyssey, founded by autonomous vehicle veterans, secures unicorn status with backing from Amazon, AMD Ventures, and prominent angels.
Allen AI releases MolmoMotion, a language-guided 3D motion forecasting model
Allen AI's new MolmoMotion model predicts object trajectories from video, text instructions, and marked 3D points—with applications in robotic manipulation and video generation.
Hollywood's AI Inflection: From Slop to Craft at Tribeca 2026
Experimental films at Tribeca Film Festival reveal how human-directed AI can create compelling cinema—a sharp contrast to the short-form content flooding the internet.
Avataar's Varya Model Targets India's Video-First Market With 20x Cost Reduction
Peak XV-backed startup releases culturally-tuned video generation model at ₹0.48/second, positioning open-weights AI for mass adoption across India.
Decart's Oasis 3 world model brings photorealistic driving simulation to API, priced at $0.02 per second
Decart launches Oasis 3, an interactive world model for autonomous vehicle simulation, as the startup positions itself as the 'OpenAI of world models' with a $4B valuation.
Google Unveils Gemini Omni Video Generation and Gemini 3.5 Flash for Agentic AI
Google announced Gemini Omni, a multimodal model that generates and edits video through natural language, and Gemini 3.5 Flash, optimized for complex agent workflows.
Google I/O 2026: Gemini Omni, Multimodal Search, and AI Agents debut
Google unveiled Gemini Omni for video generation, Gemini 3.5 Flash for agents and coding, and autonomous Search agents that monitor the web 24/7.
Google's Gemini Omni Raises Questions About Video Generation Quality and Consistency
Google released Omni Flash, the first model in its anything-to-anything Gemini family, but early tests reveal significant flaws in character consistency and object rendering.
Google's Gemini Avatar Tool Generates Photorealistic Video Clones—With a Catch
Google's new Gemini avatar feature lets users create AI videos of themselves, but usage limits and setup quirks raise questions about accessibility and deepfake safeguards.
Google's Gemini Omni blurs the line between text prompt and video simulation
At I/O 2026, Google DeepMind unveiled Gemini Omni, a multimodal family that generates video from combined image, audio, and text inputs, signaling a shift from generative to simulational AI.
Google's Flow Avatars Bring Self-Deepfaking to Mainstream Creators
Google's Flow platform now lets users generate AI videos featuring digital clones of themselves, powered by the new Omni Flash model—a capability that mirrors OpenAI's defunct Sora app.
Google launches Gemini 3.5 models and Omni multimodal family at I/O 2026
Google unveiled Gemini 3.5 Flash as the new default model, introduced Gemini Omni for text-to-video generation, and previewed always-on agents powered by Gemini Spark.
Google DeepMind Launches Gemini Omni Flash for AI-Powered Video Generation and Editing
Gemini Omni Flash enables users to generate and edit videos through natural language prompts, combining multimodal inputs with real-world knowledge.
Google I/O 2026: Gemini Omni and Agent-First Development Mark Shift Toward Agentic AI
Google unveiled Gemini Omni, a multimodal model capable of video creation, alongside Gemini 3.5 Flash and expanded agent capabilities across Search, Gmail, and shopping.
NVIDIA Cosmos Predict 2.5 Fine-Tuning with LoRA/DoRA Cuts Robot Video Model Training to Single GPU
Hugging Face publishes parameter-efficient fine-tuning guide for NVIDIA's 2B-parameter world model, enabling domain adaptation for robotic manipulation on consumer hardware.
Runway Pivots From Video Generation to World Models, Betting Against Language-First AI
The video-generation startup is expanding into physics-aware world models, positioning itself as an alternative to Google's language-dominated AI strategy.
Chinese Short-Drama Studios Deploy AI to Mass-Produce Content at Industrial Scale
As Chinese short-drama platforms dominate global streaming, generative AI is collapsing production timelines from months to weeks while displacing traditional crew roles.