Generative AI for images, video, and creative media — Sora, DALL·E, Midjourney, Stable Diffusion, and more.
957 articles
Importance:NewsNVIDIA LongLive-2.0 real-time video generation
NVIDIA has unveiled 'LongLive-2.0,' a real-time video generation AI that achieves lightweight and high-quality generation through training designed for FP4 quantization.
The news blog specialized in Japanese culture, odd news, gadgets and all other funny stuffs. Updated everyday. Source: gigazine.net
Importance:NewsGemini Omni video editing launch
Google introduces Gemini Omni for AI-generated video creation & editing
The latest model lets users edit videos through conversational prompts while maintaining continuity across scenes, characters, and visual elements. Source: afaqs.com
Importance:NewsGemini Omni video editing in Google Flow
Google Adds Gemini Omni Video Editing to Flow
According to Google's blog post, the company is rolling out **Gemini Omni Flash** to the **Gemini app**, **Google Flow**, and **YouTube Shorts**. Source: letsdatascience.com
Importance:NewsGoogle Gemini Omni launch coverage
Google announces launch of Gemini Omni to enhance AI capabilities
Google launches Gemini Omni to expand Gemini's AI-powered content generation capabilities; The new model enables users to create fully integrated videos... Source: entarabi.com
Importance:NewsAvatar video generation — open model release
Meituan puts avatar video startups under new pressure
Meituan has released LongCat-Video-Avatar 1.5 as an open avatar video model with local deployment potential. The release increases pressure on synthetic. Source: startupfortune.com
Importance:NewsYouTube AI video creator tools
YouTube Will Let Creators Use AI to Insert Themselves into Other People’s Videos
AI Video Generation Industry size was worth around 551.7 million in 2025 and is predicted to grow to around USD 551.7 million by 2035, CAGR 18.37% Source: sphericalinsights.com
Importance:NewsGoogle Gemini Omni multimodal video generation
Google unveils Gemini Omni, a multimodal AI model that generates video from text, images, and audio
Google DeepMind unveiled Gemini Omni at Google I/O, a multimodal AI model family for video generation with implications for decentralized compute and Web3... Source: cryptobriefing.com
Importance:NewsGoogle Gemini Omni AI video production
Google Signals AI Video’s Shift From Clip Generation To Production
Google used its latest I/O event this week to introduce Gemini Omni Flash, a new AI model that can take text, photos, video, and audio as inputs,... Source: forbes.com
Importance:OpinionGoogle I/O Gemini Omni recap
Google I/O Recap
Gemini Omni is a multi-modal video/audio/text input-output model that, as the “Nano Banana for video,” can restyle any video on command into another format. Source: patmcguinness.substack.com
Importance:OpinionGemini Omni vs Seedance 2.0 comparison
Gemini Omni VS Seedance 2.0: Who is the true king of video models?
Gemini Omni is more like a future-oriented video editor, while Seedance 2.0 is a more mature AI video generation tool for today. Source: panewslab.com
Importance:NewsAI video generation landscape
The AI Video Race Is Moving Beyond Pretty Clips
Google used its latest I/O event this week to introduce Gemini Omni Flash, a new AI model that can take text, photos, video, and audio as inputs,... Source: forbes.com
Importance:Newscreative tools integration with Gemini
Adobe, Canva, CapCut Are Coming to Gemini to Help You Edit AI Creations
The popular creative tools plan to offer software within Google's AI chatbot. Source: pcmag.com
Importance:NewsByteDance Lance multimodal model
ByteDance Releases Lance Unified Multimodal Model
ByteDance released the open-source multimodal model Lance, a native unified system that handles **image and video understanding, generation, and editing**... Source: letsdatascience.com
Tencent's L2P makes pixel-space image generation practical again
Tencent Youtu Lab and Nanjing University researchers are drawing fresh attention with L2P, a method that transfers latent diffusion models like Alibaba's. Source: startupfortune.com
Importance:NewsGemini Omni multimodal video generation
Google Introduces Gemini Omni: How To Turn Image, Text, Video And Audio Into Single Output
At Google I/O 2026, the tech giant unveiled Gemini Omni, a new multimodal AI model that can create and edit videos using text, images, audio and video... Source: ndtvprofit.com
Importance:ResearchByteDance Lance unified image/video model
One Model, Three Modalities: ByteDance Releases Lance for Image and Video Understanding, Generation, and Editing
ByteDance releases Lance, a 3B native unified multimodal model for image and video understanding, generation, and editing. Source: marktechpost.com
‘Indians among most avid users’: Team behind OpenAI’s Image 2.0 on multilingual AI image generation
India is playing a growing role in shaping how AI image generation models are developed, with OpenAI's ChatGPT Images 2.0 now capable of generating... Source: indianexpress.com
Importance:NewsGemini Omni video generation
Google launches Gemini Omni, new AI video models
Google has unveiled Gemini Omni, a new family of generative AI models capable of producing and editing video from a mix of text, images, audio, and existing... Source: thedailystar.net
Importance:OpinionGemini Omni creative workflow impact
Beyond Text-to-Video: How Gemini Omni Could Change the Creative Workflow for AI Video
The release of Gemini Omni at this year's Google I/O signals a new stage in the race to make AI video more practical, controllable, and useful for everyday... Source: ipsnews.net
Importance:NewsGemini Omni video generation
Google launches Gemini Omni to convert text and images into video, expanding the AI race.
Google launched a new model called Gemini Omni, in a move aimed at expanding generative AI capabilities within the video and visual content industry. Source: jawlah.co