AIskimIQ

Daily AI & tech news brief

Archive/image & video generation

🎨 Image & Video Generation

Generative AI for images, video, and creative media — Sora, DALL·E, Midjourney, Stable Diffusion, and more.

957 articles

Importance:NewsNVIDIA LongLive-2.0 real-time video generation

NVIDIA has unveiled 'LongLive-2.0,' a real-time video generation AI that achieves lightweight and high-quality generation through training designed for FP4 quantization.

The news blog specialized in Japanese culture, odd news, gadgets and all other funny stuffs. Updated everyday. Source: gigazine.net

Importance:NewsGemini Omni video editing launch

Google introduces Gemini Omni for AI-generated video creation & editing

The latest model lets users edit videos through conversational prompts while maintaining continuity across scenes, characters, and visual elements. Source: afaqs.com

Importance:NewsGemini Omni video editing in Google Flow

Google Adds Gemini Omni Video Editing to Flow

According to Google's blog post, the company is rolling out **Gemini Omni Flash** to the **Gemini app**, **Google Flow**, and **YouTube Shorts**. Source: letsdatascience.com

Importance:NewsGoogle Gemini Omni launch coverage

Google announces launch of Gemini Omni to enhance AI capabilities

Google launches Gemini Omni to expand Gemini's AI-powered content generation capabilities; The new model enables users to create fully integrated videos... Source: entarabi.com

Importance:NewsAvatar video generation — open model release

Meituan puts avatar video startups under new pressure

Meituan has released LongCat-Video-Avatar 1.5 as an open avatar video model with local deployment potential. The release increases pressure on synthetic. Source: startupfortune.com

Importance:NewsYouTube AI video creator tools

YouTube Will Let Creators Use AI to Insert Themselves into Other People’s Videos

AI Video Generation Industry size was worth around 551.7 million in 2025 and is predicted to grow to around USD 551.7 million by 2035, CAGR 18.37% Source: sphericalinsights.com

Importance:NewsGoogle Gemini Omni multimodal video generation

Google unveils Gemini Omni, a multimodal AI model that generates video from text, images, and audio

Google DeepMind unveiled Gemini Omni at Google I/O, a multimodal AI model family for video generation with implications for decentralized compute and Web3... Source: cryptobriefing.com

Importance:NewsGoogle Gemini Omni AI video production

Google Signals AI Video’s Shift From Clip Generation To Production

Google used its latest I/O event this week to introduce Gemini Omni Flash, a new AI model that can take text, photos, video, and audio as inputs,... Source: forbes.com

Importance:OpinionGoogle I/O Gemini Omni recap

Google I/O Recap

Gemini Omni is a multi-modal video/audio/text input-output model that, as the “Nano Banana for video,” can restyle any video on command into another format. Source: patmcguinness.substack.com

Importance:OpinionGemini Omni vs Seedance 2.0 comparison

Gemini Omni VS Seedance 2.0: Who is the true king of video models?

Gemini Omni is more like a future-oriented video editor, while Seedance 2.0 is a more mature AI video generation tool for today. Source: panewslab.com

Importance:NewsAI video generation landscape

The AI Video Race Is Moving Beyond Pretty Clips

Google used its latest I/O event this week to introduce Gemini Omni Flash, a new AI model that can take text, photos, video, and audio as inputs,... Source: forbes.com

Importance:Newscreative tools integration with Gemini

Adobe, Canva, CapCut Are Coming to Gemini to Help You Edit AI Creations

The popular creative tools plan to offer software within Google's AI chatbot. Source: pcmag.com

Importance:NewsByteDance Lance multimodal model

ByteDance Releases Lance Unified Multimodal Model

ByteDance released the open-source multimodal model Lance, a native unified system that handles **image and video understanding, generation, and editing**... Source: letsdatascience.com

Importance:Researchpixel-space image generation / diffusion models

Tencent's L2P makes pixel-space image generation practical again

Tencent Youtu Lab and Nanjing University researchers are drawing fresh attention with L2P, a method that transfers latent diffusion models like Alibaba's. Source: startupfortune.com

Importance:NewsGemini Omni multimodal video generation

Google Introduces Gemini Omni: How To Turn Image, Text, Video And Audio Into Single Output

At Google I/O 2026, the tech giant unveiled Gemini Omni, a new multimodal AI model that can create and edit videos using text, images, audio and video... Source: ndtvprofit.com

Importance:ResearchByteDance Lance unified image/video model

One Model, Three Modalities: ByteDance Releases Lance for Image and Video Understanding, Generation, and Editing

ByteDance releases Lance, a 3B native unified multimodal model for image and video understanding, generation, and editing. Source: marktechpost.com

Importance:NewsOpenAI multilingual image generation

‘Indians among most avid users’: Team behind OpenAI’s Image 2.0 on multilingual AI image generation

India is playing a growing role in shaping how AI image generation models are developed, with OpenAI's ChatGPT Images 2.0 now capable of generating... Source: indianexpress.com

Importance:NewsGemini Omni video generation

Google launches Gemini Omni, new AI video models

Google has unveiled Gemini Omni, a new family of generative AI models capable of producing and editing video from a mix of text, images, audio, and existing... Source: thedailystar.net

Importance:OpinionGemini Omni creative workflow impact

Beyond Text-to-Video: How Gemini Omni Could Change the Creative Workflow for AI Video

The release of Gemini Omni at this year's Google I/O signals a new stage in the race to make AI video more practical, controllable, and useful for everyday... Source: ipsnews.net

Importance:NewsGemini Omni video generation

Google launches Gemini Omni to convert text and images into video, expanding the AI race.

Google launched a new model called Gemini Omni, in a move aimed at expanding generative AI capabilities within the video and visual content industry. Source: jawlah.co