AIskimIQ

Daily AI & tech news brief

Archive/image & video generation

🎨 Image & Video Generation

Generative AI for images, video, and creative media — Sora, DALL·E, Midjourney, Stable Diffusion, and more.

963 articles

Importance:NewsGemini Omni multimodal video generation

Google Introduces Gemini Omni: How To Turn Image, Text, Video And Audio Into Single Output

At Google I/O 2026, the tech giant unveiled Gemini Omni, a new multimodal AI model that can create and edit videos using text, images, audio and video... Source: ndtvprofit.com

Importance:ResearchByteDance Lance unified image/video model

One Model, Three Modalities: ByteDance Releases Lance for Image and Video Understanding, Generation, and Editing

ByteDance releases Lance, a 3B native unified multimodal model for image and video understanding, generation, and editing. Source: marktechpost.com

Importance:NewsOpenAI multilingual image generation

‘Indians among most avid users’: Team behind OpenAI’s Image 2.0 on multilingual AI image generation

India is playing a growing role in shaping how AI image generation models are developed, with OpenAI's ChatGPT Images 2.0 now capable of generating... Source: indianexpress.com

Importance:NewsGemini Omni video generation

Google launches Gemini Omni, new AI video models

Google has unveiled Gemini Omni, a new family of generative AI models capable of producing and editing video from a mix of text, images, audio, and existing... Source: thedailystar.net

Importance:OpinionGemini Omni creative workflow impact

Beyond Text-to-Video: How Gemini Omni Could Change the Creative Workflow for AI Video

The release of Gemini Omni at this year's Google I/O signals a new stage in the race to make AI video more practical, controllable, and useful for everyday... Source: ipsnews.net

Importance:NewsGemini Omni video generation

Google launches Gemini Omni to convert text and images into video, expanding the AI race.

Google launched a new model called Gemini Omni, in a move aimed at expanding generative AI capabilities within the video and visual content industry. Source: jawlah.co

Importance:NewsGoogle Gemini Omni video creation

Google's Gemini Omni brings AI video generation and conversational editing

Google expanded Gemini into AI video creation with Gemini Omni, allowing users to generate and edit videos using text, images, audio and video inputs... Source: storyboard18.com

Importance:NewsGoogle Gemini Omni multimodal

Gemini Omni Flash adds multimodal AI video creation to Google ecosystem

Google has unveiled Gemini Omni, a new multimodal AI model designed to generate and edit videos using combinations of text, images, audio, and video prompts... Source: indianexpress.com

Importance:NewsGoogle Gemini Omni video generation

Google’s Gemini Omni Model Creates Video From Text, Images, and Audio

Last year's Nano Banana brought Gemini to image generation. Now Google is going a step further. Gemini Omni is the company's new model family that brings... Source: talkandroid.com

Importance:LaunchGemini Omni official announcement

Introducing Gemini Omni

Last year, Nano Banana brought Gemini's intelligence to image generation and editing. Since then, it's helped millions of people restore old photos,... Source: blog.google

Importance:LaunchGemini Omni model family announcement

Gemini Omni is a new family of AI models meant to ‘create anything’

Google is announcing a major new family of generative media models during Google I/O that it calls Gemini Omni. The first Omni model, Omni Flash, is able to... Source: theverge.com

Importance:NewsGemini Omni video generation availability

Gemini Omni, the ‘create anything’ model, starts today with lifelike video

Google has unveiled Gemini Omni, a new family of generative models designed to “create anything,” and you can use it today to create surprisingly realistic... Source: 9to5google.com

Importance:NewsGemini Omni and 3.5 Flash launch

Google targets AI agents and video generation with Gemini 3.5 Flash and Omni

Google LLC today introduced two new generative artificial intelligence models that push its Gemini family further into AI agents and multimodal creation:... Source: siliconangle.com

Importance:NewsGemini Omni Google I/O reveal

Gemini Omni is Google's new world model, with advanced AI video generation capabilities

At Google I/O 2026, Google unveiled a new AI model called Gemini Omni. Source: mashable.com

Importance:NewsGemini Omni enterprise implications

Google unveils Gemini Omni 'any-to-any' AI model: what enterprises should know

The model marks Google's bid to collapse the multimodal generative stack — text-to-image, image-to-video, video-to-video, audio generation — into a single... Source: venturebeat.com

Importance:NewsGemini Omni multimodal video generation

Google’s Gemini Omni turns images, audio, and text into video — and that’s just the start

Google's Gemini Omni is a new multimodal model that reasons across text, images, audio, and video to generate and edit videos through simple conversation... Source: techcrunch.com

Importance:OpinionGoogle AI tools for creatives

Inside Google’s quest to build AI products for creatives

When Google's Nano Banana image-generation tool first appeared in the wild in summer 2025, it quickly captured the internet's attention for its ability to... Source: fastcompany.com

Importance:NewsGemini Omni rollout details

Google rolls out Gemini Omni AI for video generation and editing

Google has officially introduced Gemini Omni, a multimodal AI model that integrates reasoning abilities with creative generation across video, image, audio,... Source: testingcatalog.com

Importance:NewsGemini Omni and Spark agent launch

Google rolls out Gemini 3.5 Flash, Spark agent, and new video generation tools

Google introduced Gemini Spark, a 24/7 AI agent for background tasks, alongside Gemini Omni, a new model for cinematic video generation. Source: cryptobriefing.com

Importance:Launchopen-source multimodal image/video model

ByteDance's Lance Puts Open, Efficient Multimodal AI Within Reach

ByteDance released Lance, a 3B-parameter multimodal model under Apache 2.0, offering open, commercially usable image and video generation and editing. Source: startupfortune.com