gen‑ai.news

The pulse of generative image & video AI.

Twice a week, the most important stories in image and video generation - new models, notable research, and meaningful product releases - distilled into a 2-minute read. No hype, no filler.

Free. Unsubscribe any time. No spam, ever.

Archive

You’re Trying to Spot AI-Generated Faces Wrong
Image

You’re Trying to Spot AI-Generated Faces Wrong

Spotting an AI-generated face used to be straightforward - look for a mangled ear, an extra finger, or eyes that didn't quite line up. In 2026, those obvious tells have largely disappeared, and the conventional wisdom about how to detect synthetic faces needs revisiting. PetaPixel examines what the research and current tools actually say about identifying AI-generated portraits today.

Proton's privacy-focused Lumo chatbot gets image generation
Image

Proton's privacy-focused Lumo chatbot gets image generation

Proton has updated its Lumo chatbot with image generation and editing capabilities, marking a significant expansion of the privacy-focused AI assistant. The upgrade brings visual creative tools to a platform that distinguishes itself by keeping user data away from ad networks and third-party trackers. For users already in the Proton ecosystem, Lumo now offers a more complete AI workspace without sacrificing the company's data privacy commitments.

The Gemini app is bringing personalized image creation to more users.
Image

The Gemini app is bringing personalized image creation to more users.

Google is expanding Personal Intelligence features in the Gemini app, allowing it to draw on data from Gmail, Google Photos, YouTube, and Search to generate images and content that feel more relevant to individual users. The expansion brings these personalized capabilities to a broader audience, with user permission as a prerequisite. It marks a notable step in Google's effort to make generative AI outputs feel less generic and more grounded in a person's actual context.

No image
Video

A24 Will Survive the AI Backlash, but Some Are Convinced the Company Has ‘Sold Its Soul’

A24, long regarded as a standard-bearer for artistically ambitious cinema, is facing sharp criticism from its own fanbase after announcing a $75 million partnership with Google DeepMind to develop AI workflow tools. While the backlash has been vocal, industry observers largely expect the company to weather it without lasting damage. The deal has nonetheless raised pointed questions about what the studio's brand identity actually stands for.

No image
Image

Databricks’ former AI chief thinks he can cut AI’s power bill by 1,000x

Ali Ghodsi's successor at Databricks has launched a startup claiming its approach to AI inference can reduce energy consumption by a factor of 1,000. The company's first public demonstration, an image-generation system called Un-0, aims to show that its underlying technology can match the output of conventional AI systems at a fraction of the power cost. If the claims hold up under scrutiny, the implications for data center energy demand could be significant.

No image
Multimodal

Adobe acquires image and video enhancement tool maker Topaz Labs

Adobe has acquired Topaz Labs, the company behind a suite of AI-powered image and video enhancement tools. Adobe says it plans to integrate Topaz Labs' technology across its existing applications. The deal brings a well-regarded set of upscaling, sharpening, and noise-reduction tools under Adobe's umbrella.

Less Than a Quarter of Americans Use AI to Create or Edit Images
Multimodal

Less Than a Quarter of Americans Use AI to Create or Edit Images

A new Pew Research study finds that only 24% of Americans use AI tools for creating or editing images and videos. Despite the rapid growth of generative image and video platforms, adoption remains limited across the general population. The findings offer a grounded look at where everyday use of these tools actually stands.

Figma now has AI motion graphics and shader tools
Multimodal

Figma now has AI motion graphics and shader tools

Figma used its annual Config conference to announce a set of updates aimed at tightening the loop between design and development. The additions include AI-generated motion graphics - where animations and transitions are created from text descriptions - alongside new shader tools and a reworked canvas built with full-stack workflows in mind. The changes reflect a broader push to reduce the context-switching that typically slows down creative and engineering teams.

Models Accuse Fashion Brand of Using AI to Recreate Them
Image

Models Accuse Fashion Brand of Using AI to Recreate Them

Several models are accusing fashion retailer Rainbow Shops of using AI to generate digital lookalikes of them, reportedly around the same time their bookings with the brand came to a halt. The allegations raise pointed questions about consent, likeness rights, and the growing use of generative AI in commercial fashion photography. The case is drawing attention as one of the more concrete examples of AI image tools intersecting with labor disputes in the modeling industry.

Man Traumatized After Woman Uses His Photos for AI Social Media Posts Showing Fake Family Life
Image

Man Traumatized After Woman Uses His Photos for AI Social Media Posts Showing Fake Family Life

A man in Singapore discovered that a former schoolmate had been using his personal photos as source material to generate AI images depicting a fabricated family life on social media. The case highlights how generative AI tools can be weaponized for identity-based deception, even by people with only casual access to someone's public photos. Authorities in Singapore are now involved in the investigation.

ByteDance's Seedance 2.5 breaks the 30-second barrier for AI video generation
Video

ByteDance's Seedance 2.5 breaks the 30-second barrier for AI video generation

ByteDance unveiled Seedance 2.5 at Volcano Engine's FORCE conference, a video generation model capable of producing clips longer than 30 seconds - a threshold few AI video tools have crossed. The model is expected to launch in early July alongside four other newly announced AI models from the company.

Cycling Brand is Mocked Over AI Image of Handlebars Protruding From Bike Seat
Image

Cycling Brand is Mocked Over AI Image of Handlebars Protruding From Bike Seat

REI, the outdoor and cycling retailer, drew widespread mockery this week after posting an AI-generated image on Instagram that depicted handlebars growing directly out of a bike seat - a physically impossible configuration that many followers were quick to point out. The incident adds to a growing list of public AI image blunders from brands that have skipped careful review of generated visuals before publishing them.

The Oversight Board says Meta needs to do more to protect regular people from sexualized deepfakes
Video

The Oversight Board says Meta needs to do more to protect regular people from sexualized deepfakes

Meta's Oversight Board has issued recommendations calling on the company to strengthen protections for ordinary people targeted by sexualized AI-generated deepfakes. The board's suggestions focus on making the reporting process easier and more effective for non-public figures. It marks a continued push by oversight bodies to hold major platforms accountable for harms tied to generative AI content.

Google Deepmind and A24 team up on AI filmmaking research
Video

Google Deepmind and A24 team up on AI filmmaking research

Google DeepMind and independent film studio A24 have announced a long-term research partnership focused on AI filmmaking. Google is also making a roughly $75 million investment in A24, according to the Wall Street Journal. The deal pairs one of the leading AI research labs with a studio known for distinctive, filmmaker-driven projects.

The EU doesn't really know what a deepfake is, and that's becoming a problem for retail
Image

The EU doesn't really know what a deepfake is, and that's becoming a problem for retail

A major European retail trade group is pushing back against the EU AI Act's transparency requirements, arguing that AI-generated product imagery - think a sofa in a computer-generated living room - should not be classified alongside deepfakes. The dispute exposes a genuine ambiguity in the regulation's language that has real consequences for how online retail operates. With platforms like Zalando reporting that 90 percent of their marketing content is already AI-generated, the stakes are signifi

No image
Video

Snap spins off AI video team into new company, Dotmo, due to costs

Snap is spinning off its internal AI video team into a new independent company called Dotmo, with the move driven primarily by the high costs of developing generative video technology in-house. The staff involved are departing Snap to focus solely on AI video work under the new entity. It marks another instance of Snap shedding an internal unit rather than continuing to absorb the expense of frontier AI development.

Powering the world’s first AI arts museum
Multimodal

Powering the world’s first AI arts museum

Rafik Anadol Studio has opened Dataland, billed as the world's first museum dedicated to AI arts, with Google Cloud providing the underlying infrastructure and Google Arts & Culture lending institutional support. The museum marks a notable step in bringing generative AI art into a dedicated physical and cultural space. It represents one of the more concrete attempts to treat AI-generated art as a serious curatorial discipline.

Adobe’s redesigned AI studio remembers what your creations look like
Image

Adobe’s redesigned AI studio remembers what your creations look like

Adobe is rolling out a redesigned Firefly AI studio in private beta, bringing editing and image generation into a single interface. A key addition is the ability to save named visual elements - characters, objects, and backgrounds - so they can be reused consistently across projects without drifting in appearance.

Adobe adds AI agents to Photoshop, Premiere, and more Creative Cloud apps
Multimodal

Adobe adds AI agents to Photoshop, Premiere, and more Creative Cloud apps

Adobe is integrating AI agents into its core Creative Cloud applications, including Photoshop and Premiere, allowing users to describe a desired outcome in plain language while the software carries out the underlying multi-step tasks. The rollout also extends to third-party platforms such as ChatGPT and Claude. The move reflects a broader shift in how professional creative tools are beginning to handle complex, multi-action workflows.

Midjourney goes from generating cat images to full-body ultrasound scans
Image

Midjourney goes from generating cat images to full-body ultrasound scans

Midjourney, best known for its AI image generator, has unveiled its first hardware product: a full-body ultrasound scanner designed to image muscle, fat, bone, and organs. CEO David Holz described the device as aiming for image quality comparable to MRI, and envisions it being used as frequently as once a day. The announcement comes alongside plans for a San Francisco spa where the scanner would be available to the public.

Amazon, Nvidia, and AMD bet $310 million on AI startup building 3D world models
Video

Amazon, Nvidia, and AMD bet $310 million on AI startup building 3D world models

Odyssey ML has raised $310 million from Amazon, Nvidia, and AMD, pushing its valuation to $1.45 billion. The startup is focused on building 3D world models - AI systems that can understand and generate structured representations of physical space. The round also draws in notable backers including Google chief scientist Jeff Dean and CIA-linked venture fund IQT.

June Pixel Drop: New features for creators, Gemini upgrades and more
Multimodal

June Pixel Drop: New features for creators, Gemini upgrades and more

Google's June 2026 Pixel Drop brings a set of updates focused on creative tools and productivity, including new text-to-video capabilities powered by Gemini Omni. The update also refines screen recording and improves multitasking across Pixel devices. It continues Google's pattern of rolling out AI-driven features to its hardware lineup through periodic software drops.

No image
Video

Meet Qwen-RobotSuite: Three Embodied AI Models for VLA Manipulation, Video World Modeling, and Navigation

The Qwen team has released Qwen-RobotSuite, a collection of three specialized models targeting different challenges in embodied AI: physical manipulation, world modeling, and navigation. Each model draws on existing Qwen language and vision foundations while introducing architecture and training choices tuned for robotics tasks. The release comes with benchmark results and details on the data pipelines used to train each system.

Photographer Disturbed By AI-Generated ‘Women’ in Beauty Magazine
Image

Photographer Disturbed By AI-Generated ‘Women’ in Beauty Magazine

Austin-based photographer and director of photography Cassandra Klepac recently noticed AI-generated images of women appearing in a beauty magazine, raising concerns about the implications for working photographers. The incident highlights how generative AI is quietly making its way into fashion and beauty editorial content. Her reaction has sparked a broader conversation about transparency, labor, and the future of commercial photography.

Adobe Adds More User Control to AI Features Inside Lightroom and Photoshop
Image

Adobe Adds More User Control to AI Features Inside Lightroom and Photoshop

Adobe has rolled out new Creative Cloud updates to Lightroom and Photoshop that give photographers more control over AI-assisted workflows, particularly around photo culling and selection. The changes are aimed at reducing the time photographers spend manually sorting through large batches of images. The updates reflect Adobe's ongoing effort to make AI tools feel more transparent and adjustable rather than fully automated.

Cutback launches AI tool to automate long-form video editing
Video

Cutback launches AI tool to automate long-form video editing

Cutback has introduced Selects, an AI editing assistant designed to handle the early, time-consuming stages of long-form video editing. The tool ingests raw footage, organizes it automatically, and produces a draft edit based on a single text prompt. It targets creators and editors who spend significant time just getting footage into a workable shape before any real editing begins.

Microsoft Research's Mirage gives video generation a persistent spatial memory that doesn't forget what's around the corner
Video

Microsoft Research's Mirage gives video generation a persistent spatial memory that doesn't forget what's around the corner

Mirage, a video world model developed by Microsoft Research and academic collaborators, introduces a persistent spatial memory system that stores scene information in latent space rather than relying on pixel-based point clouds. The approach keeps environments visually consistent across long camera movements while significantly reducing compute and memory costs. Moving object tracking across segments remains an open limitation.

New AI model called "Count Anything" does exactly what it says, and that's harder than it sounds
Image

New AI model called "Count Anything" does exactly what it says, and that's harder than it sounds

A new model called "Count Anything" aims to be the first general-purpose AI system capable of counting objects in virtually any image using only a text prompt. Researchers report it cuts counting error rates roughly in half compared to prior approaches. The system handles a wide range of subjects - from crowd scenes to microscopic cell samples - though very dense arrangements and vague descriptions remain challenging.

The future of Hollywood isn’t feeding prompts into vanilla gen AI models
Video

The future of Hollywood isn’t feeding prompts into vanilla gen AI models

Despite years of bold claims about AI transforming filmmaking, very few projects have emerged that feel like genuine entertainment audiences would seek out. A new short film, "Dear Upstairs Neighbors," offers a different approach - one built on custom-trained versions of Google DeepMind's Veo and Imagen models rather than off-the-shelf AI tools. It may point toward what a more serious production pipeline actually looks like.

MiniMax M3 launches on NVIDIA platform with Free Endpoint
Multimodal

MiniMax M3 launches on NVIDIA platform with Free Endpoint

MiniMax M3, a multimodal model supporting text, image, and video, is now available through NVIDIA's accelerated compute platform with a free API endpoint. The model incorporates sparse attention mechanisms, making it well-suited for long-context tasks. It marks a notable expansion of MiniMax's distribution reach via NVIDIA's infrastructure.

No image
Video

Cheaper, faster, and culturally aware, Avataar’s video AI is built for India’s scale

Avataar AI has released a distilled video generation model priced at $0.005 per second of output, positioning it as a cost-conscious option for the Indian market. The model is designed with cultural and regional awareness built in, targeting the practical demands of businesses operating at India's scale. It represents a focused approach to making AI video generation accessible beyond Western-centric platforms.

No image
Video

The Model Doesn’t Matter: Inside the Race to Be the AI Production Platform Filmmakers Want to Use

A growing number of platforms - including Artlist Studio, ComfyUI, Flora, and Amazon's Project Nara - are competing to become the go-to AI production environment for filmmakers. The race is less about which underlying model is most capable and more about which workflow best fits how film professionals actually work. IndieWire examines what these platforms are building and what the film industry is looking for.

Google DeepMind releases DiffusionGemma, a model that runs local AI 4x faster
Image

Google DeepMind releases DiffusionGemma, a model that runs local AI 4x faster

Google DeepMind has released DiffusionGemma, an open model that applies diffusion-based generation to text, promising outputs up to four times faster than conventional autoregressive approaches. While diffusion has long been the dominant technique in image generation, its application to language models is still relatively new territory. The release adds a notable open option to a field that has so far seen limited competition.

Maket debuts Auto-Complete for generating residential floor plans
Image

Maket debuts Auto-Complete for generating residential floor plans

Maket has introduced Auto-Complete, a feature that takes a rough sketch and any already-placed rooms and generates a complete residential floor plan from them. The tool then carries that layout into 3D modeling and rendered visuals. It is aimed at streamlining early-stage residential design work.

Apple’s best AI idea looks a lot like vibe coding
Image

Apple’s best AI idea looks a lot like vibe coding

Apple's WWDC announcements were largely familiar ground - chatbots, text summarization, and image generation tools that already exist elsewhere. But buried in the first developer beta of iPadOS 26 is a feature that takes a more interesting direction, drawing comparisons to the "vibe coding" approach of describing what you want and letting AI figure out the rest.

No image
Image

Apple’s Image Playground doesn’t suck anymore

Apple is updating Image Playground, its built-in AI image generator, with improvements that bring it closer to the quality users expect from competing tools. The changes address longstanding criticism that the feature produced stilted, cartoon-like results with limited practical use. It remains to be seen how the updated output holds up against third-party generators available on Apple devices.

iOS 27 gets new AI photo editing tools
Image

iOS 27 gets new AI photo editing tools

Apple's upcoming iOS 27 will bring a new set of AI-powered photo editing tools to iPhone users. The update expands on the computational photography features Apple has been building into its operating system over recent releases. Details on the specific capabilities are emerging ahead of a formal announcement.

Microsoft Research's Lens proves detailed captions matter more than raw scale for training efficient image generators
Image

Microsoft Research's Lens proves detailed captions matter more than raw scale for training efficient image generators

Microsoft Research's Lens is a 3.8-billion-parameter text-to-image model that matches much larger systems on standard benchmarks, trained at a fraction of the usual cost. The key difference is data quality: 800 million detailed captions written by GPT-4.1, rather than the noisy alt-text typically scraped from the web. Model weights and code are released under an open-source license.

No image
Image

Amazon now lets you design custom merch using AI

Amazon has added an AI-powered design tool to its Shopping app, letting users describe an image concept to Alexa and apply the resulting artwork to printable merchandise. The feature covers products such as T-shirts, hoodies, and tumblers. It connects generative image capabilities directly to Amazon's existing print-on-demand infrastructure.

AI ‘content creators’ are getting harder to spot
Image

AI ‘content creators’ are getting harder to spot

AI-generated social media personas have grown harder to distinguish from real people, raising questions about transparency and trust on platforms built around personal identity. Early virtual influencers were visually distinct enough that audiences could easily spot them, but that gap is closing fast. The Verge traces how the technology and the business models around it have matured together.

Meta made its own AI-generated clickbait news feed
Image

Meta made its own AI-generated clickbait news feed

Meta's standalone AI app has introduced a 'For You' section that serves up AI-generated news-style articles, complete with AI-produced images and text. The content follows familiar clickbait patterns, raising questions about accuracy and the platform's direction. It marks a notable shift from the app's original focus on a social feed of user-shared AI conversations and images.

Google shuts down the AI image app Pixel Studio
Image

Google shuts down the AI image app Pixel Studio

Google is closing Pixel Studio, its AI image generation app for Pixel devices, less than two years after it launched. The shutdown continues a pattern of Google retiring products that failed to gain lasting traction. Users will need to look elsewhere for on-device AI image tools.

K-pop Fans Are Calling Out Creepy Deepfakes of Idols
Multimodal

K-pop Fans Are Calling Out Creepy Deepfakes of Idols

AI-generated deepfakes of K-pop idols, including sexualized images and videos, have drawn pushback from within fan communities themselves. Fans are organizing to call out and report this content, framing it as a violation of the artists' dignity. The situation reflects a broader tension as generative tools become easier to access and misuse.

Industry leaders share new perspectives on generative media for startups
Multimodal

Industry leaders share new perspectives on generative media for startups

Google for Startups has published a new report examining how early-stage companies are approaching generative media tools and workflows. The findings draw on perspectives from founders and industry figures navigating this space. The report aims to offer practical context for startups integrating AI-generated image and video into their products.

Let us filter AI slop, you cowards
Multimodal

Let us filter AI slop, you cowards

Content labels on AI-generated images and videos have become more common across major platforms, but critics argue that labeling alone is not enough. The Verge makes the case that YouTube, Instagram, TikTok, and others should go a step further and give users the ability to actively filter AI-generated content from their feeds. Without that option, labels function more as a disclosure footnote than a meaningful tool for audience control.