gen‑ai.news
← Back
Video

Higgsfield AI ships new video features in a day with GPT-6 Astra

Higgsfield AI, a startup focused on AI-generated video for advertising, has shared how it is using OpenAI's GPT-4o Astra to dramatically shorten the time it takes to ship new features on its platform. According to the company, what previously required days or weeks of development work can now be completed within a single day, allowing its small engineering team to iterate and release at a pace that would otherwise require a much larger staff.

The core use case Higgsfield is targeting is video ad creation for small businesses - a segment that has historically been underserved because professional video production is expensive and time-consuming. By combining its own generative video models with Astra's ability to assist in writing, debugging, and reasoning through code, Higgsfield says it can bring creative tools to market quickly enough to respond to what its users actually need rather than working from long-range roadmaps.

GPT-4o Astra, OpenAI's agent-oriented capability layer, is designed to handle complex, multi-step tasks that go beyond single-turn question-and-answer interactions. For software teams, this means it can assist across the full development loop - from drafting logic and writing code to identifying errors and suggesting fixes - functioning more like a collaborative technical partner than a simple autocomplete tool. Higgsfield's use of it reflects a broader pattern of AI-native startups leaning on large language models not just for their end products but for the internal work of building those products.

The practical outcome for Higgsfield's customers is a platform that can evolve more rapidly, with new creative controls and workflow improvements arriving on a faster cadence. For the wider generative video industry, the case illustrates how smaller teams can remain competitive against larger players by using AI assistance to multiply their output - without necessarily scaling headcount in proportion to their ambitions.

Read at OpenAI →
Share:X

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

No image
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has introduced Gemini 3.8 Live, an updated multimodal model paired with a new Live Avatar feature that generates an animated, talking on-screen presence during real-time conversations. The combination allows users to interact with a responsive visual agent rather than a purely voice-based interface. The release marks another step in Google's effort to make AI interactions feel more immediate and embodied.