Sources
Stay ahead of AI art
Get the week's top AI and AI-art stories delivered to your inbox — curated, concise, free.
Free. Unsubscribe any time.

Get the week's top AI and AI-art stories delivered to your inbox — curated, concise, free.
Free. Unsubscribe any time.
Google's July 2026 AI update wave covers everything from a second-generation robotics model to smarter on-device generation — and several of the changes have direct consequences for creators who rely on Google's stack for images and video.
The headline model announcement is Gemini Robotics ER 2, according to the Google AI Blog. The ER in the name stands for embodied reasoning — the model's ability to understand and act within physical environments. For AI creators, the practical relevance isn't robots: it's what improved spatial and physical reasoning in a foundation model tends to unlock downstream. Better object permanence and scene understanding in Gemini's core architecture typically improves prompt-to-image coherence, particularly for complex compositional scenes where objects need to relate correctly in 3D space.
Veo, Google's AI video model, received updates in the July batch. Google hasn't broken out every technical detail publicly, but the July roundup frames the improvements as part of a broader push to make AI video more usable across consumer and developer surfaces. For creators already experimenting with AI video, Veo remains one of the few models with genuine cinematic quality at longer durations — any capability bump matters when you're choosing between it and competitors. If you're weighing options, the AI video generation tools landscape has shifted fast this year.
A recurring theme across Google's July announcements is on-device inference — running AI models locally on Pixel hardware and other Android devices rather than routing every request through Google's data centers. For image and video generation specifically, on-device processing can mean lower latency and fewer round-trips to the cloud, though the trade-off is model size: the most capable generation models still need server-side compute. Google is threading this needle by keeping heavy generation in the cloud while offloading lighter tasks — style suggestions, prompt refinement, upscaling — to the device. That split architecture is worth watching as it could eventually let creators iterate on prompts faster without burning API quota.
Google also announced AI tools aimed at education and helping people thrive in an AI-augmented economy — categories that might seem distant from image generation but reflect how Google is positioning Gemini as infrastructure across every surface, not just a chatbot. The more Gemini improves as a general reasoner, the better it performs as a backend for creative tools built on the Google stack.
For creators using Google's tools directly — whether through AI image generation or third-party apps built on Gemini APIs — the July updates represent incremental but consistent forward movement. No single announcement is a step-change, but the pattern is clear: Google is tightening the integration between its foundation models, its device hardware, and its consumer apps. Creators who build workflows on Google's stack should expect the on-device and cloud split to become more pronounced through the rest of 2026, with the most capable generation remaining cloud-bound while everyday creative assists move closer to the hardware.
The next test will be whether Veo's July improvements show up in third-party integrations — that's where most creators will actually encounter them.