Hedra
  • Enterprise
  • Pricing
  • Blog
  • Creators
Log inSign Up
Open Hedra
Your account
    Explore
  • Home
  • Enterprise
  • Pricing
  • Blog
  • Creators
    Log inSign Up
    Open Hedra

Imagen 4

All image models
Image modelGoogle
OverviewText → Image

Overview

Imagen 4 is a text-to-image model developed by Google DeepMind. It generates photorealistic visuals at up to 2K resolution and features improvements in typography and fine detail rendering. The model is available in multiple tiers, including a high-speed variant and a high-fidelity Ultra version. It is built for professional branding, intricate scene composition, and design tasks that require precise text integration and complex lighting.

Imagen 4 Text to Image

Text → Image generation.

Specifications

Input mode
Text → Image
Aspect ratios
16:9, 9:16, 1:1, 4:3, 3:4
Pricing
8 credits / image — higher resolutions cost more
Typical generation time
~24s
Free tier
No

Text → Image examples

A medium close-up of a distinguished older man with silver hair and a beard, wearing a tweed jacket. He sits behind a desk with maps, compasses, and a globe in the background. This image was generated using the Imagen4 model at a resolution of 1408x768.Imagen4: Distinguished Scholar in Vintage StudyA digital image generated by Imagen4 at 2816x1536 resolution, showing a male parkour runner mid-air between wet rooftops in a futuristic city at night. Bright neon signs reading FUTURESCAPE and LIVE BOLD illuminate the scene, with skyscrapers and steam vents in the background.Parkour Athlete Leaping Across Cyberpunk Rooftops — Imagen4A high-resolution 2816x1536 image of a whole, ripe watermelon with dark and light green striped patterns resting on a light wooden cutting board. The background is clean and solid white. The image was generated using the Imagen4 model.Imagen4: Whole Striped Watermelon on Cutting BoardA vertical portrait of a young woman with slicked-back, wet-look hair and futuristic makeup, featuring bleached eyebrows and shimmering silver eyelashes. She wears a black top and silver jewelry against a blurred, brightly lit corridor background. This 768x1408 image was generated using the Imagen4 model.Futuristic Portrait with Silver Makeup — Imagen4A vertical portrait of a young woman with slicked-back wet hair, bleached eyebrows, and shimmering silver eyeshadow and eyelashes. She wears a black top and a thin silver chain necklace, looking directly at the camera. The background is a softly blurred hallway. Generated at 768x1408 resolution using Imagen4.Editorial Beauty Portrait — Imagen4A high-resolution 768x1408 image generated by Imagen4 showing a close-up, wide-angle selfie of a laughing young woman with baby's breath flowers in her hair. She wears a white lace-trimmed dress inside a sunlit greenhouse filled with lush green foliage.Smiling Woman in Greenhouse, by Imagen4An overhead shot of a wooden desk where hands sketch a stylized cartoon boy with glasses in a spiral notebook. Scattered pencils and a cup of black coffee sit nearby, generated by Imagen4 at 1408x768 resolution.Artist Sketching Cartoon Boy on Paper, by Imagen4

What is Imagen 4 best used for?

Imagen 4 excels at photorealism, fine detail rendering, and strict prompt adherence. Community consensus highlights its ability to accurately render difficult materials like glass and skin tones, maintain coherent depth-of-field, and generate legible typography. It is suited for complex scene compositions, professional branding, and marketing assets where precise lighting and text are required.

When was Imagen 4 released and what is its lineage?

Developed by Google DeepMind, Imagen 4 was announced on May 20, 2025, succeeding Imagen 3. While Google later introduced the lightweight Nano Banana (based on Gemini 2.5 Flash) as the default generator in its consumer apps, Imagen 4 remains the flagship standalone API model for high-resolution (up to 2K) generation and complex prompt following.

How can I get the best results with Imagen 4?

Use the 10,000-token context window by writing highly detailed, multi-element prompts. Specify camera angles, film grain, lighting, and exact textures, as the model obeys stylistic directions strictly. If you need rapid ideation, the Fast variant generates images in under three seconds. For final production assets, the Ultra variant provides native 2K resolution. All outputs contain an invisible SynthID watermark embedded at the pixel level. For more details, consult Google's official documentation.

Similar models

Dreamina 3.1ByteDanceFlux 1.1 ProBlack Forest LabsFlux 1.1 UltraBlack Forest LabsFlux Kontext MaxBlack Forest LabsFlux Kontext ProBlack Forest LabsSeedream 4.0ByteDance

Prompt tips

  • Max out details: Take advantage of the large context window by writing descriptive, paragraph-long prompts detailing lighting, camera angles, and textures.,- Specify text clearly: When generating text, use quotes for the exact words and describe the font style clearly (e.g., bold serif typography reading "SALE").,- Counteract smoothness: If images look too artificial, explicitly add terms like film grain, raw photo, or subtle imperfections to ground the realism.,- Use seeds for consistency: Leverage seed values to maintain character or style consistency across multiple generations.
Illustration of a laptop with the Hedra spark logo in front of a city skyline at sunset

What Will You Create?

Sign up for free

Product

StudioCommunityFeedbackUse CasesModels

Legal

Privacy PolicyTerms of useAcceptable useCookie PolicyBiometric data policy

Company

AboutTeamChangelogCareersCreatorsSupportAlternatives
LinkedinInstagramDiscord
support@hedra.comHedra 2026 — All rights reserved