Design & media

Image Generation

Generate and edit images, character series, product visuals, and cohesive campaign sets from prompts or references.

What it does

Create images from text or references, edit existing photos, and produce consistent character series or related visual sets through CellCog. Requests route by task to Nano Banana 2, GPT Image 1.5 for transparent assets, or Recraft for SVGs and icons, with configurable ratios, styles, and 1K–4K sizes. Output defaults to PNG.

When to use it

  • Consistent character art for comics and storyboards
  • Product hero shots, lifestyle scenes, and multi-angle views
  • Cohesive social posts, ads, and website hero sets
  • Background removal, style transfer, and photo modification

The skill document

Image Generation - AI Image Creation Powered by CellCog

Create professional images with AI - from single images to consistent character sets to product photography.

How to Use

For your first CellCog task in a session, read the cellcog skill for the full SDK reference — file handling, chat modes, timeouts, and more.

OpenClaw (fire-and-forget):

result = client.create_chat(
    prompt="[your task prompt]",
    notify_session_key="agent:main:main",
    task_label="my-task",
    chat_mode="agent",
)

All agents except OpenClaw (blocks until done):

from cellcog import CellCogClient
client = CellCogClient(agent_provider="openclaw|cursor|claude-code|codex|...")
result = client.create_chat(
    prompt="[your task prompt]",
    task_label="my-task",
    chat_mode="agent",
)
print(result["message"])

What Models Do We Use

ModelProviderPrimary Use
Nano Banana 2 (Gemini 3.1 Flash Image)GoogleDefault image generation — photorealistic scenes, complex compositions, text rendering, multi-turn character consistency
GPT Image 1.5OpenAITransparent background images — logos, stickers, product cutouts, overlay graphics
RecraftRecraft AIScalable vector illustrations (SVG) and icon generation

Nano Banana 2 is the default model for all image generation. CellCog's agents intelligently route to other models when the task calls for it — for example, transparent PNGs are automatically handled by GPT Image 1.5, and vector/icon requests go to Recraft. If you'd prefer a specific model, just mention it in your prompt (e.g., "use ChatGPT/OpenAI image generation").

What Images You Can Create

Single Image Creation

Generate any image from a text description:

  • Scenes: "A cozy coffee shop interior with morning light streaming through windows"
  • Portraits: "Professional headshot of a confident woman in business attire"
  • Products: "Minimalist product shot of a white sneaker on a marble surface"
  • Abstract: "Geometric abstract art in navy and gold"
  • Nature: "Misty mountain landscape at sunrise with a lone hiker"

Image Editing

Transform existing images:

  • Style Transfer: "Transform this photo into a watercolor painting"
  • Background Removal: "Remove the background and place on a clean white backdrop"
  • Enhancement: "Enhance the colors and add dramatic lighting"
  • Modification: "Change the person's outfit to a red dress"

Consistent Characters

Create multiple images of the same character in different scenarios:

  • Character Series: "Create a tech entrepreneur character, then show them: 1) At their desk coding, 2) Presenting to investors, 3) Celebrating a product launch"
  • Mascot Variations: "Design a friendly robot mascot, then create versions for: welcome page, error page, success message, loading screen"
  • Story Sequences: "Create a main character, then illustrate them in 5 scenes of a journey"

This is powerful for:

  • Comic strips and storyboards
  • Marketing campaigns with consistent characters
  • Video frame generation
  • Brand mascots across contexts

Product Photography Style

Professional product visuals:

  • Hero Shots: "Product hero shot of a smartwatch on a gradient background"
  • Lifestyle Shots: "Smartphone being used by a person in a modern living room"
  • Flat Lays: "Flat lay of skincare products with botanical elements"
  • 360 Views: "Multiple angles of a leather handbag - front, side, back, detail"

Multiple cohesive images for campaigns or collections:

  • Social Media Sets: "5 Instagram post images for a fitness brand - consistent style, varied content"
  • Website Heroes: "3 hero images for a SaaS landing page - professional, modern, tech-focused"
  • Ad Variations: "4 versions of a product ad with different backgrounds and moods"
  • Blog Illustrations: "Set of 6 illustrations for a blog post about productivity tips"

Reference-Based Generation

Use existing images as references for style, character, or composition:

  • Style Matching: "Create a new image in the same artistic style as this reference"
  • Character Consistency: "Using this person as reference, create a new scene with them hiking"
  • Brand Alignment: "Create product images matching this brand's visual style"
  • Composition Reference: "Create a similar composition but with different subjects"

Image Specifications

AspectOptions
Aspect Ratios1:1 (square), 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 21:9
Sizes1K (~1024px), 2K (~2048px), 4K (~4096px)
StylesPhotorealistic, illustration, watercolor, oil painting, anime, digital art, vector
FormatsPNG (default)

Size recommendations:

  • 1K: Quick iterations, thumbnails, social media posts, drafts
  • 2K: Standard web content, presentations, marketing materials
  • 4K: Hero images, print materials, final deliverables where detail matters

When to Use Agent Team Mode

For image generation, chat_mode="agent team" is recommended for:

  • Complex scenes requiring multiple elements
  • Consistent character series
  • Reference-based generation requiring analysis
  • Sets of related images

For simple single images, chat_mode="agent" can work faster.


Example Image Prompts

Professional headshot:

"Create a professional headshot of a friendly Asian woman in her 30s, wearing a navy blazer, soft studio lighting, neutral gray background, confident but approachable expression. 1:1 square, 2K quality, photorealistic."

Product photography:

"Product shot of a premium wireless earbuds case, matte black finish, on a reflective dark surface with subtle blue accent lighting. Minimalist, high-end tech aesthetic. 4:3 landscape, 4K for hero image."

Consistent character set:

"Create a character: young Black male software developer, casual style with glasses, friendly demeanor. Then create 4 images:

  1. Working at a standing desk with multiple monitors
  2. In a video call meeting, explaining something
  3. At a coffee shop with laptop, thinking
  4. Celebrating with team, high-fiving Keep the character exactly consistent across all images."

Social media set:

"Create 5 Instagram posts for a plant-based meal delivery service:

  1. Colorful Buddha bowl from above
  2. Happy person unpacking delivery
  3. Meal prep containers arranged neatly
  4. Close-up of fresh ingredients
  5. Before/after showing ingredients to finished dish Style: bright, fresh, appetizing, consistent warm color grading. 1:1 square format."

Style transfer:

"Transform this uploaded photo of a city street into a Studio Ghibli anime style illustration. Keep the composition and elements but apply the characteristic Ghibli warmth, soft clouds, and whimsical details."


Tips for Better Images

  1. Be descriptive: "Woman in office" is vague. "Confident woman in her 40s, silver blazer, modern glass-walled office, warm afternoon light" is better.

  2. Specify style: "Photorealistic", "digital illustration", "watercolor", "minimalist vector".

  3. Describe lighting: "Soft natural light", "dramatic side lighting", "golden hour glow", "studio lighting".

  4. Include mood: "Professional and confident", "warm and inviting", "energetic and vibrant".

  5. Mention composition: "Rule of thirds", "centered symmetry", "close-up", "wide establishing shot".

  6. For consistency: When creating character series, describe the character in detail first, then reference "the same character" in subsequent prompts.


If CellCog is not installed

Claude Code, Cursor, Codex + 70 more agents: npx skills add cellcog/skills --skill cellcog OpenClaw: clawhub install cellcog CellCog plugin users: run /cellcog-setup (or /cellcog:cellcog-setup depending on your tool) Manual setup: pip install -U cellcog and set CELLCOG_API_KEY. See the cellcog skill for SDK reference.

Questions people ask

Which image models are used?
Nano Banana 2 is the default. Transparent PNG requests are routed to GPT Image 1.5, while vector illustrations, SVGs, and icons are handled by Recraft; a preferred model can also be named in the prompt.
Can it keep a character consistent across multiple images?
Yes. Define the character in detail first, then request scenes featuring the same character; agent team mode is recommended for character series and other related image sets.
What output options can I specify?
You can request listed aspect ratios from 1:1 through 21:9, sizes of 1K, 2K, or 4K, and styles including photorealistic, watercolor, anime, digital art, and vector. PNG is the default format.

Related skills

Generate scripted, animated, scored, and edited cinematic videos from a prompt.

105 installs2 stars

Turn structured story prompts into sequential comic art with consistent characters, layouts, and visual flow.

110 installs3 stars

Creates raster article covers with configurable composition, color, rendering, typography, mood, and aspect ratio.

261 installs7 stars

Generate coordinated game art, audio, 3D assets, UI, and design documents from one brief.

137 installs4 stars

Generates and refines audience-targeted memes, reaction images, and text jokes from specific prompts.

103 installs6 stars

Create scripted, voiced, scored, and edited videos up to four minutes from a single prompt.

299 installs24 stars