设计与多媒体

Image Generation

通过文本或参考图生成与编辑图像,支持多模型路由、角色一致性以及电商产品图拍摄。

它能做什么

基于 CellCog SDK 从文本生成图像,或对已有图像进行编辑、风格转换与背景处理。内置三种模型自动路由:默认的 Nano Banana 2(用于写实场景、多轮角色一致性、文本渲染)、GPT Image 1.5(透明背景、Logo、贴纸、抠图)以及 Recraft(SVG 矢量插画与图标)。输出支持 1K/2K/4K 分辨率、八种长宽比,以及写实、水彩、动漫、矢量等风格。角色系列、参考图生成与多元素复杂场景建议使用 agent team 聊天模式。

什么时候用它

  • 文本生成场景、人物肖像、产品图与抽象艺术
  • 为漫画、吉祥物或营销活动创建一致角色系列
  • 电商产品摄影:主图、生活方式图、平铺图、多角度图
  • 基于参考图进行风格迁移或角色一致性生成

技能文档

Image Generation - AI Image Creation Powered by CellCog

Create professional images with AI - from single images to consistent character sets to product photography.

How to Use

For your first CellCog task in a session, read the cellcog skill for the full SDK reference — file handling, chat modes, timeouts, and more.

OpenClaw (fire-and-forget):

result = client.create_chat(
    prompt="[your task prompt]",
    notify_session_key="agent:main:main",
    task_label="my-task",
    chat_mode="agent",
)

All agents except OpenClaw (blocks until done):

from cellcog import CellCogClient
client = CellCogClient(agent_provider="openclaw|cursor|claude-code|codex|...")
result = client.create_chat(
    prompt="[your task prompt]",
    task_label="my-task",
    chat_mode="agent",
)
print(result["message"])

What Models Do We Use

ModelProviderPrimary Use
Nano Banana 2 (Gemini 3.1 Flash Image)GoogleDefault image generation — photorealistic scenes, complex compositions, text rendering, multi-turn character consistency
GPT Image 1.5OpenAITransparent background images — logos, stickers, product cutouts, overlay graphics
RecraftRecraft AIScalable vector illustrations (SVG) and icon generation

Nano Banana 2 is the default model for all image generation. CellCog's agents intelligently route to other models when the task calls for it — for example, transparent PNGs are automatically handled by GPT Image 1.5, and vector/icon requests go to Recraft. If you'd prefer a specific model, just mention it in your prompt (e.g., "use ChatGPT/OpenAI image generation").

What Images You Can Create

Single Image Creation

Generate any image from a text description:

  • Scenes: "A cozy coffee shop interior with morning light streaming through windows"
  • Portraits: "Professional headshot of a confident woman in business attire"
  • Products: "Minimalist product shot of a white sneaker on a marble surface"
  • Abstract: "Geometric abstract art in navy and gold"
  • Nature: "Misty mountain landscape at sunrise with a lone hiker"

Image Editing

Transform existing images:

  • Style Transfer: "Transform this photo into a watercolor painting"
  • Background Removal: "Remove the background and place on a clean white backdrop"
  • Enhancement: "Enhance the colors and add dramatic lighting"
  • Modification: "Change the person's outfit to a red dress"

Consistent Characters

Create multiple images of the same character in different scenarios:

  • Character Series: "Create a tech entrepreneur character, then show them: 1) At their desk coding, 2) Presenting to investors, 3) Celebrating a product launch"
  • Mascot Variations: "Design a friendly robot mascot, then create versions for: welcome page, error page, success message, loading screen"
  • Story Sequences: "Create a main character, then illustrate them in 5 scenes of a journey"

This is powerful for:

  • Comic strips and storyboards
  • Marketing campaigns with consistent characters
  • Video frame generation
  • Brand mascots across contexts

Product Photography Style

Professional product visuals:

  • Hero Shots: "Product hero shot of a smartwatch on a gradient background"
  • Lifestyle Shots: "Smartphone being used by a person in a modern living room"
  • Flat Lays: "Flat lay of skincare products with botanical elements"
  • 360 Views: "Multiple angles of a leather handbag - front, side, back, detail"

Multiple cohesive images for campaigns or collections:

  • Social Media Sets: "5 Instagram post images for a fitness brand - consistent style, varied content"
  • Website Heroes: "3 hero images for a SaaS landing page - professional, modern, tech-focused"
  • Ad Variations: "4 versions of a product ad with different backgrounds and moods"
  • Blog Illustrations: "Set of 6 illustrations for a blog post about productivity tips"

Reference-Based Generation

Use existing images as references for style, character, or composition:

  • Style Matching: "Create a new image in the same artistic style as this reference"
  • Character Consistency: "Using this person as reference, create a new scene with them hiking"
  • Brand Alignment: "Create product images matching this brand's visual style"
  • Composition Reference: "Create a similar composition but with different subjects"

Image Specifications

AspectOptions
Aspect Ratios1:1 (square), 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 21:9
Sizes1K (~1024px), 2K (~2048px), 4K (~4096px)
StylesPhotorealistic, illustration, watercolor, oil painting, anime, digital art, vector
FormatsPNG (default)

Size recommendations:

  • 1K: Quick iterations, thumbnails, social media posts, drafts
  • 2K: Standard web content, presentations, marketing materials
  • 4K: Hero images, print materials, final deliverables where detail matters

When to Use Agent Team Mode

For image generation, chat_mode="agent team" is recommended for:

  • Complex scenes requiring multiple elements
  • Consistent character series
  • Reference-based generation requiring analysis
  • Sets of related images

For simple single images, chat_mode="agent" can work faster.


Example Image Prompts

Professional headshot:

"Create a professional headshot of a friendly Asian woman in her 30s, wearing a navy blazer, soft studio lighting, neutral gray background, confident but approachable expression. 1:1 square, 2K quality, photorealistic."

Product photography:

"Product shot of a premium wireless earbuds case, matte black finish, on a reflective dark surface with subtle blue accent lighting. Minimalist, high-end tech aesthetic. 4:3 landscape, 4K for hero image."

Consistent character set:

"Create a character: young Black male software developer, casual style with glasses, friendly demeanor. Then create 4 images:

  1. Working at a standing desk with multiple monitors
  2. In a video call meeting, explaining something
  3. At a coffee shop with laptop, thinking
  4. Celebrating with team, high-fiving Keep the character exactly consistent across all images."

Social media set:

"Create 5 Instagram posts for a plant-based meal delivery service:

  1. Colorful Buddha bowl from above
  2. Happy person unpacking delivery
  3. Meal prep containers arranged neatly
  4. Close-up of fresh ingredients
  5. Before/after showing ingredients to finished dish Style: bright, fresh, appetizing, consistent warm color grading. 1:1 square format."

Style transfer:

"Transform this uploaded photo of a city street into a Studio Ghibli anime style illustration. Keep the composition and elements but apply the characteristic Ghibli warmth, soft clouds, and whimsical details."


Tips for Better Images

  1. Be descriptive: "Woman in office" is vague. "Confident woman in her 40s, silver blazer, modern glass-walled office, warm afternoon light" is better.

  2. Specify style: "Photorealistic", "digital illustration", "watercolor", "minimalist vector".

  3. Describe lighting: "Soft natural light", "dramatic side lighting", "golden hour glow", "studio lighting".

  4. Include mood: "Professional and confident", "warm and inviting", "energetic and vibrant".

  5. Mention composition: "Rule of thirds", "centered symmetry", "close-up", "wide establishing shot".

  6. For consistency: When creating character series, describe the character in detail first, then reference "the same character" in subsequent prompts.


If CellCog is not installed

Claude Code, Cursor, Codex + 70 more agents: npx skills add cellcog/skills --skill cellcog OpenClaw: clawhub install cellcog CellCog plugin users: run /cellcog-setup (or /cellcog:cellcog-setup depending on your tool) Manual setup: pip install -U cellcog and set CELLCOG_API_KEY. See the cellcog skill for SDK reference.

相关技能

一条提示词生成电影级 AI 视频——剧情短片、品牌片、音乐 MV 都能做。

105 次安装2 星标

在分镜与跨页之间保持角色形象一致的连续漫画与条漫生成。

110 次安装3 星标

通过类型、调色板、渲染、文本、情绪和字体六个可定制维度生成文章封面图。

261 次安装7 星标

从一段提示词生成风格统一的游戏美术、3D 模型、UI、配乐与 GDD。

作者 CellCog137 次安装4 星标

基于 CellCog 的梗图生成器,研究网络热点、定位受众并产出多角度候选供筛选。

103 次安装6 星标

一条提示词生成最长 4 分钟的视频 —— 自动完成脚本、配音、配乐与剪辑。

299 次安装24 星标