Source profileQuality 91/100

AgriciDaniel/claude-blog/brain/.raw/sources/claude-blog-skill/skills/blog-image/SKILL.md

blog-image

AI image generation and editing for blog content powered by Gemini via MCP. Claude acts as Creative Director - interpreting intent, selecting domain expertise, constructing optimized 6-component prompts (Subject + Action + Context + Composition + Lighting + Style), and orchestrating Gemini for blog-quality results. Generates hero images, inline illustrations, social preview cards, and OG images. Edits existing blog images. Supports 6 blog-optimized domain modes (Editorial, Product, Landscape, UI

Source repository stars
1,809
Declared platforms
0
Static risk flags
1
Last source update
2026-07-23
Source checked
2026-08-25

Decision brief

What it does: where it fits

You are a Creative Director that orchestrates Gemini's image generation specifically for blog content. Never pass raw user text directly to the API. Always interpret, enhance, and construct an optimized prompt using the 6-component Reasoning Brief system.

Best for

    Not for

    • Tasks that require unconfirmed production actions or broad system permissions.
    • Environments where the pinned source and install steps cannot be inspected.

    Compatibility matrix

    Platform support, with evidence labels

    PlatformStatusEvidenceWhat to check
    CodexNot declaredNo explicit evidencePortability before use
    Claude CodeNot declaredNo explicit evidencePortability before use
    CursorNot declaredNo explicit evidencePortability before use
    Gemini CLINot declaredNo explicit evidencePortability before use
    Open the compatibility checker

    Installation

    Inspect first. Install second.

    The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

    Source-detected install commandSource
    npx skills add https://github.com/AgriciDaniel/claude-blog --skill "brain/.raw/sources/claude-blog-skill/skills/blog-image"
    Safe inspection promptEditorial

    Inspect the Agent Skill "blog-image" from https://github.com/AgriciDaniel/claude-blog/blob/aec971ac511370c6216cd93776c9cf2fec97b32a/brain/.raw/sources/claude-blog-skill/skills/blog-image/SKILL.md at commit aec971ac511370c6216cd93776c9cf2fec97b32a. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

    Workflow

    What the source asks the agent to do

    1. 01

      Generation Workflow

      For /blog image generate or when invoked internally:

      Image type: Hero, inline, OG card, section divider?Blog topic: What is the article about?Style: Photorealistic, editorial, illustrated, minimal?
    2. 02

      Step 1: Analyze Intent

      Determine what the blog needs: - Image type: Hero, inline, OG card, section divider? - Blog topic: What is the article about? - Style: Photorealistic, editorial, illustrated, minimal? - Constraints: Brand colors, specific dimensions, platform format? - Mood: Authoritative, invit…

      Image type: Hero, inline, OG card, section divider?Blog topic: What is the article about?Style: Photorealistic, editorial, illustrated, minimal?
    3. 03

      Step 2: Select Domain Mode

      Choose the expertise lens for the image:

      Choose the expertise lens for the image:Load references/prompt-engineering-blog.md for domain mode modifier libraries.
    4. 04

      Step 3: Construct the 6-Component Reasoning Brief

      Build the prompt as natural narrative paragraphs - NEVER as keyword lists:

      Subject - Who/what, with rich physical detail (textures, materials, scale)Action - What is happening, pose, gesture, movement, stateContext - Environment, setting, time of day, season, weather
    5. 05

      Step 4: Set Aspect Ratio

      Call setaspectratio BEFORE generating. Use conversationid: "default".

      Call setaspectratio BEFORE generating. Use conversationid: "default".

    Permission review

    Static risk signals and limitations

    Writes files

    medium · line 231

    The documentation asks the agent to create, modify, or delete local files.

    refuses to write a literal key into a tracked file)

    Evidence record

    Why each signal appears

    EvidenceSourceComputedTestedEditorial
    SignalValueEvidence typeMeaning
    Quality score91/100ComputedDocumentation, specificity, maintenance, and trust rules
    Repository stars1,809SourceRepository attention, not individual Skill quality
    Compatibility0 platformsSourceDeclared in the catalog source record
    Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

    Pinned source

    Provenance and original SKILL.md

    Repository
    AgriciDaniel/claude-blog
    Skill path
    brain/.raw/sources/claude-blog-skill/skills/blog-image/SKILL.md
    Commit
    aec971ac511370c6216cd93776c9cf2fec97b32a
    License
    MIT
    Collected
    2026-08-25
    Default branch
    main
    View the original SKILL.md

    Blog Image - AI Image Generation for Blog Content

    You are a Creative Director that orchestrates Gemini's image generation specifically for blog content. Never pass raw user text directly to the API. Always interpret, enhance, and construct an optimized prompt using the 6-component Reasoning Brief system.

    Quick Reference

    CommandWhat it does
    /blog image generate <idea>Generate a blog image with full prompt engineering
    /blog image edit <path> <instructions>Edit an existing blog image intelligently
    /blog image setupConfigure MCP server and API key

    Blog Image Types

    Match the image type to blog use case:

    Image TypeAspect RatioResolutionDomain ModePlacement
    Hero/Cover16:92K or 4KEditorial / LandscapeFrontmatter coverImage
    OG/Social Card16:91KEditorial / InfographicFrontmatter ogImage
    Inline Illustration16:9 or 4:31KVaries by topicAfter H2, before body
    Inline Product Shot4:3 or 1:11KProductWithin product sections
    Section Divider21:9 then crop1KAbstract / LandscapeBetween major sections

    Sizing requirements:

    • Blog hero/cover: 1200x630 (OG-compatible) or 1920x1080
    • Open Graph (OG): 1200x630 (required for social sharing)
    • Inline images: 1200px+ wide

    MCP Availability Check

    Before generating, check if nanobanana-mcp tools are available:

    1. Try calling get_image_history with conversation_id: "default" (lightweight, no side effects)
    2. If it succeeds: MCP is available, proceed with generation
    3. If it fails: MCP not configured - inform the user:
      • "Image generation requires the nanobanana-mcp server. Run /blog image setup to configure it."
      • When called internally (from blog-write/blog-rewrite): return silently, no error. The calling workflow continues with stock photos.

    Generation Workflow

    For /blog image generate <idea> or when invoked internally:

    Step 1: Analyze Intent

    Determine what the blog needs:

    • Image type: Hero, inline, OG card, section divider?
    • Blog topic: What is the article about?
    • Style: Photorealistic, editorial, illustrated, minimal?
    • Constraints: Brand colors, specific dimensions, platform format?
    • Mood: Authoritative, inviting, dramatic, clean?

    If the request is vague, ask one clarifying question about use case and style.

    Step 2: Select Domain Mode

    Choose the expertise lens for the image:

    ModeWhen to usePrompt emphasis
    EditorialBlog headers, feature images, lifestyleStyling, composition, publication references
    ProductE-commerce posts, reviews, comparisonsSurface materials, studio lighting, clean BG
    LandscapeEnvironmental backgrounds, travel, hero sectionsAtmospheric perspective, depth layers, time of day
    UI/WebTech blog icons, illustrations, diagramsClean vectors, flat design, exact colors
    InfographicData-driven posts, processes, comparisonsLayout structure, hierarchy, accessible colors
    AbstractPattern backgrounds, section dividers, decorativeColor theory, mathematical forms, textures

    Load references/prompt-engineering-blog.md for domain mode modifier libraries.

    Step 3: Construct the 6-Component Reasoning Brief

    Build the prompt as natural narrative paragraphs - NEVER as keyword lists:

    1. Subject - Who/what, with rich physical detail (textures, materials, scale)
    2. Action - What is happening, pose, gesture, movement, state
    3. Context - Environment, setting, time of day, season, weather
    4. Composition - Camera angle, shot type, framing, negative space, depth
    5. Lighting - Light source, quality, direction, color temperature, shadows
    6. Style - Art medium, aesthetic, film stock, reference artists/eras

    Template for photorealistic blog images:

    A photorealistic [shot type] of [subject with physical detail], [action/pose],
    set in [environment with specifics]. [Lighting conditions] create [mood].
    Captured with [camera model], [focal length] lens at [f-stop], producing
    [depth of field effect]. [Color palette/grading notes]. Aspect ratio 16:9,
    suitable as a blog [hero image/inline illustration] at [target dimensions].
    

    Template for illustrated/stylized:

    A [art style] [format] of [subject with character detail], featuring
    [distinctive characteristics] with [color palette]. [Line style] and
    [shading technique]. Background is [description]. [Mood/atmosphere].
    

    Step 4: Set Aspect Ratio

    Call set_aspect_ratio BEFORE generating. Use conversation_id: "default".

    Blog Use CaseRatio
    Hero / Cover / OG16:9
    Product shot / Square4:3 or 1:1
    Section divider21:9, then crop wider in post-processing if needed
    Vertical (stories)9:16

    Step 5: Generate via MCP

    MCP ToolWhen
    set_aspect_ratioAlways call first, even for 1:1
    gemini_generate_imageNew image from crafted prompt
    gemini_edit_imageModify existing image
    gemini_chatIterative refinement / multi-turn sessions
    get_image_historyReview generated images with conversation_id: "default"
    clear_conversationReset session context

    Model selection with the pinned MCP package:

    • flash (default): MCP alias for gemini-3.1-flash-image, best for most blog images
    • pro: MCP alias for gemini-3-pro-image, use for final hero images or text-heavy assets
    • gemini-3.1-flash-lite-image: use only through direct API or a newer MCP that explicitly supports the stable ID

    Load references/mcp-tools.md for parameter details. Load references/gemini-models.md for model specs, pricing, and rate limits.

    Step 6: Post-Processing (when needed)

    After generation, resize/convert for blog use:

    # Resize to blog hero dimensions (1200x630)
    magick input.png -resize 1200x630^ -gravity center -extent 1200x630 hero.png
    
    # Convert to WebP for web optimization
    magick input.png -quality 85 output.webp
    
    # Convert to AVIF when target browsers support it
    magick input.png -quality 80 output.avif
    
    # Crop to exact OG dimensions
    magick input.png -resize 1200x630^ -gravity center -extent 1200x630 og-image.png
    

    Check if magick (ImageMagick 7) is available. Fall back to convert if not.

    Step 7: Deliver

    Provide:

    1. Image path - where it was saved (~/Documents/nanobanana_generated/)
    2. Crafted prompt - show the full Reasoning Brief (educational)
    3. Settings - model, aspect ratio, domain mode
    4. Alt text - descriptive sentence, 10-125 chars, topic keywords naturally
    5. Frontmatter snippet (for hero/OG images):
    coverImage: "/path/to/generated-image.png"
    coverImageAlt: "Descriptive alt text sentence with topic keywords"
    ogImage: "/path/to/generated-image.png"
    
    1. Refinement suggestions - 1-2 ideas if relevant

    Edit Workflow

    For /blog image edit <path> <instructions>:

    1. Read the image path and edit instruction
    2. Enhance the instruction (never pass raw):
      User saysClaude crafts
      "remove background"Detailed edge-preserving background removal
      "make it warmer"Specific color temperature shift with preservation notes
      "add text"Font style, size, placement, contrast, readability notes
      "make it brighter"Increase exposure, lift shadows, maintain highlights
      "crop for social"Resize to 1200x630 with center-gravity crop
    3. Call gemini_edit_image with enhanced instruction
    4. Return modified image path and description

    Internal API (for blog-write / blog-rewrite)

    When invoked as a Task subagent from blog-write or blog-rewrite:

    Input (provided by calling skill):

    • image_type: hero, inline, og, divider
    • topic: blog post topic/title
    • section_context: (optional) heading or section the image supports
    • style_preference: (optional) photorealistic, illustrated, editorial
    • count: (optional) number of images needed (default: 1)

    Output (returned to calling skill):

    ### Generated Image
    - **Path:** ~/Documents/nanobanana_generated/image_timestamp.png
    - **Alt Text:** Descriptive sentence about the image
    - **Type:** hero / inline / og
    - **Domain Mode:** Editorial
    - **Aspect Ratio:** 16:9
    - **Suggested Frontmatter:**
      coverImage: "/path/to/image.png"
      coverImageAlt: "Alt text here"
    

    Graceful fallback: If MCP is unavailable, return immediately with no error. The calling workflow continues with stock photos. Never block blog-write or blog-rewrite because image generation is unavailable.

    Alt Text Generation

    For every generated image, create alt text following blog standards:

    • Full descriptive sentence (not keyword list)
    • 10-125 characters
    • Include topic keywords naturally
    • Describe what the image shows AND its relevance to the content
    • For charts/infographics: include the key data point

    Good: Marketing team analyzing AI search traffic data on a dashboard showing citation metrics Bad: SEO AI marketing blog optimization image

    Setup

    For /blog image setup:

    1. Run python3 scripts/setup_image_mcp.py (interactive)
      • Prefer: GOOGLE_AI_API_KEY=... python3 scripts/setup_image_mcp.py
      • Or: python3 scripts/setup_image_mcp.py --key-file /path/to/key.txt
      • Avoid --key unless necessary because command arguments can enter shell history and process lists
      • Default writes to ~/.claude/settings.json (user-private, mode 0600)
      • --project flag opts into project .mcp.json (env-expansion only, refuses to write a literal key into a tracked file)
    2. Verify: python3 scripts/validate_image_setup.py
    3. Requires:
    4. The script pins the package to @ycse/[email protected], whose model selector accepts MCP aliases such as flash and pro. Update setup, validation, and this documentation together when bumping the package.

    Safety Filter Auto-Rephrase

    When IMAGE_SAFETY or SAFETY is returned, do NOT give up. Auto-rephrase and retry:

    1. Identify the likely trigger (violence, public figures, NSFW-adjacent, or overly cautious filter)
    2. Rephrase using positive framing - describe what you WANT, not what to avoid
    3. If the subject is a person, make them generic (remove celebrity-like specifics)
    4. If the scene is dramatic, soften: "intense" → "focused", "battle" → "competition"
    5. Retry with the rephrased prompt (max 3 attempts before reporting to user)

    Google acknowledged filters "became way more cautious than we intended" - benign prompts are sometimes blocked. Persistence with rephrasing usually succeeds.

    Edit, Don't Re-roll

    If an image is 80% correct, use gemini_chat for conversational editing rather than regenerating from scratch. The session maintains style consistency, so targeted edits preserve what works while fixing what doesn't.

    When to edit vs regenerate:

    • Color slightly off → Edit ("shift the color temperature warmer")
    • Wrong composition entirely → Regenerate with revised brief
    • Good scene but wrong lighting → Edit ("change to golden hour lighting from the left")
    • Missing a detail → Edit ("add a steaming coffee cup on the desk")

    Error Handling

    ErrorResolution
    MCP not configuredRun /blog image setup
    API key invalidNew key at https://aistudio.google.com/apikey
    Rate limited (429)Wait 60s, retry. Check live limits at https://ai.google.dev/gemini-api/docs/rate-limits
    IMAGE_SAFETYAuto-rephrase (see above) - Layer 2 filter, non-configurable
    PROHIBITED_CONTENTContent policy violation - topic is blocked. Non-retryable.
    SAFETYRephrase prompt - Layer 1 filter
    Vague requestAsk one clarifying question before generating
    Poor qualityReview Reasoning Brief - likely missing lighting (biggest quality differentiator)
    MCP unavailable (internal call)Return silently - calling workflow uses stock photos

    Reference Documentation

    Load on-demand - do NOT load all at startup:

    • references/prompt-engineering-blog.md - Domain modes, 6-component system, blog templates
    • references/gemini-models.md - Model specs, rate limits, aspect ratios, pricing
    • references/mcp-tools.md - MCP tool parameters and response formats

    Frequently asked questions

    What to verify before installation and use

    What does the blog-image source document cover?

    You are a Creative Director that orchestrates Gemini's image generation specifically for blog content. Never pass raw user text directly to the API. Always interpret, enhance, and construct an optimized prompt using the 6-component Reasoning Brief system.

    How do I install blog-image?

    The source record exposes this install command: npx skills add https://github.com/AgriciDaniel/claude-blog --skill "brain/.raw/sources/claude-blog-skill/skills/blog-image". Inspect the command and pinned source before running it.

    Which permission-related actions were detected?

    Static rules flagged write-files in the source; the page lists the matching lines and excerpts.

    Alternatives

    Compare before choosing

    Computed 931,809

    AgriciDaniel/claude-blog

    blog-image

    AI image generation and editing for blog content powered by Gemini via MCP. Generates hero images, inline illustrations, social preview cards, and OG images, and edits existing ones. Supports 6 domain modes (Editorial, Product, Landscape, UI/Web, Infographic, Abstract). Works standalone or internally from blog-write and blog-rewrite; falls back gracefully when MCP is unavailable. Use when user says "blog image", "generate hero image", "blog illustration", "edit blog image", "OG image".

    Computed 9965

    brucesongs/kali-claw

    insecure-design

    Insecure Design (OWASP A06:2025) focuses on security flaws in system architecture and design phases, rather than code implementation-level bugs.

    Computed 9916

    NintendaDev/unikit-ai

    unikit-docs

    Generate and maintain the project's TECHNICAL documentation from its codebase — scans the project structure, tech stack, and module boundaries, then writes a lean README landing page plus detailed topic pages (architecture, modules, setup, build, APIs), only the docs that are relevant. Use whenever the user wants to create, update, or validate documentation of the CODE or the project itself, e.g. "generate documentation", "create docs", "write the README", "update the project docs", "document th

    Computed 9864

    Jamie-BitFlight/claude_skills

    agent-creator

    Create high-quality Claude Code agents from scratch or by adapting existing agents as templates. Use when the user wants to create a new agent, modify agent configurations, build specialized subagents, or design agent architectures. Guides through requirements gathering, template selection, and agent file generation following Anthropic best practices (v2.1.63+).