Muapi Nano Banana

by samuraigpt66b55e2a27eeNo license4 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 9 days ago

Reasoning-driven image generation using structured creative briefs (Gemini 3 style) — generates high-fidelity images via muapi.ai with logic-based prompting

Instructions onlyDesign & Creative
AI-generated overview

Guides agents in writing structured, reasoning-driven image generation prompts for the muapi.ai image service.

What it does
This skill provides an instruction framework for turning a user's image idea into a structured creative brief covering subject, action, context, composition, lighting and style. It also gives rules for negative constraints, identity consistency and precise text rendering in generated images. The output is an optimized prompt intended for an image generation service, not the image file itself.
When to use it
Use it when an agent needs to prepare or improve a prompt for AI image generation, especially when the request involves complex scenes, physical lighting, or legible text in the image. It is also useful for replacing keyword-style prompts with descriptive natural-language briefs.
Requirements
Requires an agent that can call an image generation service; the document references a muapi.ai image generation primitive and a generate-nano-art.sh script, but no scripts are included in this skill.

🍌 Nano-Banana Expert Skill (Gemini 3 Style)

A specialized skill for AI Agents to leverage "Reasoning-Driven" image generation. Based on the advanced prompting architecture of Google's Gemini 3 (Nano Banana Pro), this skill moves beyond keyword stuffing to structured, logic-based creative briefs.

Core Competencies

  1. Reasoning-Driven Prompting: Using natural language logic to define physics, lighting, and spatial relationships.
  2. Structured Creative Briefs: Implementing the "Perfect Prompt" formula: Subject + Action + Context + Composition + Lighting.
  3. Text Rendering Precision: Explicitly defining typography and signifiers for legible text integration.
  4. Contextual Grounding: Using "Search Grounding" logic (simulated) to anchor generations in real-world accuracy.

🏗️ Technical Specification

1. The "Perfect Prompt" Formula

ComponentDescriptionExample
SubjectDetailed entity description"A stoic robot barista with exposed copper wiring"
ActionDynamic interaction"Pouring a latte art leaf with mechanical precision"
ContextEnvironment & Atmosphere"Inside a neon-lit cyberpunk cafe at midnight"
CompositionCamera & Lens choice"Close-up, 85mm lens, f/1.8 aperture"
LightingMood & Direction"Volumetric blue rim light, warm cafe glow"
StyleAesthetic anchor"Cinematic, photorealistic, 4K production value"

2. Advanced Features

  • Negative Constraint Logic: Instead of "no blurry," use "Ensure sharp focus on the subject's eyes."
  • Identity Consistency: (Simulated) "Maintain consistent facial structure across variations."
  • Text Integration: Use double quotes for specific text: The sign reads "OPEN 24/7".

🧠 Prompt Optimization Protocol (Agent Instruction)

Before calling the script, the Agent MUST rewrite the user's prompt into a logic-driven Reasoning Brief:

  1. NO KEYWORD SOUP: Remove "8k, masterpiece, ultra-detailed." Use full, descriptive sentences.
  2. PHYSICAL CONSISTENCY: Describe how elements interact (e.g., "The light from the crystal shards casts caustic patterns across the obsidian floor").
  3. TEXT PRECISION: If the user wants text, define it precisely: featuring a sign that says "STORE NAME" in a weathered serif font.
  4. OPTICAL DIRECTIVES: Specify lens behavior: Shallow Depth of Field (f/1.8), Macro Lens, Anamorphic Flare.

🚀 Protocol: Using Nano-Banana

Step 1: Define the Creative Logic

Provide the agent with a subject and a specific scenario.

Step 2: Invoke the Script

The generate-nano-art.sh script translates the logic into a structured Gemini 3-style prompt.

bash
# Generating a reasoning-driven imagebash scripts/generate-nano-art.sh \  --subject "a glass chess piece" \  --action "shattering into liquid shards" \  --context "on a obsidian table" \  --style "macro photography"

⚠️ Constraints & Guardrails

  • No Keyword Soup: MANDATORY - Do not use "trending on artstation, masterpiece, 8k". Use natural language descriptions.
  • Physics Logic: Ensure the prompt describes physically possible lighting and reflection interactions.
  • Full Sentences: The model parses relationships; use "light reflecting off the water" instead of "water, reflection".

⚙️ Implementation Details

This skill applies a "Logic Wrapper" around the core/media/generate-image.sh primitive, converting fragmented inputs into a coherent, reasoning-ready narrative prompt.

Source and attribution

Source:samuraigpt/generative-media-skillsin.opencode/skills/muapi-nano-bananaat commit66b55e2

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal