Layer Image

layerai/skills/skills/layer-image

作者 layerai315d06db6f76MIT4 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫8 天前更新

Use when generating a still image with Layer: concept art, key art, illustrations, marketing images, icons, or any text-to-image run. Also when a prompt is not landing, when the aspect ratio or resolution is wrong, when a style or pose reference should steer the result, when readable text must appear in the image, or when several images must share one look. Keywords: txt2img, text to image, prompt, aspect ratio, negative prompt, style reference, seed.

AI 產生的概覽

指導使用 Layer 平台進行文字生圖,涵蓋提示詞、尺寸、參考圖、種子與圖中文字。

功能
說明如何執行 Layer 的文字生圖流程:依用途篩選基礎模型、檢視模型契約、估價、執行並輪詢。它詳述提示詞應指定的內容,例如主體、風格取向、取景、光照、配色與背景,並說明長寬比、解析度、負面提示詞、引導參考圖與種子在不同模型間的差異。也談到讓圖片中出現可讀文字,並提供一個主視覺圖的完整範例。
適用情境
適用於透過 Layer 產生靜態圖片的場景,例如概念圖、主視覺圖、插畫、圖示或行銷圖片。也適用於提示詞效果不理想、長寬比或解析度不對、需要用風格或姿勢參考圖引導結果、圖片中必須出現可讀文字,或需要多張圖片維持同一風格時。
執行需求
需要存取 Layer 平台及其工具,包括基礎模型的列出與檢視、價格估算、執行與輪詢,以及為引導參考圖上傳檔案。此技能不含指令碼,僅為說明文件;其中還引用了用於圖像編輯、遊戲素材與參考集的其他 Layer 技能。

Layer Image Generation

Overview

The loop is the one the layer skill teaches: list_base_models with filter.use_case: "text_to_image", get_base_model, estimate, execute, poll. What separates a usable image from a near miss is the prompt and the per-model contract, not the model choice, which the curated ranking already handles.

Editing an image that already exists is a different use case with different rules: see layer-image-editing. Sprites, icons, and game UI have their own constraints: see layer-game-assets. Holding one look across a set: see layer-reference-sets.

If a sibling skill named here is missing from your available skills, ask the user to install it (npx skills add layerai/skills --skill <name>); unattended, proceed from tool schemas and flag the gap.

What a prompt must name

A model fills every gap you leave, and it fills it with the most average answer in its training data. Each of these is a gap worth closing, in roughly this order of impact:

  1. Subject, concretely. "A dwarf blacksmith" beats "a fantasy character".
  2. Style register, as a production term rather than an artist's name: hand-painted, cel-shaded, flat vector, painterly semi-realism, stylised PBR render, gouache, 90s airbrush.
  3. Shot and framing: full body, three-quarter portrait, top-down, isometric, extreme close-up. This is the single most common omission and the most common cause of a reroll.
  4. Lighting: key direction, hardness, time of day, practical sources. "Rim-lit from behind, warm forge glow below" is a decision; "good lighting" is not.
  5. Palette, named or constrained: "muted teal and rust, two accent hues only".
  6. Background treatment: flat colour, isolated on transparent, environment implied, full scene. Assets that will be composited need this stated, not assumed.

Leave out what does not matter. A prompt that specifies everything specifies nothing, because the model weights the whole string and dilutes the parts you cared about.

Sizing, and the fields that differ

Aspect ratio and resolution are per-model contracts. Two models that both do text-to-image disagree about whether they take an aspect ratio string, explicit width and height, or a fixed set of named sizes, so read get_base_model rather than reusing the shape that worked last time.

When the deliverable has a fixed ratio, filter for it. filter.capabilities is an object of booleans, not a list of names, so it is {portrait_9_16: true}: square, portrait_9_16, landscape_16_9, or multiple_aspect_ratios when the user will want several crops from one setup. Set only what the task needs, since false filters rather than defaults.

Negatives, references, and seeds

negative_prompt is a capability, not a universal parameter. Filter for it when the run needs one, and do not send it blindly: on a model without it, it is either ignored or an error, and on a model with it, a long negative list costs more than it saves. Reach for it to suppress one specific, recurring defect.

Three reference types steer a text-to-image run, all passed through guidance_files after an upload:

  • reference_image with the image_editing capability, for "reproduce this subject" and for new art that has to match a reference. style_reference is the older style-only transfer: it moves a look across, it does not carry a subject, so reaching for it to hold a character is the usual reason a cast drifts.
  • pose with character_pose, for "put the character in this position".
  • depth, canny, lineart, or scribble with structure or outline, for "keep this layout".

A reference is worth more than three adjectives. When a user has an image and is describing it in words, upload the image.

Seeds make a run repeatable, which matters when iterating: hold the seed and change one clause to see what that clause actually does. They do not make two different prompts consistent with each other. For a look held across a set, train it: see layer-reference-sets.

Text in the image

Most image models garble text. When the deliverable needs readable words, filter for a model whose description claims typography, keep the string short, and put it in quotes in the prompt. If it still fails after two attempts, stop rerolling: generate the art clean and composite the text, which is also what makes it editable and localisable later.

Worked example

"Key art for our roguelike, 16:9, for the Steam page."

  1. list_base_models with filter.use_case: "text_to_image" and filter.capabilities: {landscape_16_9: true}. Take the first result.
  2. get_base_model, which reports the sizing fields it accepts and whether it needs a trigger word.
  3. Compose against the six points above: "Hooded rogue mid-leap over a collapsing stone bridge, seen three-quarter from below, hand-painted semi-realism, hard moonlight from the upper left with cold blue rim light, warm torchlight pooling below, muted slate and amber, storm-lit ruins receding into fog behind."
  4. estimate_forge_price with batch_size: 4, to see four compositions of one idea.
  5. Under 20 CUs, so execute, then poll at poll_interval_seconds.
  6. Present all four. Pick one, hold its seed, and iterate one clause at a time.

Common mistakes

  • Omitting shot and framing, then rerolling the same prompt hoping for a different crop.
  • Describing a reference image in words instead of uploading it.
  • Sending negative_prompt to a model that does not declare the capability.
  • Reusing the sizing fields from a different model rather than reading get_base_model.
  • Using batch_size for N different assets. It makes variations of one prompt.
  • Chasing readable text through a sixth reroll instead of compositing it.
  • Stacking five style adjectives, which averages them into none of them.

來源與署名

來源:layerai/skills位於skills/layer-image提交315d06d

授權條款: MIT

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架

更多來自 layerai/skills 的技能

Layer Workflows

layerai

Use when running a saved Layer Blueprint workflow rather than a single generation: discovering what workflows a workspace has, reading a workflow's input schema, estimating and executing a run, polling its steps, or cancelling it. Also when a repeatable multi-step pipeline exists for a task, when the user names a workflow, or when a workflow input expects a style or a file. Keywords: workflow, blueprint, pipeline, app, multi-step, run, node graph.

待分類48 天前更新

Layer Workflow Import

layerai

將其他工具中的工作流程重建為 Layer Blueprint 圖,並透過 import_workflow 匯入。

AI & Agents48 天前更新

Layer Video Timeline

layerai

將現有片段、圖像與音訊組合成一支完整影片時間軸,包含轉場、疊加層與字幕。

Design & Creative48 天前更新

Layer Video

layerai

Use when generating video with Layer: text-to-video, animating a still image, extending a clip, adding camera motion, generating native audio or lip sync, looping animations, or planning a multi-shot ad, trailer, or cutscene. Also when a video prompt produces the wrong motion or the shot drifts off the source image. Keywords: txt2vid, img2vid, image to video, camera motion, video effects, loop, seamless, trailer, cutscene, lipsync.

待分類48 天前更新

Layer Textures

layerai

指導使用 Layer 影像模型產生無縫可平鋪的紋理,涵蓋尺度、光照與重複檢查。

Design & Creative48 天前更新

Layer Reference Sets

layerai

Use when a look must hold across many Layer generations: training a custom style or LoRA on a studio's own artwork, curating the images that go into a reference set, choosing its training category, tuning reference-set weight on a run, or deciding whether to train at all rather than attach a style reference. Also when a trained style produces weak, inconsistent, or silently ignored results. Keywords: LoRA, custom model, trained style, reference set, consistency, on-model, art direction, dataset.

待分類48 天前更新