HappyHorse 1.0 — Pro Pack on RunComfy

作者 prime-skillsfca19ae084c2MIT收錄於 2026年10月8日更新於 2026年10月8日

Generate text-to-video with HappyHorse 1.0 on RunComfy. Documents HappyHorse 1.0's strengths (#1 on Artificial Analysis Video Arena, native 1080p with in-pass synchronized audio, multi-shot character consistency, 6-language prompt support), the duration / aspect-ratio / resolution schema, and when to route to Wan 2.7 / Seedance 2 / LTX 2 instead. Calls `runcomfy run happyhorse/happyhorse-1-0/text-to-video` through the local RunComfy CLI. Triggers on "happyhorse", "happy horse", "happyhorse 1.0", "happyhorse video", or any explicit ask to generate video with this model.

僅含說明Design & Creative
AI 產生的概覽

透過 RunComfy CLI 使用 HappyHorse 1.0 模型生成文字轉影片,並提供提示詞與模型選擇指引。

功能
此技能說明如何呼叫託管於 RunComfy 的 HappyHorse 1.0 文字轉影片模型,涵蓋提示詞、長寬比、解析度、時長、隨機種子與浮水印等輸入參數。它提供提示詞撰寫建議、範例提示詞、與同類模型的選用比較,以及 CLI 呼叫範例。執行後會產生影片檔案並下載到指定的輸出目錄。
適用情境
當使用者明確要求使用 HappyHorse 或 happy horse 影片,或需要多鏡頭角色一致性、同一次生成中同步音訊、原生 1080p 文字轉影片輸出時使用。它也可用來判斷何時應改用其他影片模型。
執行需求
需要全域安裝 npm 套件 @runcomfy/cli,並擁有已透過 runcomfy login 登入的 RunComfy 帳號,或在 CI 與容器中設定 RUNCOMFY_TOKEN 環境變數。需要連線至 RunComfy 相關端點的網路存取。此技能不含指令碼,僅為說明文件。

HappyHorse 1.0 — Pro Pack on RunComfy

runcomfy.com · Text-to-video · GitHub

HappyHorse 1.0 — currently #1 on Artificial Analysis Video Arena (Elo 1333 t2v / 1392 i2v) — hosted on the RunComfy Model API. Native 1080p video with in-pass synchronized audio (dialogue, ambient, Foley) and multi-shot character consistency.

bash
npx skills add agentspace-so/runcomfy-skills --skill happyhorse-1-0 -g

When to pick this model (vs siblings)

You wantUse
Multi-shot story with character / wardrobe consistencyHappyHorse 1.0
Native audio in the same generation passHappyHorse 1.0
Currently-#1 blind-vote video modelHappyHorse 1.0
Detailed lip-synced dialogue + reference videoSeedance 2.0 Pro
Fine motion control + multi-reference conditioningWan 2.7
Ultra-fast iteration (sub-second per frame)LTX 2
Cinematic motion editing on existing footageKling Video O1

If the user said "HappyHorse" / "happy horse video" explicitly, route here regardless.

Prerequisites

  1. RunComfy CLI — npm i -g @runcomfy/cli
  2. RunComfy account — runcomfy login opens a browser device-code flow.
  3. CI / containers — set RUNCOMFY_TOKEN=<token> instead of runcomfy login.

Endpoints + input schema

happyhorse/happyhorse-1-0/text-to-video

FieldTypeRequiredDefaultNotes
promptstringyes—Up to 2,500 chars. 6 languages (CN/EN/JP/KR/DE/FR).
aspect_ratioenumno16:916:9, 9:16, 1:1, 4:3, 3:4 only.
resolutionenumno1080P720P or 1080P.
durationintno53–15 seconds.
seedintno00..2^31-1. Reuse for variant comparisons.
watermarkboolnotrueProvider watermark.

How to invoke

Default (16:9 1080p 5s):

bash
runcomfy run happyhorse/happyhorse-1-0/text-to-video \  --input '{"prompt": "<user prompt>"}' \  --output-dir <absolute/path>

Vertical short (9:16, 8s, no watermark):

bash
runcomfy run happyhorse/happyhorse-1-0/text-to-video \  --input '{    "prompt": "<user prompt>",    "aspect_ratio": "9:16",    "duration": 8,    "watermark": false  }' \  --output-dir <absolute/path>

Cheaper test pass (720p):

bash
runcomfy run happyhorse/happyhorse-1-0/text-to-video \  --input '{"prompt": "<user prompt>", "resolution": "720P", "duration": 3}' \  --output-dir <absolute/path>

The CLI submits, polls every 2s until terminal, then downloads any *.runcomfy.net / *.runcomfy.com URL from the result into --output-dir. Stdout is the result JSON. Stderr is progress.

Prompting — what actually works

Describe motion over time, not a still. "A woman turns from the window, walks two paces to the desk, picks up the cup, lifts it to her face, takes a sip" beats "a woman drinking coffee".

Camera + shot in plain English. Front-load the shot: "Wide shot. ..." / "Tracking shot. ..." / "Locked tripod, low angle. ..." works as a real directive. Specify lens feel: "35mm anamorphic", "shallow DOF", "crushed shadows".

One visual beat per clip when iterating. Don't pile up "she walks AND the dog runs AND a car passes". Pick the beat, get it sharp, then layer with multi-shot prompts.

Multi-shot consistency — when describing two beats, restate the anchor at each: "Shot 1: tall woman in red wool coat, blue scarf, in a rainy alley. Shot 2: same woman in red coat / blue scarf, now ducking under an awning." HappyHorse holds the look but needs the anchor.

Audio direction — say what you want to hear: "distant temple bells, footsteps on wet pavement, no dialogue" or "warm friendly tone, English".

Anti-patterns:

  • Static-frame descriptions (no temporal verbs) → motion will be vague.
  • Conflicting style directions → cancels.
  • 2500 char prompts → degrades.

  • Aspect ratios outside the 5 supported → 422.

Where it shines

Use caseWhy HappyHorse 1.0
Multi-shot brand stories with one consistent characterNative cross-shot identity preservation
Talking-head explainers needing in-clip voiceover + ambientSynchronized audio in the same pass
Multilingual short-form ads6 prompt languages, no script-quality drop
Cinematic 1080p deliveryNative 1080p output, broadcast-ready
Blind-vote leader for general video quality#1 on Artificial Analysis Video Arena

Sample prompts (verified to produce strong results)

From the model page (cinematic scope):

Wide shot. A lone astronaut in dusty orange suit with blue-gray harnessskis across lunar plain, leaving parallel tracks in gray regolith.Mid-stride, poles planted, pushing in 1/6th gravity with subtle upwarddrift. Fine dust haze along ski tracks. Crescent Earth above lunarhorizon, blue-white glow against black sky. Raw sunlight, crushedshadows, no fill. 8K photorealistic.

Multi-shot consistency:

Shot 1: Medium close-up. A woman in a navy trench coat enters arain-slick neon-lit Tokyo alley, looks left, holds up an umbrella.Shot 2: Same woman in same navy trench, now under the awning of aramen shop, shaking water off the umbrella. Warm interior glow, softchatter, gentle rain on metal roof in the audio.

Vertical platform-native:

9:16 vertical short. A barista in a black apron pulls a singleespresso shot, steam rising into the morning sun, rich crema slowlyforming. Close-up handheld, shallow DOF, warm cafe ambience and thehiss of the steam wand.

Limitations

  • Duration cap 15s — for longer narratives, segment into multi-shot prompts and stitch.
  • Aspect ratios — only the 5 documented values; ultra-wide cinematic gets cropped or rejected.
  • Audio is in-pass only — you can't pass external audio to drive lip-sync. For audio-driven lip-sync, use Wan 2.7 (which accepts an audio_url) or Seedance 2.0 Pro.
  • No free image-to-video on this template — i2v is supported by HappyHorse via a separate pipeline; the t2v endpoint here is text-only.

Exit codes

The runcomfy CLI uses sysexits-style codes:

codemeaning
0success
64bad CLI args
65bad input JSON / schema mismatch (e.g. duration: 30 would 422)
69upstream 5xx
75retryable: timeout / 429
77not signed in or token rejected

Full reference: docs.runcomfy.com/cli/troubleshooting.

How it works

  1. The skill invokes runcomfy run happyhorse/happyhorse-1-0/text-to-video with a JSON body matching the schema.
  2. The CLI POSTs to https://model-api.runcomfy.net/v1/models/happyhorse/happyhorse-1-0/text-to-video with the user's bearer token.
  3. The Model API returns a request_id; the CLI polls GET .../requests/<id>/status every 2 seconds.
  4. On terminal status, the CLI fetches GET .../requests/<id>/result and downloads any URL whose host ends with .runcomfy.net or .runcomfy.com into --output-dir. Other URLs are listed but not fetched.
  5. Ctrl-C while polling sends POST .../requests/<id>/cancel so you don't get billed for GPU you stopped.

What this skill is not

Not a self-hosted video runner. Not a capability grant — depends on a working RunComfy account.

Security & Privacy

  • Token storage: runcomfy login writes the API token to ~/.config/runcomfy/token.json with mode 0600 (owner-only read/write). Set RUNCOMFY_TOKEN env var to bypass the file entirely in CI / containers.
  • Input boundary: the user prompt is passed as a JSON string to the CLI via --input. The CLI does NOT shell-expand the prompt; it transmits the JSON body directly to the Model API over HTTPS. No shell injection surface from prompt content.
  • Third-party content: image / mask / video URLs you pass are fetched by the RunComfy model server, not by the CLI on your machine. Treat external URLs as untrusted; image-based prompt injection is a known risk for any image-edit / video-edit model.
  • Outbound endpoints: only model-api.runcomfy.net (request submission) and *.runcomfy.net / *.runcomfy.com (download whitelist for generated outputs). No telemetry, no callbacks.
  • Generated-file size cap: the CLI aborts any single download > 2 GiB to prevent disk-fill from a malicious or runaway model output.

來源與署名

來源:prime-skills/runcomfy-agent-skills位於happyhorse-1-0提交fca19ae

授權條款: MIT

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架

更多來自 prime-skills/runcomfy-agent-skills 的技能

Wan 2.7 — Pro Pack on RunComfy

prime-skills

Generate text-to-video with Wan 2.7 (Wan-AI's flagship motion model) on RunComfy. Documents Wan 2.7's strengths (multi-reference conditioning, audio-driven lip-sync via `audio_url`, smoother transitions, prompt expansion), the duration / resolution / aspect-ratio schema, and when to route to HappyHorse 1.0 / Seedance 2.0 / Kling / LTX 2 instead. Calls `runcomfy run wan-ai/wan-2-7/text-to-video` through the local RunComfy CLI. Triggers on "wan", "wan 2.7", "wan-2-7", "wan video", or any explicit ask to generate video with this model.

待分類2026年10月8日

Video Edit — Pro Pack on RunComfy

prime-skills

把影片編輯需求路由到三個 RunComfy 影片模型之一,並透過 RunComfy CLI 呼叫。

Design & Creative2026年10月8日

Seedance 2.0 Pro — Pro Pack on RunComfy

prime-skills

指導透過 RunComfy CLI 使用位元組跳動 Seedance 2.0 Pro 生成電影感短影音。

Design & Creative2026年10月8日

Nano Banana Edit — Pro Pack on RunComfy

prime-skills

指導透過 RunComfy CLI 使用 Google Nano Banana 2 進行圖生圖編輯。

Design & Creative2026年10月8日

Nano Banana 2 — Pro Pack on RunComfy

prime-skills

Generate images with Google Nano Banana 2 (Gemini-family flash-tier text-to-image) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents Nano Banana 2's strengths (rapid iteration, in-image typography rendering, predictable framing, optional web-grounded context), the resolution-tier pricing, the safety-tolerance dial, and when to route to Nano Banana Pro / GPT Image 2 / Flux 2 / Seedream instead. Calls `runcomfy run google/nano-banana-2/text-to-image` through the local RunComfy CLI. Triggers on "nano banana", "nano-banana-2", "nano banana 2", "google image gen", "gemini image", or any explicit ask to generate with this model.

待分類2026年10月8日

Image-to-Video — Pro Pack on RunComfy

prime-skills

將圖生影片請求路由到合適的 RunComfy 模型,並透過 RunComfy CLI 呼叫。

Design & Creative2026年10月8日