Fal Models Catalog

fal-ai-community/skills/skills/fal-models-catalog

作者 fal-ai-community9ca850412943251fc9a466c4c29fdaf7a303a3d8無授權條款收錄於 2026年10月9日更新於 2026年10月9日

Choose the right fal.ai endpoint for a given task. Modality-organized catalog of production endpoint defaults, text-to-image, image-to-image, text-to-video, image-to-video, and more. Use when the user has not named a specific model, or asks "which model for X", "best endpoint for Y", "what should I use for Z".

僅含說明AI & Agents
AI 產生的概覽

此技能是一份目錄,協助為媒體生成或編輯任務挑選合適的 fal.ai 端點。

功能
此技能是依模態整理的 fal.ai 正式環境端點目錄,涵蓋文字轉圖像、圖像轉圖像、文字轉影片、圖像轉影片、影片轉影片、文字轉 3D、圖像轉 3D、文字轉音訊、音訊轉文字與圖像轉文字。它會把任務導向對應的參考檔案,檔案中依使用情境(例如頂級寫實、快速低成本、4K 或特殊用途)列出精選端點。它也說明使用 genmedia CLI 確認、檢視與查詢端點價格的驗證流程,並指向另一個技能以取得工作流程公用端點。
適用情境
當使用者未指定特定模型,詢問哪個模型或端點適合某項任務,或詢問某項媒體生成或編輯工作該用什麼時使用。它的用途是在呼叫 fal.ai 端點之前進行選擇。
執行需求
需要 genmedia CLI 來呼叫端點,若尚未安裝須先執行一次 genmedia init;此技能本身不含指令碼,只有參考文件。呼叫端點意味著需要連線至 fal.ai 服務的網路存取。

fal.ai Models Catalog

Endpoint-first navigation for fal.ai production work. Each modality reference lists curated picks organized by use case (premium realism / fast & cheap / 4K / specialized). Before reaching for free-text search, consult the modality reference that matches the task.

Runtime: All endpoint calls run via the genmedia CLI. See the genmedia skill for command syntax; run genmedia init once if not yet installed.

Endpoint-first rule

  1. Pick the endpoint ID from the right modality reference.
  2. Verify it: genmedia models --endpoint_id <endpoint_id> --json.
  3. Inspect it: genmedia schema <endpoint_id> --json.
  4. Check cost when relevant: genmedia pricing <endpoint_id> --json.
  5. Use text search only if the routed endpoint is missing, deprecated, rejected, or the role is not covered here:
bash
genmedia models "<task description>" --jsongenmedia docs "<topic>" --json

Do not invent endpoint IDs.

Modality references

Load the reference matching the user's task:

  • text-to-image.md [blocked], image generation from prompt (text-heavy, premium still, fast draft)
  • image-to-image.md [blocked], image editing, inpainting, background removal, upscaling
  • text-to-video.md [blocked], video generation from prompt (highest quality, fast/economical, multi-shot storytelling)
  • image-to-video.md [blocked], video from a reference frame (including audio-driven and lip-sync variants)
  • video-to-video.md [blocked], video edit, restyle, upscale, background removal
  • text-to-3d.md [blocked], 3D model generation from text
  • image-to-3d.md [blocked], 3D model generation from reference images
  • text-to-audio.md [blocked]. TTS, music, SFX generation
  • audio-to-text.md [blocked], speech-to-text (Whisper, ElevenLabs Scribe with diarization)
  • image-to-text.md [blocked]. OCR, captioning, VQA, detection, segmentation

Utility endpoints

Workflow utility endpoint IDs (resize, composite, mask, audio merge, subtitle, etc.) live in the fal-workflow skill: fal-workflow/references/utility-endpoints.md.

Utility endpoints are explicit because they are deterministic tools, not creative model choices. Always inspect schema before use.

來源與署名

來源:fal-ai-community/skills位於skills/fal-models-catalog提交9ca8504

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架