Run Models

replicate/skills/skills/run-models

作者 replicate2f36e415965a無授權條款98 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫4 個月前更新

Run AI models on Replicate via predictions, webhooks, and streaming.

僅含說明AI & Agents
AI 產生的概覽

指導透過預測、Webhook、輪詢與串流在 Replicate 上執行 AI 模型。

功能
此技能提供呼叫 Replicate API 執行 AI 模型的說明。內容涵蓋選擇模型、取得其結構描述、建立預測、透過輪詢或 Webhook 取得結果、串流輸出、處理檔案,以及將多個模型串接為多模型工作流程。它產出的是預測結果與輸出 URL,而非檔案或程式碼。
適用情境
當代理需要在 Replicate 上呼叫託管的 AI 模型(例如影像、視訊或語言模型)時使用。適合需要建立預測、非同步取得結果,或串接多個模型輸出的任務。
執行需求
需要連線至 Replicate API 的網路存取,以及用於身分驗證的 API 權杖。不包含指令碼,僅為說明文件。

Docs

Workflow

  1. Choose the right model - Search with the API or ask the user.
  2. Get model metadata - Fetch input and output schema via API.
  3. Create prediction - POST to /v1/predictions.
  4. Poll for results - GET prediction until status is "succeeded".
  5. Return output - Usually URLs to generated content.

Three ways to get output

  1. Create a prediction, store its id from the response, and poll until completion.
  2. Set a Prefer: wait header when creating a prediction for a blocking synchronous response. Only recommended for very fast models. Max 60 seconds.
  3. Set an HTTPS webhook URL when creating a prediction, and Replicate will POST to that URL when the prediction completes.

Guidelines

  • Use the POST /v1/predictions endpoint, as it supports both official and community models.
  • Every model has its own OpenAPI schema. Always fetch and check model schemas to make sure you're setting valid inputs. Even popular models change their schemas.
  • Validate input parameters against schema constraints (minimum, maximum, enum values). Don't generate values that violate them.
  • When unsure about a parameter value, use the model's default example or omit the optional parameter.
  • Don't set optional inputs unless you have a reason to. Stick to the required inputs and let the model's defaults do the work.
  • Use HTTPS URLs for file inputs whenever possible. You can also send base64-encoded files, but they should be avoided.
  • Fire off multiple predictions concurrently. Don't wait for one to finish before starting the next.
  • Output file URLs expire after 1 hour, so back them up if you need to keep them, using a service like Cloudflare R2.
  • Webhooks are a good mechanism for receiving and storing prediction output.

Predictions

  • A prediction goes through these states: starting -> processing -> succeeded / failed / canceled.
  • Official models use owner/name format. Community models require owner/name:version_id.
  • The POST /v1/predictions endpoint handles both.

Webhooks

  • Set webhook to an HTTPS URL when creating a prediction. Replicate POSTs the full prediction object when it completes.
  • Filter events with webhook_events_filter: start, output, logs, completed.
  • Validate webhook signatures using the Webhook-ID, Webhook-Timestamp, and Webhook-Signature headers. Get the signing secret from GET /v1/webhooks/default/secret.

Prediction lifetime

  • Set lifetime to auto-cancel predictions that run too long (e.g. 30s, 5m, 1h). Measured from creation time.

Streaming

  • Language models that support streaming include a stream URL in the response. Use SSE to receive incremental output.

File handling

  • Prefer HTTPS URLs for file inputs. Output URLs from one prediction can be passed directly as file inputs to the next model.
  • Output file URLs expire after 1 hour. Download and store them immediately if you need to keep them.

Multi-model workflows

  • Chain models by passing output URLs as file inputs to the next model.
  • Start all independent predictions in parallel, then collect results.
  • Output URLs are valid for 1 hour, which is enough for pipeline steps.

來源與署名

來源:replicate/skills位於skills/run-models提交2f36e41

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架