
Mapleai Mcp
io.github.veenrekv0.3.1Updated Oct 5, 2026
Pay-per-request GPT chat, embeddings, agents and prepaid keys with local x402 USDC signing
Overview
Lets an assistant call paid GPT chat, image, audio, embeddings and structured-decision APIs, paying per request in USDC via local x402 signing.
- What it does
- A local stdio MCP server for the MapleAI gateway. It exposes a free model catalog, free embeddings and prepaid key status, plus paid chat, Jev structured decisions, agent execution and prepaid key packs. Payments are signed locally with an EVM or Solana wallet key over the x402 protocol, and bought prepaid keys can be spent in-server without a wallet.
- When to use it
- Use it when an assistant should reach GPT, image, audio or embedding models without an account or subscription, and per-request USDC payment is acceptable. It also suits agents that want a prepaid token budget instead of per-call settlement.
- Requirements
- Runs locally as an npm package (mapleai-mcp) over stdio, desktop only. Needs Node.js, network access to the MapleAI endpoints, and a funded wallet key: EVM_PRIVATE_KEY for Base, Polygon or Arc, or SVM_PRIVATE_KEY for Solana. Optional caps MCP_MAX_PAYMENT_USDC, MCP_MAX_PREPAID_USDC, MCP_NETWORK and the MCP_PAY_TO_* recipient variables configure spending and network.
Installation
In SourceWeft
- Open Mapleai Mcp in the dashboard and add it to a workspace.
- Enable the server for the chats that should use its tools.
Desktop only via STDIO. STDIO servers start a local process, so they need the SourceWeft desktop host.
Other MCP clients
Follow the launch instructions in the repository.
README
MapleAI — GPT and Image API with x402 Pay-Per-Request
OpenAI-compatible GPT, image, embeddings and structured-decision APIs behind one endpoint. No accounts, no subscriptions, no API keys for the pay-per-request tier — an agent pays per call in USDC over the x402 protocol. Prepaid API keys are available for clients that prefer a token budget over per-request payments.
Live endpoints
Prepaid key endpoint (shared): https://mapleai.shop/v1
Models and pricing
GPT (per 1M tokens)
A request is quoted as counted input tokens × input rate + requested output tokens × output rate, plus the network settlement fee (about $0.001; the exact amount is always in the 402 challenge). Failed calls (HTTP ≥ 400) are cancelled, not settled.
Images (per image)
POST /api/v1/images/generations and POST /api/v1/images/image2image
(edit a PNG/JPEG/WebP data URI, max 10 MB). On Solana these routes settle
upfront (image generation outlives a Solana blockhash) and retry transient
upstream failures inside the same paid request.
Audio (per request)
TTS takes JSON {model, input, voice?, response_format?} and returns a WAV
stream; OpenAI voice presets map to Gemini voices, the orpheus models use
their own emotive voices (en: autumn/diana/hannah/austin/daniel/troy, ar:
fahad/sultan/noura/lulwa/aisha/abdullah; [laughs]-style tags and real speed
supported). STT takes multipart file + model or JSON base64, up to 25 MB;
whisper-large-v3* add verbose_json word timestamps and srt/vtt.
Responses carry x-audio-upstream and x-fallback-used headers — when the
primary vendor fails the request fails over to a secondary one inside the
same payment (for TTS the voice differs then).
Jev structured decisions
POST /jev — jev-latest, $0.06 per 1M input tokens, output free. SystemOne
protocol: send model, state and named questions (noul, choice or
score with instructions), read answers from the response.
Free embeddings
POST /v1/embeddings — free NVIDIA nvidia/nemotron-3-embed-1b (2048
dimensions). Send input as a string or an array of up to 128 strings;
model is optional and ignored.
Quickstart
A runnable end-to-end demo (free embeddings → paid GPT answer) lives in
examples/embeddings-to-chat/ and online at
https://base.mapleai.shop/examples/embeddings-to-chat/.
Building an agent? Fork the agent-starter/ template —
repo instructions for agent clients (CLAUDE.md/AGENTS.md), MCP wiring,
a full paid scenario script (free embeddings → prepaid tap → prepaid chat),
and a free CI smoke for all four gateways.
Prepaid API keys
Buy a prepaid bearer key for one GPT model with a single x402 payment:
Budgets run in 100 000-token steps from 100 000 to 1 000 000, priced at the model input rate plus the settlement fee:
The key works at https://mapleai.shop/v1 (OpenAI-compatible). Check usage
and status for free:
Discovery
GET /v1/models— free model catalog with live pricingGET /.well-known/x402— x402 resource manifestGET /openapi.json— OpenAPI 3.1 specificationGET /llms.txt,GET /AI-AGENTS.md— agent-oriented docsGET /developers— developer guide with examples
Repository layout
src/— the x402 gateway (Express, TypeScript, runs undertsx)admin/— OmniRoute-based admin app: provider connections, combos, prepaid key issuance and wallet-only (SIWE) dashboardexamples/— runnable client examplespackages/mapleai-mcp/— local stdio MCP server (free model catalog, embeddings and prepaid status; paid chat, Jev, agent execution and prepaid key packs with local x402 signing; bought keys spend in-server via prepaid_chat with no wallet)tools/— payment and upstream test scriptsdeploy/— systemd units and Apache vhosts used on the VDS
Development
Required env: UPSTREAM_API_KEY, PAY_TO, MODEL_PRICES, MODEL_MAPPING.
Optional features are enabled by their own keys (IMAGE_UPSTREAM_API_KEY,
NVIDIA_API_KEY, JEV_UPSTREAM_API_KEY, PREPAID_ISSUER_TOKEN, …) — see
.env.example for the full list.
Source: README.md at commit 520ec70
Tools
0Version history
1- v0.3.1LatestOct 5, 2026


