gpusmarket

gpusmarket

MCP server for renting real GPUs from the terminal. Enables browsing, renting, chatting, managing, and pooling GPU instances with per-second billing, designed for AI agents.

Category
访问服务器

README

gpusmarket

Rent real GPUs from the terminal. Built for AI agents.

npm version MCP — 32 tools Agent-native smithery badge

One command to sign up, one to rent, one to run inference. Per-second billing, no subscriptions, $0.50 minimum per rental. The CLI your coding agent can drive end-to-end: GPUsMarket.

npm install for GPUs. One command in, one command out.

Renting a GPU usually means a signup form, a quota request, a support ticket, and a $50 minimum — before you've run a single token. GPUsMarket collapses that to three commands: everything from consumer RTX 4090s to datacenter H100s and MI300Xs, rentable by the second. And it's built so your agent can sign up and rent by itself — no human, no browser, no dashboard.

Works with: Claude Code · Cursor · Cline · Windsurf · Aider · Codex · any MCP client

<!-- HERO: money loop — install → signup → search → rent → chat (real inference) → stop → per-second bill --> gpusmarket demo

Install to real inference in under a minute — watch the full-res demo.

npm install -g gpusmarket
gpusmarket init

Quick start

# Create an account (returns your API key, auto-saved)
gpusmarket signup --email you@example.com --tenant my-agent

# Find a GPU
gpusmarket search --gpu "RTX 4090"

# Rent the cheapest match — returns an OpenAI/Ollama-compatible endpoint + key
gpusmarket rent --gpu "RTX 4090"

# Run inference straight through your rental
gpusmarket chat "What is 2+2?" --model gemma3:4b

# Stop when done — you pay for exactly the seconds you used
gpusmarket stop RENTAL-ID

Built for agents

Built so an agent can go from zero to running inference with no human in the loop — sign up, rent, infer, and stop, all from the terminal or over MCP. This CLI is designed to be driven by AI coding agents (Claude Code, Cursor, Cline, Windsurf, Codex, Copilot), not just humans.

  • gpusmarket init — in-terminal quickstart: detects your agent's config file, prints a ready-to-paste skill block, MCP setup, and next steps.
  • gpusmarket init --agent-schema — the full command contract as JSON: every command, every flag, plus the rules agents must follow. Load it into context and your agent never hallucinates a flag.
  • --json everywhere — every command has machine-readable output.
  • Zero dependencies — a single self-contained Node.js CLI. Nothing to audit but us.

Watch an agent drive it end-to-end (click for the full-res video):

agent drives gpusmarket via MCP

Claude Code searches, rents, runs inference, and stops the GPU — no human in the loop.

MCP server

gpusmarket mcp-serve exposes 32 tools over stdio — search, rent, status, stop, chat, host, pools, the whole surface. Add it to Claude Code:

claude mcp add gpusmarket -- gpusmarket mcp-serve

…or drop it into any MCP client's .mcp.json. The MCP server runs outside your project directory, so pass your API key via the GPUSMARKET_API_KEY environment variable:

{
  "mcpServers": {
    "gpusmarket": {
      "type": "stdio",
      "command": "gpusmarket",
      "args": ["mcp-serve"],
      "env": { "GPUSMARKET_API_KEY": "gpu_your_key_here" }
    }
  }
}

No key yet? Run gpusmarket signup to provision one, or grab it from your dashboard.

Core capabilities

Each demo links to the full-resolution MP4.

Browse the marketplace Manage your rentals Load-balance a pool
browse manage pools
search · cheapest · models rentals · status · stop pool create · add · list
# Browse
gpusmarket search --gpu "RTX 4090"      # filter by GPU, price, model
gpusmarket cheapest --model gemma3:4b   # the single cheapest node that serves a model
gpusmarket models                       # every model available across the marketplace

# Manage
gpusmarket rentals                      # your active rentals
gpusmarket status RENTAL-ID             # live status + running cost
gpusmarket stop RENTAL-ID               # stop, pay for the seconds you used

# Pool — one endpoint, several rentals behind it
gpusmarket pool create my-pool
gpusmarket pool add my-pool RENTAL-ID
gpusmarket pool list

Honest billing

  • Per-second billing. A 90-second experiment costs 90 seconds of GPU time.
  • A hold, not a charge. Renting places a pre-auth hold (default $10) — it's released in full when you stop. You are never charged the hold.
  • $0.50 minimum per rental. You pay per second for the compute you consume, subject to a $0.50 minimum charge per rental (disclosed when you rent). When you stop, your card is charged the exact usage cost, or the $0.50 minimum — whichever is greater.
  • Receipts for everything, with a full "where your money goes" breakdown.

Earn with your GPU

Your GPU earns nothing while it sits idle. List it in one command — the CLI auto-detects your hardware, prints your host key, and the listing goes live in minutes. You keep 90% of every rental, paid out to your Stripe account automatically. No port forwarding, no static IP, no open inbound ports.

Three ways to host (click any demo for the full-res video):

Mac (Apple Silicon) NVIDIA GPU Docker sidecar
host on mac host on nvidia gpu host via docker
gpusmarket host setup auto-detects CUDA + VRAM docker compose up -d
gpusmarket host setup          # auto-detects your GPU, creates the listing
npx gpusmarket-host --key KEY  # connect your machine — no port forwarding needed

Commands

Command Description
init Quickstart guide (--agent-schema for the JSON command contract)
signup / login / logout / whoami Account + API key management
search / models / cheapest Browse the GPU marketplace
rent / rentals / status / stop Rent, monitor, and stop GPUs
chat Inference through your rental (streams, or pipe from stdin)
pool create/list/add/remove Load-balance several rentals behind one key
host setup List your own GPU for rent (auto-detects hardware)
benchmark submit Run the gmbench/1 suite on your local Ollama + publish measured speed to your listing
mcp-serve Start the MCP server (stdio)

Run gpusmarket --help for the full reference, flags included.

Features

  • Search — filter the marketplace by GPU, price, or the exact model you need to run.
  • Rent — one command returns an OpenAI/Ollama-compatible endpoint + key, ready for inference.
  • Chat — stream inference straight through your rental, or pipe from stdin.
  • Pools — put several rentals behind one endpoint and load-balance across them.
  • Host — list your own GPU (Mac, NVIDIA, or Docker) and keep 90% of every rental.
  • Per-second billing — receipts for everything, $0.50 minimum per rental, no subscriptions.
  • MCP server — 32 tools over stdio for Claude Code, Cursor, Cline, Windsurf, any MCP client.
  • Agent-native — --json everywhere plus init --agent-schema so agents never guess a flag.

Why this exists

Renting a GPU should be as fast as npm install. Not a signup form, a quota request, a support ticket, and a $50 minimum. GPUsMarket is a marketplace where anyone can list an idle GPU and anyone — or any agent — can rent it by the second. This CLI is the agent-native front door to that marketplace. It's early and moving fast: if something's rough or missing, open an issue — we read every one.

Documentation

Full docs at gpusmarket.com. Run gpusmarket init for the interactive quickstart.

License

Proprietary — Copyright (c) 2026 Tyga.Cloud Ltd. All rights reserved. See LICENSE file.

推荐服务器

Baidu Map

Baidu Map

百度地图核心API现已全面兼容MCP协议,是国内首家兼容MCP协议的地图服务商。

官方
精选
JavaScript
Playwright MCP Server

Playwright MCP Server

一个模型上下文协议服务器,它使大型语言模型能够通过结构化的可访问性快照与网页进行交互,而无需视觉模型或屏幕截图。

官方
精选
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

一个由人工智能驱动的工具,可以从自然语言描述生成现代化的用户界面组件,并与流行的集成开发环境(IDE)集成,从而简化用户界面开发流程。

官方
精选
本地
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

通过模型上下文协议启用与 Audiense Insights 账户的交互,从而促进营销洞察和受众数据的提取和分析,包括人口统计信息、行为和影响者互动。

官方
精选
本地
TypeScript
VeyraX

VeyraX

一个单一的 MCP 工具,连接你所有喜爱的工具:Gmail、日历以及其他 40 多个工具。

官方
精选
本地
graphlit-mcp-server

graphlit-mcp-server

模型上下文协议 (MCP) 服务器实现了 MCP 客户端与 Graphlit 服务之间的集成。 除了网络爬取之外,还可以将任何内容(从 Slack 到 Gmail 再到播客订阅源)导入到 Graphlit 项目中,然后从 MCP 客户端检索相关内容。

官方
精选
TypeScript
Kagi MCP Server

Kagi MCP Server

一个 MCP 服务器,集成了 Kagi 搜索功能和 Claude AI,使 Claude 能够在回答需要最新信息的问题时执行实时网络搜索。

官方
精选
Python
e2b-mcp-server

e2b-mcp-server

使用 MCP 通过 e2b 运行代码。

官方
精选
Neon MCP Server

Neon MCP Server

用于与 Neon 管理 API 和数据库交互的 MCP 服务器

官方
精选
Exa MCP Server

Exa MCP Server

模型上下文协议(MCP)服务器允许像 Claude 这样的 AI 助手使用 Exa AI 搜索 API 进行网络搜索。这种设置允许 AI 模型以安全和受控的方式获取实时的网络信息。

官方
精选