A URL for every agent
Each agent answers at https://<name>.run.agntspark.com, with its certificate issued and renewed automatically.
AgntSpark runs your agent in managed containers and gives it its own HTTPS endpoint. Callers need one of the agent's access keys, rate limits are on from the start, and your model provider key stays encrypted.
Free plan available · Pro $29/month · Bring your own OpenAI, Anthropic or Google key
# 1. Deploy a template curl -X POST https://agntapi.agntspark.com/v1/agents \ -H "Authorization: Bearer $AGNTSPARK_API_KEY" \ -d '{"name": "support", "model": "gpt-4o", "api_key": "sk-…", "deploy": {"image": "agntspark/template-customer-support"}}' # → {"url": "https://support-k3v9qa.run.agntspark.com", # "access": "private", "access_key": "agk_…"} # 2. Call it with its key curl -X POST https://support-k3v9qa.run.agntspark.com/invoke \ -H "Authorization: Bearer agk_…" \ -d '{"input": "I was charged twice this month."}' # → {"output": "Sorry about that — I've opened ticket # T-1042 with billing…", "session_id": "5f0c…"}
from agntspark import Client, DeployConfig with Client() as client: agent = client.agents.create( "support", model="gpt-4o", api_key=OPENAI_API_KEY, deploy=DeployConfig( image="agntspark/template-customer-support", ), ) # Private by default: the first key is returned once. key = agent.access_key reply = client.agents.invoke( agent, "I was charged twice this month.", access_key=key, ) print(reply.output)
$ agntspark deploy agent.yaml --wait ✓ Agent created: agt_7Hq2… URL https://support-k3v9qa.run.agntspark.com Access private Access key (shown once): agk_… $ agntspark invoke agt_7Hq2… "Where is my refund?" --key agk_… Refunds go back to the original card within 5–10… $ agntspark access agt_7Hq2… private --rpm 30 agt_7Hq2…: private, 30 requests/min per caller $ agntspark keys create agt_7Hq2… --label website
You don't run a cluster, write ingress rules or renew certificates. You choose what the agent is; the platform keeps it reachable and fenced in.
An official template, a blank agent with your own system prompt, or any public container image that serves /health and /invoke.
From the console, the API, the Python SDK or the CLI. Choose the model, add your provider key, and set replicas, CPU and memory.
The agent answers at its own URL. Give each app its own access key, revoke keys independently, and cap how fast any one caller can go.
A backend, a chatbot widget, a CI job — anything that can send HTTPS.
Every request to *.run.agntspark.com passes through here first.
Its own containers with CPU, memory and process limits, cut off from other agents and the platform's database.
Called with the key you supplied. Usage is billed by them, not marked up by us.
Everything below is in the platform today.
Each agent answers at https://<name>.run.agntspark.com, with its certificate issued and renewed automatically.
New agents are private. Create a key per app, see previews of each, and revoke one without touching the rest.
A per-caller limit you choose, checked before the key, plus a per-agent ceiling from your plan — so a leaked URL can't run up your bill.
OpenAI, Anthropic or Google. Stored encrypted, masked everywhere, and only decrypted when your agent's container starts.
Run several replicas, scale by hand or let the platform adjust between your minimum and maximum — never past your plan.
Container logs, live CPU and memory, and metered replica-, vCPU- and memory-hours plus request counts, by the hour.
Each template is a prompt, a default model and a set of tools on top of the platform runtime. Any model works — pick one and bring that provider's key.
Answers from your knowledge base with citations, notices when a customer is upset, and opens tickets for what it can't resolve — optionally posting them to your helpdesk or Slack webhook.
Reads a GitHub pull request's diff, runs quick security and bug checks, and writes a review ordered by severity with file and line references — or reviews code you paste.
Any framework, any language. Serve GET /health and POST /invoke on your port — or build on the open-source runtime and just write your tools in Python.
What we do so one customer's agent can't reach another's — and what we'd rather tell you than have you discover.
Limits are checked before anything starts, so you're never billed for going over — a request past a limit is simply refused.
For trying the platform and small side projects.
For agents in front of real users.
When you need more than Pro.
Prices in US dollars, billed monthly through Stripe, plus applicable tax. Cancel any time; you keep Pro until the end of the period. Every plan includes the API, SDK, CLI, logs, metrics and usage reports.
Every component of the platform is public on GitHub under the MIT license.
AgntSpark was founded in 2026 by Xinning Wu and is registered in Sheridan, Wyoming. We started it because putting an agent in front of users still means running servers, TLS, routing, auth and scaling that have nothing to do with the agent itself. We run that part so developers and small teams can ship the agent.
The platform is live in an invite-only alpha: agents deploy and serve traffic today, and we're onboarding a small number of teams at a time.
The alpha is invite-only. Email admin@agntspark.com with what you want to build, and we'll send an invite code. Sign up at agntapi.agntspark.com with that code.
No. You bring your own OpenAI, Anthropic or Google key, and your provider bills you directly. AgntSpark charges only for hosting, per the plans above.
An official template, a blank agent with your own prompt on the platform runtime, or any public container image that listens on its port and serves GET /health and POST /invoke. Building from source isn't supported yet — push an image first.
Not by default. New agents are private: requests without one of the agent's access keys are refused at the edge before they reach your container. You can make an agent public, and rate limits apply either way.
On Google Cloud. The platform currently runs on a single host; there's no multi-region deployment or uptime SLA during the alpha. Database backups run nightly.
You keep Pro until the end of the billing period, then return to the free plan's limits. Running agents keep running; you just can't add resources past the free limits.
Tell us what you're building and we'll send you an invite.