Permanent API rent
You build your product on an endpoint whose price, limits, and rules can change the moment it's inconvenient for you.
A few corporations hoard the GPUs, meter every request, and keep your private data on their servers. Koinos AI puts the first layer back in your hands — private AI on hardware you already own — and turns the rest into a decentralized network no single company controls. The alpha is downloadable today, and the network is live on Koinos testnet (harbinger): chatting locally is free and private, and earning pays KAI on testnet. Two ways to take it back:
Windows 10/11 · free · alpha earnings are testnet KAI — no monetary value yet.
Chat privately on your own machine. Turn unused compute into KAI.
Start earning → 02 · buildersAn OpenAI-compatible endpoint you control that overflows to the network.
Tester quickstart →Intelligence is becoming the most valuable resource on Earth — and it's being locked inside a handful of hyperscale data centers owned by a few of the largest companies that have ever existed. That isn't a market. It's a monopoly on thought.
Koinos AI is the counter-move: millions of machines, owned by the people who run them, coordinated by an open network no single company controls — with the private machine on your desk as the default, not the exception.
The dominant model concentrates the GPUs, the pricing power, and your data inside a handful of data centers. Meanwhile the capable machine on your desk sits idle most of the day.
You build your product on an endpoint whose price, limits, and rules can change the moment it's inconvenient for you.
Prompts, files, source code, and business context leave hardware you control every time you hit a cloud model.
Millions of capable GPUs do nothing all night while all AI demand is funneled to hyperscale clusters.
Every request is checked against your privacy and spending policy before anything about quality, speed, or price. Local always wins first. The network is a fallback you opt into.
Every app built on a hyperscale API is one price hike, rate-limit, or deprecation away from breaking — and it hands your users' data to the same firm racing to ship your product itself. Koinos AI is the exit: self-host an OpenAI-compatible gateway on hardware you own, keep your GPUs as the primary inference layer, and overflow into a decentralized network by policy, not a rewrite. Same SDK. Scoped app keys, never wallet keys. Nobody holding your roadmap hostage.
# same SDK. your endpoint. your hardware first. client = OpenAI( base_url="http://localhost:8181/v1", api_key="kai_proj_live_••••", ) r = client.chat.completions.create( model="koinos-smart", messages=[{"role":"user","content":"…"}], # overflow policy, not a different API extra_body={"routing":"local-first", "max_spend_kai":2.5}, )
Streaming · project keys · per-request/day/month budgets · hard stops · model pinning · usage logs. Embeddings, tools & batch on the roadmap.
The same GPU that runs your games can run real AI work for the network. Install, let it benchmark your hardware, and press one button — Koinos AI handles models, safe resource limits, and pricing. You earn KAI for verified useful work, not idle uptime, and intelligence gets a little less centralized every time you do. During alpha, earnings are testnet KAI on Koinos testnet (harbinger) — they have no monetary value yet.
Illustrative model — a relative earning weight from GPU class × hours × demand, not a yield promise. Real KAI reward depends on verified work, live network demand, and the KAI reference price, all still in economic simulation. Alpha earnings are testnet KAI — no monetary value yet.
Koinos AI finds your CPU, GPU, VRAM, and drivers and configures a compatible runtime and model automatically.
Download for Windows — free →No KAI, KOIN, MANA, or manual pricing. Foreground use always wins; battery and thermal limits protect the hardware.
Completed, verified AI jobs settle in KAI through Koinos — with gasless onboarding sponsored underneath.
Same familiar API surface. A completely different relationship to your compute, your data, and your bill.
| Koinos AI | Centralized AI API | |
|---|---|---|
| Where inference runs | Your hardware first, network on demand | Always their data center |
| Your prompts & files | Stay local unless you opt out | Leave your control by default |
| Cost floor | $0 for local — you own the compute | Per-token rent, forever |
| Idle GPU | Earns KAI for verified work | Earns nothing |
| Elastic scale | Overflow into a decentralized network | Only theirs, at their price |
| Lock-in | Open, self-hostable, forkable | Their SDK, their terms |
To Big AI, your prompts and files leave your control the moment you hit send. Koinos AI flips the default: local execution never leaves your device, and when you do reach for the network you pick the exact privacy mode — and see precisely where your request is allowed to run.
KAI settles the network: providers earn it for verified compute, model creators can earn bounded royalties, and you spend it only when you use resources someone else supplied. Local inference has no KAI charge, ever.
Benchmark, route, protect foreground use, and settle verified AI work automatically. Providers can earn KAI without first buying KAI or KOIN.
An optional module runs a full node and, if you choose, Proof-of-Burn block production. A separate earning path with its own economics — no GPU mining required.
Clean by design: KAI rewards useful AI work. KOIN rewards Koinos consensus. We never double-pay KAI just for producing blocks. Token supply, fees, and bootstrap pricing stay provisional pending economic simulation and legal review.
Prove the core loop first — install, run local AI, earn, settle, overflow — then earn the harder layers. Gated by readiness, not a calendar.
Be first to run private AI on your own hardware — and first to earn from it when the network opens.
No. Local AI runs on your own hardware with no KAI required. Network services can be paid through familiar payment and credit experiences, with Koinos handling settlement underneath.
No. Entry-level providers start with compatible hardware and sponsored (gasless) Koinos resource usage. Higher-value provider classes may later use KAI collateral — but you never have to buy KAI to begin earning KAI.
Not automatically — and we won't pretend otherwise. Local mode never leaves your device; private pools restrict execution to trusted hardware; verified-network mode uses encrypted transport; confidential compute (future) adds hardware attestation. You choose the fallback order.
It keeps your own GPUs as the primary inference layer while giving you an elastic path to decentralized capacity — instead of building your product permanently around a hyperscale API whose terms you don't control.
No. You're not rewarded for burning power on hashes — you earn KAI for completing verified, useful AI work: inference, embeddings, evaluation, and later fine-tuning. Passive uptime isn't the reward.