Reach AI Demand
Serve inference for people, applications, and autonomous agents through one network.
Katara Cortex
Access AI through chat, our API, or autonomous agents. Powered by a decentralized network of compute providers.
Paid in USDC. Settled Onchain.
Earn with Katara Cortex
Put your Apple silicon Mac or supported GPU to work. Run approved AI models with Katara Cortex and get paid instantly in USDC, settled on Avalanche 🔺.
You bring the compute and set your prices. Katara Cortex connects you with demand from people, applications, and agents, and handles routing, verification, and payment.
Serve inference for people, applications, and autonomous agents through one network.
Choose what to charge for input and output tokens. Compete for requests on your terms.
Find approved models that fit your machine’s chip, memory, and supported runtime.
For Everyday AI
Find a recipe for what’s in your fridge, plan your weekend, or make sense of something new. Bring everyday questions to Katara Cortex and work through them in chat.
Choose a supported model in your browser and pay as you go.
Start ChattingFor Developers
Use the OpenAI client you already have. Point it at Katara Cortex and access supported models through the same request format.
Katara Cortex handles provider selection, routing, validation, metering, and settlement. You get the response.
katara/llama-3.1-8b-instruct@1Illustrative response from the API documentation.
For x402 Agents
An agent can delegate work to new agents, which can spawn agents of their own. Each can call Katara Cortex for inference and pay on the spot with x402.
Define a task, budget, and spending rules for each agent. It can buy inference as the workflow runs, paying in USDC for the work it needs to complete.
Recursive Agents.
Inference Paid on Demand.
POST /v1/chat/completions402 Payment RequiredPAYMENT-SIGNATURE200 OKInside the Network
Katara Cortex combines approved model bundles, provider validation, and clear usage records so you can understand what serves your requests and what you’re paying for.
Pinned model weights, tokenizers, and runtimes define what each provider must run.
Providers are tested with known prompts and checked against the bundle’s requirements.
Request records report token usage and the cost in USDC, so you can see what you paid for.
From the Blog
Explore ideas for building applications and agents, providing compute, and getting more from everyday AI.
Explore the Blog
Application Developers
OpenAI Compatiblex402 NativeUSDC · Settled Onchain