Easy Migration - Change Just One line code
Migrate seamlessly by simply updating your baseURL no code refactoring needed.

Always-On Decentralized Agents and AI
SaaS with USDC Payments on X402 Rails
Gemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeEasy Migration with Just a BaseURL Change and X402 Logic for Agents Unified in One SDK
Migrate seamlessly by simply updating your baseURL no code refactoring needed.

All X402 payment logic and agent capabilities consolidated in a single SDK.
Reduce manual tasks with AI‑driven workflows (like expense categorization or time tracking).
Monitor API call volume by type with detailed usage analytics and cost breakdown.
Migrate instantly by changing just the baseURL—zero refactoring needed.
Access complete agent capabilities and X402 payment rails through a single SDK.
Decentralized agents that pay as they go, managing their own LLM expenses seamlessly.
Autonomous payment handling: request, pay, and retry seamlessly.
Issue unlimited API keys to track and limit costs for every individual agent precisely.
X402 lets agents fund AGIHALO service usage and retry supported routes automatically.
Gemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeGemini 3 pro
Gemini 2.5 Flash Image Preview
Calls by TypeEnable real-time micro-payments based on your actual usage on the x402 rail for maximum cost efficiency.
Empower Agents to autonomously manage and refuel their own LLM credits through integrated sub-LLM systems.
Use one AGIHALO service layer for supported models, spend controls, metering, and agent memory.
Shift payment authority to Agents and DAOs using USDC, bypassing the need for traditional credit card dependencies.
Agent Friendly LLM ROUTER WITH HALO
integrations and growing
Sync your code, issues, and deployments directly into your workflow.
Connect support and CRM tools to turn conversations into action.
Capture conversations per end user, organize them into summaries and topics, and retrieve only the context needed for the next response.
Keep every customer memory isolated by project and user.
Return relevant summaries and linked conversation context.
Add memory to Node or Python agents without changing providers.
Product builder evaluating agent infrastructure. Prefers concise, code-first answers with predictable implementation costs.
Concise, code-first responses
Autonomous support agent
Weekly cost limit
One membership for long-term memory and distinct Auth users per UTC month.
For prototypes with managed memory and authentication.
For production apps growing memory and signed-in users.
For multi-agent products with durable context and Auth.
For large-scale identity and memory workloads.
Published Google standard rates for comparison only. AGIHALO service credits are not Google or Gemini API credits.
Official pricing| Model | Input rate | Output rate | Search after free quota | Source |
|---|---|---|---|---|
| Gemini 3.5 Flashgemini-3.5-flash | $1.50Standard | $9.00Standard | $14 / 1K queries | Google list |
| Gemini 3.1 Pro Previewgemini-3.1-pro-preview | $2.00 / $4.00<=200K / >200K | $12.00 / $18.00<=200K / >200K | $14 / 1K queries | Google list |
| Gemini 3.1 Flash-Litegemini-3.1-flash-lite | $0.25Standard | $1.50Standard | $14 / 1K queries | Google list |
| Gemini 3 Flash Previewgemini-3-flash-preview | $0.50Standard | $3.00Standard | $14 / 1K queries | Google list |
| Gemini 3.1 Flash Imagegemini-3.1-flash-image | $0.501K image output | $3 text / $0.067 image1K image output | $14 / 1K queries | Google list |
| Gemini 3 Pro Imagegemini-3-pro-image | $2.001K-2K image output | $12 text / $0.134 image1K-2K image output | $14 / 1K queries | Google list |
| Gemini 2.5 Progemini-2.5-pro | $1.25 / $2.50<=200K / >200K | $10.00 / $15.00<=200K / >200K | $35 / 1K grounded prompts | Google list |
| Gemini 2.5 Flashgemini-2.5-flash | $0.30Standard | $2.50Standard | $35 / 1K grounded prompts | Google list |
| Gemini 2.5 Flash Imagegemini-2.5-flash-image | $0.30Up to 1024px | $0.039 / imageUp to 1024px | Not available | Google list |
OpenAI Standard rates in USD per 1M tokens. Paired values show short and long context rates; long-context pricing applies above 272K input tokens.
Official pricing| Model | Input | Cached input | Cache write | Output | Rate tier |
|---|---|---|---|---|---|
| GPT-5.6 Solgpt-5.6 → gpt-5.6-sol | $5.00 / $10.00per 1M tokens | $0.50 / $1.00per 1M tokens | $6.25 / $12.50per 1M tokens | $30.00 / $45.00per 1M tokens | Short / long context |
| GPT-5.6 Terragpt-5.6-terra | $2.00 / $4.00per 1M tokens | $0.20 / $0.40per 1M tokens | $2.50 / $5.00per 1M tokens | $12.00 / $18.00per 1M tokens | Short / long context |
| GPT-5.6 Lunagpt-5.6-luna | $0.20 / $0.40per 1M tokens | $0.02 / $0.04per 1M tokens | $0.25 / $0.50per 1M tokens | $1.20 / $1.80per 1M tokens | Short / long context |
| GPT-5.5gpt-5.5 | $5.00 / $10.00per 1M tokens | $0.50 / $1.00per 1M tokens | —per 1M tokens | $30.00 / $45.00per 1M tokens | Short / long context |
| GPT-5.5 Progpt-5.5-pro | $30.00 / $60.00per 1M tokens | —per 1M tokens | —per 1M tokens | $180.00 / $270.00per 1M tokens | Responses API only |
| GPT-5.4gpt-5.4 | $2.50 / $5.00per 1M tokens | $0.25 / $0.50per 1M tokens | —per 1M tokens | $15.00 / $22.50per 1M tokens | Short / long context |
| GPT-5.4 Minigpt-5.4-mini | $0.75per 1M tokens | $0.075per 1M tokens | —per 1M tokens | $4.50per 1M tokens | Standard |
| GPT-5.4 Nanogpt-5.4-nano | $0.20per 1M tokens | $0.02per 1M tokens | —per 1M tokens | $1.25per 1M tokens | Standard |
| GPT-5.4 Progpt-5.4-pro | $30.00 / $60.00per 1M tokens | —per 1M tokens | —per 1M tokens | $180.00 / $270.00per 1M tokens | Responses API only |
| GPT-5.2gpt-5.2 | $1.75per 1M tokens | $0.175per 1M tokens | —per 1M tokens | $14.00per 1M tokens | Standard |
| GPT-5.2 Progpt-5.2-pro | $21.00per 1M tokens | —per 1M tokens | —per 1M tokens | $168.00per 1M tokens | Responses API only |
| GPT-5.1gpt-5.1 | $1.25per 1M tokens | $0.125per 1M tokens | —per 1M tokens | $10.00per 1M tokens | Standard |
Standard rates in USD per 1M text or image tokens. Actual per-image cost varies by size and quality.
Official pricing| Model | Text tokens | Image tokens | Tier |
|---|---|---|---|
| GPT Image 2gpt-image-2 | $5.00 input$1.25 cached | $8.00 input · $30.00 output$2.00 cached | Latest |
| GPT Image 1.5gpt-image-1.5 | $5.00 input · $10.00 output$1.25 cached | $8.00 input · $32.00 output$2.00 cached | Standard |
| GPT Image 1 Minigpt-image-1-mini | $2.00 input$0.20 cached | $2.50 input · $8.00 output$0.25 cached | Economy |
| GPT Image 1gpt-image-1 | $5.00 input$1.25 cached | $10.00 input · $40.00 output$2.50 cached | Previous |
| ChatGPT Image Latestchatgpt-image-latest | $5.00 input · $10.00 output$1.25 cached | $8.00 input · $32.00 output$2.00 cached | ChatGPT alias |
Publish an OpenAI-compatible inference endpoint backed by your own compute. Halo tests the connection, checks the advertised models, and lets you prepare model-specific marketplace pricing.
Start providingCustom-provider registration tests every advertised Halo model with a real minimal request and shows its response time. Buyer requests reserve their maximum cost before dispatch, then settle with Halo text units calculated from the canonical request and visible assistant text, without switching to another seller. Individual KYC is required, and the default split is a 10% Halo fee and 90% provider earnings before upstream costs.Hosted seller credentials may appear during rollout, but they are not currently connected to buyer routing or seller earnings. Do not register one expecting traffic or revenue.
Publish an OpenAI-compatible chat endpoint backed by your GPU or inference server, then enable the models you can serve.
Get answers to commonly asked questions.
AGIHALO keeps the underlying model capabilities while adding unified routing, API spend governance, long-term memory, and X402 payment automation for agents.
No. Account credits can be purchased by card or with USDC. X402 enables autonomous agent payments over supported routes.
No. AGIHALO credits pay only for AGIHALO routing, memory, metering, and related infrastructure. They cannot be transferred to or redeemed with Google or another model vendor.
Requests are processed to provide model routing, usage metering, and billing. Memory content is stored only when you enable long-term memory, as described in our Privacy Policy.
AGIHALO gives agents one API for model access, per-key cost controls, durable end-user memory, provider monetization, and automated X402 payment flows.
Routing, long-term memory, and autonomous USDC payments for always-on agents
Get started