# Halo

> Halo is a projects-and-protocols that operates a permissionless peer-to-peer marketplace for AI inference, routing requests to operators, verifying execution via Statistical Proof of Execution (SPEX), and settling payments in USDC on Base.

- Canonical URL: https://iq.wiki/wiki/halo
- Categories: Projects & Protocols
- Tags: Identity, AIPlatform, AIInfrastructure, Infrastructure, Privacy
- Created: 2026-09-21T16:26:53.815Z
- Last updated: 2026-09-21T16:36:42.958Z
- Source: IQ.wiki — the world's largest blockchain and crypto encyclopedia (https://iq.wiki)

---

**Halo** is a peer-to-peer marketplace for AI inference that connects users and autonomous [Agents](https://iq.wiki/wiki/ai-agents) with operators providing access to AI models. Its infrastructure includes onchain payments, model-serving interfaces, and Statistical Proof of Execution (SPEX) for verifying that inference was performed as claimed. [\[1\]](#cite-id-worg6sjrq5)&#x20;

## Overview

Halo is a permissionless peer-to-peer marketplace for AI inference on [Base](https://iq.wiki/wiki/base), connecting users and autonomous [Agents](https://iq.wiki/wiki/ai-agents) with operators that provide access to AI models. Consumers and [Agents](https://iq.wiki/wiki/ai-agents) can deposit [USDC](https://iq.wiki/wiki/usdc) into Halo vaults and use compatible models through an OpenAI-compatible endpoint, while operators can serve models through APIs, self-hosted open-weight models, or local hardware and receive [USDC](https://iq.wiki/wiki/usdc) payments for completed inference requests. The network is designed without a centralized access gatekeeper, allowing participants to access and provide models through a distributed marketplace, and supports confidential inference through [trusted execution environments (TEEs)](https://iq.wiki/wiki/trusted-execution-environments) on [NEAR](https://iq.wiki/wiki/near-protocol) for workloads requiring additional privacy. HALO serves as the network's coordination asset, with [staking](https://iq.wiki/wiki/staking), protocol fees, buybacks, [token burns](https://iq.wiki/wiki/token-burn), and usage-based issuance forming part of its economic and incentive mechanisms. [\[2\]](#cite-id-n2y7ub88y4)  [\[4\]](#cite-id-wl6z3f5gmy)&#x20;

### Launch

HALO launches through [Virtuals Protocol](https://iq.wiki/wiki/virtuals-protocol) using a launch format for established teams that requires committed liquidity at token generation. HALO is paired with VIRTUAL in a [liquidity pool](https://iq.wiki/wiki/liquidity-pool), with the liquidity position locked for ten years. At generation, a portion of the community allocation is released, including a genesis airdrop distributed over the first 7 to 15 days rather than being claimable at once, with eligibility based on verified protocol usage and additional incentives for liquidity provision. The remainder of the community allocation is distributed through recurring League seasons based on settled [USDC](https://iq.wiki/wiki/usdc) volume generated by participants serving or consuming inference, with consistent activity weighted more heavily than short-term bursts. The launch also activates the protocol's buyback mechanism from the first trade, with [trading fees](https://iq.wiki/wiki/trading-fee) directed toward the buyback allocation, while the complete token release schedule, including team-controlled addresses, is published at generation and trackable onchain. [\[4\]](#cite-id-wl6z3f5gmy)&#x20;

## Architecture

![](https://ipfs.everipedia.org/ipfs/QmcuVMU17BkLkXbHZf9c4EkViifZpmcXeaJCyNRwgrH444)

Halo's marketplace consists of consumers, operators, a relay, a facilitator, and an indexer. Consumers, including people and autonomous [Agents](https://iq.wiki/wiki/ai-agents), pay for AI inference in [USDC](https://iq.wiki/wiki/usdc), while operators provide models through APIs, locally hosted open-weight models, or their own hardware and receive payment for completed jobs. The relay routes requests between consumers and operators, the facilitator verifies payments and submits gas-sponsored transactions without taking custody of user funds, and the indexer records inference events and operator activity while providing reputation data and public network information. For payments, Halo primarily uses a batched, receipt-based settlement system in which consumers deposit [USDC](https://iq.wiki/wiki/usdc) into the HaloVault contract and jobs are covered by operator-specific reservations and cumulative offchain receipts, allowing operators to settle multiple inference requests in a single onchain transaction while the protocol fee is deducted at settlement. Halo also supports a Permit2-based budget system for users who prefer authorization without a prefunded deposit, while [x402](https://iq.wiki/wiki/x402) is used for one-time payments to external services such as metered APIs and tool calls. [\[4\]](#cite-id-wl6z3f5gmy)&#x20;

### SPEX

![](https://ipfs.everipedia.org/ipfs/QmQ8oAJdde8nyR34L7vZGGaGypABrrcnBizqNHCJQ3nvYt)

Verifiable AI inference provides evidence that an AI model was executed as requested, rather than relying solely on the provider's claims. Halo uses Statistical Proof of Execution (SPEX), which generates a statistical fingerprint of an inference result using Bloom filters and allows an independent verifier to compare its own model execution against that fingerprint. Honest executions are expected to produce substantially higher overlap than fabricated outputs, with the network applying an acceptance threshold to determine whether a result passes verification. Halo also uses swarm verification, which distributes verification tasks among multiple micro-verifiers so different parts of an inference can be checked independently. Verification results are recorded against an operator's pseudonymous [ERC-8004](https://iq.wiki/wiki/erc-8004) identity, with accurate work improving reputation, incorrect verification reducing it, and incorrect verification temporarily delaying settlement. This system is intended to allow [AI Agents](https://iq.wiki/wiki/ai-agents) and other users to obtain inference from distributed operators without relying entirely on centralized providers or trusted hardware, while making operator performance and verification history independently observable. [\[3\]](#cite-id-9vjjwyofba) [\[5\]](#cite-id-pipdihd41) &#x20;

## HALO

![](https://ipfs.everipedia.org/ipfs/QmQBguzveHDSN8vfFcQSWtN7xn7vNPCrDQxc4UohLu1KgQ)

HALO is the coordination token for the Halo network and is used primarily by participants who maintain and secure the protocol, rather than by consumers or operators paying for inference, who use [USDC](https://iq.wiki/wiki/usdc) instead. Verified operators, SPEX verifiers, and, as the network becomes more decentralized, federated relayers can [stake](https://iq.wiki/wiki/staking) HALO as part of their network roles. The protocol collects a 10% fee on settled inference volume, with 80% initially allocated to an onchain buyback mechanism and 20% directed to a [USDC](https://iq.wiki/wiki/usdc) treasury for operations. When the buyback threshold is reached, the accumulated [USDC](https://iq.wiki/wiki/usdc) can be used to purchase HALO on a [decentralized exchange](https://iq.wiki/wiki/decentralized-exchange), with the purchased tokens initially divided between staker distributions and [token burns](https://iq.wiki/wiki/token-burn). The program also issues new HALO based on verified network usage within a declining issuance budget, with the initial supply designed to support network growth during its first two years. [\[4\]](#cite-id-wl6z3f5gmy)&#x20;

### Tokenomics

![](https://ipfs.everipedia.org/ipfs/QmVFmDKUrA3tqthaB5cuDPKxp8SCaWe9j1R2qS5iSaL9in)

HALO has a total supply of 1B tokens and has the following allocation: [\[4\]](#cite-id-wl6z3f5gmy)&#x20;

* **Community & Ecosystem**: 35%
* **Core Contributors**: 25%
* **Treasury**: 25%
* **Liquidity & Market-Making**: 15%

## Partnerships

* [Venice](https://iq.wiki/wiki/venice-ai)
* [NEAR](https://iq.wiki/wiki/near-protocol)
* [0G Labs](https://iq.wiki/wiki/0g-labs)
* [Warden Protocol](https://iq.wiki/wiki/warden-protocol)
* [Base](https://iq.wiki/wiki/base)
