Two ways in
Build with the API or host a Mac. Real hardware, real models, real settlement.
The OpenAI SDK is the Malibu SDK.
Point any OpenAI-compatible client at api.malibu.tech/v1. Streaming, tool calling, structured output, sticky conversations, and signed inference receipts — all live today.
client = OpenAI(
base_url="https://api.malibu.tech/v1",
api_key="mp_…"
)
Cheaper by design
Distributed Apple Silicon undercuts hyperscaler GPU pricing on open models. Pay per token, not per H100-hour.
Private by architecture
No central prompt collector. Requests route to individual providers, not a data-hoarding platform.
Signed receipts
Every response carries an Ed25519 receipt you can verify offline with malibu-verify.
Open model catalog
Llama, Qwen, GPT-OSS, and more — one endpoint, no per-provider onboarding.
Drop-in OpenAI SDK
Change base_url, keep your code. Streaming, tool calling, structured output all work.
Prepaid credits
Top up once, no subscription. GitHub OAuth for the API key; no wallet required to start.
Your Mac already has the hardware.
Download Malibu for macOS, click Launch Provider, and your Apple Silicon Mac joins the pool. Malibu picks the right open model for your memory tier, serves when you're idle, and pays you in USDC + $MALIBU.
Paid in USDC + $MALIBU
Every served prompt credits USDC and $MALIBU in real time. Payouts settle to your wallet weekly.
Runs on pennies
Apple Silicon draws 20–40W under inference load. Passive income, not a power bill.
Your Mac stays yours
Cooperative scheduling — serving interrupts for milliseconds. No datacenter, no rack.
Zero configuration
Malibu picks the right open model for your memory tier at install time. No tuning.
Early providers earn 5×
The first 90 days mint 5× more $MALIBU per prompt served. Multiplier tapers over 4 years.
Pause anytime
Toggle serving on when idle, off when you're working. No commitment, no schedule.
Frontier open models, MLX-native.
A curated catalog of the strongest open-weight models — quantized for Apple Silicon, matched to provider hardware at install time. Pricing tracks the OpenRouter frontier — typically 20–40% below hyperscaler rates, no subscription, no seat license.
A single Mac Studio Ultra already runs frontier open models — gpt-oss 120B — from unified memory, no datacenter. Pooling Macs to serve models beyond any single machine is next on the roadmap; a verifiable receipt is worth the most exactly at that scale. Not live yet.
The network is the token
Every dollar of buyer spend flows through a fixed split: providers get paid, supply burns, reserves grow. Network activity is token demand.
Hold $MALIBU, own the network. Every dollar of buyer spend flows through providers, burn, and reserve.
Three things happen the moment a prompt clears.
The Mac gets paid.
70¢ of every dollar goes to the person whose Mac ran the prompt, in USDC.
Supply shrinks.
18¢ buys $MALIBU on the open market and burns it. Every prompt leaves fewer tokens in circulation.
The floor rises.
12¢ goes into a reserve that backs every new $MALIBU. Macs keeping the network up earn the new tokens.