Tool-native
Trained on hand-written tool traces, not scraped chat. It learns the shape of a call, so new tools appended at inference route correctly without a retrain.
Supply is the tools HYRE settles. Demand is the calls agents send. Price is set per call and quoted before execution, so an agent can decide whether the answer is worth the money before it spends any.
Model distribution
3 public builds · Apache-2.0
Counted103,643resolve requests
Hugging Face API · Homura-30B since 2026-08-13
Hugging Face's figure, which counts range-probes as well as real pulls. Read it with the arithmetic below, not on its own.
18 HYRE endpoints and 17 partner services, across 13 families.
The floor on the tools table. Priced per call, quoted before anything executes.
Solana, Base, BSC and Robinhood Chain. One balance behind them, and the chain only picks the stablecoin.
35 tools across 13 families, read from the production catalogue.
Homura is a thin LoRA on somebody else's frontier model. We say which one, and we say which part of the work was not ours.
Every build is Apache-2.0 and free to download. Counts read from the Hugging Face API as this page renders.
Every page that quotes a Hugging Face download count is quoting resolve requests. We would rather do the division in public than let someone else do it for us.
Across 3 builds, read from the API as this page rendered.
Served, if every count were a full pull of a 16.9 GB file. It was not.
A genuinely popular GGUF repo runs 50–200:1. This is an order of magnitude out.
LM Studio, Ollama, Jan and mirror bots each issue header range-probes before pulling weights, and Hugging Face counts every one of them the same as a download. So the honest reading is that real human users are in the low thousands, not the tens of thousands — which is still roughly a hundred times the reach our gateway has ever had, and the only demand-side asset we own.
GGUF and 4-bit MLX, so it loads in whatever you already have open. Nothing leaves the machine unless it is a paid call you chose.
One balance, any endpoint. The chain only decides which stablecoin the call is priced in.
Homura is the model HYRE built to drive its own rails. It is not tuned to chat with people — it is tuned to turn intent into the right paid call, then get out of the way. Open weights, Apache-2.0, running on hardware you already own.
Trained on hand-written tool traces, not scraped chat. It learns the shape of a call, so new tools appended at inference route correctly without a retrain.
A refusal mid-chain is a system failure — there is nobody at 3am to click retry. On the internal bench Homura refused 0 of 7 legitimate technical prompts. The stock base model refused 2.
Ships as GGUF and 4-bit MLX. LM Studio, Ollama, or bare llama.cpp on a Mac. No key, no vendor, no request leaving your machine unless it is a paid call you chose.
Download the weights, load them in LM Studio or Ollama, and the agent loop is yours. 560 hand-written training examples, 527 tool calls, zero malformed JSON — and the whole tune cost $12.10.
16.9 GB GGUF for llama.cpp and Ollama, 21.3 GB 4-bit MLX for Apple Silicon. Both quantised to fit a 32-64 GB machine.
Commercial use allowed, modification allowed, attribution required. No gated repo, no waitlist, no acceptance form.
The tune touches the language tower only — all 832 adapted tensors. The vision tower was never modified and still scores 6/6 on the internal check.
11 tools are in the trained surface — resolve_token, get_token_price, swap_quote, bridge_quote and the rest. Anything beyond them is appended to the system prompt at inference and the model still routes it, because it learned the pattern rather than the names.
Trench data, token prices, whale flow, wallet PnL, Meteora DLMM pools, Nansen smart money and cross-chain yield — 13 families, each call priced on its own.
Install once in Claude Desktop or Cursor and the same tools show up in your editor, settled the same way.
bazaar_search finds an x402 endpoint the model has never seen, quotes it, then pays it. Quote first, execute second — always.
There is no API key to issue, rotate or leak. The agent presents payment, the endpoint answers. That is the whole authorisation model, and it is the reason an agent can use HYRE with no human in the loop.
Solana through x402/MPP, Base through PayAI, BSC through b402, and Robinhood Chain through r402 — the first x402 rail on that chain, where calls settle in USDG with no gas.
USDC, quoted before execution and capped server-side. No subscription, no minimum, no seat.
No private key, no send authority. Spend limits live in the payment layer, where a limit can actually be enforced.
Every settled call leaves an on-chain receipt, every agent can carry a portable identity, and a share of revenue buys $HYRE back and burns it. All three are verifiable without asking us.
Each paid call writes a receipt on its settlement chain. Read it yourself — no dashboard required.
Agents carry a .ME passport and ERC-8004 registry entry, so the thing calling your endpoint can be identified across runs.
Ecosystem revenue buys $HYRE from the open market and destroys it. Every burn is verifiable on-chain.
Homura runs on hardware we own, so a call it can handle has no marginal cost. That turns frontier compute into something the agent buys on purpose — from the UsePod marketplace, over x402, one call at a time. Move the controls and the router below answers exactly as the backend would — same policy, same reasons.
Nothing forced a purchase, so the model we host answers and the call costs nothing.
A real 402 from UsePod for a model at — max tokens. Quoting is free, so nothing was paid to show you this.
These are the fields, not a sample transaction. Signatures appear here once calls settle on mainnet — we are not going to print one that never happened.
18 HYRE endpoints and 17 partner services, read from the production catalogue when this page rendered. The price column is what settles.
| Tool | Family | Returns | Served by | Rail | Per call |
|---|---|---|---|---|---|
| Token Intelligence | Detect sniper wallets that bought at launch | HYRE | $0.004 | ||
| Token Intelligence | Holders + snipers + bonding curve + risk assessment | HYRE | $0.015 | ||
| Token Intelligence | Top holders with % of supply | HYRE | $0.003 | ||
| Token Intelligence | First-60s birth-window rug-risk score (bundle/dev-dump/concentration) + clean|suspect|rigged band | HYRE | $0.010 | ||
| Yield Finder | Top yield pools sorted by APY | HYRE | $0.002 | ||
| Yield Finder | Cross-chain yield comparison + break-even math | HYRE | $0.008 | ||
| Wallet Analyzer | Realized + unrealized PnL, portfolio value | HYRE | $0.005 | ||
| Wallet Analyzer | 30-day trading patterns + success rate + archetype | HYRE | $0.012 | ||
| LP Analytics | Top DLMM pools with APR/TVL/volume | HYRE | $0.001 | ||
| LP Analytics | Best pool + entry strategy (spot/curve/bidask) | HYRE | $0.008 | ||
| Trench Watch | Latest PumpFun token launches with snipe signal | HYRE | $0.008 | ||
| Trench Watch | Tokens near bonding curve graduation (>70% progress) | HYRE | $0.003 | ||
| DeFi Overview | Total Value Locked across DeFi chains | HYRE | $0.001 | ||
| Ask Anything | AI-orchestrated multi-source DeFi answer | HYRE | $0.025 | ||
| Nansen Smart Money | Net-flow breakdown: Smart Traders vs Whales vs Public Figures for one token | HYRE | $0.020 | ||
| Nansen Smart Money | Find tokens that smart money is actively accumulating | HYRE | $0.025 | ||
| Nansen Smart Money | Top wallets ranked by realized PnL on one specific token | HYRE | $0.030 | ||
| Nansen Smart Money | Full wallet profile: archetype, win rate, top tokens, recent smart trades | HYRE | $0.050 | ||
| External Services | Perpetual contract screening data — direct from Nansen, no enrichment | api.nansen.ai | $0.025 | ||
| External Services | Top perpetual traders by PnL — direct from Nansen, no enrichment | api.nansen.ai | $0.075 | ||
| External Services | Generate a short AI video from a text prompt (Grok Imagine Video). Returns video URL. | api.xona-agent.com | $0.600 | ||
| Travel | Amadeus GDS flight offers — real bookable fares by route and date | stabletravel.dev | $0.110 | ||
| Travel | Google Flights offers + price insights for a route and date | stabletravel.dev | $0.040 | ||
| Travel | Hotels and villas by location, dates, guest count, and budget | orbonomy-xyz.vercel.app | $0.100 | ||
| Web Search | Neural web search — semantic results for research-style queries | stableenrich.dev | $0.020 | ||
| Web Search | Direct AI-generated answer with source citations | stableenrich.dev | $0.020 | ||
| Web Search | Classic Google Search results via Serper | stableenrich.dev | $0.040 | ||
| Web Search | Scrape a single URL with JS rendering — returns page content | stableenrich.dev | $0.025 | ||
| Web Search | Google News results via Serper — latest headlines for a query | stableenrich.dev | $0.080 | ||
| Maps & Places | Free-text Google Maps search — places matching a query string | stableenrich.dev | $0.040 | ||
| Maps & Places | Places around a lat/lng point, filtered by type and radius | stableenrich.dev | $0.040 | ||
| Maps & Places | Full details for one place — hours, rating, contact (needs placeId) | stableenrich.dev | $0.040 | ||
| People & Company | B2B prospect search by title, company, and location filters | stableenrich.dev | $0.040 | ||
| People & Company | Enrich one person by email, name, or company domain | stableenrich.dev | $0.100 | ||
| People & Company | Person enrichment by PID, LinkedIn URL, email, or phone | stableenrich.dev | $0.100 |
A 30B open-weight model, fine-tuned by HYRE to call HYRE's tools. The base is Meta's Muse Glimmer 30B under Apache-2.0, run through a community abliteration, then given a thin LoRA trained on hand-written tool traces. We did not pretrain a model, and we say so on the model card.
Because for an agent a refusal is a bug, not a safety feature. An automated chain that stops halfway because the model declined to name a number has failed, and there is no human there to retry. The real limit is that the model holds no key and cannot move funds — the brake is in the payment layer, not the personality.
The LoRA and the dataset. 560 conversations written by hand — roughly 70% tool-call traces and 30% house voice — producing 527 tool calls with zero malformed JSON. Training loss went 2.4 to 0.048 over three epochs, for about $12.10 of GPU time.
No. Every tool is plain paid HTTP, so any model or script that can present payment can call it. Homura is what we run ourselves, and what we recommend if you want the whole loop to stay on your machine.
The endpoint answers a bare request with a 402 and a price. Your agent settles it on whichever of the four chains it holds a balance on — USDC on Solana or Base, USD1/USDT/USDC on BSC, USDG gaslessly on Robinhood Chain — then retries and gets the data. There is nothing to sign up for, because the payment is the authentication.
Weights and download counts on Hugging Face, settlement receipts on the chain the call settled on, burns on-chain. The numbers on this page come from those sources and nowhere else.