Skip to content
Harogo
Models Request volume Pricing Usage FAQ 中文版 Sign in Get Harogo
100 discounted subscription slots open now

More usage, lower cost

Harogo is a multi-model subscription built for coding agents. Stop juggling separate accounts, API keys, balances and rate limits for every model provider — subscribe once, get one API key, and switch models per task.

$9.90 $4.90 first month Limited-time launch price

$12 every 5 hours · $30 every week · $60 every month — one shared quota, all models

Request volume
316,290req / 5h on DeepSeek V4 Flash

The same $12 / 5-hour quota buys a very different number of requests depending on the model — here's the estimated breakdown, highest first.

See the full breakdown, all 13 models →
13 models, one entry point
One account, one API key
Real-time usage across 3 windows
Up to 1.58M requests/mo on one model
Pricing

One plan. One quota. 13 models.

100 slots

Harogo

Everything a Coding Agent workflow needs, on one subscription.

$9.90$4.90first month
  • $9.90/month after the first month
  • $12 / $30 / $60 shared quota — 5-hour, weekly, monthly
  • 13 models across GLM, Qwen, Kimi, MiniMax, DeepSeek, Claude and GPT
  • One account, one API key, one Base URL
  • Switch models per task, no separate sign-ups
  • Real-time usage and remaining-quota dashboard
  • Cancel anytime
Get Harogo

Limited to the first 100 discounted subscription slots. Remaining slots and settlement price are shown on the subscription page. Model list and pricing may change — check the console for current status.

Discounted slots Limited-time · first come, first served
26/ 100example data

discounted slots used so far

26 used (example)74 remaining (example)

Illustrative only — this isn't a live counter. See the actual, current remaining slots on the subscription page.

Compared to subscribing to each model API separately:

4 separate model API subscriptions4 accounts · 4 bills
Harogo$9.90/mo · 1 bill
See the live remaining slots →
Why Harogo

One subscription instead of a spreadsheet of accounts

Different models are good at different coding tasks. Using several providers directly usually means managing several accounts, keys, balances, rate limits and bills at once.

Subscribing to each provider directly

  • A separate account per model provider
  • A separate API key to store and rotate per provider
  • A separate balance to top up and monitor
  • A separate bill and rate limit to track

With Harogo

  • One Harogo account for all 13 models
  • One API key, one Base URL for every model
  • One shared quota, settled across every call
  • One subscription price, one dashboard

You still choose the model per task — Harogo doesn't route automatically. It just removes the account, key and billing overhead of doing that across multiple providers.

Models

13 models, grouped by what they're good at

These categories are for reference when choosing a model. Actual results depend on the task, context and how you use the tools. Availability may change — check your dashboard for current status.

Large codebases & hard reasoning

Large codebases, long-context tasks and high-difficulty reasoning.

Kimi K3
kimi-k3

Complex coding, long context, agent tasks

Quota cap $151,040 req/5h
Claude Opus 5
claude-opus-5

Complex agent coding and hard engineering tasks

Quota cap $60300 req/5h
GPT-5.6 Sol
gpt-5.6-sol

Hard code generation, review and reasoning

Quota cap $601,130 req/5h

Harogo is an independent service and is not affiliated with, endorsed by, or sponsored by OpenAI, Anthropic, Zhipu AI, Alibaba, Moonshot AI, DeepSeek or MiniMax. Model availability may change; see your dashboard for the current list.

Request volume

The quota goes a long way — that's the point

Limits are set in USD value, not a fixed request count, so what that buys you depends entirely on the model. Here's what the same $12 / 5-hour window is estimated to give you, model by model.

Bars use a log scale for readability across a wide range — compare the numbers on the right, not the bar lengths, for exact values. Estimates assume typical request patterns for each model (input length, cache hit rate, reasoning effort, output length) and are not a fixed-count guarantee. Actual counts vary by prompt, cached context, reasoning effort and output length; a single coding task can involve multiple model calls.

Choosing a model

You pick the model — here's a starting point

Harogo doesn't auto-route between models. You choose per task; these four groups are a quick way to decide.

1

Hard problem?

Large codebase, long context, or a genuinely difficult reasoning task — reach for Kimi K3, Claude Opus 5 or GPT-5.6 Sol.

2

Regular work?

Day-to-day feature work, generation and debugging — GLM-5.2, Qwen3.8 Max, Claude Sonnet 5, GPT-5.6 Terra, MiniMax M3 or DeepSeek V4 Pro.

3

High frequency?

Quick, repeated, cost-sensitive calls — DeepSeek V4 Flash, Claude Haiku 4.5 or GPT-5.6 Luna stretch the quota the furthest.

4

Long agent run?

A workflow with many steps over a long session — Claude Fable 5 is built for that shape of task.

Switching between them costs nothing extra to set up — same account, same key, same Base URL. See the full model list for IDs and per-model quota caps.

Usage

Limits you can actually understand

Three rolling windows apply simultaneously. Every model call draws from the same shared quota, priced in USD value rather than a fixed request count.

Every 5 hours
$12
shared usage
Every week
$30
shared usage
Every month
$60
shared usage

All three limits apply at the same time; hitting any one of them pauses further paid usage until that window resets. Cheaper models allow far more requests per dollar than premium models — see the request volume breakdown above.

Look up a model's real numbers

Pick a model to see its estimated request counts and per-token pricing — pulled directly from the current model table, not a guess.

Actual request counts vary with prompt length, cached context, reasoning effort and output length. Figures are estimates, not a fixed-count guarantee.
Use cases

Built for real development work

Ship features

Build new features, endpoints, pages and database work end to end.

CoreGLM-5.2, Claude Sonnet 5, GPT-5.6 Terra

Debug at high frequency

Run many quick edit-and-test loops without burning through your quota.

FastDeepSeek V4 Flash, Claude Haiku 4.5

Review hard changes

Bring in a reasoning-heavy model for security review and edge cases before you ship.

DeepClaude Opus 5, Kimi K3, GPT-5.6 Sol

Run long agent sessions

Multi-step workflows that need to stay coherent across a long run.

LongClaude Fable 5
Compatibility

Use Harogo where you already code

Harogo exposes an OpenAI-compatible Base URL at https://harogo.ai, so it works with any client that lets you set a custom base URL and API key.

Terminal coding agents
OpenAI-compatible clients
Editor plugins & custom scripts
Account & usage

Clear rules, one dashboard

Usage and remaining quota are shown in real time in your dashboard.
One subscription is for one individual account.
Model list, IDs and pricing are published and versioned in the console.
Keep your API key private — don't commit it or share it.

On model access

Harogo is an independent multi-model subscription service and is not affiliated with, endorsed by, or sponsored by any of the model providers listed on this page. Model names and IDs shown here are the routing identifiers Harogo uses — see console.harogo.ai for the current, authoritative list.

FAQ

Questions, answered plainly

One Harogo account, one API key, and access to all 13 supported models through a single Base URL and a single shared usage quota. You switch models per task — no separate sign-ups.
$4.90 for the first month, then $9.90/month. Currently limited to 100 discounted subscription slots — remaining slots and settlement price are shown on the subscription page.
No. Three rolling usage windows apply at the same time: $12 every 5 hours, $30 every week, $60 every month. Limits are defined in USD value, not a fixed request count. You can see current usage and remaining quota in the console.
It depends entirely on which model you use — cheaper, faster models allow far more requests per dollar. See the request volume breakdown for per-model estimates, from roughly 220 requests/5h on the heaviest model up to roughly 316,290 requests/5h on the lightest. These are estimates based on typical usage patterns, not a fixed-count guarantee, and actual counts vary with context length, cache hits, reasoning effort and output length.
No. You choose the model per task via the model ID (e.g. deepseek-v4-flash or claude-opus-5). The model guide above is a starting point for which model fits which kind of work.
Reaching any one of the three rolling windows pauses further paid usage on that window until it resets. Your dashboard shows exactly how much is remaining and when each window resets.
No. A Harogo subscription is for a single individual account.
No. Harogo is an independent service and is not endorsed or sponsored by OpenAI, Anthropic, Zhipu AI, Alibaba, Moonshot AI, DeepSeek or MiniMax.

One key. 13 models. Real request volume.

One subscription, one shared quota, no juggling separate provider accounts.

$9.90$4.90first month

Cancel anytime. Usage limits apply. 100 discounted slots available now.