↓ Skip to main content
  1. Agents/

Model access

The layer that sells access to models themselves: gateways and routers metering a percentage, vendor coding plans selling a quota, and flat subscriptions selling a ceiling. Editors and harnesses live in their own categories; this is where the token bill gets paid.

  • Cerebras Code - wafer-scale inference sold as speed, $50/$200 per month, currently sold out.
  • Chutes - decentralized inference with pay-as-you-go plus $10/$20 subscriptions capped at 5x pay-as-you-go value.
  • GLM Coding Plan - Z.AI’s flat monthly quota for the GLM line, from $18, restructured twice since launch.
  • Kimi Code - Moonshot’s membership ladder for coding, $19 to $199 monthly with Code from the second tier.
  • MiniMax Coding Plan - token-quota subscriptions for the MiniMax line, $22 to $132 per month, born from a silent plan replacement.
  • NanoGPT - the community aggregator: hundreds of routes pay-as-you-go plus a $12 open-weight subscription.
  • OpenCode Go - the OpenCode team’s $10/month open-model token pack, usable from any agent.
  • OpenCode Zen - the OpenCode team’s curated pay-per-use gateway over benchmarked endpoints.
  • OpenRouter - the largest model gateway, passthrough tokens plus a 5.5% credit fee, now joining Stripe.
  • Requesty - EU-residency gateway charging a flat 5% markup on upstream spend.
  • Synthetic - a flat $30/month subscription for open-weight coding LLMs aimed at agent users.

Its members are compared on shared rows in the Model Access Feature Matrix.

Changes
#

  • 2026-09-26 - Added Cerebras Code.
  • 2026-09-26 - Added Chutes.
  • 2026-09-26 - Added GLM Coding Plan.
  • 2026-09-26 - Added Kimi Code.
  • 2026-09-26 - Added MiniMax Coding Plan.
  • 2026-09-26 - Added NanoGPT.
  • 2026-09-26 - Added OpenCode Go.
  • 2026-09-26 - Added OpenCode Zen.
  • 2026-09-26 - Added OpenRouter.
  • 2026-09-26 - Added Requesty.
  • 2026-09-26 - Added Synthetic.