2 models listed; Claude 3.5 Sonnet and GPT-4o are the only options.

Model Breakdown

ClaudeCN is not a multi-model relay. It’s a focused proxy for two models: Claude 3.5 Sonnet and GPT-4o. That’s the entire catalog. If you need Gemini, DeepSeek, or any open-weight model, this isn’t the site for you.

Claude 3.5 Sonnet

This is clearly the flagship. The service describes itself as “Claude-focused with MAX quality,” which suggests they’re prioritizing bandwidth and stability specifically for Sonnet traffic. Given the 100K max token limit matches Anthropic’s own context window, you’re getting the full Sonnet experience — not a quantized or truncated version. The 94% uptime is below what I’d expect from a dedicated single-provider relay. For comparison, most multi-model services hover around 96-98% for their top models. That 4% downtime gap matters if you’re building production workflows around Claude.

GPT-4o

Included, but feels like an afterthought. The “Claude-only” pro/con note confirms the platform’s DNA is Anthropic-first. I’d expect GPT-4o routing to be secondary priority here — likely lower caching priority and slower fallback times during congestion. No speed benchmarks are provided, but the 100K token cap applies to both models.

Pricing

No free trial is available. Payment goes through Alipay or WeChat Pay — standard for Chinese users. There’s no promo code and no minimum recharge amount specified. The composite score of 75.2/100 suggests the value proposition is average compared to competitors.

Model Context Window Uptime Pricing Tier
Claude 3.5 Sonnet 100K tokens 94.0% Premium (Claude-focused)
GPT-4o 100K tokens 94.0% Premium (included)

Pros & Cons

Pros

  • Single-provider focus means Claude 3.5 Sonnet gets full routing priority
  • China-optimized routing avoids VPN overhead
  • 100K context window matches official Anthropic limits Cons
  • Only 2 models — no Gemini, DeepSeek, or open-source options
  • 94% uptime is below average for a paid relay service
  • No free trial to test quality before committing
  • Premium pricing with no multi-model flexibility

Verdict

ClaudeCN is for one specific developer: the Claude 3.5 Sonnet power user in China who doesn’t want to manage VPNs and doesn’t care about any other model. If that’s you, the China-optimized routing and full 100K context are legit advantages. The 94% uptime is concerning — you’ll hit downtime roughly 22 hours per month. For everyone else: you’re paying premium prices for a two-model catalog with below-average reliability. Multi-model relays like Helpaio or OpenRouter give you 50+ models at similar or lower cost with better uptime. Skip this unless you’re all-in on Claude and willing to accept the reliability trade-off.

FAQ

Q: Can I use ClaudeCN for GPT-4o-heavy workloads? A: Technically yes, but the platform is Claude-optimized. GPT-4o traffic likely gets secondary priority. If GPT-4o is your primary model, look elsewhere. Q: Does the 100K token limit apply per request or per session? A: It’s the per-request maximum context window. Both Claude 3.5 Sonnet and GPT-4o support up to 100K tokens in a single API call. Q: What happens when uptime drops below 94%? A: No refund policy is specified. You’re paying for access with no guaranteed SLA. The 94% figure is an average — individual months may be worse.

Data provenance: any figures from hands-on checks are author-reported and tested on the recorded date. Independent verification is unavailable unless a source is linked.