87.5% uptime and 1,134ms average latency make Terminal.Pub a budget experiment, not a production backbone for Chinese developers needing GPT-4o or Claude 3.5 Sonnet without a VPN. This is a pure price-play relay. You get two capable models for what looks like pennies, but you pay for it in reliability. Let’s cut through the marketing and look at the raw numbers.
Pricing & Payment: The Only Reason to Look
The pricing is aggressively low. Terminal.Pub’s data shows Claude Opus 4.6 priced at ¥0.5 (input) / ¥2.5 (output) per million tokens. For reference, that’s a fraction of the official API cost. This is the primary draw.
- Free Trial: The platform offers a free_trial: True flag, meaning you can test the water without recharging.
- Payment Methods: They accept Alipay and WeChat Pay, which is standard for domestic Chinese users. No crypto or international cards here.
- No Promo Code: The platform data confirms Promo code: None available. What you see is what you get.
The catch? There’s no monthly subscription, and the Min recharge (CNY): Not specified. This lack of transparency on the minimum top-up is a minor red flag for a new service.
Plan Price (per million tokens) Notes GPT-4o Input Included in base pricing Check platform for exact rate GPT-4o Output Included in base pricing Check platform for exact rate Claude 3.5 Sonnet Input Included in base pricing Check platform for exact rate Claude 3.5 Sonnet Output Included in base pricing Check platform for exact rate Claude Opus 4.6 Input ¥0.5 Extremely low Claude Opus 4.6 Output ¥2.5 Extremely low Bottom line: The pricing is the hook. If you are purely cost-sensitive and can tolerate downtime, this is your only reason to sign up.
Models & Capabilities: Limited but Focused
Terminal.Pub only lists GPT-4o and Claude 3.5 Sonnet as available models. Do not expect a 200-model router like OpenRouter. This is a narrow, focused offering.
- Max tokens: 100,000. This is decent for long-context tasks like document analysis or code review, but not bleeding-edge (Claude’s 200k context is not available here).
- Safety rating: 3/5. This is a neutral score. It’s not actively censoring, but it’s not a safe harbor for uncensored model behavior either.
- Speed rating: 3.5/5. Combined with the 1,134ms latency, expect noticeable lag on each request. This is not for real-time chat. Key takeaways:
- Only two flagship models are available. No niche or open-source models.
- The 100k token context window is adequate but not best-in-class.
- The platform is designed for cost-sensitive batch processing, not interactive applications.
Performance & Reliability: The Hard Truth
Here is the data that should make you pause. Terminal.Pub’s Uptime: 87.5%. In a 30-day month, that means nearly 4 full days of downtime. For any production service, this is catastrophic.
- Latency: 1,134ms (over 1 second). This is measured from the relay, not your Chinese ISP. In practice, expect this to be higher due to the Great Firewall routing.
- Composite score: 52.5/100. This is a low overall rating, dragged down by the uptime. The Cons in the data are clear: “Very new (Feb 2026) — unproven long-term stability” and “Such low pricing raises sustainability concerns.” This is a platform launched less than a year ago. There is no track record. During peak hours, expect throttling. The data doesn’t specify rate limits per tier, but with an 87.5% uptime, the service is likely overwhelmed or under-provisioned. For a Chinese developer, this means your batch jobs might fail silently, and you’ll need retry logic. Key takeaways:
- 87.5% uptime is unacceptable for production. Use only for testing or non-critical tasks.
- 1,134ms latency is slow. Add Chinese ISP routing delay, and you’re looking at 2-3 seconds per request.
- The platform is too new to trust for long-term projects.
Pros & Cons
Pros
- Extremely low pricing, especially for Claude Opus 4.6.
- Free API groups available to test without cost.
- Supports Chinese payment methods (Alipay, WeChat Pay). Cons
- 87.5% uptime is unreliable for production workloads.
- Very new (Feb 2026) — no long-term stability data.
- No monthly subscription option limits budget predictability.
- Limited to only 2 models (GPT-4o, Claude 3.5 Sonnet).
- High latency (1,134ms) makes it unsuitable for real-time apps.
Verdict
Terminal.Pub is a budget relay for tinkering. If you are a Chinese developer who needs to test a few prompts on GPT-4o or Claude 3.5 Sonnet without paying full price, this is a valid option. Use the free trial, run your experiments, and move on. Do not use this for any customer-facing application, automated workflow, or task that requires consistent uptime. The 87.5% uptime and high latency are deal-breakers for production. The platform’s long-term sustainability is a question mark. It launched in Feb 2026 and is already offering prices that seem unsustainable. You are betting that the service will still be online next month. Final recommendation: Use for testing. Have a backup relay like OpenRouter or a direct API key ready for when Terminal.Pub goes down.
FAQ
Q: Can I use Terminal.Pub for a production chatbot application? A: No. The 87.5% uptime means your service will be down for approximately 4 days per month. This is not suitable for production. Use it for testing or batch processing where failures are acceptable. Q: How do I pay for Terminal.Pub as a Chinese user? A: The platform accepts Alipay and WeChat Pay, which are standard domestic Chinese payment methods. There is no support for international credit cards or crypto. Q: What is the actual latency I should expect from China? A: The platform’s base latency is 1,134ms. From a Chinese ISP, add the Great Firewall routing overhead. Expect 1.5 to 3 seconds per request in practice. This is too slow for real-time chat. Q: Does Terminal.Pub offer a free trial? A: Yes.
Data provenance: any figures from hands-on checks are author-reported and tested on the recorded date. Independent verification is unavailable unless a source is linked.