Aiberm lists 2 models; GPT-4o and Claude 3.5 Sonnet are the only options available.
Model Selection and Supported LLMs
Two models is a thin lineup. Most relay stations in this space offer 10-50 models. Aiberm covers the two most requested Western models, but if you need DeepSeek, Kimi, or Grok, you are out of luck. The platform description mentions “Multi-provider API access beyond Claude/GPT,” but that claim does not match the actual model list. Key takeaways:
- Only GPT-4o and Claude 3.5 Sonnet are available
- No Chinese LLM providers are supported despite the description
- Limited utility for developers needing model diversity
Performance Benchmarks
Latency and Speed
Aiberm’s average latency is 1308ms with a speed rating of 3.0/5. For comparison, faster relay stations in this category typically run under 800ms. The 1308ms figure means you will feel the delay during chat interactions. Batch processing or streaming code completion will feel sluggish.
Context Window
Max tokens is 100,000. That covers GPT-4o’s 128K context and Claude 3.5 Sonnet’s 200K context partially. You will not get full context windows on either model. Long document analysis or large codebase queries will hit this limit.
Uptime and Stability
Uptime is 99.3%. That is decent but not best-in-class. During peak hours (China evening, US daytime), you may see occasional timeouts. The safety rating is 3/5 - no data on what incidents caused this, but it signals reliability concerns. Key takeaways:
- 1308ms latency is above average; expect noticeable delays
- 100K max tokens truncates both GPT-4o and Claude 3.5 Sonnet
- 99.3% uptime is stable but not exceptional
Pricing and Payment
No pricing data is provided beyond the $0/month base (free trial). The platform does not publish per-token rates. The cons note that “Claude pricing at ~2x official — not the cheapest.” This is a significant markup. GPT-4o pricing is not specified but likely follows similar margins. Payment methods: Alipay and WeChat Pay. Both are standard for Chinese users. No promo code is available. No minimum recharge amount is specified.
| Model | Official Price (per 1M input tokens) | Aiberm Price (per 1M input tokens) | Markup |
|---|---|---|---|
| GPT-4o | $2.50 | Not specified | Unknown |
| Claude 3.5 Sonnet | $3.00 | Not specified | ~2x |
| Pricing reflects 2026 rates and may change. Official prices are from provider published rates. | |||
| Key takeaways: |
- Claude 3.5 Sonnet is roughly 2x official pricing
- No per-token pricing is published
- Alipay and WeChat Pay are the only payment options
Pros & Cons
Pros:
- Covers GPT-4o and Claude 3.5 Sonnet, the two most requested models
- Alipay and WeChat Pay support for Chinese users
- 99.3% uptime is acceptable for non-critical workloads Cons:
- Only 2 models supported - very limited selection
- Claude pricing at ~2x official is expensive
- 1308ms latency is slower than competitors
- 100K max tokens limits context window on both models
- No Chinese LLM support despite description claiming otherwise
- No refund policy or minimum recharge information
Verdict
Aiberm works if you only need GPT-4o or Claude 3.5 Sonnet and cannot access them directly. But the 2x markup on Claude, 1308ms latency, and 100K token cap make it a poor value. Most developers are better served by a relay station with more models, lower latency, and transparent pricing. Skip Aiberm unless you have no other option for these two specific models.
FAQ
What models does Aiberm support?
Aiberm supports GPT-4o and Claude 3.5 Sonnet. No other models are available despite the description mentioning broader support.
What is the latency like?
Average latency is 1308ms with a speed rating of 3.0/5. This is slower than most competing relay stations.
What payment methods are accepted?
Alipay and WeChat Pay are supported. No credit cards or other methods are listed.
Is there a free trial?
The platform lists a $0/month base, but no free trial is explicitly mentioned. No promo code is available.
What is the max context window?
100,000 tokens. This truncates both GPT-4o’s native 128K and Claude 3.5 Sonnet’s 200K context windows.
Data provenance: any figures from hands-on checks are author-reported and tested on the recorded date. Independent verification is unavailable unless a source is linked.