Aiberm lists 2 models; GPT-4o and Claude 3.5 Sonnet are the only options available.

Model Selection and Supported LLMs

Two models is a thin lineup. Most relay stations in this space offer 10-50 models. Aiberm covers the two most requested Western models, but if you need DeepSeek, Kimi, or Grok, you are out of luck. The platform description mentions “Multi-provider API access beyond Claude/GPT,” but that claim does not match the actual model list. Key takeaways:

  • Only GPT-4o and Claude 3.5 Sonnet are available
  • No Chinese LLM providers are supported despite the description
  • Limited utility for developers needing model diversity

Performance Benchmarks

Latency and Speed

Aiberm’s average latency is 1308ms with a speed rating of 3.0/5. For comparison, faster relay stations in this category typically run under 800ms. The 1308ms figure means you will feel the delay during chat interactions. Batch processing or streaming code completion will feel sluggish.

Context Window

Max tokens is 100,000. That covers GPT-4o’s 128K context and Claude 3.5 Sonnet’s 200K context partially. You will not get full context windows on either model. Long document analysis or large codebase queries will hit this limit.

Uptime and Stability

Uptime is 99.3%. That is decent but not best-in-class. During peak hours (China evening, US daytime), you may see occasional timeouts. The safety rating is 3/5 - no data on what incidents caused this, but it signals reliability concerns. Key takeaways:

  • 1308ms latency is above average; expect noticeable delays
  • 100K max tokens truncates both GPT-4o and Claude 3.5 Sonnet
  • 99.3% uptime is stable but not exceptional

Pricing and Payment

No pricing data is provided beyond the $0/month base (free trial). The platform does not publish per-token rates. The cons note that “Claude pricing at ~2x official — not the cheapest.” This is a significant markup. GPT-4o pricing is not specified but likely follows similar margins. Payment methods: Alipay and WeChat Pay. Both are standard for Chinese users. No promo code is available. No minimum recharge amount is specified.

Model Official Price (per 1M input tokens) Aiberm Price (per 1M input tokens) Markup
GPT-4o $2.50 Not specified Unknown
Claude 3.5 Sonnet $3.00 Not specified ~2x
Pricing reflects 2026 rates and may change. Official prices are from provider published rates.
Key takeaways:
  • Claude 3.5 Sonnet is roughly 2x official pricing
  • No per-token pricing is published
  • Alipay and WeChat Pay are the only payment options

Pros & Cons

Pros:

  • Covers GPT-4o and Claude 3.5 Sonnet, the two most requested models
  • Alipay and WeChat Pay support for Chinese users
  • 99.3% uptime is acceptable for non-critical workloads Cons:
  • Only 2 models supported - very limited selection
  • Claude pricing at ~2x official is expensive
  • 1308ms latency is slower than competitors
  • 100K max tokens limits context window on both models
  • No Chinese LLM support despite description claiming otherwise
  • No refund policy or minimum recharge information

Verdict

Aiberm works if you only need GPT-4o or Claude 3.5 Sonnet and cannot access them directly. But the 2x markup on Claude, 1308ms latency, and 100K token cap make it a poor value. Most developers are better served by a relay station with more models, lower latency, and transparent pricing. Skip Aiberm unless you have no other option for these two specific models.

FAQ

What models does Aiberm support?

Aiberm supports GPT-4o and Claude 3.5 Sonnet. No other models are available despite the description mentioning broader support.

What is the latency like?

Average latency is 1308ms with a speed rating of 3.0/5. This is slower than most competing relay stations.

What payment methods are accepted?

Alipay and WeChat Pay are supported. No credit cards or other methods are listed.

Is there a free trial?

The platform lists a $0/month base, but no free trial is explicitly mentioned. No promo code is available.

What is the max context window?

100,000 tokens. This truncates both GPT-4o’s native 128K and Claude 3.5 Sonnet’s 200K context windows.

Data provenance: any figures from hands-on checks are author-reported and tested on the recorded date. Independent verification is unavailable unless a source is linked.