Infer by Flow7 gateway details

Flow7's prepaid multi-model API gateway provides Responses, Chat Completions, and Anthropic Messages compatibility, request spend limits, locked price versions, and completed-response charge receipts. Public model-family labels do not establish verified upstream supplier or model-weight provenance.

Small amounts
Trial first

Overview

Flow7's prepaid multi-model API gateway provides Responses, Chat Completions, and Anthropic Messages compatibility, request spend limits, locked price versions, and completed-response charge receipts. Public model-family labels do not establish verified upstream supplier or model-weight provenance.

Best for: For coding-agent and API developers who need a Responses-compatible endpoint, request-level spend controls, and charge evidence.

Availability and latency

Current status
Test yourself
Latency reference
Depends on your deployment
/models check
/api/public/catalog
Last verified
2026-10-01
Status notes
The public catalog, status, docs, and OpenAPI contract were accessible; the status API reported operational. AI Beyond has not tested paid inference or latency. Check model and policy availability immediately before use.

Supported models

OpenAIClaudeGeminiDeepSeekGrokMoonshot

Pricing and multipliers

Pricing mode
Pay-as-you-go from a prepaid USD wallet. First funding starts at $20; later reloads start at $50, with applicable tax potentially added. The maximum estimated cost is reserved before dispatch, then settled with usage, the locked price version, final charge, and a receipt. Current model and policy-selector prices come from the public catalog.
Price tags
Pay-as-you-go
Free quota
No free live inference quota

Payment methods

Credit card

How to call the API

Usually you create an API key in the platform console, then replace the SDK's baseURL, model name, and auth details per the docs. Test streaming, function calling, error codes, and rate limits before launch.

import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AI_GATEWAY_API_KEY,
  baseURL: "see the platform docs"
});

const response = await client.chat.completions.create({
  model: "gpt-4.1-mini",
  messages: [{ role: "user", content: "Test one call through an AI gateway." }]
});

Pros

  • Supported models: OpenAI, Claude, Gemini, DeepSeek, Grok, Moonshot.
  • Capabilities: OpenAI-compatible, Streaming, Function calling.
  • Documentation completeness: High; stability not tested.

Cons

  • Public paid beta with no SLA and an incomplete production-readiness review. Serving suppliers remain private and unattested; model labels and policy selectors do not prove supplier identity, weights, or snapshot provenance. Current docs describe live text streaming, while older client records used buffered SSE; compatibility applies only to the listed versions and tested scope. Wallet balances and receipts are available in the authenticated dashboard; this does not establish a public balance-query or log-query API. Check status before funding or paid requests, and match privacy policies and terms to sensitive workloads.
  • Third-party relay pricing, models, and availability can change, so re-check them periodically.

Alternatives

OpenRouter

Details
Aggregator

A multi-model routing platform for developers; switch models by changing the base_url with the OpenAI SDK, ideal for teams comparing the quality and price of many models at once.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Stripe, Free trial
Price
Billed per model and provider; some models offer free routing or free quota.

Vercel AI Gateway

Details
Cloud vendorHigh stability

Vercel's official AI Gateway, emphasizing unified model calls, failover, caching, and observability; ideal for apps already deployed on Vercel.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Free trial
Price
Settled via your Vercel account and model usage; the docs stress the gateway itself adds no markup.

302.AI

Details
Aggregator

A tool and API aggregation platform covering text, image, video, and audio AI capabilities; ideal for individuals and small teams wanting one-stop access to many models.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Free trial
Price
Billed per tool or model usage; the site lists separate prices per model.

OhMyGPT

Details
AggregatorPrepaidLow cost

A common API relay service among Chinese users; the site highlights aggregation of OpenAI, Claude, and Gemini plus multiple top-up methods.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Alipay, WeChat, USDT, PayPal, Stripe, Free trial
Price
Pay-as-you-go after top-up; the site lists per-model token prices.

FAQ

Is Infer by Flow7 compatible with the OpenAI API?
This entry is marked as supporting the OpenAI-compatible format, but the exact base_url, model names, and parameters still follow the official docs.
Is Infer by Flow7 ready for production?
Run a small-traffic load test, rate-limit test, and error-code test first, and confirm the billing rules; keep a backup provider for important workloads.

Risk

Risk notice

This site only organizes information about AI API relay services; it makes no guarantee about third-party platforms' stability, price, data security, or availability. Judge the risk yourself before use, and avoid uploading sensitive data, personal privacy, trade secrets, or core API keys to unknown platforms.