$ uprouter — directory
Low-risk AI API routers & providers
118 AI API routers pass UPROUTER.ONLINE's low-risk bar: an established operator, clear terms and a stable operating history. Among them, Nscale, Vertex AI, Mistral AI carry the biggest free-credit allowances.
Risk is an evidence-based rating from community research — it weighs how long the operator has run, how transparent its terms are, and its reliability record, and the reasoning is shown on each entry. It is a starting point, not a guarantee, and it is never influenced by affiliate relationships.
This list is ordered by free-credit value, never by sponsorship. Treat the rating as one input alongside your own due diligence, and verify current terms with the provider before purchase.
- Nscaleinference.api.nscale.comOfficial free tier
Nscale is a European AI hyperscaler providing Serverless Inference for large-scale AI deployment. It offers an OpenAI-compatible API to leading open models like Llama and DeepSeek. New users receive a $5 signup credit to start prototyping without initial cost.
models: Llama family, DeepSeek family, gpt-oss-120b
up · 279ms$1,089.78estconnectLow risk - Vertex AIcloud.google.comPaid API
Google Cloud's enterprise-grade AI platform (Vertex AI) providing managed access to Gemini models and third-party models like Llama and Mistral. Offers advanced features like grounding with Google Search, model tuning, and provisioned throughput for predictable performance.
models: Gemini 2.5 Flash, Gemini 2.5 Pro, Meta: Llama 3.1 405B Instruct +1
unknown$300.00Low risk - Amazon Bedrockaws.amazon.comPaid API
Amazon Bedrock is a fully managed service that offers a choice of high-performing foundation models from leading AI companies like Anthropic, Meta, and Mistral via a single API. It is primarily pay-as-you-go with various service tiers (Standard, Flex, Priority).
models: Nova Lite 1.0, Nova Pro 1.0, Claude family +2
up · 65msno free tierLow risk - Deepgramdeepgram.comSearch / Audio API
Deepgram is a specialized speech-to-text and text-to-speech API provider. It offers high-performance audio models like Nova-3 and Flux, with real-time streaming and batch processing capabilities. Authentication is via API key created in its own dashboard.
unknown$200.00Low risk - Mistral AIconsole.mistral.aiOfficial free tier
Mistral AI's official platform (La Plateforme) provides access to their open and proprietary models via an OpenAI-compatible API. It features a free 'Experiment' tier for developers and pay-as-you-go pricing for production workloads.
models: Mistral Small 4, Mistral Medium 3.5, Mixtral 8x22B Instruct +3
down$200.00estconnectLow risk - SenseNovatoken.sensenova.cnOfficial free tier
token.sensenova.cn is the official Token Plan for SenseTime's SenseNova multimodal AI platform. During its public beta, it offers generous free quotas for models like SenseNova 6.8 Flash Lite and U1 Fast, suitable for complex office workflows.
models: SenseNova 6.7 Flash-Lite, SenseNova U1 Fast, DeepSeek V4 Flash +1
up · 718ms$200.00estconnectLow risk - LongCatlongcat.chatFree product
LongCat is Meituan's large model API platform (longcat.chat/platform), offering OpenAI/Anthropic-compatible endpoints. During its public beta phase in 2026, it provides generous daily free token allowances for models including LongCat-Flash and LongCat-2.0, with a context window of 131K tokens.
models: LongCat-Flash-Chat, LongCat-Flash-Thinking, LongCat-Flash-Lite
up · 1223ms$150.00estconnectLow risk - xAI Consoleconsole.x.aiOfficial free tier
The xAI Console is the official developer interface for Elon Musk's Grok models. It offers a usage-based API with prepaid credits, prioritizing performance and direct access to their flagship large language models.
models: grok-beta, grok-2, grok-3 +1
up · 19ms$150.00estconnectLow risk - You.com Searchyou.comSearch / Audio API
You.com provides a specialized Web Search API for AI agents and LLM applications. It offers real-time web intelligence and structured search results (snippets, full text). Features a consumption-based pricing model and generous onboarding credits for new developers.
unknown$100.00Low risk - ModelScopemodelscope.cnOfficial free tier
ModelScope (modelscope.cn), backed by Alibaba Cloud, is an open-source model community and inference platform. It provides a free API tier allowing users to make up to 2,000 OpenAI-compatible calls per day across a wide range of open-source models (Qwen, DeepSeek, GLM).
models: Qwen family (most stable), DeepSeek-V3.1, Kimi +2
up · 1972ms$60.00estconnectLow risk - Gladiagladia.ioSearch / Audio API
Gladia is a speech-to-text API provider offering high-accuracy asynchronous and real-time transcription across 100+ languages. It features speaker diarization and automatic language detection, billed via a prepaid credit wallet system with a significant free starting grant.
unknown$55.00estLow risk - Context7 (library docs)context7.comMonitor / Directory
Context7 (powered by Upstash) is a documentation search API for LLMs, providing up-to-date context from libraries and codebases. It offers an anonymous free tier and a free API key for higher limits to support coding agents.
unknown$50.00Low risk - Groqconsole.groq.comOfficial free tier
Groq Cloud provides ultra-fast LLM inference using its proprietary LPU (Language Processing Unit) hardware. It offers an OpenAI-compatible API with a substantial free tier for developers to test and build applications at scale.
models: Qwen family, Whisper Large v3, Llama-3.1-8B-Instruct +1
up · 200ms$50.00estconnectLow risk - Basetenapp.baseten.coOfficial free tier
Baseten is a robust model inference and deployment platform that enables enterprises to run custom and open-weights models in production. It offers high scalability, dedicated GPU infrastructure, and a developer-friendly API for deploying models from a comprehensive library.
models: Any supported model in their library (billed by compute; no fixed free-model list)
up$30.00connectLow risk - Cerebrascerebras.aiOfficial free tier
Cerebras Inference is powered by wafer-scale WSE chips, offering world-leading inference speeds. It provides an OpenAI-compatible API with $5 in free credits and generous rate limits for developers. Details were checked against the provider's own page at www.cerebras.ai.
models: gpt-oss-120b, Qwen3 32B, Llama 4 Scout +1
up$30.00estconnectLow risk - Cerebras APIapi.cerebras.aiOfficial free tier
Cerebras Inference offers ultra-fast AI inference powered by its Wafer-Scale Engine (WSE-3) technology, achieving record-breaking speeds up to 3000 tokens/s. The official API is OpenAI-compatible and features a generous free tier alongside professional pay-as-you-go access.
models: Llama 3.1/3.3 family, Qwen 3 family (per official model list), Llama 4 Scout +1
up · 43ms$30.00estconnectLow risk - Exa Searchexa.aiSearch / Audio API
Exa is an internet-scale search and content extraction API designed for AI agents. It features embeddings-based search, deep research capabilities, and structured content retrieval. Details were checked against the provider's own page at exa.ai.
unknown$30.00Low risk - Z.aiapi.z.aiOfficial free tier
Z.ai is Zhipu AI's international developer platform, offering access to the GLM (General Language Model) family. It provides a high-performance alternative to Western models, with the GLM-Flash series offered permanently for free to developers to encourage global adoption of its flagship reasoning models.
models: GLM-4.5-Flash, GLM-4.6V-Flash, GLM-4.7-Flash
up · 1650ms$30.00estconnectLow risk - Voyage AIwww.voyageai.comPaid API
Voyage AI specializes in state-of-the-art embedding models and rerankers. It provides a managed API for high-quality vector embeddings (voyage-4 family) and efficient reranking, optimized for RAG applications. Offers a highly generous free tier for developers.
unknown$24.00Low risk - Alibaba Cloud Bailianbailian.console.alibabacloud.comOfficial free tier
Alibaba Cloud Model Studio (Bailian) is the official international platform for Qwen and other foundation models. It offers 1 million free tokens per model for new users and a 50% discount on batch inference.
up$20.00estconnectLow risk - DeepInfradeepinfra.comPaid API
DeepInfra is a high-speed inference cloud for open-source AI models. It provides an OpenAI-compatible API for text generation, embeddings, and image generation. Authentication uses a standard API key from its dashboard.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct +3
unknown$20.00estconnectLow risk - Google AI Studioai.google.devOfficial free tier
Google AI Studio (ai.google.dev) is the official developer platform for Google's Gemini model family. It offers both a generous free tier for prototyping and a paid production tier with higher rate limits, context caching, and data privacy guarantees.
models: Gemma 4 26B A4B , Gemma 4 31B, Gemini 2.5 Flash Lite +3
up · 215ms$15.00estconnectLow risk - Lightning AIlightning.aiOfficial free tier
Lightning AI (LitAI) provides a unified gateway to frontier models through an OpenAI-compatible API. New users receive 40 million free tokens (approx. $15 in credits) to test models like Claude, GPT, and Gemini. Usage beyond the initial credit is billed as pay-as-you-go.
models: gpt-5 / gpt-5.5 family, gemini-3.5-flash / gemini-2.5-pro, nemotron-3-ultra-550b +3
up · 87ms$15.00estconnectLow risk - NLP Cloudnlpcloud.comOfficial free tier
NLP Cloud is a commercial provider hosting various open-model APIs (Llama, Mixtral, Dolphin) with both pay-as-you-go and subscription options. It grants new users $15 in free credit upon phone verification.
models: Llama 3.1 405B, gpt-oss-120b, Llama 3.3 70B Instruct
up$15.00estconnectLow risk - StepFunplatform.stepfun.comOfficial free tier
StepFun Open Platform is the official developer platform of a Chinese AI unicorn, serving its Step-series multimodal models via an OpenAI-compatible API. It supports text, vision, and end-to-end speech models with tiered rate limits and a credits-based free allowance for new accounts.
models: Step-3, Step-series text LLMs, Step-series multimodal/vision models +1
up · 1157ms$15.00estconnectLow risk - AI21 Labsapi.ai21.comOfficial free tier
AI21 Labs is an established LLM vendor known for its Jamba model family. The developer platform (api.ai21.com) offers OpenAI-style endpoints for its high-performance models, with a trial credit for new accounts to explore its capabilities.
models: Jamba Large (1.7), Jamba Mini (1.7), Jamba 1.6 Large
up · 89ms$10.00connectLow risk - Eden AIedenai.coCommercial aggregator
Eden AI is a commercial AI aggregation gateway providing a single API to 500+ models. It features cost tracking, automatic fallback, and a unified format for LLMs, vision, and translation. Details were checked against the provider's own page at www.edenai.co.
up · 188ms$10.00connectLow risk - Upstageconsole.upstage.aiOfficial free tier
Upstage Console provides an OpenAI-compatible API for their high-performance Solar models. It features a developer-friendly $10 signup credit and commitment-based tiers for businesses requiring higher rate limits and support.
models: Solar Pro, Solar Mini, Solar Pro 3
up$10.00estconnectLow risk - Inception Labsapi.inceptionlabs.aiOfficial free tier
Inception Labs offers a cloud inference platform for its in-house diffusion-based LLMs, such as Mercury 2 and Mercury Coder. Designed for extreme efficiency, it provides an OpenAI-compatible API that significantly undercuts GPU-based inference costs while maintaining high reasoning quality.
models: mercury (general chat diffusion model), mercury-coder (code, with fim/completions), mercury-2
up · 98ms$7.00estconnectLow risk - Ollamaapi.ollama.comOfficial free tier
Ollama Cloud is the official managed inference platform for the Ollama ecosystem. It allows developers to deploy and call popular open-weights models (Llama, DeepSeek, Qwen) via a unified OpenAI-compatible API, providing a bridge between local development and cloud scale.
models: deepseek-v3.1:671b-cloud, gpt-oss-120b, gpt-oss:20b +1
up · 70ms$7.00estconnectLow risk - Requestyrequesty.aiCommercial aggregator
Requesty is an AI model gateway and router that aggregates over 600 models (Claude, GPT, Gemini, DeepSeek) through a single OpenAI-compatible API. It offers cost optimization, unified billing, and advanced routing features for enterprise and individual developers.
up · 198ms$6.00estconnectLow risk - Firecrawlfirecrawl.devSearch / Audio API
Firecrawl is a search and scrape API designed to turn websites into LLM-ready markdown. It supports crawling, scraping, and real-time search with a usage-based credit system. Details were checked against the provider's own page at www.firecrawl.dev.
unknown$5.00estLow risk - InternAIchat.intern-ai.org.cnOfficial free tier
InternLM (Shanghai AI Laboratory) provides an open platform with an OpenAI-compatible API. It offers a monthly free token allowance for its InternVL and long-thinking reasoning models. Details were checked against the provider's own page at internlm.intern-ai.org.cn.
models: intern-latest, intern-s1, intern-s1-mini +2
up · 1077ms$5.00estconnectLow risk - Modalmodal.comOfficial free tier
Modal (modal.com) is a serverless compute platform optimized for AI infrastructure. Rather than a per-token API, it bills by the GPU-second, allowing developers to deploy custom models. A 'Starter' tier provides $30/month in free compute credits.
models: No fixed free-model list (self-deploy any supported model, billed by compute)
up$5.00estLow risk - Pollinationspollinations.aiFree product
Pollinations.ai is a Berlin-based open-source platform providing a signup-free, OpenAI-compatible API for text, image, and audio generation. It aggregates multiple open-weights models (DeepSeek, Qwen, Mistral) and offers them for free to foster creative AI development.
models: OpenAI GPT-5 (Mini/Nano/5.2), Google Gemini 3 Flash / 2.5 Flash Lite, Mistral Small 3.2 +3
up · 60ms$5.00estconnectLow risk - SambaNovacloud.sambanova.aiOfficial free tier
SambaNova Cloud provides fast LLM inference on specialized RDU hardware. It offers an OpenAI-compatible API with a persistent free tier (Developer Tier) and pay-as-you-go credits for higher rate limits.
models: Llama 3.1 405B, DeepSeek-V3.1/V3.2, gpt-oss-120b +2
up · 293ms$5.00connectLow risk - Vercel AI Gatewayvercel.comOfficial free tier
Vercel AI Gateway is a unified API gateway that allows routing to hundreds of models from various providers through a single OpenAI-compatible endpoint. It provides observability, budgets, and zero-data-retention (ZDR) options with no token markup over upstream list prices.
models: Access to OpenAI/Anthropic/Google/Meta/DeepSeek and more via the gateway (at upstream list prices)
up · 40ms$5.00connectLow risk - v0 (Vercel)v0.devPaid API
v0 by Vercel is an AI-powered full-stack web application builder that generates React components and backend logic. Usage is metered on input and output tokens which convert to credits. It provides access to various models including v0 Mini, Pro, and Max tiers.
unknown$5.00Low risk - AnyAPIanyapi.aiCommercial aggregator
AnyAPI is a commercial multi-model API aggregation gateway providing access to 400+ models via a single key. It offers a generous free tier for developers and scales to high-volume enterprise plans, supporting models from OpenAI, Anthropic, and open-weights families.
models: QwQ 32B, Gemma 4, Qwen3 Coder +3
up · 55ms$3.00estconnectLow risk - Difydify.aiPaid API
Dify is an open-source LLM app development platform. Its cloud service (Dify Cloud) provides hosted model access via an API, with credits consumable across OpenAI, Anthropic, Gemini, and others. Details were checked against the provider's own page at dify.ai.
unknown$3.00estconnectLow risk - Qinius.qiniu.comOfficial free tier
Qiniu Cloud AI LLM Inference is a managed MaaS platform that provides a unified, compliant API for over 50 mainstream LLMs. It is compatible with both OpenAI and Anthropic protocols, offering new users substantial free tokens and flexible resource packages for production scaling.
models: DeepSeek-V4 series, Qwen, GLM (Zhipu) +3
up · 1823ms$3.00estconnectLow risk - SiliconFlowapi.siliconflow.cnOfficial free tier
SiliconFlow (SiliconCloud) is a premier Chinese AI model platform offering high-performance inference for a wide range of open-source models (Qwen, Llama, DeepSeek). It is notable for providing free, permanent API access to models with parameters under 9B, alongside competitive pay-as-you-go rates for larger flagship
models: Qwen2.5-7B-Instruct, Llama-3.1-8B-Instruct
up · 1005ms$2.38estconnectLow risk - Kimiplatform.kimi.comOfficial free tier
The official Moonshot AI developer platform for Kimi models. Offers pay-as-you-go access to Kimi K3, K2.7-Code, and K2.6 over an OpenAI-compatible API. One-off trial credits available upon verification.
up · 745ms$2.23estconnectLow risk - Moonshot AIplatform.moonshot.cnOfficial free tier
Official Moonshot AI developer platform for Kimi models. Offers pay-as-you-go access to Kimi K3, K2.7-Code, and K2.6 over an OpenAI-compatible API. One-off trial credits available upon verification. Details were checked against the provider's own page at platform.moonshot.cn.
models: moonshot-v1-8k, moonshot-v1-32k, moonshot-v1-128k +1
up · 1947ms$2.23estconnectLow risk - Cohereapi.cohere.aiOfficial free tier
Cohere provides enterprise-ready LLMs optimized for search, RAG, and agentic workflows. Its official API offers free Trial keys for development and evaluation, alongside high-performance production endpoints for models like Command R+ and the Aya multilingual series.
models: Command R (08-2024), Command R+ (08-2024), Command R+ +2
up · 49ms$2.00estLow risk - ElevenLabselevenlabs.ioSearch / Audio API
ElevenLabs is a leading AI audio platform providing high-fidelity text-to-speech, speech-to-speech, and voice cloning APIs. It supports multiple languages and low-latency streaming for voice agents. Details were checked against the provider's own page at elevenlabs.io.
unknown$2.00estLow risk - GitHub Modelsmodels.github.aiOfficial free tier
GitHub Models (models.github.ai) was an official model-inference platform by GitHub for prototyping. As of July 30, 2026, the service has been fully retired and the catalog/API are no longer available. Users are directed to Azure AI for production workloads.
models: Llama family, GPT-4o, GPT-4o-mini
up · 28msno free tierconnectLow risk - OpenRouteropenrouter.aiCommercial aggregator
OpenRouter is a commercial LLM gateway aggregating 500+ models from every major provider. It offers a single OpenAI-compatible API with a platform fee (5.5%) and a selection of free models for development.
models: Auto Router, Free Models Router, qwen/qwen3-next-80b-a3b-instruct:free +3
up · 49ms$1.50estconnectLow risk - Hvoy.aihvoy.aiMonitor / Directory
Hvoy AI is a specialized diagnostic and directory site for third-party AI API gateways. It performs continuous latency, uptime, and model-adulteration testing to help developers benchmark and select relay providers. Mirrored at hvoyai.com.
models: sonnet-4.6 (reverse-proxied, limited freebie), gpt5.3 (reverse-proxied, limited freebie)
up · 144ms$1.49estLow risk - Electron Hubwww.electronhub.aiCommercial aggregator
Electron Hub provides a unified API to 600+ models, including premium frontier models (GPT, Claude, Gemini). It offers weekly credit refills on both free and paid subscription plans, plus permanent top-ups.
models: Meta: Llama 3.1 405B Instruct, claude-sonnet-5, Google Gemini Pro Latest +1
unknown$1.00connectLow risk - FastRouterfastrouter.aiCommercial aggregator
FastRouter is a high-performance gateway for AI model routing, comparison, and operations. It supports 200+ models via an OpenAI-compatible API with intelligent cost optimization and failover. Details were checked against the provider's own page at fastrouter.ai.
models: gemini-3.5-flash-lite, gpt-5.6-luna, DeepSeek V4 Flash +3
unknown$1.00estconnectLow risk - Fireworks AIfireworks.aiOfficial free tier
Fireworks AI is a high-speed serverless inference platform for open-weights models (Llama, Qwen, DeepSeek). It provides an OpenAI-compatible API optimized for low latency and high throughput. Details were checked against the provider's own page at docs.fireworks.ai.
models: Llama family, DeepSeek family, Qwen family
up · 171ms$1.00connectLow risk - Hyperbolichyperbolic.aiOfficial free tier
Hyperbolic (hyperbolic.ai) is an open-access AI cloud providing high-performance serverless inference and GPU rentals. It offers an OpenAI-compatible API to 25+ open-source models including DeepSeek, Qwen, and Llama 3.1 405B. New users receive roughly $1 in trial credits upon phone verification.
models: DeepSeek V3, Llama 3.1 405B (BF16, currently the only public 405B Base), Qwen family +2
up · 173ms$1.00estconnectLow risk - Hyperbolic XYZhyperbolic.xyzOfficial free tier
Hyperbolic (hyperbolic.xyz) is the API-specific endpoint for Hyperbolic's AI cloud. It provides serverless inference for over 25 open-source models through an OpenAI-compatible /v1 endpoint. The platform supports text, image, and audio generation with pay-as-you-go pricing.
models: DeepSeek V3, Llama 3.1 405B, Qwen family +1
up · 55ms$1.00estconnectLow risk - Nebiusnebius.comOfficial free tier
Nebius AI Studio is a European AI cloud offering an OpenAI-compatible API to 60+ open-source LLMs; it provides high-performance inference-as-a-service with European data residency and transparent pay-as-you-go pricing.
models: DeepSeek family, Llama/Qwen and 60+ open-source models, Qwen3-235B-A22B +1
up · 1321ms$1.00estconnectLow risk - Nebius Studioapi.studio.nebius.comOfficial free tier
Nebius AI Studio is a European-based GPU cloud platform providing managed inference for top open-source LLMs like Llama 3.1 and DeepSeek R1. It leverages its own EU data centers to offer high-performance, privacy-compliant inference with transparent pay-as-you-go pricing tailored for both developers and enterprises.
models: DeepSeek family, Qwen family, Llama-3.1-8B-Instruct
up · 387ms$1.00estconnectLow risk - OVHcloudovhcloud.comOfficial free tier
OVHcloud AI Endpoints is an inference API from the established European cloud provider OVHcloud. It offers OpenAI-compatible access to open-weight models with anonymous and registered free tiers for developers.
models: Qwen, Mistral, Llama +1
up · 86ms$1.00estconnectLow risk - iFlyTek Sparkxfyun.cnOfficial free tier
iFlytek Spark (Xunfei Xinghuo) is a large model platform providing access to the Spark LLM family. It offers a permanently free Spark Lite version and significant one-time token allowances for newer users. The API is OpenAI-compatible and supports various modalities including text, image, and voice.
models: Spark Lite (permanently free), Spark Max (limited/campaign free), Spark Pro/other versions (limited free tokens)
up · 2034ms$1.00estconnectLow risk - AionLabsapi.aionlabs.aiFree product
Aion Labs provides specialized fine-tuned models for immersive roleplay and storytelling. Its API is OpenAI-compatible and features a long-term daily free allowance for developers, with higher tiers available upon account top-up.
models: aion-2.5, aion-1.0, aion-1.0-mini +2
up · 190ms$0.50estconnectLow risk - Novita AInovita.aiOfficial free tier
Novita AI is a developer-focused inference cloud offering an OpenAI-compatible API to models like DeepSeek, Qwen, and Llama. New users receive a $0.50 trial credit; service is otherwise pay-as-you-go with very low latency.
models: DeepSeek family, Qwen family, Llama family
up · 341ms$0.50estconnectLow risk - Scalewayconsole.scaleway.comOfficial free tier
Scaleway Generative APIs offer a serverless, OpenAI-compatible endpoint for deploying open LLMs. Hosted in Europe, it emphasizes data sovereignty and grants 1,000,000 free tokens to all customers to jumpstart development.
models: Mistral family, Qwen family, Pixtral 12B (2409) +3
up$0.22estconnectLow risk - Poepoe.comSubscription router
Poe by Quora is a major AI aggregator and bot-building platform. It provides an OpenAI-compatible API that allows developers to access hundreds of models (GPT, Claude, Gemini, Llama) using a unified credits-based billing system. Poe also supports custom server bots via the Poe Protocol.
up · 81ms$0.18estconnectLow risk - Cloudflare Workers AIdevelopers.cloudflare.comOfficial free tier
Cloudflare Workers AI allows running AI models on Cloudflare's global network. It provides an OpenAI-compatible API for inference, embeddings, and image generation, billed in compute 'Neurons'. Details were checked against the provider's own page at developers.cloudflare.com.
models: Gemma family, Llama family, Qwen family +3
up$0.11connectLow risk - Hugging Facehuggingface.coOfficial free tier
Hugging Face is the leading hub for open-source AI, offering model hosting, datasets, and serverless inference. It provides a permanent free tier for basic CPU spaces and community GPU grants, with professional-grade Inference Endpoints available on a pay-as-you-go basis.
models: Llama family, Qwen family, Gemma family +2
up · 17ms$0.10connectLow risk - Nomicnomic.aiOfficial free tier
Nomic AI provides the Atlas Embedding API and open-weight models like nomic-embed. It offers a monthly free tier for its embedding API and is transitioning to focus on document AI for specialized industries.
models: nomic-embed-text-v1.5, nomic-embed-text, nomic-embed-vision
up · 305ms$0.10estconnectLow risk - All API Huball-api-hub.qixing1217.topMonitor / Directory
All API Hub is an open-source browser extension and account manager for the New-API/Sub2API ecosystem. It provides a unified dashboard for tracking balances, usage, and auto-checking across multiple relay providers, rather than serving as an API endpoint itself.
up · 52msfree tier · not quantifiedLow risk - Apishopapishop.orgCommercial aggregator
Apishop is a professional AI API gateway designed for developers in mainland China. It provides direct, low-latency access to Anthropic, OpenAI, and Gemini models with zero code changes required. Billing is transparently mapped 1:1 to official USD rates but paid in local currency.
up · 751msfree tier · not quantifiedconnectLow risk - AssemblyAIassemblyai.comSearch / Audio API
AssemblyAI is a leading speech-to-text API provider offering high-accuracy transcription (Universal-3.5 Pro) and a multi-model LLM Gateway. It supports OpenAI-compatible routing for its language model features. Access is pay-as-you-go after a free trial.
unknownfree tier · not quantifiedconnectLow risk - Azure AI Foundrylearn.microsoft.comPaid API
Azure AI Foundry (formerly Azure AI Studio) provides access to a vast catalog of models including OpenAI, Anthropic, and Llama. It features a unified API surface and enterprise-grade management. Pricing is unified under Azure's pay-as-you-go model.
models: Claude 3 Haiku, DeepSeek Chat, Phi 4 +1
unknownfree tier unknownconnectLow risk - Azure OpenAIazure.microsoft.comPaid API
Azure OpenAI Service provides REST API access to OpenAI's powerful language models including the GPT-4 and GPT-5 series with Azure's enterprise capabilities. It uses a strictly pay-as-you-go or provisioned throughput model.
models: gpt-5.6-luna, GPT-4o-mini, GPT-4o +2
unknownfree tier unknownconnectLow risk - Black Forest Labsapi.bfl.aiOfficial free tier
Black Forest Labs (api.bfl.ai) provides official API access to the FLUX diffusion model family. It is a credit-based image generation API where users purchase balance to generate high-fidelity images using FLUX.2 and FLUX.3 models.
up · 290msno free tierLow risk - Black Forest Labs (ML)api.bfl.mlOfficial free tier
BFL ML (api.bfl.ml) is the technical endpoint for Black Forest Labs' FLUX model API. It shares the same credit-based billing as the main bfl.ai domain, providing high-performance access to diffusion models for production workloads.
downno free tierLow risk - Brave Searchbrave.comSearch / Audio API
Brave Search API offers privacy-preserving web search results for AI grounding. It features a $5 monthly free credit and is priced per 1,000 queries plus token usage for the LLM-powered context features.
unknownfree tier · not quantifiedLow risk - Cartesiacartesia.aiSearch / Audio API
Cartesia offers ultra-fast text-to-speech (Sonic) and speech-to-text (Ink) APIs. It uses a credit-based subscription model starting at $5/mo, with a free tier providing 20,000 credits monthly for development.
unknownfree tier · not quantifiedLow risk - Codex Cloudopenai.comOfficial free tier
OpenAI Codex is the specialized model family for coding tasks, now integrated into the flagship GPT models. It provides a specialized developer tier for high-performance code generation, planning, and task automation.
models: gpt-5.6-sol
unknownfree tier · not quantifiedconnectLow risk - DataRobotdocs.datarobot.comEnterprise
DataRobot is an enterprise AI platform that provides centralized access to a wide range of LLMs through its AI Gateway. It emphasizes model governance, security, and unified billing for large-scale enterprise AI deployments.
models: Google: Gemma 3 9B, Meta-Llama-3.1-70B-Instruct, claude-sonnet-5
unknownfree tier unknownLow risk - Databrickswww.databricks.comPaid API
Databricks Foundation Model Serving provides enterprise-grade access to open foundation models like Llama 3.3 and DBRX. It is billed per token using Databricks Units (DBUs), integrated into the broader Databricks Data Intelligence Platform.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Mistral: Mixtral 8x7B Instruct +1
unknownfree tier unknownconnectLow risk - DigitalOceandocs.digitalocean.comPaid API
DigitalOcean Inference provides a unified control plane for AI model inference. It offers serverless access to foundation models (Anthropic, OpenAI, DeepSeek, Kimi) at provider-aligned rates, plus dedicated GPU deployments.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Mistral: Mixtral 8x7B Instruct
unknownfree tier unknownconnectLow risk - Doubaodoubao.comPaid API
Doubao (by Bytedance/Volcengine) provides a suite of models including the high-performance Doubao-Pro and Lite series. It is a major Chinese cloud provider with native API endpoints. Details were checked against the provider's own page at www.volcengine.com.
models: ByteDance: Doubao Lite 32k, ByteDance: Doubao Pro 32k
unknownfree tier unknownLow risk - Factoryfactory.aiEnterprise
Factory provides the 'Autonomy Stack' for enterprise engineering teams, featuring autonomous 'Droids' that code, review PRs, and manage wikis. It offers model routing across 200+ models. Details were checked against the provider's own page at factory.ai.
models: minimax-m2.7, [次]kimi-k2.5, GLM-5.1
unknownfree tier unknownconnectLow risk - Fal.aifal.aiPaid API
Fal.ai provides serverless inference and GPU compute for media generation models (Wan, Kling, Flux). It offers output-based pricing for image/video and time-based pricing for dedicated GPU fleets. Details were checked against the provider's own page at fal.ai.
unknownfree tier · not quantifiedLow risk - Featherless AIfeatherless.aiPaid API
Featherless AI provides serverless access to LLMs and agent runtimes. It offers subscription-based 'Chat' plans with unlimited tokens and credit-based 'Developer' plans for API usage. Details were checked against the provider's own page at featherless.ai.
models: Ling 3.0 Flash, Laguna S 2.1, Qwen3.8 27B +3
unknownfree tier · not quantifiedconnectLow risk - FreeInferencefreeinference.orgOfficial free tier
FreeInference.org is a research-focused AI API gateway built at Harvard SEAS MadSys Lab. It provides free OpenAI-compatible access to frontier open models for the research and education community. No credit card is required for access.
models: Qwen3 Coder, MiniMax-M3, DeepSeek V4 Flash +1
unknownfree tier · not quantifiedconnectLow risk - FriendliAIfriendli.aiOfficial free tier
FriendliAI is a high-performance AI inference platform specializing in serving frontier open-weight models like GLM, DeepSeek, and Llama with industry-leading speed. It offers serverless Model APIs, dedicated GPU endpoints, and on-premise containers.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Mistral Nemo +1
unknownfree tier · not quantifiedconnectLow risk - GLM Coding (China)open.bigmodel.cnOfficial free tier
BigModel.cn (智谱AI) is the official platform for GLM series models, offering high-performance text and multimodal (image/video) inference. It features flagship models like GLM-5.3 and provides a permanent free tier for Flash models along with significant token grants for new users.
models: glm-5.3-flash, glm-5.3, GLM 5 Turbo
unknownfree tier · not quantifiedconnectLow risk - GitLab Duo PATdocs.gitlab.comPaid API
GitLab Duo is a suite of AI-powered features integrated into the GitLab DevSecOps platform. It uses models like Claude and Gemini to provide code suggestions and chat. API access is powered by 'GitLab Credits', which can be purchased or included in premium plans.
models: claude-sonnet-5, gemini-3.1-pro-preview
unknownfree tier unknownLow risk - Google Julesjules.googlePaid API
Google Jules is an AI-driven cloud coding agent platform. It allows developers to automate complex coding tasks and manage cloud resources via a dedicated API. It is part of Google's broader AI ecosystem for developers.
models: Gemini 2.5 Pro, gemini-3.1-pro-preview
unknownfree tier unknownLow risk - Heroku AIwww.heroku.comPaid API
Heroku AI is a suite of managed AI services on the Salesforce Heroku platform. It enables developers to integrate and scale AI features using vector databases and third-party AI add-ons. It focuses on simplifying the transition from prototype to production for AI applications.
models: Nova Lite 1.0, MiniMax M2, Qwen: Qwen3 72B
unknownfree tier unknownLow risk - HvoyAIhvoyai.comMonitor / Directory
HvoyAI.com (formerly Hvoy.ai) is the primary domain for the Hvoy AI diagnostics platform. It provides API key testing tools, gateway rankings, and comparative performance data for AI relay providers across the global ecosystem.
up · 433msfree tier · not quantifiedLow risk - IBM watsonx.ai Gatewaywww.ibm.comPaid API
IBM watsonx.ai Gateway provides managed access to foundation models from IBM and third parties. It offers an OpenAI-compatible endpoint and enterprise-grade tools for model governance and scaling. Access is available via pay-as-you-go billing per million tokens.
models: Meta-Llama-3.1-70B-Instruct, Mistral Large
unknownfree tier · not quantifiedconnectLow risk - Jina AI (Foundation API)jina.aiPaid API
Jina AI (Foundation API) provides specialized AI models for search and RAG, including high-quality embeddings, reranking, and a Reader API that converts web content to LLM-friendly markdown. The Reader API is currently free for developers, while other APIs follow token-based usage billing.
unknownfree tier · not quantifiedLow risk - Lambda AIlambda.aiPaid API
Lambda AI (Lambda Labs) is a premier GPU cloud provider offering serverless inference for large open-source models like Llama 3.1 405B. It provides a robust, developer-centric platform for high-performance AI tasks with usage-based billing.
models: Meta: Llama 3.1 405B Instruct, Mistral Large
unknownfree tier unknownconnectLow risk - Liquid AIliquid.aiPaid API
Liquid AI develops Liquid Foundation Models (LFMs), a new class of efficient, hybrid AI models. The platform offers serverless inference for models like LFM-2.5-2.6B, which are free for use by companies with annual revenue under $10 million.
models: LiquidAI: LFM2.5-2.6B (free)
unknownfree tier · not quantifiedLow risk - Magnificwww.magnific.comPaid API
Magnific (formerly Freepik) is a professional AI creative platform offering image, video, and audio generation and upscaling. Its API provides credit-based access to proprietary models like Nano Banana 2 and Seedream 5.0, with subscription tiers ranging from a limited free plan to professional unlimited tiers.
unknownfree tier · not quantifiedconnectLow risk - Meta Llama APIllama.developer.meta.comPaid API
Meta Model API (developer.meta.com) is the official hosted inference service for Llama and Muse models. It provides OpenAI- and Anthropic-compatible endpoints with pay-as-you-go pricing for frontier models like Muse Spark and specialized models for transcription and reasoning.
models: Muse Glimmer 30B, Muse Spark 1.3, Muse Spark 1.2 +3
unknownfree tier unknownconnectLow risk - Mixedbread AIwww.mixedbread.comPaid API
Mixedbread AI is a specialized provider of knowledge-retrieval models and agents. Its platform offers an OpenAI-compatible API for its 'Toast' agent models, alongside managed search and indexing services billed by content tokens. It features a $5 one-time free credit for new users.
unknownfree tier · not quantifiedconnectLow risk - NVIDIA NIMintegrate.api.nvidia.comOfficial free tier
NVIDIA NIM (Build) is NVIDIA's official hosted model-inference service, providing API access to over 100 frontier open-source models (Llama 3.3, Nemotron, DeepSeek). It features a permanent free tier with a 40 RPM limit, authenticated via a standard API key and accessible through an OpenAI-compatible endpoint.
models: Nemotron 3 Nano 30B A3B, Nemotron 3.5 Lightning, Nemotron 3 Super +3
up · 14msfree tier · not quantifiedconnectLow risk - OCI Generative AIwww.oracle.comPaid API
OCI Generative AI is Oracle's enterprise-grade inference service hosting Llama and Cohere models. It provides on-demand and dedicated hosting with native integration into the Oracle Cloud ecosystem. Details were checked against the provider's own page at www.oracle.com.
models: Command R (08-2024), Meta-Llama-3.1-70B-Instruct, Command R+ (08-2024)
unknownfree tier unknownLow risk - OpenAIplatform.openai.comPaid API
OpenAI is the developer of GPT-4o, o1, and other leading LLMs. It offers a usage-based API for developers with tiered processing (Fast/Batch) and extensive multimodal capabilities including audio and video generation.
models: GPT-4o-mini, o3 Mini, o4 Mini +3
unknownfree tier unknownLow risk - Perplexityperplexity.aiOfficial free tier
Perplexity provides the Sonar API for usage-based AI search and reasoning. It offers a limited free tier for end-users and a usage-based API for developers, optimized for real-time grounded answers. Details were checked against the provider's own page at perplexity.ai.
models: Perplexity default free search model (basic Sonar class)
up · 39msfree tier · not quantifiedconnectLow risk - Regolo AIregolo.aiPaid API
Regolo AI provides scalable serverless AI infrastructure with a focus on privacy and zero data retention. It serves popular open-source models (Llama, Mistral, Qwen) via an OpenAI-compatible API, offering monthly token capacity plans and a flexible free trial.
models: Qwen3.8 27B, Gemma 4 31B, gpt-oss-120b +3
unknownfree tier · not quantifiedconnectLow risk - SearchAPIwww.searchapi.ioSearch / Audio API
SearchAPI is a real-time SERP scraping API providing structured data from Google, Bing, Baidu, and more. It is designed for AI agents and developers, offering an OpenAI-compatible interface and a focus on 'pay-per-success' results.
unknownfree tier · not quantifiedconnectLow risk - Snowflake Cortexwww.snowflake.comPaid API
Snowflake Cortex is a managed service within the Snowflake Data Cloud that provides access to LLMs (Llama, Mistral, Gemma) and AI functions. It uses a credit-based system where usage is billed per million tokens processed, independent of the Snowflake edition.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Gemma 2 27B +1
unknownfree tier unknownLow risk - Stability AIstability.aiPaid API
Stability AI provides a developer platform for its generative models, including Stable Diffusion 3.5 and Stable Audio. The API uses a credit-based system where different tasks (image generation, upscaling, editing) consume varying amounts of credits.
unknownfree tier · not quantifiedLow risk - Sunosuno.aiPaid API
Suno is a leading generative AI platform for music and audio. It offers subscription-based access to its music generation models (v3.5, v4) with daily or monthly credit allowances. Credits are used to generate songs, extend tracks, and extract stems.
unknownfree tier · not quantifiedLow risk - Tavily Searchtavily.comSearch / Audio API
Tavily is an AI-optimized search engine designed for LLMs and RAG systems. It provides structured, clean web data and content extraction without the need for manual scraping. It offers an API with a free monthly credit allowance.
unknownfree tier · not quantifiedLow risk - Tencent Hunyuanhunyuan.tencent.comPaid API
Tencent Hunyuan (混元) is a series of large-scale AI models integrated with the Tencent ecosystem (WeChat, Tencent Cloud). It offers strong Chinese language support and multimodal capabilities, including vision and code generation.
models: Hunyuan A13B Instruct, Hy4 preview, Hy3
unknownfree tier · not quantifiedLow risk - Together AIwww.together.aiPaid API
Together AI is a leading cloud provider for open-source AI models, offering serverless inference, provisioned throughput, and dedicated clusters. It hosts over 100 models, including Llama, Qwen, and Mistral, with highly competitive per-token pricing.
models: Meta-Llama-3.1-70B-Instruct, DeepSeek Chat, Qwen2.5 72B Instruct +2
unknownfree tier unknownconnectLow risk - Topaztopazlabs.comPaid API
Topaz Labs specializes in AI-powered image and video enhancement software (Photo AI, Video AI). While primarily a desktop application suite, it offers cloud-based rendering and is associated with AI model APIs for media processing.
unknownfree tier unknownLow risk - Typhoondocs.opentyphoon.aiOfficial free tier
OpenTyphoon is a Thai-focused AI model platform by SCB 10X. It provides free access to models like Typhoon 2.5 via its hosted API at opentyphoon.ai. Production-grade access is offered through partners like Together AI and Float16.
unknownfree tier · not quantifiedLow risk - Udioudio.comPaid API
Udio is an AI music generation platform that allows users to create high-fidelity audio tracks from text descriptions. It offers both a free tier with daily credits and subscription plans with higher quotas and advanced features (Voice Control, stem extraction).
unknownfree tier · not quantifiedLow risk - Volcenginewww.volcengine.comPaid API
Volcengine is ByteDance's enterprise cloud platform. Its Ark (Huoshan Fangzhou) platform provides access to the Doubao (Doubao) model family and other high-performance LLMs. Supports OpenAI-compatible API calls and various billing plans including token-based monthly packages.
models: Qwen-Plus, zai-glm-4.7
unknownfree tier · not quantifiedconnectLow risk - Volcengine Ark Agent Planconsole.volcengine.comPaid API
The Volcengine Ark Agent Plan is a subscription-based billing option for the Ark platform, offering monthly token quotas for Doubao and other models. It is designed for developers requiring predictable monthly costs and higher quotas than standard PAYG.
models: Qwen-Plus, zai-glm-4.7
unknownfree tier unknownLow risk - Weights & Biases Inferencewandb.aiPaid API
Weights & Biases Weave provides an inference router and observability tools for AI development. It integrates with OpenRouter and other providers to offer a unified, OpenAI-compatible endpoint for model testing and production deployments.
models: gpt-oss-120b, gpt-oss:20b, Qwen3 235B A22B Thinking 2507 +2
unknownfree tier · not quantifiedconnectLow risk - Writerdev.writer.comPaid API
Writer is an enterprise generative AI platform featuring its own family of Palmyra models. It provides a consumption-based API for Palmyra X6, X5, and X4 models, optimized for business writing, data security, and precision. It does not natively use OpenAI-compatible request shapes.
models: Palmyra X5
unknownfree tier unknownLow risk - Xiaomi MiMomimo.xiaomi.comOfficial free tier
Xiaomi MiMo (mimo.mi.com) is the official AI open platform for Xiaomi's in-house large models. It offers an OpenAI-compatible API for the MiMo-V2.5 series, featuring a high-concurrency token plan (¥60 for 50B tokens) and competitive per-million token pricing for Pro and Flash variants.
models: MiMo-V2.5, MiMo-V2-Flash, MiMo-V2-Pro (time-limited/quota-limited free) +1
up · 795msfree tier · not quantifiedconnectLow risk - Yi (01.AI)01.aiPaid API
Yi is the model family by 01.AI, offering high-performance LLMs including Yi-Lightning and Yi-Large. The developer platform provides OpenAI-compatible API access with usage-based billing. It is optimized for both Chinese and English language tasks.
unknownfree tier unknownconnectLow risk - iFlyTek Spark Openspark-api-open.xf-yun.comOfficial free tier
spark-api-open.xf-yun.com is the OpenAI-compatible HTTP endpoint for iFlyTek's Spark (讯飞星火) LLM. It provides access to multimodal models via the iFlyTek Open Platform. A permanently free 'Lite' version is available for developers.
models: Spark Lite (model=lite), Spark Pro/Max/4.0 Ultra (one-off free token packs only; paid after depletion)
up · 822msfree tier · not quantifiedconnectLow risk
free-credit values are normalized USD estimates from public sources — see methodology · ordering is data (free-credit value), never sponsorship
What is an AI API router?
An AI API router is a service that exposes AI model APIs (often OpenAI- or Claude-compatible endpoints) so you can call many models — from different vendors or free and paid tiers — through one interface, with a single key and one billing surface.
How is the free-credit value calculated?
Free-credit values are normalized estimates in USD from public sources (signup bonuses, monthly quotas and usage allowances). Estimates are flagged on each entry; values we could not quantify are shown honestly as "Unknown" or "Not quantified" instead of invented numbers.
Is the ranking sponsored?
No. Every list on UPROUTER.ONLINE is ordered by data (free-credit value, price, risk rating or community score). There is no pay-to-rank — affiliate and referral links never influence data, scores or sorting.
How do I know a router is safe to use?
Each entry carries a risk rating (low / medium / high / unrated) with the reasoning behind it, a data-confidence score and live status. Risk is an evidence rating from community research — treat it as a starting point and always review the provider terms yourself.