Sağlayıcı APIleri
- Google Gemini Gemini 2.5 Pro, Flash, Flash-Lite +4 more. 10 RPM, 20 RPD
- Cohere Command A, Command R+, Aya Expanse 32B +9 more. 20 RPM, 1K req/mo
- Mistral AI Mistral Large 3, Small 3.1, Ministral 8B +3 more. 1 req/s, 1B tok/mo
- Zhipu AI GLM-4.7-Flash, GLM-4.5-Flash, GLM-4.6V-Flash. Limitler belirtilmemiş
(Inference) Sağlayıcıları
- GitHub Models GPT-4o, Llama 3.3 70B, DeepSeek-R1 +daha fazlası. 1015 RPM, 50150 RPD
- NVIDIA NIM Llama 3.3 70B, Mistral Large, Qwen3 235B +daha fazlası . 40 RPM
- Groq Llama 3.3 70B, Llama 4 Scout, Kimi K2 +17 daha fazlası . 30 RPM, 14,400 RPD
- Cerebras Llama 3.3 70B, Qwen3 235B, GPT-OSS-120B +3 daha fazlası . 30 RPM, 14,400 RPD
- Cloudflare Workers AI Llama 3.3 70B, Qwen QwQ 32B +47 daha fazlası . 10K neurons / günlük
- LLM7.io DeepSeek R1, Flash-Lite, Qwen2.5 Coder +27 daha fazlası . 30 RPM (120 with token)
- Kluster AI DeepSeek-R1, Llama 4 Maverick, Qwen3-235B +2 daha fazlası . Limits undocumented
- OpenRouter DeepSeek R1, Llama 3.3 70B, GPT-OSS-120B +29 daha fazlası . 20 RPM, 50 RPD
- Hugging Face Llama 3.3 70B, Qwen2.5 72B, Mistral 7B +daha fazlası . $0.10 aylık ücretsiz kredi veriyor.
RPM = dakika başına istek · RPD = günlük istek sayısı. Tüm endpointler OpenAI SDK ile uyumludur.
