Please confirm you are human
This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.
A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.
News
sonar: Compare 1 Provider, API Pricing & Performance
8+ hour, 46+ min ago (271+ words) Lightweight offering with search grounding, quicker and cheaper than Sonar Pro. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists,…...
glm-4.5: Compare 2 Providers, API Pricing & Performance
8+ hour, 45+ min ago (276+ words) Which id to call 2 endpoints / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists, and every row links to that provider's endpoint page. This model…...
DeepInfra Inc. glm-5.2:flex API Pricing & Cost: Context Window & Benchmarks
2+ day, 20+ hour ago (109+ words) Artificial Analysis Intelligence Index — a composite of multiple evaluations measuring overall model capability. Scores are sourced from official model cards, Artificial Analysis, and public leaderboards. Benchmarks measure specific skills and do not capture every aspect of model quality. Always test…...
Nebius AI kimi-k3 API Pricing & Cost: Context Window & Benchmarks
3+ week, 3+ day ago (140+ words) Kimi K3 is Moonshot AI's flagship reasoning model with a 1M token context window, strong agentic tool use, and long horizon task execution. Hosted on Nebius. Artificial Analysis Intelligence Index — a composite of multiple evaluations measuring overall model capability. Scores are sourced…...
sference deepseek-v4-flash API Pricing & Cost: Context Window & Benchmarks
4+ week, 2+ day ago (155+ words) DeepSeek V4 Flash is an efficiency-focused MoE model with 284B total parameters (13B active) and a 1M-token context window. It's tuned for fast inference and high-throughput use cases while still holding up on reasoning and coding tasks. Artificial Analysis Intelligence Index — a composite…...
Fireworks AI glm-5.2-fast API Pricing & Cost: Context Window & Benchmarks
1+ mon, 6+ day ago (104+ words) Benchmarks haven't been published yet for this exact variant. Some variants (region-specific deployments, highspeed tiers) share benchmarks with their base model. Check the base model page or the Fireworks AI models overview. Requesty charges exactly what the upstream provider charges,…...
sference glm-5.2 API Pricing & Cost: Context Window & Benchmarks
1+ mon, 2+ week ago (155+ words) GLM-5.2 is the latest model in the GLM series from Z.ai, continuing its focus on coding and long horizon agentic work. Building on GLM-5.1, it plans, executes, and iterates autonomously on extended engineering grade tasks. Served via Sference. Artificial Analysis…...
Google LLC (Vertex AI) claude-sonnet-5 API Pricing & Cost: Context Window & Benchmarks
1+ mon, 2+ week ago (158+ words) Claude Sonnet 5 is the latest model in the Sonnet family and an upgrade to Sonnet 4.6, with gains in agentic coding and professional work. It delivers top-tier intelligence at Sonnet pricing, well suited to production agents, high-volume pipelines, and everyday professional…...
Free AI Models & Free LLM API: Run LLMs Free for Coding
1+ mon, 3+ week ago (222+ words) A free AI API with 200 requests per day: run free LLMs for coding, chat and agents. No credit card, OpenAI-compatible, and it works in Claude Code, Cline and Cursor. Routing, caching and EU data residency are built in. Outgrow free?...
Free AI Models: Run LLMs at $0 via One API
2+ mon, 5+ day ago (104+ words) Every free AI model on Requesty, available at zero cost per token. Open-weight LLMs from Google Gemma, NVIDIA Nemotron, DeepSeek, Meta Llama and more, with 200 free requests per day through one OpenAI-compatible API. Compare context windows, capabilities and specs. 8 free…...