Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

@RequestyAI
requesty.ai > models > fireworks > glm-5.3

Fireworks AI glm-5.3 API Pricing & Cost: Context Window & Benchmarks

2+ week, 22+ hour ago   (415+ words) Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay against these rates, not…...

Requesty
requesty.ai > blog > open-weight-frontier-august-2026-glm-qwen-hy4

Five open weight releases in nine days: GLM-5.3-Flash, Qwen3.8-Flash, Hy4 and the collapse of the capability premium

2+ week, 3+ day ago   (846+ words) Model launch chatter in our social listening corpus went from 506 mentions in the week of 15 to 21 August to 978 in the week of 22 to 28 August. That is 1.93x, and it is not a scraping artifact. Five labs shipped open weight models with…...

@RequestyAI
requesty.ai > models > fireworks > nemotron-lightning-3.5-30b-a3b

Fireworks AI nemotron-lightning-3.5-30b-a3b API Pricing & Cost: Context Window & Benchmarks

3+ week, 2+ day ago   (344+ words) Fireworks AI/🇺🇸 US/chat10% off Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay…...

@RequestyAI
requesty.ai > models > novita > inclusionai-ling-3.0-tiny

Novita AI inclusionai/ling-3.0-tiny API Pricing & Cost: Context Window & Benchmarks

3+ week, 3+ day ago   (432+ words) Novita AI/🇺🇸 US/chat10% off Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay…...

@RequestyAI
requesty.ai > model > meta > llama-3.1-8b-instruct

llama-3.1-8b-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (268+ words) Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists, and every row links to that provider's endpoint page. This model…...

@RequestyAI
requesty.ai > model > meta > meta-llama-3.1-8b-instruct-turbo

meta-llama-3.1-8b-instruct-turbo: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (286+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...

@RequestyAI
requesty.ai > model > meta > meta-llama-3.1-405b-instruct

meta-llama-3.1-405b-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (286+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...

@RequestyAI
requesty.ai > model > perplexity > sonar

sonar: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (271+ words) Lightweight offering with search grounding, quicker and cheaper than Sonar Pro. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists,…...

@RequestyAI
requesty.ai > model > poolside > laguna-xs.2

laguna-xs.2: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (301+ words) Poolside Laguna XS.2 — a small, fast code-focused LLM from Poolside, served via an OpenAI-compatible API. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where…...

@RequestyAI
requesty.ai > model > meta > llama-4-maverick-17b-128e-instruct

llama-4-maverick-17b-128e-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 3+ day ago   (294+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...