Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

@RequestyAI
requesty.ai > models > fireworks > glm-5.3

Fireworks AI glm-5.3 API Pricing & Cost: Context Window & Benchmarks

1+ week, 6+ day ago   (415+ words) Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay against these rates, not…...

@RequestyAI
requesty.ai > models > fireworks > nemotron-lightning-3.5-30b-a3b

Fireworks AI nemotron-lightning-3.5-30b-a3b API Pricing & Cost: Context Window & Benchmarks

3+ week, 1+ day ago   (344+ words) Fireworks AI/🇺🇸 US/chat10% off Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay…...

@RequestyAI
requesty.ai > models > novita > inclusionai-ling-3.0-tiny

Novita AI inclusionai/ling-3.0-tiny API Pricing & Cost: Context Window & Benchmarks

3+ week, 1+ day ago   (432+ words) Novita AI/🇺🇸 US/chat10% off Which id to call These are the upstream provider rates. Pay as you go adds 5%, or 0% if you bring your own keys, and there is no per-request fee. Prompt caching and routing change what you pay…...

@RequestyAI
requesty.ai > model > meta > llama-3.1-8b-instruct

llama-3.1-8b-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (268+ words) Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists, and every row links to that provider's endpoint page. This model…...

@RequestyAI
requesty.ai > model > meta > meta-llama-3.1-8b-instruct-turbo

meta-llama-3.1-8b-instruct-turbo: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (286+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...

@RequestyAI
requesty.ai > model > meta > meta-llama-3.1-405b-instruct

meta-llama-3.1-405b-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (286+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...

@RequestyAI
requesty.ai > model > perplexity > sonar

sonar: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (271+ words) Lightweight offering with search grounding, quicker and cheaper than Sonar Pro. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists,…...

@RequestyAI
requesty.ai > model > poolside > laguna-xs.2

laguna-xs.2: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (301+ words) Poolside Laguna XS.2 — a small, fast code-focused LLM from Poolside, served via an OpenAI-compatible API. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where…...

@RequestyAI
requesty.ai > model > meta > llama-4-maverick-17b-128e-instruct

llama-4-maverick-17b-128e-instruct: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (294+ words) A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most. Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank…...

@RequestyAI
requesty.ai > model > thinkingmachines > inkling

inkling: Compare 1 Provider, API Pricing & Performance

3+ week, 2+ day ago   (325+ words) Which id to call 1 endpoint / 1 region Provider prices, per 1M tokens. Pay as you go adds 5%, or 0% on your own keys. A column is blank where no qualifying sample exists, and every row links to that provider's endpoint page. Requesty routes…...