Skip to content

docs: rebuild the free-model lineup after NVIDIA's 2026-08-30 sweep - #68

Merged
VickyXAI merged 4 commits into
mainfrom
docs/free-tier-2026-08-30
Aug 31, 2026
Merged

docs: rebuild the free-model lineup after NVIDIA's 2026-08-30 sweep#68
VickyXAI merged 4 commits into
mainfrom
docs/free-tier-2026-08-30

Conversation

@VickyXAI

Copy link
Copy Markdown
Contributor

Docs half of BlockRunAI/blockrun#447.

A live two-pass --real probe found four of the five VISIBLE free models gone at once — step-3.7-flash, nemotron-nano-9b-v2 and nemotron-nano-12b-v2-vl return a published 410, and mistral-nemotron is the quiet kind: still listed by NVIDIA, but a completion never comes back (>150s zero bytes, both passes). nemotron-super-49b (hidden, but the free cascade's tertiary rung) is 410 too.

Every reference to those five is repointed at a probe-verified replacement:

  • nvidia/nemotron-3.5-lightning — thinking-mode reasoning, 131K, ~35 tok/s
  • nvidia/nemotron-3-nano-30b — ~121 tok/s, the fastest free model in the catalog
  • nvidia/llama-3.2-11b-vision — the only Llama NVIDIA still serves; 128K, image input

Counts move with the catalog: 71 → 70 chat, 95 → 94 total, 5 → 4 free. Dated changelog entries keep their original numbers.

1bcMax added 4 commits August 30, 2026 21:39
BlockRun's 402 responses now spread x402Version/accepts at the top level
of the JSON body, not just the signed headers — v1-era x402 clients that
only parse the body were silently failing to auto-pay. Updates every
endpoint doc's 402 example plus the x402 protocol pages to match.
A live two-pass --real probe found four of the five VISIBLE free models
gone at once: step-3.7-flash, nemotron-nano-9b-v2 and nemotron-nano-12b-v2-vl
return a published 410, and mistral-nemotron is the quiet kind — still listed
by NVIDIA, but a completion never comes back (>150s zero bytes, both passes).
nemotron-super-49b (hidden, but the free cascade's tertiary rung) is 410 too.

Every reference to those five is repointed at a probe-verified replacement:
nemotron-3.5-lightning, nemotron-3-nano-30b, and llama-3.2-11b-vision — the
only Llama NVIDIA still serves.

Counts move with the catalog: 71 -> 70 chat, 95 -> 94 total, 5 -> 4 free.
Two free models (nemotron-3.5-lightning, nemotron-3-nano-omni) now serve from
OpenRouter's $0 ":free" pool with the direct-NVIDIA path as fallback: same
models, larger pool, 4.9s median against 16.3s, and Lightning gains a 1M
context. That pool also made three more models listable — nemotron-3-ultra-550b
(550B/55B, 1M ctx, unreachable on our own key), cohere/north-mini-code and
poolside/laguna-xs-2.1.

Counts: 70 -> 73 chat, 94 -> 97 total, 4 -> 7 free.
@VickyXAI
VickyXAI merged commit 6b4c84f into main Aug 31, 2026
1 check passed
@VickyXAI
VickyXAI deleted the docs/free-tier-2026-08-30 branch August 31, 2026 03:44
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant