inferenstack — the market read on AI inference · last observation 2026-09-20. See the boards
Price spread

modelsdev/nvidianemotron3nano30ba3b

modelsdev· 128K ctx
open weightstool callstructured quant: unknown

spread (output)

1.3×

$0.240 – $0.300 · 2 providers

2 providers priced
ProviderInputOutputCache rdBlended
Cortecs
relay router
$0.060$0.240—$0.076*
Venice AI
unknown
$0.075$0.300—$0.095*

Data: models.dev (MIT), snapshot 2026-09-20. Prices are vendor-published; verify before purchase. Methodology.