PhaseoPhaseo
PhaseoPhaseo
Checking statusChecking statusVisit status page
Component-level status is unavailable.

Explore

  • Models
  • Chat
  • Compare
  • Providers
  • Apps
  • Rankings
  • Monitor

Build

  • Documentation
  • API Reference
  • Quickstart
  • SDKs
  • Methodology
  • Gateway Comparisons

Company

  • About
  • Mission
  • Blog
  • Pricing
  • Works With
  • Acknowledgements
  • Support
  • Privacy
  • Terms

Community

  • Discord
  • GitHub
  • LinkedIn
  • Reddit
  • X

© 2025 • Phaseo

Report:Issue·Support

Spotted a data issue or broken page?Open an issueorcontact support

PhaseoPhaseo
ModelsChatCompareProvidersAppsRankings
ModelsChatCompareProvidersAppsRankings
Sign Up
Compare
Nvidia
Nemotron 3.5 Lightning 30B A3B
Add modelChat
Nvidia
Nemotron 3.5 Lightning 30B A3B

Overview

Input modalities
T
Output modalities
T
Providers
5 providers
Input context
1,000,000
Max output
-
Release
Aug 2026
Capabilities
ReasoningWebFine-tune

Pricing

Provider
DeepInfra
DeepInfra
Input
$0.05 / M tokens
Output
$0.20 / M tokens
Cached input
- / M tokens
Plan
standard
Source
Pricing source

Performance

Latency (p50)
-
Throughput (p50)
-
Provider latency
-
Provider throughput
-
Visualize performance
View charts

Activity

30d tokens
0
Total requests
0
Requests in 30m
0

Benchmarks

Shared wins
0
Comparable tests
0
Total results
14
Benchmark charts
View detail

Simulate a response

Estimated input
21 tokens
Context fit
Fits
Estimated cost
$0.0002
Est. response time
-
Pricing basis
$0.05 in / $0.20 out

Overview

Input/output modalities and key model metadata from the catalog.

Nvidia
Nemotron 3.5 Lightning 30B A3B
Nvidia
Active4 priced providers
Input Modalities
Text
Output Modalities
Text
ReleaseAug 2026
Knowledge Cutoff-
Context1,000,000
Max Output-
LicenseOpenMDW License Agreement 1.1

Gateway Usage

30-day activity plus recent runtime. Text-first models use token volume; other modalities fallback to request activity.

Last 30d
Nvidia
Nemotron 3.5 Lightning 30B A3B
Nvidia
0
tokens · last 30 days
Token data up to 12 Aug 2026
Requests
0
Latency
-
Throughput
-
Request activity · 24h0 in 30m

Benchmarks Comparison

Only benchmarks with comparable results across every selected model are shown.

Benchmark Scores (%)

Switch benchmark type to compare percent and numerical families separately.

Nvidia
Nemotron 3.5 Lightning 30B A3B
AA-LCR
%Lower is better
Nvidia
Nemotron 3.5 Lightning 30B A3B
52%
BrowseComp
%Lower is better
Nvidia
Nemotron 3.5 Lightning 30B A3B
36.97%
GPQA Diamond
%Lower is better
Nvidia
Nemotron 3.5 Lightning 30B A3B
75.44%
IFBench
%Lower is better
Nvidia
Nemotron 3.5 Lightning 30B A3B
71.88%

Pricing

Per-1M normalized pricing from observed provider tiers. Blended total uses 90% input + 10% output.

Nvidia
Nemotron 3.5 Lightning 30B A3B
Input $/M
$0.05
DeepInfra
Output $/M
$0.20
DeepInfra
Blended $/M
$0.07
90/10 input-output
Pricing
Blended (90/10)Input $/MOutput $/M
Pricing by meter

All unique meters observed across the selected models.

Meter
Nemotron 3.5 Lightning 30B A3B
Best option
Input Text Tokens$0.05
Output Text Tokens$0.20
Cached Read Text Tokens$0.01

Availability

API provider availability and subscription plans.

API Availability

Nvidia
Nemotron 3.5 Lightning 30B A3BNvidia
Providers
DeepInfra
Fireworks
Nebius Token FactoryNebius Token Factory
Weights & Biases

Subscription Plans