PhaseoPhaseo
PhaseoPhaseo
Checking statusChecking statusVisit status page
Component-level status is unavailable.

Explore

  • Models
  • Chat
  • Providers
  • Apps
  • Rankings
  • Tools
  • Monitor

Resources

  • Compare
  • Migration Guides
  • Methodology
  • Blog

Community

  • Discord
  • GitHub
  • LinkedIn
  • Reddit
  • X

Build

  • Documentation
  • API Reference
  • Quickstart
  • SDKs

Company

  • About
  • Trust Centre
  • Mission
  • Pricing
  • Works With
  • Acknowledgements
  • Support
  • Privacy
  • Terms

Explore

  • Models
  • Chat
  • Providers
  • Apps
  • Rankings
  • Tools
  • Monitor

Build

  • Documentation
  • API Reference
  • Quickstart
  • SDKs

Resources

  • Compare
  • Migration Guides
  • Methodology
  • Blog

Company

  • About
  • Trust Centre
  • Mission
  • Pricing
  • Works With
  • Acknowledgements
  • Support
  • Privacy
  • Terms

Community

  • Discord
  • GitHub
  • LinkedIn
  • Reddit
  • X

© 2025 • Phaseo

Report:Issue·Support

Spotted a data issue or broken page?Open an issueorcontact support

PhaseoPhaseo
ModelsChatCompareProvidersAppsRankings
ModelsChatCompareProvidersAppsRankings
Sign Up
Chat
Thinking Machines Lab
Inkling Small

Overview

Input modalities
TextImageAudio
Output modalities
Text
Providers
6 providers
Input context
1,000,000
Max output
-
Release
Jul 2026
Capabilities
ReasoningWebFine-tune

Pricing

Provider
OpenRouterOpenRouter
OpenRouter
Input
$0.00 / M tokens
Output
$0.00 / M tokens
Cached input
$0.10 / M tokens
Plan
standard

Performance

Latency (p50)
-
Throughput (p50)
-
Provider latency
-
Provider throughput
-
Visualize performance
View charts

Activity

30d tokens
0
Total requests
0
Requests in 30m
0

Benchmarks

Shared wins
0
Comparable tests
0
Total results
1
Benchmark charts
View detail

Simulate a response

Estimated input
21 tokens
Context fit
Fits
Estimated cost
$0.00
Est. response time
-
Pricing basis
$0.00 in / $0.00 out

Overview

Input/output modalities and key model metadata from the catalog.

Thinking Machines Lab
Inkling Small
Thinking Machines Lab
Active6 priced providers
Input Modalities
TextImageAudio
Output Modalities
Text
ReleaseJul 2026
Knowledge Cutoff-
Context1,000,000
Max Output-
LicenseApache 2.0

Gateway Usage

30-day activity plus recent runtime. Text-first models use token volume; other modalities fallback to request activity.

Last 30d
Thinking Machines Lab
Inkling Small
Thinking Machines Lab
0
tokens · last 30 days
Token data up to 30 Aug 2026
Requests
0
Latency
-
Throughput
-
Request activity · 24h0 in 30m
No activity points

Benchmarks Comparison

Only benchmarks with comparable results across every selected model are shown.

Benchmark Scores (Numerical)

Switch benchmark type to compare percent and numerical families separately.

Thinking Machines Lab
Inkling Small
Artificial Analysis Intelligence Index v4.1.1
NumericalLower is better
Thinking Machines Lab
Inkling Small
41.2

Pricing

Per-1M normalized pricing from observed provider tiers. Blended total uses 90% input + 10% output.

Thinking Machines Lab
Inkling Small
Input $/M
$0.00
Baseten
Output $/M
$0.00
OpenRouter
Blended $/M
$0.00
90/10 input-output
Pricing
Blended (90/10)Input $/MOutput $/M
Pricing by meter

All unique meters observed across the selected models.

Meter
Inkling Small
Best option
Input Text Tokens$0.00
Output Text Tokens$0.00
Cached Read Text Tokens$0.06
Cached Write Text Tokens$0.50
Input Image Tokens$0.00

Availability

API provider availability and subscription plans.

API Availability

Thinking Machines Lab
Inkling SmallThinking Machines Lab
Providers
Arcee AI
Baseten
OpenRouterOpenRouter
Thinking Machines
Together

Subscription Plans