PhaseoPhaseo
PhaseoPhaseo
Checking statusChecking statusVisit status page
Component-level status is unavailable.

Explore

  • Models
  • Chat
  • Compare
  • Providers
  • Apps
  • Rankings
  • Monitor

Build

  • Documentation
  • API Reference
  • Quickstart
  • SDKs
  • Methodology

Company

  • About
  • Mission
  • Blog
  • Pricing
  • Works With
  • Support
  • Privacy
  • Terms

Community

  • Discord
  • GitHub
  • LinkedIn
  • Reddit
  • X

© 2025 • Phaseo

Report:Issue·Support

Spotted a data issue or broken page?Open an issueorcontact support

PhaseoPhaseo
ModelsChatCompareProvidersAppsRankings
ModelsChatCompareProvidersAppsRankings
Sign Up
Add modelChat
Thinking Machines Lab
Inkling

Overview

Input modalities
TImgAud
Output modalities
T
Providers
4 providers
Input context
1,000,000
Max output
-
Release
Jul 2026
Capabilities
ReasoningWebFine-tune

Pricing

Provider
Baseten
Baseten
Input
$1.00 / M tokens
Output
$4.05 / M tokens
Cached input
$0.17 / M tokens
Plan
standard
Source
Pricing source

Performance

Latency (p50)
-
Throughput (p50)
-
Provider latency
-
Provider throughput
-
Visualize performance
View charts

Activity

30d tokens
11.8K
Total requests
0
Requests in 30m
0

Benchmarks

Shared wins
0
Comparable tests
0
Total results
0
Benchmark charts
View detail

Simulate a response

Estimated input
21 tokens
Context fit
Fits
Estimated cost
$0.0033
Est. response time
-
Pricing basis
$1.00 in / $4.05 out

Overview

Input/output modalities and key model metadata from the catalog.

Thinking Machines Lab
Inkling
Thinking Machines Lab
Active4 priced providers
Input Modalities
TextImageAudio
Output Modalities
Text
ReleaseJul 2026
Knowledge CutoffApr 2026
Context1,000,000
Max Output-
LicenseApache 2.0

Gateway Usage

30-day activity plus recent runtime. Text-first models use token volume; other modalities fallback to request activity.

Last 30d
Thinking Machines Lab
Inkling
Thinking Machines Lab
11.83K
tokens · last 30 days
Token data up to 02 Aug 2026
Requests
0
Latency
-
Throughput
-
Request activity · 24h0 in 30m

Pricing

Per-1M normalized pricing from observed provider tiers. Blended total uses 90% input + 10% output.

Thinking Machines Lab
Inkling
Input $/M
$1.00
Baseten
Output $/M
$4.05
Baseten
Blended $/M
$1.31
90/10 input-output
Pricing
Blended (90/10)Input $/MOutput $/M
Pricing by meter

All unique meters observed across the selected models.

Meter
Inkling
Best option
Input Text Tokens$1.00
Output Text Tokens$4.05
Cached Read Text Tokens$0.17

Availability

API provider availability and subscription plans.

API Availability

Thinking Machines Lab
InklingThinking Machines Lab
Providers
Baseten
DeepInfra
Thinking Machines
Together

Subscription Plans