Mercatus AI pulled all 17 third-party hosts serving the same V4 weights three weeks later and ran the workload math, finding sixteen of seventeen hosts now undercut DeepSeek's own API.
Public source
Publisher name
Public post
New on Mercatus: every DeepSeek model, every host, one page. DeepSeek raised its API prices on August 16. We pulled all 17 third-party hosts serving the same V4 weights…
Company
Mercatus AI
The AI economy runs on tokens. Mercatus discovers their price.
- Industry
- Information Services
- Location
- Bellevue, US
- Company size
- 1 employee
About Mercatus AI
Mercatus is the first exchange for AI inference tokens. AI companies increasingly budget around one metric: the cost of inference, measured in dollars per million output tokens. Yet there has never been a market where that price could be traded directly. Mercatus changes that. Model providers and inference platforms sell forward token capacity. Enterprises, traders, and investors buy it to hedge future inference costs or express market views. Orders match continuously in a 24/7 central limit order book, with physical settlement of LLM tokens and no intermediaries. Mercatus doesn't publish prices—it discovers them. Every physically settled forward establishes a market-clearing price for AI inference. That price becomes the benchmark for pricing, hedging, and financing compute across the AI economy. The AI economy runs on tokens. Mercatus is where those tokens are priced.
See moreLatest activity
Latest activity from Mercatus AI
3 signals
Research & Knowledge
Mercatus AI published research showing four open-weight models served by 7 to 21 hosts running identical weights, with prices ranging from a rounding error to a 6.5x range.
Products & Services
Mercatus AI raised DeepSeek's API prices on August 16, pulling all 17 third-party hosts serving the same V4 weights three weeks later and running the workload math.
Discover more
Similar signals
Similar public activity from other companies.
Products & Services
Ollama
Ollama introduced off-peak token rates for DeepSeek-V4-Flash and Pro, offering half price outside of 12:00 to 18:00 UTC on weekdays and all day on weekends.
Products & Services
Vals AI
Vals AI ran the Meta Muse Spark 1.3 (Max) model with max reasoning, 131k max output tokens, default temperature and top p, a 1M context window, and priced at $1.25 per MTok or $4.25 per MTok.
Products & Services
ScitiX
ScitiX reports 1s average time to first token, 93.9% KV-cache hit rate, 99.9% uptime, and 72% average cost savings.
Products & Services
Brand Nexus AI
Brand Nexus AI reports that Fable 5.1 and Mythos 5.1 cut cache reads 75% to $0.25 per million tokens, about 25% off typical workloads.
Products & Services
Perceptron AI