Skip to main content
DeepSeekCustomer feedback

A user running the DeepSeek V4 Flash 0423 model via OpenRouter was surprised by the final cost, which was much higher than the model's attractive headline pricing due to provider routing and...

What happened

A user running the DeepSeek V4 Flash 0423 model via OpenRouter was surprised by the final cost, which was much higher than the model's attractive headline pricing due to provider routing and cache-read fees on the platform.

Source

RedditSep 27, 2026By u/matrixoar

r/openrouter

OpenRouter can charged 3x more because of provider routing ,be careful!

upvotes
8
comments
19

Post

Highlighted: the lines this signal was extracted from

I think OpenRouter needs to make provider-specific pricing much more obvious, especially for cache-heavy coding agents. I exported my Activity CSV after running DeepSeek V4 Flash 0423 through a coding agent. My results for roughly 8 hours 15 minutes: 1,175 requests 123.86M prompt tokens 115.75M cached tokens 93.45% cache ratio $7.82 actually charged The surprising part was the provider routing: Parasail: 687 requests - $6.15 NextBit: 482 requests - $1.65 Baidu: 5 requests StreamLake: 1 request I had been looking at the attractive DeepSeek/OpenRouter pricing and assumed the actual cost would be somewhere around that level. But Parasail charges $0.07/M for cache reads, while StreamLake is currently around $0.017/M for the same DeepSeek V4 Flash 0423 model. Using the exact token profile from my Activity export, I calculate that the same workload pinned to StreamLake would have cost roughly $2.78 instead of $7.82. That's about almost 3× the cost simply because of provider routing. This matters enormously for coding agents because almost all of the context gets repeatedly read from cache. In my case, more than 93% of input tokens were cached. What makes it even more striking is that newer DeepSeek V4 Flash 0731 endpoints currently have cache-read prices as low as ~$0.00182/M on StreamLake. I'm not saying OpenRouter is adding a hidden markup. I understand that...

Keep reading with a free account

The rest of this post, and every signal for DeepSeek, is in your free account.

Comments on the post

5 of 19 comments
  • “OpenRouter already does provider sticky routing. https://openrouter.ai/docs/guides/best-practices/prompt-caching#provider-sticky-routing. They also offer the auto routing that optimizes on cost. As someone mentioned - you can also control which providers each workspace can use if cost management is important. Ultimately the whole field of cost / token management is an emerging area of developm”

    u/bikesandboots10 points · Sep 27, 2026View

  • “Just allow the providers you want for each model so every other provider is blocked.”

    u/AndoniFdez5 points · Sep 27, 2026View

  • “Parasail CEO here. Openrouter will route to the lowest cost provider if you ask it to and it calculates cost of cache reads and writes when determining "lowest cost". The real missing piece here is capacity. We are not the lowest cost provider on openrouter, probably more in the middle AND we throwing a lot of capacity at Openrouter and its still mostly maxed out. This means in general demand is a”

    u/Dizzy-Bad44234 points · Sep 27, 2026View

  • “You can add guardrails to limit the providers served by opencode”

    u/albertortilla3 points · Sep 27, 2026View

  • “We do something similar but you can specially what provider and the limits to help avoid any unwanted routing.”

    u/RogerAI-fm2 points · Sep 27, 2026View

Extracted by Autobound

From the Signal API record
Signal
Customer feedback

What this signalsUser posts often show product pain before it reaches reviews or churn.

Subreddit
r/openrouter
Event date
Sep 2026

Companies

  • OpenRouterAlso named

The full record

From the Signal API record

Numbers

Mentions
13

Details

Timing
Completed
Category
Pricing
Virality
Somewhat high
Post kind
Text
Prominence
Aside
Company's role
Vendor

Topics and mentions

Topics

  • pricing
  • api
  • ai
  • llm

Products named

  • DeepSeek V4 Flash 0423
  • DeepSeek V4 Flash 0731

Extraction

Sentiment
Neutral
Detected
Sep 27, 2026
signal_type
reddit-company
signal_subtype
customerFeedback

Use this data

Get every Reddit signal for DeepSeek and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at DeepSeek this week?”

  2. Send it to your own tools

    The Signal API returns Reddit signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full reddit-company record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/d0411735-9f71-58c1-a4ae-5ecb24e0ce65 returns this record as JSON. POST /v1/companies/enrich returns every signal for deepseek.com.

{
  "signal_id": "d0411735-9f71-58c1-a4ae-5ecb24e0ce65",
  "signal_type": "reddit-company",
  "signal_subtype": "customerFeedback",
  "detected_at": "2026-09-27T15:15:01+00:00",
  "company": {
    "name": "DeepSeek",
    "domain": "deepseek.com"
  },
  "data": {
    "nsfw": false,
    "stage": "none",
    "awards": 0,
    "timing": "completed",
    "topics": [
      "pricing",
      "api",
      "ai",
      "llm"
    ],
    "post_id": "1wrmqsr",
    "summary": "A user running the DeepSeek V4 Flash 0423 model via OpenRouter was surprised by the final cost, which was much higher than the model's attractive headline pricing due to provider routing and cache-read fees on the platform.",
    "category": "pricing",
    "comments": [
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pcdwdew/",
        "depth": 0,
        "score": 10,
        "author": "bikesandboots",
        "excerpt": "OpenRouter already does provider sticky routing. https://openrouter.ai/docs/guides/best-practices/prompt-caching#provider-sticky-routing. They also offer the auto routing that optimizes on cost.\n\n As someone mentioned - you can also control which providers each workspace can use if cost management is important.\n\n Ultimately the whole field of cost / token management is an emerging area of developm",
        "posted_at": "2026-09-27T15:44:23.000Z",
        "author_url": "https://www.reddit.com/user/bikesandboots/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pce006l/",
        "depth": 0,
        "score": 5,
        "author": "AndoniFdez",
        "excerpt": "Just allow the providers you want for each model so every other provider is blocked.",
        "posted_at": "2026-09-27T15:59:49.000Z",
        "author_url": "https://www.reddit.com/user/AndoniFdez/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pcgt4j6/",
        "depth": 0,
        "score": 4,
        "author": "Dizzy-Bad4423",
        "excerpt": "Parasail CEO here. Openrouter will route to the lowest cost provider if you ask it to and it calculates cost of cache reads and writes when determining \"lowest cost\". The real missing piece here is capacity. We are not the lowest cost provider on openrouter, probably more in the middle AND we throwing a lot of capacity at Openrouter and its still mostly maxed out. This means in general demand is a",
        "posted_at": "2026-09-27T22:53:23.000Z",
        "author_url": "https://www.reddit.com/user/Dizzy-Bad4423/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pcdub14/",
        "depth": 0,
        "score": 3,
        "author": "albertortilla",
        "excerpt": "You can add guardrails to limit the providers served by opencode",
        "posted_at": "2026-09-27T15:35:32.000Z",
        "author_url": "https://www.reddit.com/user/albertortilla/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pcevyq5/",
        "depth": 0,
        "score": 2,
        "author": "RogerAI-fm",
        "excerpt": "We do something similar but you can specially what provider and the limits to help avoid any unwanted routing.",
        "posted_at": "2026-09-27T18:09:57.000Z",
        "author_url": "https://www.reddit.com/user/RogerAI-fm/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pce2das/",
        "depth": 0,
        "score": 2,
        "author": "not420guilty",
        "excerpt": "It’s a bit of a scam. Advertising one price but route to another higher price.",
        "posted_at": "2026-09-27T16:09:44.000Z",
        "author_url": "https://www.reddit.com/user/not420guilty/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pcle8cb/",
        "depth": 0,
        "score": 1,
        "author": "PoppaBear1950",
        "excerpt": "use a direct API to deepseek, use code-review-graph for context control use JEV to map structural context then off to the coding harness... you cost will hit a downward trend quickly.",
        "posted_at": "2026-09-28T15:25:39.000Z",
        "author_url": "https://www.reddit.com/user/PoppaBear1950/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/comment/pci8j8l/",
        "depth": 0,
        "score": 0,
        "author": "funbike",
        "excerpt": "Read the docs. They clearly explain you can prioritize cost, speed, or price, or you can choose a specific set of providers.\n\n This is your fault, not openrouters. They are open about how their service works, how pricing works, and how to to use it. You just didn't do your due diligence in reading the docs.",
        "posted_at": "2026-09-28T03:30:42.000Z",
        "author_url": "https://www.reddit.com/user/funbike/"
      }
    ],
    "evidence": [
      "[post] I exported my Activity CSV after running DeepSeek V4 Flash 0423 through a coding agent.",
      "[post] I had been looking at the attractive DeepSeek/OpenRouter pricing and assumed the actual cost would be somewhere around that level."
    ],
    "virality": "somewhat_high",
    "post_date": "2026-09-27T15:15:01.000Z",
    "post_kind": "text",
    "post_text": "I think OpenRouter needs to make provider-specific pricing much more obvious, especially for cache-heavy coding agents.\n\nI exported my Activity CSV after running DeepSeek V4 Flash 0423 through a coding agent.\n\nMy results for roughly 8 hours 15 minutes:\n\n1,175 requests\n\n123.86M prompt tokens\n\n115.75M cached tokens\n\n93.45% cache ratio\n\n$7.82 actually charged\n\nThe surprising part was the provider routing:\n\nParasail: 687 requests - $6.15\n\nNextBit: 482 requests - $1.65\n\nBaidu: 5 requests\n\nStreamLake: 1 request\n\nI had been looking at the attractive DeepSeek/OpenRouter pricing and assumed the actual cost would be somewhere around that level.\n\nBut Parasail charges $0.07/M for cache reads, while StreamLake is currently around $0.017/M for the same DeepSeek V4 Flash 0423 model.\n\nUsing the exact token profile from my Activity export, I calculate that the same workload pinned to StreamLake would have cost roughly $2.78 instead of $7.82.\n\nThat's about almost 3× the cost simply because of provider routing.\n\nThis matters enormously for coding agents because almost all of the context gets repeatedly read from cache. In my case, more than 93% of input tokens were cached.\n\nWhat makes it even more striking is that newer DeepSeek V4 Flash 0731 endpoints currently have cache-read prices as low as ~$0.00182/M on StreamLake.\n\nI'm not saying OpenRouter is adding a hidden markup. I understand that...",
    "sentiment": "neutral",
    "subreddit": "openrouter",
    "event_date": "2026-09",
    "post_title": "OpenRouter can charged 3x more because of provider routing ,be careful!",
    "prominence": "aside",
    "source_url": "https://www.reddit.com/r/openrouter/comments/1wrmqsr/openrouter_can_charged_3x_more_because_of/",
    "entity_role": "vendor",
    "post_author": "matrixoar",
    "upvote_ratio": 0.7,
    "mention_count": 13,
    "mention_surge": true,
    "subreddit_url": "https://www.reddit.com/r/openrouter/",
    "total_upvotes": 8,
    "comments_total": 19,
    "total_comments": 19,
    "other_companies": [
      {
        "name": "OpenRouter",
        "role": "partner",
        "domain": "openrouter.ai"
      }
    ],
    "post_author_url": "https://www.reddit.com/user/matrixoar/",
    "signal_category": "feedback",
    "comments_included": 8,
    "products_mentioned": [
      "DeepSeek V4 Flash 0423",
      "DeepSeek V4 Flash 0731"
    ]
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.