Skip to main content
DeepSeekCustomer feedback

A commenter recommends using DeepSeek's models directly from the source company, noting great pricing and quality compared to using them through third-party providers on OpenRouter.

What happened

Post: "Has anyone else had endlessly nonsense output from Open Inference? (GLM 5.3 Flash / DeepSeek V4.1 Flash)"

Source

Post

Hi all, I want to know whether anyone else has run into this, or whether it's just me. Yesterday I was using GLM 5.3 Flash and DeepSeek V4.1 Flash through OpenRouter in Pi. After a few turns, both started vomiting a seemingly never ending list of unrelated words, until i stopped them. Most of the tokens were billed as reasoning. It happened with two unrelated models, which seemed odd, so digging deeper i found that every request had gone to the same provider, Open Inference, which serves both models in fp4. Once I blocked Open Inference in my settings, the same model in the same session went straight back to normal (on Wafer and DekaLLM). I also noticed that, for GLM, it was also the most expensive provider by a long shot (and one of the slowest), which makes me question why i was routed to Open Inference in the first place (I have no particular routing config/setup). I don't know how Open Inference runs its deployments or exactly how OpenRouter decides where to route. It could just be a bad deployment or a bug, looking at the token volume (https://openrouter.ai/provider/open-inference) there's been a big drop after Sep 26th, so maybe it is something happening systematically. Has anyone experienced this? am I missing something obvious? but also: shouldn't there be something on openrouter gauging provider output quality? this could have been very expensive, both for the...

Keep reading with a free account

The rest of this post, and every signal for DeepSeek, is in your free account.

Extracted from these lines

  • [comment u/Brilliant-Hall1387] I recommend using DeepSeek and GLM directly from the source companies in China, great pricing and you know you get the real thing at really good prices. At least try it, top up an account with 10 USD each and evaluate the difference. DeepSeek offers 2 USD top up if you want to start really low.

Comments on the post

5 of 9 comments
  • “yeah, blocking Open Inference is a useful workaround. i'd put a low max output/reasoning cap on those requests while testing, so a bad backend can't run away with tokens before you notice.”

    u/locbuilds6 points · Oct 2, 2026View

  • “The providers for the popular open source ones can be very hit or miss, and the ones that charge less especially so. I've pinned the providers who are generally consistent, but even then I have problems. This isn't as much of an OpenRouter thing, but a provider and model thing.”

    u/MaybeLiterally3 points · Oct 2, 2026View

  • “OpenInference is lately cheapest among providers, so I guess they're compensating it by making the output verbose. I usually pin StreamLake or the original provider for most.”

    u/squirrelscrush3 points · Oct 2, 2026View

  • “Yeah I have had this as well. Probably some provider quantizing the model secretly so that it becomes retarded”

    u/Beeschurger_xd3 points · Oct 2, 2026View

  • “YES I was coming here to post this exact thing. I banned them by putting them on my block list. Not sure what shenanigans they did to the GLM model. I took it personally and felt like I was being played. If anyone has their favorite providers with strong Zdr policies please post here.”

    u/generic-d-engineer3 points · Oct 2, 2026View

Extracted by Autobound

From the Signal API record
Signal
Customer feedback

What this signalsUser posts often show product pain before it reaches reviews or churn.

Subreddit
r/openrouter
Stage
Considering

The full record

From the Signal API record

Numbers

Mentions
12

Details

Timing
Ongoing state
Category
General
Virality
Medium
Post kind
Multi media
Prominence
Aside
Company's role
Vendor

Topics and mentions

Topics

  • reliability
  • direct sourcing
  • pricing

Extraction

Sentiment
Positive
Detected
Oct 2, 2026
signal_type
reddit-company
signal_subtype
customerFeedback

Use this data

Get every Reddit signal for DeepSeek and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at DeepSeek this week?”

  2. Send it to your own tools

    The Signal API returns Reddit signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full reddit-company record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/e4a433c6-273a-523d-ac05-225428e5124c returns this record as JSON. POST /v1/companies/enrich returns every signal for deepseek.com.

{
  "signal_id": "e4a433c6-273a-523d-ac05-225428e5124c",
  "signal_type": "reddit-company",
  "signal_subtype": "customerFeedback",
  "detected_at": "2026-10-02T16:27:55+00:00",
  "company": {
    "name": "DeepSeek",
    "domain": "deepseek.com"
  },
  "data": {
    "nsfw": false,
    "stage": "considering",
    "awards": 0,
    "timing": "ongoing_state",
    "topics": [
      "pricing",
      "reliability",
      "direct sourcing"
    ],
    "post_id": "1wvy9ua",
    "summary": "A commenter recommends using DeepSeek's models directly from the source company, noting great pricing and quality compared to using them through third-party providers on OpenRouter.",
    "category": "general",
    "comments": [
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdg2zda/",
        "depth": 0,
        "score": 6,
        "author": "locbuilds",
        "excerpt": "yeah, blocking Open Inference is a useful workaround. i'd put a low max output/reasoning cap on those requests while testing, so a bad backend can't run away with tokens before you notice.",
        "posted_at": "2026-10-02T16:59:52.000Z",
        "author_url": "https://www.reddit.com/user/locbuilds/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdg0n1d/",
        "depth": 0,
        "score": 3,
        "author": "MaybeLiterally",
        "excerpt": "The providers for the popular open source ones can be very hit or miss, and the ones that charge less especially so. I've pinned the providers who are generally consistent, but even then I have problems. This isn't as much of an OpenRouter thing, but a provider and model thing.",
        "posted_at": "2026-10-02T16:49:52.000Z",
        "author_url": "https://www.reddit.com/user/MaybeLiterally/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdgkxpa/",
        "depth": 0,
        "score": 3,
        "author": "squirrelscrush",
        "excerpt": "OpenInference is lately cheapest among providers, so I guess they're compensating it by making the output verbose.\n\n I usually pin StreamLake or the original provider for most.",
        "posted_at": "2026-10-02T18:16:23.000Z",
        "author_url": "https://www.reddit.com/user/squirrelscrush/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdh75yw/",
        "depth": 0,
        "score": 3,
        "author": "Beeschurger_xd",
        "excerpt": "Yeah I have had this as well. Probably some provider quantizing the model secretly so that it becomes retarded",
        "posted_at": "2026-10-02T19:50:43.000Z",
        "author_url": "https://www.reddit.com/user/Beeschurger_xd/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdhhjzf/",
        "depth": 0,
        "score": 3,
        "author": "generic-d-engineer",
        "excerpt": "YES\n\n I was coming here to post this exact thing. I banned them by putting them on my block list. Not sure what shenanigans they did to the GLM model. I took it personally and felt like I was being played.\n\n If anyone has their favorite providers with strong Zdr policies please post here.",
        "posted_at": "2026-10-02T20:35:19.000Z",
        "author_url": "https://www.reddit.com/user/generic-d-engineer/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdfy5jo/",
        "depth": 0,
        "score": 2,
        "author": "Ok-Lobster-919",
        "excerpt": "I had to stop using Wafer for the same reason actually. Just pin the providers you want.",
        "posted_at": "2026-10-02T16:39:16.000Z",
        "author_url": "https://www.reddit.com/user/Ok-Lobster-919/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdimnr8/",
        "depth": 0,
        "score": 2,
        "author": "JeffBezosHater",
        "excerpt": "Buzzword buzzword buzzword",
        "posted_at": "2026-10-03T00:00:54.000Z",
        "author_url": "https://www.reddit.com/user/JeffBezosHater/"
      },
      {
        "url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/comment/pdi0ygz/",
        "depth": 0,
        "score": 2,
        "author": "queso184",
        "excerpt": "It's actually unusable, i was using deepseek through them and it was getting into terrible reasoning loops",
        "posted_at": "2026-10-02T22:05:34.000Z",
        "author_url": "https://www.reddit.com/user/queso184/"
      }
    ],
    "evidence": [
      "[comment u/Brilliant-Hall1387] I recommend using DeepSeek and GLM directly from the source companies in China, great pricing and you know you get the real thing at really good prices. At least try it, top up an account with 10 USD each and evaluate the difference. DeepSeek offers 2 USD top up if you want to start really low."
    ],
    "virality": "medium",
    "post_date": "2026-10-02T16:27:55.000Z",
    "post_kind": "multi_media",
    "post_text": "Hi all, I want to know whether anyone else has run into this, or whether it's just me.\n\nYesterday I was using GLM 5.3 Flash and DeepSeek V4.1 Flash through OpenRouter in Pi. After a few turns, both started vomiting a seemingly never ending list of unrelated words, until i stopped them. Most of the tokens were billed as reasoning.\n\nIt happened with two unrelated models, which seemed odd, so digging deeper i found that every request had gone to the same provider, Open Inference, which serves both models in fp4. Once I blocked Open Inference in my settings, the same model in the same session went straight back to normal (on Wafer and DekaLLM).\n\nI also noticed that, for GLM, it was also the most expensive provider by a long shot (and one of the slowest), which makes me question why i was routed to Open Inference in the first place (I have no particular routing config/setup).\n\nI don't know how Open Inference runs its deployments or exactly how OpenRouter decides where to route. It could just be a bad deployment or a bug, looking at the token volume (https://openrouter.ai/provider/open-inference) there's been a big drop after Sep 26th, so maybe it is something happening systematically.\n\nHas anyone experienced this? am I missing something obvious?\n\nbut also: shouldn't there be something on openrouter gauging provider output quality? this could have been very expensive, both for the...",
    "sentiment": "positive",
    "subreddit": "openrouter",
    "post_title": "Has anyone else had endlessly nonsense output from Open Inference? (GLM 5.3 Flash / DeepSeek V4.1 Flash)",
    "prominence": "aside",
    "source_url": "https://www.reddit.com/r/openrouter/comments/1wvy9ua/has_anyone_else_had_endlessly_nonsense_output/",
    "entity_role": "vendor",
    "post_author": "icandela",
    "upvote_ratio": 0.9047619047619048,
    "mention_count": 12,
    "mention_surge": true,
    "subreddit_url": "https://www.reddit.com/r/openrouter/",
    "total_upvotes": 17,
    "comments_total": 9,
    "total_comments": 9,
    "post_author_url": "https://www.reddit.com/user/icandela/",
    "signal_category": "feedback",
    "comments_included": 9
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.