Skip to main content
DeepSeekLaunch

DeepSeek-V4.1-Flash Launches on Qwen AI Platform, API and Token Plan Opened Simultaneously

What happened

DeepSeek has launched its new lightweight flagship model, DeepSeek-V4.1-Flash, which is now available on the Qwen AI platform.

Source

Article excerpt

Highlighted: the sentence this signal was extracted from

DeepSeek-V4.1-Flash has recently been launched on the Qwen AI platform, with API services and Token Plan now available. Developers can integrate the model into their systems via standard APIs, or directly use the Token Plan in tools such as Qoder, Qwen APP, and Codex for code, documentation, visual understanding, and agent tasks. According to the platform announcement, this model is DeepSeek's new lightweight flagship, featuring a MoE architecture with 552B total parameters, using a Causal-Encoder-Decoder asymmetric structure, with input activation of about 8B and output activation of about 16B; it natively supports text and image understanding, with a maximum context length of 1 million tokens, and a maximum output of approximately 393K. The official stated that it has improved significantly over its predecessor in several Agent and code benchmarks, achieving high throughput and low latency with lower activation parameters. The cost side is a major focus of this release. The new generation of cache compression reduces the KV Cache demand for HBM to one-quarter of the previous generation and for SSD storage to one-eighth, saving more resources for long contexts and multi-turn tool calls. The Qwen page provides time-based pricing: 1 yuan per million tokens during off-peak hours for input and 4 yuan per million tokens for output; 2 yuan and 8 yuan respectively during peak...

Keep reading with a free account

The rest of this article, and every signal for DeepSeek, is in your free account.

Extracted by Autobound

From the Signal API record
Event
Launch

What this signalsA launch often needs new go-to-market and support spend.

Product
DeepSeek-V4.1-Flash

The full record

From the Signal API record

Details

Release type
Model

Topics and mentions

Product tags

  • future tech
  • general technology

Extraction

Confidence
95%
Detected
Sep 14, 2026
signal_type
news
signal_subtype
launches

Use this data

Get every launch signal for DeepSeek and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at DeepSeek this week?”

  2. Send it to your own tools

    The Signal API returns launch signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full news record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/16dffe4f-9326-eff3-8b0a-f22c50f3a986 returns this record as JSON. POST /v1/companies/enrich returns every signal for deepseek.com.

{
  "signal_id": "16dffe4f-9326-eff3-8b0a-f22c50f3a986",
  "signal_type": "news",
  "signal_subtype": "launches",
  "detected_at": "2026-09-14T08:18:06+00:00",
  "company": {
    "name": "DeepSeek",
    "domain": "deepseek.com"
  },
  "data": {
    "url": "https://news.aibase.com/news/31028",
    "title": "DeepSeek-V4.1-Flash Launches on Qwen AI Platform, API and Token Plan Opened Simultaneously - AIBase",
    "excerpt": "DeepSeek-V4.1-Flash has recently been launched on the Qwen AI platform, with API services and Token Plan now available. Developers can integrate the model into their systems via standard APIs, or directly use the Token Plan in tools such as Qoder, Qwen APP, and Codex for code, documentation, visual understanding, and agent tasks. According to the platform announcement, this model is DeepSeek's new lightweight flagship, featuring a MoE architecture with 552B total parameters, using a Causal-Encoder-Decoder asymmetric structure, with input activation of about 8B and output activation of about 16B; it natively supports text and image understanding, with a maximum context length of 1 million tokens, and a maximum output of approximately 393K. The official stated that it has improved significantly over its predecessor in several Agent and code benchmarks, achieving high throughput and low latency with lower activation parameters. The cost side is a major focus of this release. The new generation of cache compression reduces the KV Cache demand for HBM to one-quarter of the previous generation and for SSD storage to one-eighth, saving more resources for long contexts and multi-turn tool calls. The Qwen page provides time-based pricing: 1 yuan per million tokens during off-peak hours for input and 4 yuan per million tokens for output; 2 yuan and 8 yuan respectively during peak hours...",
    "product": "DeepSeek-V4.1-Flash",
    "summary": "DeepSeek has launched its new lightweight flagship model, DeepSeek-V4.1-Flash, which is now available on the Qwen AI platform.",
    "planning": false,
    "confidence": 0.95,
    "product_data": {
      "name": "DeepSeek-V4.1-Flash",
      "full_text": "DeepSeek-V4.1-Flash",
      "fuzzy_match": false,
      "release_type": "model"
    },
    "product_tags": [
      "future_tech",
      "general_technology"
    ],
    "published_at": "2026-09-14T08:18:06Z",
    "article_sentence": "DeepSeek-V4.1-Flash has recently been launched on the Qwen AI platform, with API services and Token Plan now available."
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.