Skip to main content
BaiduLaunch

Baidu Releases MIT-Licensed 3B OCR Model for long documents.

What happened

Baidu, Inc. launches MIT-Licensed 3B OCR Model.

Source

Article excerpt

Highlighted: the sentence this signal was extracted from

Baidu Releases MIT-Licensed 3B OCR Model for long documents. Tl;dr. * Baidu's Unlimited-OCR is a 3-billion-parameter MIT-licensed model that processes multi-page PDFs in a single inference pass. * The model has a 32,768-token context window and supports vLLM, SGLang, Ollama, llama.cpp, and Hugging Face Transformers. * No benchmark results are included in the release; training data and language coverage are also not disclosed. Baidu published Unlimited-OCR to Hugging Face, a 3-billion-parameter model for document parsing released under an MIT license. The headline capability is what the model card calls "One-shot Long-horizon Parsing": processing multi-page PDFs and image stacks in a single inference pass rather than requiring documents to be pre-sliced page by page. A 32,768-token context window supports that approach on longer documents. Most open-source OCR pipelines force you to cut input into individual pages, run each through a model separately, then stitch the outputs back together. The architectural bet here is that a sufficiently long context window lets the model handle that coherence itself. According to the model card, Unlimited-OCR builds on DeepSeek-OCR and DeepSeek-OCR-2, and deploys via Hugging Face Transformers, vLLM, SGLang, Docker, Ollama, and llama.cpp, meaning it should slot into most existing inference setups without significant rework. The MIT...

Keep reading with a free account

The rest of this article, and every signal for Baidu, is in your free account.

Extracted by Autobound

From the Signal API record
Event
Launch

What this signalsA launch often needs new go-to-market and support spend.

Product
MIT-Licensed 3B OCR Model

The full record

From the Signal API record

Details

Ticker
OTC:BAIDF

Extraction

Confidence
59%
Detected
Jun 22, 2026
signal_type
news
signal_subtype
launches

Use this data

Get every launch signal for Baidu and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at Baidu this week?”

  2. Send it to your own tools

    The Signal API returns launch signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full news record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/e9a20f3a-3592-48e7-b5ef-6fb6a3428dfc returns this record as JSON. POST /v1/companies/enrich returns every signal for baidu.com.

{
  "signal_id": "e9a20f3a-3592-48e7-b5ef-6fb6a3428dfc",
  "signal_type": "news",
  "signal_subtype": "launches",
  "detected_at": "2026-06-22T00:00:00+00:00",
  "company": {
    "name": "Baidu",
    "domain": "baidu.com"
  },
  "data": {
    "url": "https://aiweekly.co/alerts/baidu-releases-mit-licensed-3b-ocr-model-for-long-documents",
    "title": "Baidu Releases MIT-Licensed 3B OCR Model for long documents.",
    "ticker": "OTC:BAIDF",
    "excerpt": "Baidu Releases MIT-Licensed 3B OCR Model for long documents.\n\nTl;dr.\n\n* Baidu's Unlimited-OCR is a 3-billion-parameter MIT-licensed model that processes multi-page PDFs in a single inference pass.\n* The model has a 32,768-token context window and supports vLLM, SGLang, Ollama, llama.cpp, and Hugging Face Transformers.\n* No benchmark results are included in the release; training data and language coverage are also not disclosed.\n\nBaidu published Unlimited-OCR to Hugging Face, a 3-billion-parameter model for document parsing released under an MIT license. The headline capability is what the model card calls \"One-shot Long-horizon Parsing\": processing multi-page PDFs and image stacks in a single inference pass rather than requiring documents to be pre-sliced page by page. A 32,768-token context window supports that approach on longer documents.\n\nMost open-source OCR pipelines force you to cut input into individual pages, run each through a model separately, then stitch the outputs back together. The architectural bet here is that a sufficiently long context window lets the model handle that coherence itself. According to the model card, Unlimited-OCR builds on DeepSeek-OCR and DeepSeek-OCR-2, and deploys via Hugging Face Transformers, vLLM, SGLang, Docker, Ollama, and llama.cpp, meaning it should slot into most existing inference setups without significant rework.\n\nThe MIT...",
    "product": "MIT-Licensed 3B OCR Model",
    "summary": "Baidu, Inc. launches MIT-Licensed 3B OCR Model.",
    "planning": false,
    "image_url": "https://aiweekly.co/themes/custom/aiweekly/images/logo.png",
    "confidence": 0.5858,
    "product_data": {
      "full_text": "MIT-Licensed 3B OCR Model",
      "fuzzy_match": true
    },
    "published_at": "2026-06-22T00:00:00Z",
    "article_sentence": "Baidu Releases MIT-Licensed 3B OCR Model for long documents."
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.