Skip to main content
Thinking Machines LabIn development

Former OpenAI CTO does what Altman won't: releases a frontier AI model that's actually open

What happened

Thinking Machines Lab is developing Inkling-Small, a 276-billion-parameter MoE model, and plans to release its weights once testing is complete.

Source

Article excerpt

AI and ml Thinking Machines' first open weights model is a 975 billion parameter alternative to Chinese LLMs If you’re in the market for a frontier-class open weights model, your options are few and far between outside of the Chinese model houses. With the Wednesday release of a new model code-named "Inkling", an outfit called Thinking Machines Lab aims to change that. Founded in early 2025 by former OpenAI CTO Mira Murati, Thinking Machines' first model is a big one. Weighing in at 975 billion parameters, the model requires more than two terabytes of GPU memory - a quantity present in around eight of Nvidia's B300 accelerators, or sixteen H200s - to run at its native 16-bit precision. If that’s asking too much of your hardware, Thinking Machines has also released a NVFP4 quantized version of the model capable of running on half the GPUs. This makes it the largest American open weights model to date, and comparable to Chinese models like DeepSeek V4, GLM 5.2, and Kimi K2.6 in terms of size and capabilities. Take these claims with a grain of salt - gaming AI benchmarks isn’t exactly difficult - but Thinking Machines says Inkling is competitive with these models in a variety of workloads, although its benchmark charts also show it trailing proprietary models like Anthropic’s Claude and OpenAI’s GPT. Thinking Machines describes the model as being highly adaptable, intended for...

Keep reading with a free account

The rest of this article, and every signal for Thinking Machines Lab, is in your free account.

Extracted from this sentence

Alongside its flagship model, the company is also previewing Inkling-Small, a 276-billion-parameter MoE model with 12 billion active parameters for those prioritizing latency over throughput and quality.

Extracted by Autobound

From the Signal API record
Event
In development

What this signalsWork in development often leads to new buying for tools, data and services.

Product
Inkling-Small

More in development signals at other companies

The full record

From the Signal API record

Details

Release type
Model

Topics and mentions

Product tags

  • future tech
  • general technology

Extraction

Confidence
90%
Detected
Jul 16, 2026
signal_type
news
signal_subtype
is_developing

Use this data

Get every in development signal for Thinking Machines Lab and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at Thinking Machines Lab this week?”

  2. Send it to your own tools

    The Signal API returns in development signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full news record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/495bc499-8364-0d1b-fcea-e856f1de90f6 returns this record as JSON. POST /v1/companies/enrich returns every signal for thinkingmachines.ai.

{
  "signal_id": "495bc499-8364-0d1b-fcea-e856f1de90f6",
  "signal_type": "news",
  "signal_subtype": "is_developing",
  "detected_at": "2026-07-16T00:14:41+00:00",
  "company": {
    "name": "Thinking Machines Lab",
    "domain": "thinkingmachines.ai"
  },
  "data": {
    "url": "https://www.theregister.com/ai-and-ml/2026/07/16/former-openai-cto-does-what-altman-wont-releases-a-frontier-ai-model-thats-actually-open/5272177",
    "title": "Former OpenAI CTO does what Altman won't: releases a frontier AI model that's actually open",
    "excerpt": "AI and ml Thinking Machines' first open weights model is a 975 billion parameter alternative to Chinese LLMs If you’re in the market for a frontier-class open weights model, your options are few and far between outside of the Chinese model houses. With the Wednesday release of a new model code-named \"Inkling\", an outfit called Thinking Machines Lab aims to change that. Founded in early 2025 by former OpenAI CTO Mira Murati, Thinking Machines' first model is a big one. Weighing in at 975 billion parameters, the model requires more than two terabytes of GPU memory - a quantity present in around eight of Nvidia's B300 accelerators, or sixteen H200s - to run at its native 16-bit precision. If that’s asking too much of your hardware, Thinking Machines has also released a NVFP4 quantized version of the model capable of running on half the GPUs. This makes it the largest American open weights model to date, and comparable to Chinese models like DeepSeek V4, GLM 5.2, and Kimi K2.6 in terms of size and capabilities. Take these claims with a grain of salt - gaming AI benchmarks isn’t exactly difficult - but Thinking Machines says Inkling is competitive with these models in a variety of workloads, although its benchmark charts also show it trailing proprietary models like Anthropic’s Claude and OpenAI’s GPT. Thinking Machines describes the model as being highly adaptable, intended for...",
    "product": "Inkling-Small",
    "summary": "Thinking Machines Lab is developing Inkling-Small, a 276-billion-parameter MoE model, and plans to release its weights once testing is complete.",
    "planning": true,
    "image_url": "https://image.theregister.com/5272215.jpg?imageId=5272215&x=0&y=0&cropw=100&croph=100&panox=0&panoy=0&panow=100&panoh=100&width=1200&height=683",
    "confidence": 0.9,
    "product_data": {
      "name": "Inkling-Small",
      "full_text": "Inkling-Small, a 276-billion-parameter MoE model",
      "fuzzy_match": false,
      "release_type": "model"
    },
    "product_tags": [
      "future_tech",
      "general_technology"
    ],
    "published_at": "2026-07-16T00:14:41Z",
    "article_sentence": "Alongside its flagship model, the company is also previewing Inkling-Small, a 276-billion-parameter MoE model with 12 billion active parameters for those prioritizing latency over throughput and quality."
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.