AIThis post was created with the assistance of artificial intelligence (AI).

TL;DR

Prime Big Deal Days · Oct 6–7Offer from Amazon

Get the latest gadgets delivered free — and shop member deals

  • Fast, free delivery on millions of items
  • Access to Prime Big Deal Days deals on October 6–7
  • Prime Video, Amazon Music and more included
Start your free Prime trial Free trial for eligible customers · Cancel anytime
As an affiliate, we earn on qualifying purchases.

Thinking Machines Lab released its first foundation model, Inkling, with full weights under Apache 2.0 and immediate support from major inference frameworks. The lab concedes that Inkling is not the strongest available model, while emphasizing ownership, deployment flexibility and adjustable reasoning costs.

Thinking Machines Lab, the 17-month-old company founded by former OpenAI chief technology officer Mira Murati, released its first foundation model, Inkling, on July 15 with full weights available immediately on Hugging Face under an Apache 2.0 license. The open-first release gives organizations a path to modify and operate the model themselves, although its hardware demands place the flagship beyond most individual developers and smaller teams.

Inkling is a mixture-of-experts model with 975 billion total parameters and 41 billion active parameters, according to specifications published by Thinking Machines. The company says it supports a 1-million-token context window and was pretrained on 45 trillion tokens spanning text, images, audio and video.

The model accepts text, images and audio and produces text. Thinking Machines released BF16 and NVFP4 checkpoints with day-zero support in Transformers, vLLM, SGLang and llama.cpp, among other tools. The weights and model card carry an Apache 2.0 license, which generally permits modification and commercial use.

The company did not claim overall market leadership. Its announcement said Inkling is “not the strongest model available today”, whether compared with open or closed systems. Vendor-published results place it ahead on some mathematics, audio and adversarial tests but behind models including GLM-5.2 and Fable 5 on several coding and agent-oriented evaluations. Those results have not yet received broad independent replication, and some reportedly used a prerelease checkpoint.

At a glance
announcementWhen: announced July 15, 2026; benchmark and…
The developmentThinking Machines Lab released Inkling’s full weights on July 15 before offering a closed API, making open distribution central to its first foundation-model launch.

Open Distribution Changes Model Control

Releasing the weights first makes customer ownership, rather than API access, a central part of the product. Organizations can inspect, modify, fine-tune and host Inkling within their own infrastructure, reducing dependence on a provider that can change prices, access rules or model versions.

Inkling also includes an adjustable reasoning-effort setting from 0.2 to 0.99. Thinking Machines says this lets operators trade reasoning tokens against latency and cost. The company reports that Inkling matched Nemotron 3 Ultra on Terminal-Bench 2.1 while using about one-third as many tokens, but that comparison remains a vendor claim pending outside testing.

Amazon

AI model training hardware

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Murati’s Lab Chooses Open First

Thinking Machines was founded by Murati and former OpenAI employees, including people who worked on ChatGPT. Many frontier-model developers distribute their strongest systems through controlled APIs, while open weights arrive later, apply only to smaller models or are released under more restrictive terms. Inkling reverses that sequence by putting downloadable checkpoints first.

The release also enters a field where Chinese developers have supplied several leading open-weight models. Inkling is positioned as a US-developed alternative, yet the source material reports that its post-training used synthetic data from Kimi K2.5. Thinking Machines has not published Inkling’s training dataset or full pipeline, meaning the release is open-weight rather than fully open-source.

“Inkling is not the strongest model available today, closed or open.”

— Thinking Machines Lab

Amazon

high performance GPU for AI development

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

License Limits and Benchmarks Need Checks

It is not yet clear whether a separate Model Acceptable Use Policy restricts the parameters and modified versions beyond the Apache 2.0 license. The source material reports possible bans covering surveillance, deception and fully automated decisions affecting rights, but says the policy was not independently verified. Prospective users will need to examine the current repository documents before deployment.

Independent evaluators have also not fully tested the company’s benchmark, efficiency and multimodal claims. Inkling’s real-world reliability, fine-tuning behavior and operating costs across different hardware configurations remain uncertain.

Amazon

AI model fine-tuning hardware

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Independent Tests and Smaller Weights Follow

Researchers and prospective customers will now test Inkling’s released checkpoints against GLM-5.2, Kimi K2.6 and closed systems on production workloads. Thinking Machines is also testing Inkling-Small, a 276-billion-parameter model with 12 billion active parameters, and says its full weights will follow after testing is complete.

Amazon

AI inference server

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Key Questions

What did Thinking Machines release?

The company released Inkling, its first foundation model, including BF16 and NVFP4 weights hosted on Hugging Face.

Is Inkling fully open-source?

No. Its model weights are available under Apache 2.0, but the training data and complete training pipeline have not been published.

Can Inkling run on a personal workstation?

Not in its standard released forms. The source estimates that BF16 needs at least 2 terabytes of aggregate VRAM, while NVFP4 still requires about 600 gigabytes. Quantized versions may reduce the requirement, with possible quality losses.

Is Inkling the best-performing open model?

Thinking Machines says it is not the strongest model overall. Vendor results show competitive performance on selected tests, but Inkling trails some rivals on coding, agent and general reasoning benchmarks.

When will Inkling-Small be available?

Thinking Machines has shown a preview and says the weights will be released after testing. The company has not provided a firm release date.

Source: Thorsten Meyer AI

FALL

Fall Picks

As an affiliate, we earn on qualifying purchases.

You May Also Like

Pentagon AI Goes Explicit: The Frontier Labs Move Inside the Classified Stack

The Pentagon has announced agreements with major AI firms to deploy advanced AI systems within classified networks, marking a shift toward AI-first military operations.

Alphabet has its worst day in over a year on AI concerns after high-profile exits

Alphabet experiences its worst day in over a year amid fears over AI development following a high-profile executive departure.

Technology Missouri Surges In Global Coverage

Technology Missouri has seen a significant surge in international coverage, with 43 mentions in recent media analysis, highlighting its rising prominence.

Mastering MW4: Essential Tips For Dominating The Gaming Scene

Learn essential strategies and tips confirmed by experts to excel in MW4 and elevate your gaming performance.