Product Thumbnail

Inkling

Open weights 975B multimodal model built for fine-tuning

Artificial Intelligence
Development
Visit WebsiteSee on Product HuntHugging FaceTwitter

Hunted byZac ZuoZac Zuo

Inkling is Thinking Machines’ first open-weights model, a 975B MoE with 41B active parameters, 1M context, native reasoning across text, images, and audio, and controllable thinking effort. Fine-tune it on Tinker or download the Apache 2.0 weights.

Top comment

Hi everyone!

Inkling is the first model from Thinking Machines.

It is a 975B MoE with 41B active parameters, a 1M-token context window, and native reasoning across text, images, and audio. The full weights are available under Apache 2.0.

Thinking Machines is clear about what Inkling is for. They say directly that it is not the strongest model available today. It is meant to be a broad base that can be adapted to a specific product or workflow.

You can control how much thinking it uses, fine-tune it on @Tinker, and deploy the resulting checkpoints through several inference providers.

They even had Inkling write and run its own fine-tuning job, turning itself into a model that avoids the letter “e”.

The bet is that a model shaped around your own work can be more useful than a slightly higher score on a temporary leaderboard.

Comment highlights

the framing of being explicitly not the strongest model but the best base to adapt is refreshing, most labs oversell the raw benchmark score. with 41B active params out of 975B total, what's the realistic hardware floor for someone fine-tuning this through Tinker, is this something a well-funded startup can do on rented compute or does it really need lab-scale infrastructure?

LoRA is great, but supporting full fine-tuning or QLoRA as a built-in option would make this way more useful for folks working on smaller models where LoRA just doesn't cut it.

finally an option that lets me skip the gpu setup dance and just get to the actual tuning. ran a small lora job last night and it just worked, which honestly surprised me

one thing that would make tinker a lot more useful for me is built in support for evaluating models right after fine tuning, like running a small benchmark suite automatically so you can see if your lora actually helped without wiring up a separate eval pipeline

Really interesting approach. What's the minimum dataset size you'd recommend for effective domain fine-tuning?

As someone who's never trained an AI before, it seems pretty approachable. You fine-tune open models on your own data, and they handle the heavy infrastructure stuff. Inkling also feels like it's meant to be adapted to your workflow. Good idea

Built a quick LoRA job on Tinker yesterday and the setup was honestly painless. One thing that would be a huge help though: a built-in diff viewer or summary that shows what changed in the merged adapter weights so I can sanity-check before pushing to prod without having to script it myself.

Would love to see built-in support for evaluating checkpoints mid-training, so we can compare LoRA adapters on a validation set without writing custom eval loops. A simple callback or webhook when a checkpoint saves would go a long way for experiment tracking.

Love that Thinking Machines put the honest framing up front, "this isn't the strongest model today, it's a base you shape around your own work." Most open-weights drops oversell the benchmark line, so leading with adaptability instead is a cleaner pitch, and Apache 2.0 on a 1M-context multimodal MoE is a real gift.


The thing I keep bumping on is that "open weights" and "actually touchable" aren't the same at 975B. Realistically, who fine-tunes a model this size outside of Tinker? Wondering whether the openness is meaningful in practice, or whether downloadable weights are mostly a trust signal and Tinker is the real on-ramp most people will have to take to do anything with it.

Most of our pain when we LoRA an open model comes from the post-training rather than the weights. You tune for a narrow extraction job and the model keeps sliding back into the chatty assistant phrasing it picked up in RLHF, and you burn epochs beating that out. Is there a raw pre-trained checkpoint alongside the released one, or is the Apache 2.0 drop the post-trained model only? For something billed as a base to build on, that distinction decides a lot.

About Inkling on Product Hunt

Open weights 975B multimodal model built for fine-tuning

Inkling launched on Product Hunt on July 20th, 2026 and earned 189 upvotes and 11 comments, placing #6 on the daily leaderboard. Inkling is Thinking Machines’ first open-weights model, a 975B MoE with 41B active parameters, 1M context, native reasoning across text, images, and audio, and controllable thinking effort. Fine-tune it on Tinker or download the Apache 2.0 weights.

Inkling was featured in Artificial Intelligence (474.1k followers) and Development (6k followers) on Product Hunt. Together, these topics include over 111.5k products, making this a competitive space to launch in.

Who hunted Inkling?

Inkling was hunted by Zac Zuo. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.

Want to see how Inkling stacked up against nearby launches in real time? Check out the live launch dashboard for upvote speed charts, proximity comparisons, and more analytics.