Inkling
Thinking Machines Lab's Apache-licensed 975B MoE with 1M context and day-one Tinker fine-tuning.
About
Thinking Machines Lab, the startup Mira Murati founded after leaving OpenAI, shipped its first from-scratch model on July 15, 2026 and put the full weights on Hugging Face under Apache 2.0. Inkling is a 975B parameter mixture-of-experts network with 41B active per token that accepts text, images, and audio, and it debuted as the leading US open-weights model at 41 on the Artificial Analysis Intelligence Index. Context reaches 1M tokens when self-hosting the raw weights, while the company's Tinker fine-tuning platform serves it at 64K and 256K windows with day-one support for post-training, the release being explicitly framed as a base for customization rather than a finished chat product. A smaller sibling, Inkling-Small, followed at the end of July as a 276B MoE with 12B active that matches the flagship on several reasoning and agentic benchmarks at a quarter of the size. Running the full model locally requires a multi-GPU cluster; the permissive license allows unrestricted commercial use, modification, and redistribution.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Large Language Models (LLMs)
- Price
- Freemium
- Platform
- Hybrid
- Difficulty
- Advanced (4/5)
- License
- Apache-2.0
- Added
- Aug 24, 2026
Related Tools
Lightweight open-weight LLM by Google available in 1B to 27B sizes.
Open-source code LLM family by IBM for enterprise code generation.
Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.
High-performance open-weight MoE LLM with 671B total parameters.
Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.
Open-weight code LLM trained on 2 trillion tokens of code and natural language.