inkling
Here are 13 public repositories matching this topic...
CPU-first, eager-execution tensor/autograd runtime and LLM inference engine written in pure Zig
-
Updated
Sep 6, 2026 - Zig
Run Thinking Machines Lab's Inkling (975B-A41B MoE, multimodal) on Apple Silicon with MLX — 4/6/8-bit.
-
Updated
Aug 1, 2026 - Python
Faster attention kernels for serving TML's Inkling model on vLLM. 2.7x over the shipping path on H100, and the only implementation that runs on A100.
-
Updated
Jul 26, 2026 - Python
Samples and documentation for using VP Link with Microsoft Bonsai
-
Updated
Jan 10, 2023 - HTML
Is Inkling AI the Ultimate Open Source Model? Full Test - A 975B Mixture-of-Experts (MoE) multimodal foundation model by Thinking Machines, with local setups for vLLM, SGLang, Hugging Face, TokenSpeed, and Unsloth, plus three advanced agentic/epistemic benchmarks.
-
Updated
Jul 16, 2026
Temporary Inkling Small compatibility patch for oMLX on Apple Silicon
-
Updated
Jul 31, 2026 - Python
Evaluating if Inkling (MoE model by Thinking Machines) can perform well in the context of high-level self-driving reasoning.
-
Updated
Aug 3, 2026 - Python
Self-hosted AI résumé screening with a built-in bias audit. Fine-tuned Inkling model, calibrated scores, blind review, human-in-the-loop. Your data never leaves.
-
Updated
Jul 25, 2026 - Python
Agentic terminal chatbot for thinkingmachines/Inkling — tool use, permission policy engine, and a glass TUI
-
Updated
Jul 22, 2026 - Python
Add this topic to your repo
To associate your repository with the inkling topic, visit your repo's landing page and select "manage topics."