inkling
Here are 11 public repositories matching this topic...
CPU-first, eager-execution tensor/autograd runtime and LLM inference engine written in pure Zig
-
Updated
Jul 26, 2026 - Zig
Samples and documentation for using VP Link with Microsoft Bonsai
-
Updated
Jan 10, 2023 - HTML
Faster attention kernels for serving TML's Inkling model on vLLM. 2.7x over the shipping path on H100, and the only implementation that runs on A100.
-
Updated
Jul 26, 2026 - Python
Run Thinking Machines Lab's Inkling (975B-A41B MoE, multimodal) on Apple Silicon with MLX — 4/6/8-bit.
-
Updated
Jul 17, 2026 - Python
Self-hosted AI résumé screening with a built-in bias audit. Fine-tuned Inkling model, calibrated scores, blind review, human-in-the-loop. Your data never leaves.
-
Updated
Jul 25, 2026 - Python
Is Inkling AI the Ultimate Open Source Model? Full Test - A 975B Mixture-of-Experts (MoE) multimodal foundation model by Thinking Machines, with local setups for vLLM, SGLang, Hugging Face, TokenSpeed, and Unsloth, plus three advanced agentic/epistemic benchmarks.
-
Updated
Jul 16, 2026
Agentic terminal chatbot for thinkingmachines/Inkling — tool use, permission policy engine, and a glass TUI
-
Updated
Jul 22, 2026 - Python
Improve this page
Add a description, image, and links to the inkling topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the inkling topic, visit your repo's landing page and select "manage topics."