Pilot it5,737 · 552 forks

Rust WebGPU inference engine

Worth a timeboxed spike before you bet on it.

Who it's for

Small teams with GPU-based ML needs

What it replaces

Paid inference services or Python-based engines

The catch

Immature project with limited community support

Your first hour

Evaluate shimmy's compatibility with existing ML models

The numbers

Stars5,737
Forks552
Stars added (7d)measuring…
Open issues11
LanguageRust
LicenceApache-2.0
Last pushLast push 22 days ago
Project age1.0 years old

Maintainers describe it as: ⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.

api-servercommand-line-tooldeveloper-toolsggufhuggingfacehuggingface-modelshuggingface-transformersinference-serverllamallamacpp

shimmy, in short

Should a small team use shimmy?
Pilot it. Worth a timeboxed spike before you bet on it. Small teams with GPU-based ML needs
What does shimmy actually do?
Rust WebGPU inference engine
What does shimmy replace?
Paid inference services or Python-based engines
What is the downside of shimmy?
Immature project with limited community support
Can shimmy be used in a commercial product?
Its licence is Apache-2.0, which is permissive and generally fine for commercial use. Confirm against the LICENSE file in the repository.

Weighed against

Which of these actually matters to your company?

Tell us what you build and we will screen the week's open-source moves and the week's research against it — and say which ones are worth your time. One email, Monday, free.

Or run a free brief on your own company right now — takes about 30 seconds, no signup.

Stars, forks, licence and last-push data from the public GitHub API, refreshed August 11, 2026. The verdict is NoizeOff's editorial opinion for a team of 2–20, not advice from the project's maintainers, and not legal advice on licensing. We are not affiliated with Michael-A-Kuykendall.

Adoption Radar · Company briefs · Home