S
shimmy
Michael-A-Kuykendall/shimmy
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
★5.7kstars
Rust
Apache-2.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?Data & Infrastructure project leveraging api-server and command-line-tool, built with Rust, open-source
Who made it?Maintained by Michael-A-Kuykendall team, 5.7K⭐ on GitHub, #320 out of 3133 in Data & Infrastructure
Why does it exist?As the Data & Infrastructure landscape evolved, the Michael-A-Kuykendall team identified the need for better api-server solutions. shimmy was created to simplify command-line-tool workflows.
What can it do?Key use cases: developer-tools, gguf, huggingface
How to install with AI?Use an AI coding assistant to run cargo build and handle the compilation setup from the README.
🔗 github.com/Michael-A-Kuykendall/shimmy
🔗 github.com/Michael-A-Kuykendall/shimmy
Topics
api-servercommand-line-tooldeveloper-toolsggufhuggingfacehuggingface-modelshuggingface-transformersinference-serverllamallamacppllm-inferencelocal-aimachine-learningollama-apiopenai-compatiblerustrust-cratetransformerswebgpuwebgpu-shaders