π¦
Rai
CPU-only LLM inference engine in pure Rust β 4-bit quantized models, hand-written AVX2 kernels, speculative decoding, and a local HTTP/MCP server. No GPU, no Python runtime.
0 installs
Trust: 34 β Low
Ai
Ask AI about Rai
Powered by Claude Β· Grounded in docs
I know everything about Rai. Ask me about installation, configuration, usage, or troubleshooting.
0/500
Loading tools...
Reviews
Documentation
No README available
