← Explore
TOPIC

#gemma

Open source repositories tagged with #gemma, ranked by health score.

xorbitsai
xorbitsai/inference
Python
89
health

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

9.5k
avifenesh
avifenesh/memra
Rust
89
health

Rust + CUDA inference engine for NVIDIA RTX PRO 6000 Blackwell and RTX 5090. Serves safetensors and GGUF over an OpenAI-compatible API, with per-device tuned defaults and speculative decode gated byte-identical to plain decode. Hosted instance: inference.tiyuvta.ai

321
zhongkaifu
zhongkaifu/TensorSharp
C#
89
health

A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based chatbot interface, and Ollama/OpenAI-compatible HTTP APIs for programmatic access. It supports Windows/MacOS/Linux with full GPU capability

376
drumih
drumih/turbo-fieldfare
Swift
88
health

Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook

6.3k