gemma4
Here are 398 public repositories matching this topic...
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
-
Updated
Sep 21, 2026 - Swift
A powerful Zotero AI and MCP plugin with ChatGPT, Gemini 3.7, Claude Fable 5, Claude Opus 5, DeepSeek V4, Grok, OpenRouter, Kimi k3, GLM 5.3, SiliconFlow, GPT-oss, Gemma 4, Qwen 3.8
-
Updated
Sep 21, 2026 - JavaScript
StenographAI is the secure privacy-first AI notepad & notetaker for confidential conversations in government & defence sectors. On Windows & MacOS.
-
Updated
Sep 15, 2026 - Python
OpenClaw alternative in your pocket
-
Updated
Sep 18, 2026 - Kotlin
Open-source AI browser agent for Chrome and Firefox (monorepo) 🧠
-
Updated
Sep 21, 2026 - JavaScript
PokeClaw (PocketClaw) — first on-device AI that controls your Android phone. Gemma 4, no cloud, no API key. Poke is short for Pocket.
-
Updated
Jun 2, 2026 - Kotlin
Gemma Gem runs Google's Gemma 4 model entirely on-device via WebGPU — no API keys, no cloud, no data leaving your machine.
-
Updated
May 29, 2026 - TypeScript
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
-
Updated
Sep 21, 2026 - Python
Local LLM, image&video&music generator, vibecode like cursor with local models on your phone
-
Updated
Sep 19, 2026 - C++
An open-source Cotypist with macOS system wide AI autocomplete
-
Updated
Aug 29, 2026 - Swift
A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based chatbot interface, and Ollama/OpenAI-compatible HTTP APIs for programmatic access. It supports Windows/MacOS/iOS/Linux with full GPU capability
-
Updated
Sep 21, 2026 - C#
Downloadable models and conversion recipes for Apple's Core AI on iPhone and Mac. Chat, vision, speech and generative models with per-model validation records, Swift examples through CoreAIKit, and a downloadable Mac app.
-
Updated
Sep 21, 2026 - Python
Run local LLMs like Gemma, Qwen, and LLaMA on Android for offline, private, real-time chat and question answering with LiteRT and ONNX Runtime.
-
Updated
Sep 20, 2026 - Kotlin
llama.cpp fork with TurboQuant WHT-rotated KV cache & weight compression + Gemma 4 MTP and Qwen 3.6 NextN speculative decoding (+30-50% throughput).
-
Updated
Sep 10, 2026 - C++
Private on-device AI chat for Android — runs any GGUF model locally via llama.cpp with ARM-optimised SIMD. Zero network permissions, encrypted settings, biometric lock, tamper detection. + GPU Acceleration
-
Updated
Aug 18, 2026 - Kotlin
This is end to end course on AI Agents and Agentic AI with 15+ AI Agent Projects with real time use cases and industry expertise.
-
Updated
Apr 17, 2026 - Jupyter Notebook
Agentic ✧ Gemma Inference for Android System Intelligence
-
Updated
Sep 21, 2026 - Kotlin
Real-time multimodal AI pipelines on Apple Silicon
-
Updated
Sep 19, 2026 - C++
Add this topic to your repo
To associate your repository with the gemma4 topic, visit your repo's landing page and select "manage topics."