A Modder Repurposed a Used V100 For LLM Acceleration
["google-gemma"]
Filtered by topic: Google Gemma Clear ✕
Tracked topic
Gemma is a family of open-weight language models released by Google for text generation and related NLP tasks.
Gemma Translator: offline DIY translation device built with Gemma 4 + Antigravity
Gemma 4 12B Demo: Native Audio Processing in Google AI Edge Eloquent
Bring the power of on-device AI to life with Google AI Edge and Gemma
Gemma 4 Hits 200M+ Downloads: 3 Amazing Local Builds
Building android apps with Gemma 4 for AI coding assistance
Google just casually disrupted the open-source AI narrative…
How to build on-device AI with Gemma 4
Build intelligent Android apps with Google's AI
Top 3 AI on Android updates for building intelligent experiences (Google I/O 2026)
Google Cloud Live: Accelerate data science and analytics with GPUs
["google-gemma"]
…Gemma 4: DiffusionGemma is built on Gemma 4, a 26-billion-parameter mixture-of-experts model that activates just 3.8 billion parameters per step, pairing a diffusion head with Google’s…
["google-gemma","lm-studio"]
…My Mac slowed down system-wide while it was running, and it didn't feel faster than running Google's regular Gemma 4 26B-A4B locally. Google warns that Apple Silicon Macs…
Google's new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM
Gemma 4 for Telephony: From Two AI Models to One – Until I Switched to Chinese
I wanted to know how fast a 26B mixture-of-experts model could run on a desktop CPU with no GPU. Got ~40 tok/s single-stream (lossless) and ~124 batched. The surprising part was the byte budget: for this model you compre…
…8 MIN READ 2026년 4월 11일 Gemma 4로 에지·온디바이스 AI 실현 — NVIDIA 전 플랫폼 완전 지원 Google Gemma 4 멀티모달·다국어 모델 패밀리가 출시됐습니다. 데이터센터의 NVIDIA Blackwell부터 에지의 Jetson까지 전 플랫폼을…
…Meet the small models that punch above their weight They're built for regular hardware Gemma 4 E2B is part of Google's Gemma 4 family that dropped at the end of…
["google-gemma","adguard-home"]
["deepseek","qwen3","alibaba","google-gemma"]
…Google's on-device optimization showing up at runtime. I noticed the difference on the first prompt itself. Where DeekSeek was giving me vague guidance and Qwen an unnecessary essay, Gemma actually…
…Slave explains that the idea is Google’s Per-Layer Embeddings used in Gemma 3n and Gemma 4. This is what the memory layout on the ESP32-S3 looks like: 1 2…