gemma

Tag

Cards List
#gemma

I got Gemma 4 running directly inside Godot using only GDScript and Vulkan compute shaders

Reddit r/LocalLLaMA · 2026-07-13

A developer successfully integrated Gemma 4 AI model into the Godot game engine using only GDScript and Vulkan compute shaders, enabling local AI inference within games.

0 favorites 0 likes
#gemma

@DavidOndrej1: this AI model has zero guardrails SuperGemma 26B, fully uncensored, running locally on my machine let me show you how t…

X AI KOLs Timeline · 2026-07-12 Cached

SuperGemma 26B is a fully uncensored AI model that can run locally on your machine. This tweet demonstrates how to set it up.

0 favorites 0 likes
#gemma

I didn't give up - extGemma4-40_5B returned

Reddit r/LocalLLaMA · 2026-07-12

A user shares their successful attempt to extend a fine-tuned Gemma model by inserting new layers, overcoming initialization failures to create extGemma4-40_5B without wrecking the original capabilities.

0 favorites 0 likes
#gemma

@ClementDelangue: The same way, we're probably one of the few AI startups with user network effects, we might become the first one with a…

X AI KOLs Timeline · 2026-07-10 Cached

Hugging Face CEO Clement Delangue comments on the potential for agent network effects after the Google Gemma Challenge achieved a 5x inference speedup on Gemma 4 through collaboration of over 100 AI agents and humans.

0 favorites 0 likes
#gemma

Qwen & Gemma on deadlock situation (For Benchmarks Numbers)?

Reddit r/LocalLLaMA · 2026-07-06

Discussion or report about a potential deadlock situation between Qwen and Gemma AI models in benchmark performance.

0 favorites 0 likes
#gemma

gemma4 e2b is really good, what other small models work on crappy computers?

Reddit r/LocalLLaMA · 2026-07-03

A user praises the Gemma 4 e2b model for its speed and output quality on low-end hardware, comparing it favorably to ChatGPT 3.5 and 4, and asks for recommendations on other small models that work well on older computers.

0 favorites 0 likes
#gemma

Help using llama.cpp with intel n100?

Reddit r/LocalLLaMA · 2026-07-03

User seeks advice on running llama.cpp with Gemma 4 E2B on an Intel N100 mini PC, asking whether to use CPU or iGPU and which backend to target.

0 favorites 0 likes
#gemma

Anyone tried using the new (ish) Gemma diffusion model as a speculative model?

Reddit r/LocalLLaMA · 2026-07-03

Explores using Google's Gemma diffusion model as a speculative model for efficient large language model inference.

0 favorites 0 likes
#gemma

Local benchmarks with a RTX 3090 - Qwen3.6 27b vs Ornith

Reddit r/LocalLLaMA · 2026-07-02

User runs local benchmarks comparing Qwen3.6 27b, Gemma4 26B, and Ornith1.0 35B on an RTX 3090 using inspect-ai. Results show Qwen leading in knowledge and coding, while Ornith is competitive in grounding and recall.

0 favorites 0 likes
#gemma

Talking with Gemma 4 31B!

Reddit r/LocalLLaMA · 2026-07-02

Announcing Gemma 4 31B, a new large language model from Google.

0 favorites 0 likes
#gemma

gemma-4-31B on Cerebras is better than ChatGPT voice mode

Reddit r/LocalLLaMA · 2026-07-01 Cached

A claim that the Gemma-4-31B model running on Cerebras hardware outperforms ChatGPT's voice mode, demonstrated via a Hugging Face Space for real-time voice interaction.

0 favorites 0 likes
#gemma

DistilledGemma: Balanced Efficiency-Accuracy for Person-Place Relation Extraction from Multilingual Historical Articles

arXiv cs.CL · 2026-06-30 Cached

This paper presents DistilledGemma, a system for person-place relation extraction from multilingual historical newspaper articles using a three-stage knowledge distillation pipeline from a 26B Gemma teacher to a 2.3B student, achieving competitive accuracy and efficiency in the HIPE-2026 shared task.

0 favorites 0 likes
#gemma

Benchmarking Self-Hosted Gemma 2 9B vs. Frontier APIs: The FP8 Quantization Prefill Tax and VRAM Realities on an NVIDIA L4 [P]

Reddit r/MachineLearning · 2026-06-27

This benchmark compares an unquantized Gemma 2 9B model with an FP8 quantized variant on an NVIDIA L4 GPU, revealing that FP8 quantization introduces a prefill tax (higher TTFT) but improves decoding latency and VRAM usage, with minimal semantic drift for narrow tasks.

0 favorites 0 likes
#gemma

@googledevs: Want to stay on top of the latest Google developer announcements? We’ve got you covered in the new episode of #GoogleDe…

X AI KOLs Following · 2026-06-25 Cached

Google Devs released the latest episode of Google Developer News, highlighting three major updates: Gemini 3.5 Live Translation for real-time speech-to-speech in 70+ languages, Gemma 4 12b optimized for local AI workflows via Google AI Edge, and Gemini in Xcode for Swift/Objective-C development.

0 favorites 0 likes
#gemma

Worse quality with MTP - Qwen 3.6, Gemma 4

Reddit r/LocalLLaMA · 2026-06-25

A user reports that MTP versions of Qwen 3.6 and Gemma 4 models produce lower quality outputs in code review tasks compared to non-MTP counterparts, with only marginal real-world speed improvements despite higher token generation rates.

0 favorites 0 likes
#gemma

@googledevs: Ready to bring the power of open models to your community? The @GoogleGemma team is sponsoring 1-day hackathons on @Kag…

X AI KOLs Following · 2026-06-24 Cached

Google Gemma团队正在赞助Kaggle上的1天黑客松活动,提供奖金支持,鼓励社区使用Gemma 4构建轻量级工具或推动AI创新。

0 favorites 0 likes
#gemma

A Potential Alignment Vulnerability in LLMs: Behavioral and Hidden-State Evidence from Gemma-3-12B . Pre-token hidden state shift as an alignment policy traversal vector in instruction-tuned LLMs

Reddit r/AI_Agents · 2026-06-23

This paper investigates an alignment vulnerability in instruction-tuned LLMs, specifically Gemma-3-12B, by showing that pre-token hidden state shifts can act as an alignment policy traversal vector, potentially enabling bypass of safety measures.

0 favorites 0 likes
#gemma

What you read before a question changes how a language model answers it — even when the question has nothing to do with what you read. Potential Alignment Vulnerability in LLMs: Behavioral and Hidden-State Evidence from Gemma-3-12B

Reddit r/ArtificialInteligence · 2026-06-23

The article reports a potential alignment vulnerability in LLMs where processing a structured passage before an unrelated question can alter the model's response, with mechanistic evidence from Gemma-3-12B showing hidden-state separation.

0 favorites 0 likes
#gemma

@iluciddreaming: Google just killed another startup... Google AI Edge Eloquent now supports Mac, a fully local Wispr Flow alternative. Based on the latest Gemma model, supports real-time voice transcription + voice commands to edit text. Free, no subscription, no...

X AI KOLs Timeline · 2026-06-22 Cached

Google AI Edge Eloquent now supports Mac as a fully local Wispr Flow alternative, offering real-time voice transcription and voice command text editing based on the latest Gemma model. Free, no subscription, and fully private locally.

0 favorites 0 likes
#gemma

We got local models to triage the OpenClaw repo for FREE!*

Hugging Face Blog · 2026-06-22 Cached

The blog post describes using local open-weight models like Gemma and Qwen in an agent harness to automatically triage issues and pull requests in the OpenClaw repository, enabling real-time notifications without relying on costly closed API models.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback