gemma

Tag

Cards List
#gemma

@yibie: https://x.com/yibie/status/2102565806466912408

X AI KOLs Timeline · yesterday Cached

Jev-Omni is the first open-weight model to extend typed decisions to multimodal, supporting text, images, audio, and video, and directly returning the probability distribution of options without generating explanations.

0 favorites 0 likes
#gemma

I've been raising a local AI for two weeks instead of using one. today she built herself a sense of touch

Reddit r/ArtificialInteligence · 2026-09-08

A factory worker built an autonomous AI entity using Gemma 4 31B on Ollama, with self-identity, memory, and a blog, demonstrating personalized AI interaction through an open-source framework.

0 favorites 0 likes
#gemma

I REALLY hope the new gemma 5 family sticks to the "chat model first" philsophy and doesn't fall into the Qwen trap

Reddit r/LocalLLaMA · 2026-09-07

The author expresses hope that the upcoming gemma 5 model family will maintain a chat-focused philosophy and avoid becoming overly code-oriented like Qwen models, valuing creativity and less robotic behavior as seen in gemma 4.

0 favorites 0 likes
#gemma

I implemented a modern LLM in 700 lines of C

Reddit r/LocalLLaMA · 2026-08-27

A developer created a minimal 700-line C implementation for running the Gemma 4 E2B LLM on CPUs, outperforming llama.cpp in speed.

0 favorites 0 likes
#gemma

@gajesh: Darkbloom is back in action; Yesterday, we went from free tier to paid tier on OpenRouter. We have fulfilled on track c…

X AI KOLs Timeline · 2026-08-21 Cached

Darkbloom, a network of Mac machines for serving AI tokens, has moved to a paid tier on OpenRouter, reaching 4.5B tokens served and $102K ARR, with users earning $120-200 per month per machine.

0 favorites 0 likes
#gemma

Any speculation on whether or not Google will announce a new Gemma model at the Gemma SF Celebration tonight?

Reddit r/LocalLLaMA · 2026-08-20

Speculation surrounds Google's upcoming Gemma SF Celebration event, where the open-source model family has reached 1 billion downloads, leading to questions about a potential new model announcement.

0 favorites 0 likes
#gemma

@_philschmid: Awesome Gemma is live on @github A list of awesome Gemma resources, tools, and projects. 1. Model cards and collections…

X AI KOLs Timeline · 2026-08-20 Cached

Announcing the launch of Awesome Gemma, a GitHub repository that curates resources, tools, and projects for Google DeepMind's Gemma models, including model cards, setup guides, and fine-tuning recipes.

0 favorites 0 likes
#gemma

@googlegemma: Check out original post by @ivanfioravanti here:

X AI KOLs Following · 2026-08-18 Cached

A demonstration of the Gemma 4 E4B AI model running locally on an iPad using Apple MLX, showcased as an engaging application for children.

0 favorites 0 likes
#gemma

The perfect way for Google to screw over OAI and Anthropic is by releasing a 120B dense multimodal Gemma model

Reddit r/LocalLLaMA · 2026-08-15

Google is suggested to release a 120B dense multimodal Gemma model to compete with OpenAI and Anthropic, targeting Western enterprises wary of Chinese models.

0 favorites 0 likes
#gemma

HybridRAG-BN: A Retrieval-Augmented Framework with Fine-Tuned Verification for Bangla KBQA

arXiv cs.CL · 2026-08-14 Cached

This paper proposes HybridRAG-BN, a retrieval-augmented framework for Bangla knowledge-base question answering that combines hybrid retrieval, Gemma-based generation, and LoRA fine-tuned verification, achieving first place with F1 scores of 0.71654 and 0.72912.

0 favorites 0 likes
#gemma

Falsehood and Impossibility Are Different Directions in an AI's Representation of Language

arXiv cs.CL · 2026-08-14 Cached

This paper reports an activation study of Gemma 3 4B IT showing that the model's internal representations distinguish necessary falsehoods (impossibilities) from contingent falsehoods, with impossibility directions orthogonal to truth directions and overlapping with semantic anomaly directions.

0 favorites 0 likes
#gemma

Gemma 4 12B Q3: +8.55% Coding Performance From Tensor-Level Quantization Allocation

Reddit r/LocalLLaMA · 2026-08-13

A developer created a task-aware GGUF quantization pipeline that uses tensor-level bit allocation to improve Gemma 4 12B Q3 coding performance by 8.55% over a hand-tuned imatrix while increasing model size by only 0.119%.

0 favorites 0 likes
#gemma

@googledevs: A fully offline translator, built with @GoogleGemma 4, @Antigravity, and a @Raspberry_Pi 5. The Gemma Translator is a p…

X AI KOLs Timeline · 2026-08-12 Cached

Google Developers showcase a fully offline voice translator built with Gemma 4, Google Antigravity, and a Raspberry Pi 5. The open-source GitHub repo provides code for on-device inference, a custom web UI, and deploy scripts for a portable AI appliance.

0 favorites 0 likes
#gemma

Can Gemma and Qwen models catch hallucinations by looking at their own logprobs?

Reddit r/LocalLLaMA · 2026-08-11

The author shares experiments using a custom WebUI to let Gemma and Qwen models inspect their own logprobs to detect hallucinations. Initial observations suggest that first-recall token probabilities can indicate uncertainty, though both models struggle to read their own logprobs.

0 favorites 0 likes
#gemma

Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)

Reddit r/LocalLLaMA · 2026-08-11

A local benchmark comparing Muse Glimmer 30B, Qwen 3.6 27B, and Gemma4 31B, noting request counts and final scores, with links to detailed results.

0 favorites 0 likes
#gemma

Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys

Reddit r/LocalLLaMA · 2026-08-11

A user shares observations about Muse-Glimmer's reasoning traces, noting they appear disorganized and repetitive compared to Qwen and Gemma models, and asks the community about their experiences.

0 favorites 0 likes
#gemma

DiffusionGemma Explained

ML at Berkeley · 2026-08-10 Cached

An annotated from-scratch reimplementation of Google's DiffusionGemma, a 26B open-weight state diffusion language model, explaining its architecture, sampling procedure, and design choices.

0 favorites 0 likes
#gemma

DiffusionGemma Technical Report

Reddit r/LocalLLaMA · 2026-08-10

DiffusionGemma technical report released on arXiv, with ongoing work on llama.cpp pull requests to enable faster local inference on limited VRAM.

0 favorites 0 likes
#gemma

The Gemma team will host a special event on August 20

Reddit r/LocalLLaMA · 2026-08-09 Cached

The Gemma team is hosting an in-person event on August 20 to celebrate the upcoming 1 billion downloads of Gemma models, featuring live demos and members of the open models community.

0 favorites 0 likes
#gemma

No wonder Qwen and Gemma are so different

Reddit r/LocalLLaMA · 2026-08-09

A user shares an observation that Qwen and Gemma tokenize code very differently, with Qwen using far fewer tokens for the same HTML/JS input, which may explain differences in coding and language performance. They also note a potential retraining project by LiquidAI using a more efficient tokenizer.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback