encoder-free

Tag

Cards List
#encoder-free

@Marktechpost: Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With a Trained Effort Dial. No RoP…

X AI KOLs Timeline · 2026-07-16 Cached

Thinking Machines Lab releases Inkling, a 975B-parameter open-weights multimodal MoE model with 41B active parameters and a controllable thinking effort feature. It uses an encoder-free approach for multimodality and achieves state-of-the-art on several benchmarks.

0 favorites 0 likes
#encoder-free

@andimarafioti: Can a VLM see without a vision encoder? We trained one for $100, inspired by Gemma 4 12B. Latency on an M3 Pro MacBook:…

X AI KOLs Timeline · 2026-06-18 Cached

Researchers trained a vision-language model without a vision encoder for only $100, inspired by Gemma 4 12B, achieving a 30% reduction in end-to-end latency on an M3 Pro MacBook.

0 favorites 0 likes
#encoder-free

Gemma 4 12B native encoder free voice input utilization suggest?

Reddit r/LocalLLaMA · 2026-06-14

Discusses leveraging Gemma 4 12B's encoder-free architecture for native voice input, seeking out-of-the-box solutions for low-latency streaming audio ingestion.

0 favorites 0 likes
#encoder-free

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

Google DeepMind Blog · 2026-06-09 Cached

Google DeepMind announces Gemma 4 12B, a novel encoder-free multimodal AI model that integrates vision and audio directly into the LLM backbone, delivering advanced reasoning and agentic capabilities on laptops with 16GB of RAM, released under Apache 2.0 license.

0 favorites 0 likes
#encoder-free

Gemma 2B multimodal model matches larger models without encoder

Reddit r/singularity · 2026-06-04

Google's Gemma 4 12B introduces an encoder-free multimodal architecture that competes with larger models, though benchmark comparisons show it trailing Qwen 2.5 9B on most tasks. The article also covers related developments including open-weight model security risks, Uber's Claude Code spending caps, and NeurIPS's misuse of an uncalibrated AI detector.

0 favorites 0 likes
#encoder-free

@analogalok: i just ran Google's brand new Unsloth Gemma4 12B dense GGUF on my RTX 4060 using llama.cpp + CUDA 13.2 21 tokens per se…

X AI KOLs Timeline · 2026-06-03 Cached

Google's new Gemma 4 12B is a single decoder-only transformer with encoder-free multimodal input, achieving strong benchmarks while being small enough to run locally on a budget GPU. It is released under Apache 2.0 license.

0 favorites 0 likes
#encoder-free

@mtschannen: For the past years my research focus was on unifying models and training paradigms across modalities. Today I'm excited…

X AI KOLs Timeline · 2026-06-03 Cached

Google DeepMind researcher announces the release of Gemma 4 12B, a dense encoder-free model that processes text, image, and audio inputs, continuing work on unifying models across modalities.

0 favorites 0 likes
#encoder-free

Google Gemma 4 12B

Product Hunt · 2026-06-03

Google's Gemma 4 12B model enables local multimodal AI using an encoder-free architecture.

0 favorites 0 likes
#encoder-free

@googleaidevs: We’re launching Gemma 4 12B: Our unified, encoder-free model that brings powerful multimodal intelligence straight to y…

X AI KOLs Timeline · 2026-06-03 Cached

Google launches Gemma 4 12B, an encoder-free multimodal model with native audio support, optimized for local execution on laptops under Apache 2.0.

0 favorites 0 likes
← Back to home

Submit Feedback