deepseek

Tag

Cards List
#deepseek

@joey00072fp4: deepseek v4.1 flash, two stage decoder arch, 20+20=40 layers first 20 layers build global kv of csa2 and swa, (called e…

X AI KOLs Timeline · 2026-09-10 Cached

DeepSeek-V4.1-Flash introduces a two-stage decoder architecture with 40 layers, activating only 8B parameters during prefill and 16B during decode, and includes 196B Engram memory for significant efficiency gains over previous versions.

0 favorites 0 likes
#deepseek

Deepseek V4.1 Flash Release Video [Made with Deepseek V4.1 Flash]

Reddit r/LocalLLaMA · 2026-09-10

The author benchmarks Deepseek V4.1 Flash on motion video generation, finding it has improved to nearly match Opus class models compared to earlier versions like Kimi K3.

0 favorites 0 likes
#deepseek

@tianyi: DeepSeek has open-sourced some new code repositories, making it easier to deploy V4.1 Flash as well as subsequent open-…

X AI KOLs Following · 2026-09-10 Cached

DeepSeek has open-sourced new code repositories, including libraries and tools, to facilitate the deployment of V4.1 Flash and subsequent open-source models.

0 favorites 0 likes
#deepseek

Deepseek V4.1 Flash is 748B, not 552B

Reddit r/LocalLLaMA · 2026-09-10

The article clarifies the parameter count of the Deepseek V4.1 Flash AI model, detailing its components like FFN experts and vision encoder, and highlights the substantial hardware requirements for deployment.

0 favorites 0 likes
#deepseek

@jakevin7: Holy shit, that's impressive! DeepSeek, is this really just a minor tweak to a small version??? V4.1 Flash made a ton o…

X AI KOLs Timeline · 2026-09-10 Cached

DeepSeek V4.1 Flash introduces architectural improvements for long context handling, including causal encoder-decoder, CSA2, hierarchical sparse indexer, and quantization, leading to major memory and compute optimizations.

0 favorites 0 likes
#deepseek

DeepSeek V4.1 Flash is getting surprisingly close to GPT-5.6 Sol territory, while being absurdly cheap

Reddit r/singularity · 2026-09-10

DeepSeek has released V4.1 Flash, a 552B MoE model with efficient active parameters, achieving performance close to GPT-5.6 Sol at a much lower cost.

0 favorites 0 likes
#deepseek

Deepseek v4.1 flash finally has engrams, what do you expect from 4.1 pro?

Reddit r/LocalLLaMA · 2026-09-10

Speculation on Deepseek v4.1 pro's specifications and future developments, including parameter counts and engram technology.

0 favorites 0 likes
#deepseek

What's the next big breakthrough after attention mechanism? My bet is not on Engrams.

Reddit r/LocalLLaMA · 2026-09-10

The article speculates that the next major breakthrough after the attention mechanism may involve AI architectures with input-dependent weights, potentially building on DeepSeek's Engram mechanism.

0 favorites 0 likes
#deepseek

Deepseek v4.1 Flash reaches 98% of Astra’s score at 1.4% of cost on OpenDesign Arena

Reddit r/singularity · 2026-09-09 Cached

DeepSeek V4.1 Flash achieves 98% of GPT-6 Astra's design score at only 1.4% of the cost, as reported in OpenDesign Arena benchmarks.

0 favorites 0 likes
#deepseek

DeepSeek launching v4.1 flash cheaper and more capable than v4 pro

Hacker News Top · 2026-09-09 Cached

DeepSeek plans to officially release the V4.1 Flash model around September 10, 2026, which surpasses V4 Pro in performance, cost, and speed, with adjusted pricing for off-peak and peak hours.

0 favorites 0 likes
#deepseek

Deepseek Has Soft Retired Deepseek V4 Pro

Reddit r/LocalLLaMA · 2026-09-09

Deepseek has soft retired its V4 Pro model, indicating that the AI model is being phased out or deprecated.

0 favorites 0 likes
#deepseek

Tested DeepSeek V4 vs V4.1 Flash Vision Beta in 5 visual tests

Reddit r/ArtificialInteligence · 2026-09-08

The article compares DeepSeek V4 and V4.1 Flash Vision Beta through 5 visual tests, highlighting significant improvements in reliability and lower API pricing.

0 favorites 0 likes
#deepseek

@HuggingModels: Ever seen an AI that reads images AND writes text? DeepSeek-V4-Flash-Vision-Exp does exactly that. It's a vision-langua…

X AI KOLs Timeline · 2026-09-08

DeepSeek-V4-Flash-Vision-Exp is a vision-language model that processes images and generates text, with over 313k downloads, useful for tasks like image description and visual question answering.

0 favorites 0 likes
#deepseek

@LotusDecoder: DeepSeek-V4.1-Flash-0910 decode 400 token/s 😋 Would deploying this to my home DGX spark also achieve this speed?

X AI KOLs Timeline · 2026-09-08

A user on X/Twitter asks if deploying the DeepSeek-V4.1-Flash-0910 model on a home DGX Spark could achieve a decode speed of 400 tokens per second.

0 favorites 0 likes
#deepseek

@Fenng: Who said there are no opportunities in the AI era? Isn't Teacher Cui hiring on a large scale right now? It's a shame I …

X AI KOLs Timeline · 2026-09-08 Cached

The article discusses AI industry job opportunities, specifically highlighting DeepSeek's large-scale hiring for backend and server engineers.

0 favorites 0 likes
#deepseek

DeepSeek-V4-Flash-Vision-Exp is amazing at creating game worlds!

Reddit r/LocalLLaMA · 2026-09-07

The DeepSeek-V4-Flash-Vision-Exp model was used to create a compelling game world in about two days, leveraging its vision capabilities for tasks like generating models, fixing glitches, and play-testing.

0 favorites 0 likes
#deepseek

@sgl_project: We just added recipes for DeepSeek-V4-Flash-Vision & DeepSeek-V4-Flash-0731 on 2x DGX Spark. https://docs.sglang.io/coo…

X AI KOLs Timeline · 2026-09-03 Cached

SGLang has added deployment recipes for DeepSeek-V4-Flash-Vision and DeepSeek-V4-Flash-0731 models on 2x DGX Spark hardware, with support for various configurations and optimizations.

0 favorites 0 likes
#deepseek

@tianyi: Looks very useful. Please also try the creator mode in DeepSeek Harness, which has been open source in MIT license for …

X AI KOLs Timeline · 2026-09-03 Cached

The tweet recommends the creator mode in DeepSeek Harness, an open source tool under MIT license, and references ClaudeCode's Function Hooks for extending Claude Code.

0 favorites 0 likes
#deepseek

Research on AI Harnesses

Reddit r/AI_Agents · 2026-09-03

The release of DeepSeek Harness has sparked debate on the importance of AI harnesses versus models, with the author highlighting the lack of scientific evidence and calling for research and benchmarks to define what makes a good harness.

0 favorites 0 likes
#deepseek

@ViC305: I DID IT!! DeepSeek-V4-Flash-Vision EXL3 MixedK is now running VISION + DSpark speculative decoding together on ONE DGX…

X AI KOLs Timeline · 2026-09-03 Cached

User @ViC305 successfully runs DeepSeek-V4-Flash-Vision with EXL3 MixedK and DSpark speculative decoding on a single DGX Spark, achieving improved performance and fixing technical issues for multimodal AI deployment.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback