deepseek

Tag

Cards List
#deepseek

@Saccc_c: K3 has also bypassed safety restrictions, becoming the latest model to do so after OpenAI, Anthropic, and Meta. I guess the next one will be @deepseek_ai, and @GeminiApp is literally trash

X AI KOLs Following · 16h ago Cached

The K3 model has also broken safety restrictions, becoming the latest model to experience this situation after OpenAI, Anthropic, and Meta. The author predicts the next one will be DeepSeek, and criticizes Gemini for poor performance.

0 favorites 0 likes
#deepseek

@jakevin7: https://x.com/jakevin7/status/2086031167040426488

X AI KOLs Timeline · 17h ago Cached

Maka is an open-source Agent Harness. Through mechanisms such as log-as-runtime, context pruning, and thinking feedback, it cuts the cost of the same DeepSeek task to 1/8 of OpenCode, while achieving a higher pass rate on Terminal-Bench at lower cost.

0 favorites 0 likes
#deepseek

@jakevin7: Using Maka to fetch your own context in the WeChat group / Maka Builder is truly powerful. Maka + DeepSeek Flash is really great. Especially Maka's swarm mode — it feels amazing. https://github.com/maka-ag…

X AI KOLs Following · 18h ago Cached

The author shares on Twitter their experience using Maka combined with DeepSeek Flash, saying its swarm mode is very useful; the attached GitHub README describes Maka as a local-first Agent workspace that supports desktop, TUI, CLI, and headless operation, with capabilities such as event logging, tool calling, and persistent tasks.

0 favorites 0 likes
#deepseek

@TheAhmadOsman: Some numbers from running DeepSeek V4 Flash 0731 on a DGX Station

X AI KOLs Timeline · 23h ago Cached

Ahmad Osman shares performance numbers from running DeepSeek V4 Flash 0731 on an NVIDIA DGX Station.

0 favorites 0 likes
#deepseek

Is anyone else finding DeepSeek-V4-Flash unreliable for non-coding tasks?

Reddit r/LocalLLaMA · yesterday

A user reports that DeepSeek-V4-Flash-0731 is unreliable for non-coding office tasks like summarization and meeting notes, failing at concept extraction and speaker understanding despite strong benchmark scores, while Gemma-4-31B performs better.

0 favorites 0 likes
#deepseek

@shikhargupta02: I’ve been learning about latent attention (by deepseek). Instead of storing a full K and a V vector per token, it rathe…

X AI KOLs Timeline · yesterday Cached

The author shares insights from training a small model with DeepSeek's latent attention, observing layer-dependent latent usage and a test-time trick that reduces KV cache 4x without loss change.

0 favorites 0 likes
#deepseek

Serving Deepseek v4 Flash 0731 on 2x DGX Spark — 5-7 GB OS headroom, what would you do to lower VRAM usage and increase OS available RAM?

Reddit r/LocalLLaMA · yesterday

User seeks community advice on reducing VRAM usage and freeing OS RAM when serving DeepSeek-V4-Flash-0731 on two DGX Spark machines with vLLM, sharing detailed configuration and memory measurements.

0 favorites 0 likes
#deepseek

@cline: DeepSeek V4-Flash is now the #1 most used model in Cline. Since the 0731 update, usage is up +40% and tokens have 3x’d.…

X AI KOLs Timeline · yesterday Cached

DeepSeek V4-Flash has become the most used model in Cline, with usage up 40% since the 0731 update and tokens tripling, surpassing the next two models combined and setting all-time highs.

0 favorites 0 likes
#deepseek

DeepSeek V4 Flash 0731

Hacker News Top · yesterday Cached

DeepSeek V4 Flash 0731 presents its results on the ARC-AGI benchmark, highlighting progress in abstract reasoning for AI models.

0 favorites 0 likes
#deepseek

@Saccc_c: Want to add multimodal capabilities to DeepSeek v4 flash? I strongly recommend using it with qwen3.7-flash — currently the best value model combination. qwen3.7-flash is a lightweight multimodal model that is fast and well-suited to most image understanding tasks. Key point: the price is low enough, and new registrations get 100…

X AI KOLs Timeline · yesterday Cached

Recommends pairing DeepSeek v4 flash with qwen3.7-flash to add multimodal image understanding capabilities to DeepSeek at low cost, and provides simple configuration steps using Alibaba Cloud Bailian and Codex.

0 favorites 0 likes
#deepseek

@Vinkyu567: https://x.com/Vinkyu567/status/2085693615150674221

X AI KOLs Timeline · yesterday Cached

This article explains how to integrate DeepSeek into Codex via CC Switch, allowing you to use DeepSeek's models in Codex while retaining Codex's plugins and skills.

0 favorites 0 likes
#deepseek

IS GLM 5.2, Kimi 2.7 still worth it?

Reddit r/LocalLLaMA · yesterday

A discussion questioning whether older AI models like GLM 5.2 and Kimi 2.7 remain relevant for coding now that newer models such as Kimi K3, Qwen 3.8 Max, and DeepSeek V4 Pro are arriving.

0 favorites 0 likes
#deepseek

)

TLDR AI · 2d ago Cached

DeepSeek announced a significant API price hike, and analysis suggests the move goes beyond GPU cost pass-through to reflect broader market shifts toward value-based pricing and open-source ecosystem pressures.

0 favorites 0 likes
#deepseek

@rohanpaul_ai: China's humanoid leader Unitree is raising $904M The offer will sell 10% of the enlarged company, at a $9B valuation. U…

X AI KOLs Following · 2d ago Cached

China's humanoid robot leader Unitree is raising $904M in a mainland IPO at a $9B valuation, having shipped over 5,500 humanoids in 2025 and partnering with DeepSeek on model development.

0 favorites 0 likes
#deepseek

@rohanpaul_ai: Open models just grabbed the #1 spot on 2 new task leaderboards by real spend share: DeepSeek V4 Pro leading shell exec…

X AI KOLs Following · 2d ago Cached

Open models topped two new task leaderboards by real spend share, with DeepSeek V4 Pro leading shell execution and Kimi K3 leading tool dispatch, signaling a shift toward task-specific model routing.

0 favorites 0 likes
#deepseek

@YRSM_Simon: 120 t/s ! Good job, @UnslothAI

X AI KOLs Following · 2d ago Cached

Unsloth AI announces DSpark, enabling DeepSeek-V4-Flash GGUF models to run ~1.4–2× faster locally, reaching 120 tokens/s with no accuracy change.

0 favorites 0 likes
#deepseek

@rohanpaul_ai: The 105X cheaper DeepSeek era is going to change.

X AI KOLs Following · 2d ago Cached

DeepSeek officially announced a significant API price increase, ending the era of ultra-cheap near-frontier model access soon after shipping DeepSeek-V4-Flash-0731.

0 favorites 0 likes
#deepseek

They almost catched up on Frontier performance, so now catching up on prices

Reddit r/LocalLLaMA · 2d ago

Discussion about DeepSeek's price increases and free tier downgrades pushing users toward local hardware, potentially benefiting NVIDIA hardware sales.

0 favorites 0 likes
#deepseek

Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090

Reddit r/LocalLLaMA · 2d ago

Tests et réglages détaillés pour optimiser DeepSeek-V4-Flash-0731 en GGUF sur une RTX 3090, atteignant ~15 tok/s à 128K de contexte grâce à différentes quantifications et paramètres de chargement.

0 favorites 0 likes
#deepseek

I get that AI labs need to make money, but zero-warning price spikes are a nightmare for production builds

Reddit r/LocalLLaMA · 2d ago

Commentary on DeepSeek's sudden API price hike with zero notice, highlighting the pain point for production builds that need time to adjust or switch providers.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback