coding

Tag

Cards List
#coding

@cline: Qwen3.8-Max is Alibaba’s largest model yet at 2.4T params, and shows a 2% higher benchmark result on Terminal-Bench tha…

X AI KOLs Following · 6d ago Cached

Alibaba unveils Qwen3.8-Max, its largest model at 2.4T parameters, showing a 2% higher Terminal-Bench result than Fable 5, with open weights to be released next week.

0 favorites 0 likes
#coding

Qwen3.8-Max: A New Bar for Coding and Cowork

Hacker News Top · 6d ago

Qwen3.8-Max sets a new benchmark for coding and collaborative work capabilities in AI models.

0 favorites 0 likes
#coding

Qwen3.8-Max: A New Bar for Coding and Cowork (25 minute read)

TLDR AI · 6d ago

Qwen 3.8-Max, a 2.4 trillion parameter model, is now available with open weights coming next week, delivering improvements in coding, work, research, and long-horizon tasks.

0 favorites 0 likes
#coding

Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper

Reddit r/LocalLLaMA · 2026-07-31

DeepSeek's new V4 Flash model is reportedly the #2 open-weight model behind Kimi K3, offering strong performance at over 50x lower cost ($0.09/$0.18 per 1M tokens) with solid coding and reasoning capabilities.

0 favorites 0 likes
#coding

DeepSeek V4 Flash GA ranks the same as Sonnet 5 and Grok 4.5 on DeepSWE

Reddit r/LocalLLaMA · 2026-07-31

DeepSeek announces V4 Flash GA, claiming it matches Sonnet 5 and Grok 4.5 on the DeepSWE benchmark, though the claims are not yet verified.

0 favorites 0 likes
#coding

Amazon accidentally spent $1.8 million using Claude for menial coding task, went 860% over budget — 'catastrophically expensive' coding blunders discovered in internal Amazon AI usage metrics

Reddit r/ArtificialInteligence · 2026-07-30 Cached

Amazon spent $1.8 million on a Claude AI project that went 860% over budget, highlighting the cost risks of deploying AI agents for coding tasks.

0 favorites 0 likes
#coding

My experience working with LLM

Reddit r/ArtificialInteligence · 2026-07-30

A VP/PM with coding background shares hands-on experience using LLMs like Claude Opus and Fable, highlighting limitations in memory, hallucination, and originality while emphasizing the irreplaceable value of human intuition and domain expertise.

0 favorites 0 likes
#coding

@PrajwalTomar_: I gave Grok 4.5 one sentence and it built me a working app before I finished my coffee. Cursor just launched a plan onl…

X AI KOLs Following · 2026-07-28 Cached

Grok 4.5, a powerful coding model, is now available in Cursor's new India-specific plan at ₹649 per month, allowing developers to build complete apps from a single sentence with cloud-based agents.

0 favorites 0 likes
#coding

@kentcdodds: Claude Code is useful, but Annie Sexton says it's also made engineering problem-solving kind of boring - and she's stil…

X AI KOLs Timeline · 2026-07-27 Cached

Annie Sexton and Kent Dodds discuss how Claude Code, while useful, has made engineering problem-solving feel boring, with Sexton still searching for where the joy went.

0 favorites 0 likes
#coding

Kimi K3’s open weights drop today — is anyone actually using Chinese AI models instead of Claude or Codex?

Reddit r/ArtificialInteligence · 2026-07-27

Moonshot AI releases open weights for Kimi K3, a 3T-parameter frontier model focused on long-horizon coding, repository-scale context, and tool use, allowing self-hosting and fine-tuning.

0 favorites 0 likes
#coding

Macaron-V1 family, built on Qwen3.6-35B-A3B

Reddit r/LocalLLaMA · 2026-07-26 Cached

MindLab Research releases Macaron-V1-Tall, a Mixture of LoRA model built on Qwen3.6-35B-A3B, featuring four specialists for chat, agent, coding, and GenUI tasks with a context length of 262K.

0 favorites 0 likes
#coding

Llama.cpp now has full MCP support!

Reddit r/LocalLLaMA · 2026-07-25

llama.cpp now fully supports the Model Context Protocol (MCP) for all protocols, enabling agentic chat and tool integration directly in its WebUI without external dependencies.

0 favorites 0 likes
#coding

2x, not 10x: coding with LLMs in 2026

Hacker News Top · 2026-07-25 Cached

The author argues that LLMs currently provide about a 2x productivity boost for coding due to their ability to handle easily verifiable tasks, but fundamental limitations prevent a 10x improvement; further gains will come from retooling around existing capabilities rather than model improvements.

0 favorites 0 likes
#coding

I'm impressed by Laguna S 2.1

Reddit r/LocalLLaMA · 2026-07-24

Laguna S 2.1, a 120B-class model, impressed by solving a complex coding problem in Julia with long thinking tokens, outperforming Qwen models on a memory-constrained rearrangement task.

0 favorites 0 likes
#coding

Anthropic's Opus 5 is about token efficiency, not a capability leap

Ars Technica · 2026-07-24 Cached

Anthropic released Opus 5, focusing on token efficiency and cost reduction rather than a major capability leap, offering performance close to Fable at half the cost.

0 favorites 0 likes
#coding

@bcherny: Opus 5 is a great model for coding, data analysis, design, biology, knowledge work. More than any of these eval scores,…

X AI KOLs Following · 2026-07-24 Cached

Anthropic's Claude Opus 5 is highlighted as a state-of-the-art model for coding, data analysis, and knowledge work, with unprecedented resistance to prompt injection attacks. The system card reveals that combined defenses reduce prompt injection success rates to near zero.

0 favorites 0 likes
#coding

Your agent’s action timed out. Does your code retry it?

Reddit r/AI_Agents · 2026-07-23

Technical article discussing the importance of retry logic when agent actions time out, highlighting a common pitfall in agent-based systems.

0 favorites 0 likes
#coding

Tested (the updated) Gemma 4 locally on coding with OpenCode

Reddit r/LocalLLaMA · 2026-07-23

Tested the updated Gemma 4 locally using llama.cpp on an M5 Pro, achieving 60 tokens/s for coding tasks with OpenCode; good for backend but poor for UI/UX.

0 favorites 0 likes
#coding

I used to be proud of these skills. Now AI agents do them better.

Reddit r/artificial · 2026-07-23

The author reflects on how AI agents now outperform them in code navigation, debugging, and report drafting, and asks others about experiences with multi-agent workflows like MCP, Anvita Flow, and Agent Protocol.

0 favorites 0 likes
#coding

Gemini and antigravity are underrated

Reddit r/singularity · 2026-07-21

A developer shares a hot take that Google's Gemini Flash model, when used in the Antigravity platform, outperforms GPT 5.6 for small coding tasks due to its speed and simplicity, despite GPT's higher intelligence ceiling.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback