coding-model

Tag

Cards List
#coding-model

Cognition's SWE-2

Product Hunt · 4d ago Cached

Cognition's SWE-2 is a new coding model post-trained from Kimi K3 with reinforcement learning, achieving 50.0% on FrontierCode 1.1 Main and offering cost savings of 64% compared to Fable 5.1 while being competitive with GPT-6 Astra. It is now available in Devin Desktop and CLI.

0 favorites 0 likes
#coding-model

Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra

Hacker News Top · 4d ago Cached

Cognition launches SWE-2, an advanced coding model that achieves competitive performance with Fable 5.1 and GPT-Astra at a fraction of the cost, leveraging novel reinforcement learning to optimize the cost-performance frontier.

0 favorites 0 likes
#coding-model

@baseten: Today, we're releasing GLM-5.3 Fast: one of the most intelligent open-weight models ever at an even higher TPS. Designe…

X AI KOLs Timeline · 2026-09-03 Cached

Z.AI releases GLM-5.3 Fast, an advanced open-weight AI model optimized for agentic coding and cybersecurity, featuring a 744B-A40B MoE architecture with substantial benchmark improvements.

0 favorites 0 likes
#coding-model

Training a coding model to paint watercolours with TRL and OpenEnv

Hugging Face Blog · 2026-09-03 Cached

The article details an open-source reproduction of training a coding model to generate watercolour art using TRL and OpenEnv, with full pipeline artifacts published on Hugging Face for reproducibility.

0 favorites 0 likes
#coding-model

Who’s behind the new ‘stealth model’ Ox Alpha?

TechCrunch AI · 2026-08-23 Cached

A mysterious AI model named Ox Alpha has been released on OpenRouter, leading to widespread speculation about its developers, with theories pointing to Chinese AI companies like Z.ai or other entities.

0 favorites 0 likes
#coding-model

MAI-Code-1.1-Flash: Better, faster, at a quarter of the cost (2 minute read)

TLDR AI · 2026-08-12 Cached

Microsoft introduces MAI-Code-1.1-Flash, a faster, cheaper, and more token-efficient coding model now in production in GitHub Copilot, with improvements in CLI and .NET tasks.

0 favorites 0 likes
#coding-model

@github: @MicrosoftAI's MAI-Code-1.1-Flash is now rolling out in GitHub Copilot. It has native vision support for image understa…

X AI KOLs Timeline · 2026-08-11 Cached

Microsoft's MAI-Code-1.1-Flash coding model is rolling out in GitHub Copilot, adding native vision support and improved coding performance at a 73% lower list price than its predecessor.

0 favorites 0 likes
#coding-model

KAT Coder 2.5 dev: Do yourself a favor and try it!

Reddit r/LocalLLaMA · 2026-08-03

A developer enthusiastically recommends KAT Coder 2.5 dev, claiming it is faster, more accurate, and uses fewer tokens than Qwen 3.6 35b a3b, and outperforms Gemma 4 models on their setup, with a GitHub repo containing detailed benchmarks.

0 favorites 0 likes
#coding-model

Kimi K3-256k

Hacker News Top · 2026-07-29 Cached

Kimi Code releases Kimi K3-256k, a 256k-context version of its flagship K3 coding model, offering reduced quota consumption while maintaining similar performance for most tasks.

0 favorites 0 likes
#coding-model

@JiaZhihao: Excited to share Lithos’ serving stack for Kimi K2.7 Code, a 1T-parameter frontier coding model. On a single 8×B200 nod…

X AI KOLs Timeline · 2026-07-10 Cached

Lithos announces its inference engine serving Kimi K2.7 Code, achieving over 1,000 tokens/sec per user on a single 8×B200 node at native precision, 3.4–5.7× faster than major providers.

0 favorites 0 likes
#coding-model

Meta enters the crowded AI coding battle with Muse Spark 1.1

TechCrunch AI · 2026-07-09 Cached

Meta launched Muse Spark 1.1, a multimodal AI model for agentic coding, competing with OpenAI and Anthropic at a competitive price.

0 favorites 0 likes
#coding-model

@rohanpaul_ai: Surprising and such a good news for open source coding model, and also that there are lots of hidden chances to reduce …

X AI KOLs Following · 2026-07-09 Cached

Databricks tested GLM-5.2, an open-source coding model, and found it competes with top closed models like Claude Opus 4.8 on real enterprise code tasks while being cheaper ($1.28/task vs $1.94/task). The evaluation also highlighted Pi, a harness that reduces costs by sending less context per turn.

0 favorites 0 likes
#coding-model

I made a tool that chains a small local model into a big coding model and auto-unloads VRAM between them

Reddit r/LocalLLaMA · 2026-07-07

A developer created a tool that chains a small local model with a larger coding model, automatically offloading VRAM between them to optimize memory usage.

0 favorites 0 likes
#coding-model

I added MTP to local SoTA Agentic Coding Model Ornith 35B FP8 E4M3

Reddit r/LocalLLaMA · 2026-07-02 Cached

Introduced MTP speculative decoding to the Ornith 35B coding model in FP8 precision, achieving approximately 18% faster inference with minimal extra VRAM.

0 favorites 0 likes
#coding-model

@nickfrosst: now seems like a good day to remind people we have an apache 2.0 coding model you can run with 20 gigs of ram locally f…

X AI KOLs Following · 2026-06-26 Cached

Cohere Labs releases North Mini Code, a 30B parameter (3B active) open-source coding model under Apache 2.0, optimized for code generation and agentic tasks, capable of running locally with 20GB RAM via 4-bit quantization.

0 favorites 0 likes
#coding-model

@seclink: Compared to Mimo, Kimi is slower and more expensive ...

X AI KOLs Timeline · 2026-06-26 Cached

Kimi released the K2.7 Code model and its high-speed version, and announced API pricing. Compared to rival Mimo, it is more expensive and slower.

0 favorites 0 likes
#coding-model

@nathanhabib1011: while we cannot use fable 5 and gpt 5.6 will be screened before allowed to use the model, open source model builders ar…

X AI KOLs Following · 2026-06-26 Cached

Open-source model builders have released Ornith-1.0-397B, a top-tier coding model, while larger models like Fable 5 and GPT-5.6 face restrictions.

0 favorites 0 likes
#coding-model

@ClementDelangue: Kog open-sourced on @huggingface the 2B model that they used to show a model running at 3,000+ tokens per second. Very …

X AI KOLs Timeline · 2026-06-24 Cached

Kog has open-sourced the Laneformer 2B model, a 2.3B parameter instruction-tuned coding model designed for high-speed decoding, achieving over 3,000 tokens per second by prioritizing latency from the architecture stage.

0 favorites 0 likes
#coding-model

Jackrong/Qwopus3.6-27B-Coder-Compat-MTP-GGUF

Hugging Face Models Trending · 2026-06-20 Cached

Jackrong releases Qwopus3.6-27B-Coder-Compat-MTP-GGUF, a GGUF quantization of the Qwopus3.6-27B-Coder model with an expanded chat template for better interoperability with tool-using runtimes and OpenAI-compatible agent frameworks.

0 favorites 0 likes
#coding-model

@PatrickToulme: I ran GLM 5.2 with OpenCode harness against Claude Opus this week deployed locally. Bottom line: It is a real frontier …

X AI KOLs Following · 2026-06-20 Cached

GLM 5.2 is a frontier open-source coding model that performs near Claude Opus quality on coding tasks, with excellent tool calling, planning, and local deployment capabilities, at no cost.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback