software-engineering

Tag

Cards List
#software-engineering

@nerdearla: @kentcdodds in Grand Hall with "The Last Software Engineer". What place is left for those who develop software in the a…

X AI KOLs Following ↗ · 6h ago Cached

Kent C. Dodds gives a talk titled 'The Last Software Engineer' at Nerdearla 2026, discussing the future of software development in the AI era.

0 favorites 0 likes
#software-engineering

@simonw: The more time I spend working with coding agents, the more convinced I am that they make software engineering even hard…

X AI KOLs Following ↗ · 18h ago Cached

The author argues that coding agents, despite their capabilities, make software engineering more challenging and require extraordinary discipline, referencing recent AI models like Opus 4.6 and GPT-5.2 in the discussion.

0 favorites 0 likes
#software-engineering

How Good Are LLMs at Decision Forking? (GitHub Repo)

TLDR AI ↗ · 20h ago Cached

Taste-Bench is a benchmark that evaluates LLMs' ability to choose optimal paths at decision forks in long-horizon tasks, using trajectories from software engineering and machine learning research, with a leaderboard indicating current top models like GPT-5.6 Sol achieving 59.7% accuracy.

0 favorites 0 likes
#software-engineering

Note on 24th September 2026

Simon Willison's Blog ↗ · 20h ago Cached

Simon Willison notes that coding agents, while powerful, add complexity to software engineering, requiring significant discipline and knowledge to unlock their full potential.

0 favorites 0 likes
#software-engineering

@kentcdodds: Stop the slop machine without having to read the code. Let me show you how I do it:

X AI KOLs Timeline ↗ · yesterday Cached

Kent C. Dodds demonstrates a technique to prevent low-quality code generation without needing to manually review the code.

0 favorites 0 likes
#software-engineering

Escaping Python Dependency Hell: A Hybrid Replay-and-Repair Pipeline for Python Dependency Resolution

arXiv cs.AI ↗ · yesterday Cached

This paper introduces PLLM+, a hybrid pipeline that combines deterministic replay of historical dependency configurations with LLM-based repair to resolve Python dependency conflicts, showing improved success rates and reduced runtime on the HG2.9K benchmark.

0 favorites 0 likes
#software-engineering

@shao__meng: Stanford University CS329Z: Engineering AI Agents course has officially started, with the first lecture slides now publ…

X AI KOLs Timeline ↗ · yesterday Cached

Stanford University's CS329Z course on Engineering AI Agents has started, with lecture slides now public, covering topics from models to autonomous systems and featuring instructors like Diyi Yang and Michael Ryan.

0 favorites 0 likes
#software-engineering

@yibie: https://x.com/yibie/status/2102619356874117594

X AI KOLs Timeline ↗ · 2d ago Cached

This article discusses the importance of evaluations in AI systems, explains why traditional testing is insufficient, introduces three main types of evaluations, and provides implementation suggestions.

0 favorites 0 likes
#software-engineering

@jiayq: TLA+ is also (one of the) method that we at Intent Lab used to build mission critical infra software like Agent FS. Key…

X AI KOLs Following ↗ · 2d ago Cached

The tweet discusses using formal verification methods like TLA+ and Lean to ensure the correctness and scalability of mission-critical AI infrastructure software, with references to Intent Lab and Boris Cherny's work on the Claude Agent SDK.

0 favorites 0 likes
#software-engineering

A clean git merge of my two agents' worktrees that failed its own tests

Reddit r/AI_Agents ↗ · 2d ago

The author describes an experiment where merging two AI agents' git worktrees led to test failures despite clean merges, highlighting the challenges of parallel agent development without mutual awareness.

0 favorites 0 likes
#software-engineering

SWE-Bench Pro V2 (9 minute read)

TLDR AI ↗ · 2d ago Cached

SWE-Bench Pro V2 is an updated benchmark for evaluating AI agents in software engineering, featuring 642 tasks across 11 repositories with improved evaluation protocols and contamination controls.

0 favorites 0 likes
#software-engineering

@GergelyOrosz: Casey is right, again

X AI KOLs Timeline ↗ · 3d ago Cached

Casey Muratori emphasizes the importance of learning assembly language for software optimization in a discussion shared by Gergely Orosz.

0 favorites 0 likes
#software-engineering

AI Has No Wisdom and Neither Will You

Lobsters Hottest ↗ · 3d ago Cached

This article critiques the over-reliance on AI in software development, arguing that AI lacks wisdom for code maintainability and architecture, which can hinder developer expertise and lead to long-term issues.

0 favorites 0 likes
#software-engineering

@MaiYangAI: Watched Lauren (@poteto)'s talk again on how she merged about 2000 PRs over the past month. This was originally a share…

X AI KOLs Timeline ↗ · 3d ago Cached

A tweet sharing Lauren's talk on merging 2000 pull requests in a month, emphasizing creativity and practical workflow in software development.

0 favorites 0 likes
#software-engineering

I built a router to cut my agent bill. Then found out it only knows how to spend up.

Reddit r/AI_Agents ↗ · 3d ago

A developer built a router to cut AI agent costs but found it only escalates requests, increasing spending; effective savings came from caching rather than routing.

0 favorites 0 likes
#software-engineering

@dabit3: Steer Cloud Agents from the CLI (not just on Web). ssh directly into it. PRs tested in a real VM, real browser / device…

X AI KOLs Timeline ↗ · 3d ago Cached

Cognition introduces new CLI and SSH access for Devin Cloud, enabling developers to manage AI coding sessions directly from the terminal and interact with a dedicated VM.

0 favorites 0 likes
#software-engineering

@GergelyOrosz: Heard a software engineer say: "my cofounder is Claude" about a project they built + shipped. I am baffled by people hu…

X AI KOLs Timeline ↗ · 4d ago Cached

A software engineer describes Claude as his cofounder, sparking a discussion on humanizing AI tools in software development.

0 favorites 0 likes
#software-engineering

@GergelyOrosz: I’m doing a lecture at my Alma mater to sophomore (2nd year) students on software engineering and originally wanted to …

X AI KOLs Timeline ↗ · 4d ago Cached

Gergely Orosz is preparing a lecture for sophomore students on software engineering, revamping his 2024 talk with an angle emphasizing proactive learning.

0 favorites 0 likes
#software-engineering

SWE-Proof: Can Language Models Resolve Real-World Issues with Machine-Checked Proofs?

arXiv cs.LG ↗ · 4d ago Cached

This paper introduces SWE-Proof, a benchmark of formally verified code patches for real-world software issues, demonstrating that formal verification improves error detection in LLM-generated code and identifies specification synthesis as a key open problem.

0 favorites 0 likes
#software-engineering

It feels like AIs are getting worse at following instructions

Reddit r/AI_Agents ↗ · 4d ago

A software engineer reports that AI models seem to be getting worse at following instructions in recent updates, often making unasked changes and ignoring contracts, leading to increased manual work.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback