architecture

Tag

Cards List
#architecture

@levidiamode: Day 248/365 of GPU Programming The Cerebras CTO has some other great talks online. For example, one from two years ago …

X AI KOLs Timeline ↗ · 2026-09-08 Cached

A social media post highlighting talks by Cerebras CTO on GPU programming and AI hardware architecture, noting their low view counts despite being informative.

0 favorites 0 likes
#architecture

Watch Los Angeles get built, one building at a time (1880–2026)

Hacker News Top ↗ · 2026-09-07 Cached

A visualization project displays every existing building in Los Angeles, showing when each was built from 1880 to 2026 to illustrate urban development.

0 favorites 0 likes
#architecture

Deep Persona: A Psychologically Grounded Architecture and Evaluation Framework for Role-Playing Agents and Simulations

Hugging Face Daily Papers ↗ · 2026-09-06 Cached

This paper introduces Deep Persona, a psychologically grounded architecture for role-playing agents, and proposes an evaluation framework. It evaluates LLMs and finds systematic limitations in emotional expression despite high pragmatic fluency.

0 favorites 0 likes
#architecture

Why multi-agent RAG pipelines choke on production databases (and the architecture that saved us)

Reddit r/AI_Agents ↗ · 2026-09-03

The article discusses why multi-agent RAG pipelines suffer from high latency in production due to synchronous tool calls and context bloat, and presents solutions like micro-agents, caching with Redis, and asynchronous processing to improve performance.

0 favorites 0 likes
#architecture

Do agents actually need memory, or are we using it to compensate for bad architecture?

Reddit r/AI_Agents ↗ · 2026-09-03

The post questions whether memory in AI agents is essential or a workaround for poor architecture, and asks practitioners what needs to be persisted in production systems.

0 favorites 0 likes
#architecture

LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails

arXiv cs.AI ↗ · 2026-09-03 Cached

This paper identifies failure modes in LLM-as-a-Judge systems for self-improving agents and introduces PROCTOR, an architecture with deterministic guardrails to mitigate these issues.

0 favorites 0 likes
#architecture

Does an AI agent really need its own inbox? That's a very dangerous and architecturally wrong trend.

Reddit r/AI_Agents ↗ · 2026-09-02

The author criticizes the trend of giving AI agents broad access to personal accounts via a single SDK, arguing it is architecturally unsound and poses security risks, and advocates for using specific, expiring permissions instead.

0 favorites 0 likes
#architecture

It's OK to hardcode feature flags (2025)

Hacker News Top ↗ · 2026-09-02 Cached

The article argues that hardcoding feature flags is often simpler and safer than using complex management software, recommending a basic implementation until scaling is truly needed.

0 favorites 0 likes
#architecture

Safin-1: Safety from Within through Memory-Native State Evolution

arXiv cs.LG ↗ · 2026-09-02 Cached

Safin-1 introduces a foundation model family with memory-native state evolution for intrinsic safety capabilities, using the MARCH architecture for selective memory retrieval and persistent state adaptation.

0 favorites 0 likes
#architecture

@rohanpaul_ai: The information reports Astra reportedly uses "recurrent depth," or a "looped transformer," which helped its performanc…

X AI KOLs Timeline ↗ · 2026-09-02 Cached

OpenAI's Astra model reportedly uses a looped transformer architecture for enhanced performance, though this may reduce the readability of internal reasoning. It is noted for reaching a critical cybersecurity capability threshold.

0 favorites 0 likes
#architecture

Created a new architecture for Large Language Models.

Reddit r/artificial ↗ · 2026-08-31

A new open-source architecture called Mixture of Models (MoM) for Large Language Models, which bundles multiple AI models to function like a Mixture of Experts model, is released on GitHub.

0 favorites 0 likes
#architecture

Best approach for adding AI agents to an existing product?

Reddit r/AI_Agents ↗ · 2026-08-31

A discussion seeking advice on the best approaches to integrate AI agents and LLM features into an existing product, focusing on architecture, reliability, and maintenance lessons.

0 favorites 0 likes
#architecture

No, Engrams won't let you run 1T models locally. It does something even better.

Reddit r/LocalLLaMA ↗ · 2026-08-27

Engrams are an architectural innovation that uses N-gram tables to offload memorization from transformer models, allowing smaller models to reason better by freeing up parameters for computation.

0 favorites 0 likes
#architecture

@rauchg: CMA is a remarkably elegant architecture. The 'hard parts' of the agent are managed, but it doesn't take away from your…

X AI KOLs Timeline ↗ · 2026-08-27 Cached

The tweet praises the Claude Managed Agent architecture for elegantly handling agent complexities while enabling customization, exemplified by integration with Vercel's Chat SDK for a universal chat layer.

0 favorites 0 likes
#architecture

Software engineering is about managing complexity

Hacker News Top ↗ · 2026-08-27 Cached

The article argues that software engineering is primarily about managing complexity through architectural decisions and tradeoffs, with AI being effective at code generation but not at handling these higher-level aspects. It highlights how building software involves critical choices about constraints, costs, and evolution beyond mere syntax.

0 favorites 0 likes
#architecture

WebSockets vs. SSE should be about ordering and correctness

Lobsters Hottest ↗ · 2026-08-26 Cached

This article argues that comparisons between WebSockets and Server-Sent Events should focus on event ordering and correctness to avoid inconsistent user interfaces, rather than just latency or simplicity.

0 favorites 0 likes
#architecture

Qwen 3.8 Flash Next: Beating DS V4 Flash at half the parameters, stronger than Opus 4.6

Reddit r/singularity ↗ · 2026-08-26

Qwen team has released an open-weight model, Qwen 3.8 Flash Next, which surpasses DS V4 Flash in performance with half the parameters and is stronger than Opus 4.6, offering a preview of the Qwen 4 architecture.

0 favorites 0 likes
#architecture

Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency

Hacker News Top ↗ · 2026-08-26

The Qwen3.8-Flash-Next introduces a new AI architecture focused on achieving ultimate cost-efficiency in model performance.

0 favorites 0 likes
#architecture

@anirudhbv_ce: 1/30 Days of Inference Transformers Architecture (Attention is All You Need) Transformers killed the RNN by making sequ…

X AI KOLs Timeline ↗ · 2026-08-26 Cached

This article is an educational piece explaining the Transformer architecture, its core components like self-attention and multi-head attention, and its significance in modern AI, as part of a 30-day inference series.

0 favorites 0 likes
#architecture

Agentic Context Management: Memory and Cost as Architecture Problems

Hacker News Top ↗ · 2026-08-26 Cached

This paper argues that managing context in AI agents should be treated as a lifecycle architecture problem, proposing Agentic Context Management (ACM) with five primitives and a reference implementation that achieves high benchmark scores.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback