continual-learning

Tag

Cards List
#continual-learning

SkillEvoReg: Regularizing Agent Skill Evolution Against Overfitting

arXiv cs.AI ↗ · 13h ago Cached

SkillEvoReg introduces a regularization framework to prevent overfitting in the skill evolution of language-model agents, combining dropout, local regularization, and causal validation to maintain performance while controlling skill growth.

0 favorites 0 likes
#continual-learning

Estimating and Orthogonalizing Unknown Pre-training Gradients for Continual Fine-tuning of Large Language Models

arXiv cs.CL ↗ · 13h ago Cached

This paper introduces EoupCT, a novel framework that estimates and orthogonalizes unknown pre-training gradients to mitigate catastrophic forgetting in continual fine-tuning of large language models.

0 favorites 0 likes
#continual-learning

@VraserX: OpenAI DevDay on Tuesday could be one of the biggest and most important AI events of the year. An always-on agent with …

X AI KOLs Timeline ↗ · 23h ago Cached

OpenAI DevDay is anticipated to be a major AI event, with potential announcements of an always-on AI agent featuring persistent memory and continual learning, as well as new hardware.

0 favorites 0 likes
#continual-learning

Beyond Class Marginals: Bounding Rehearsal Gaps without Freezing Class Co-occurrence

arXiv cs.LG ↗ · 5d ago Cached

This paper introduces randomized-pass replay (RPR) to bound rehearsal gaps in online continual learning, showing improved accuracy over independent class-balanced retrieval in experience replay methods like ER-ACE.

0 favorites 0 likes
#continual-learning

Brain-Inspired Hierarchical Modularity for General Continual Learning

arXiv cs.LG ↗ · 5d ago Cached

The paper proposes a brain-inspired hierarchical modular approach for continual learning to handle online and uncertain data streams, achieving significant performance gains in tasks like embodied manipulation by leveraging pretrained foundation models.

0 favorites 0 likes
#continual-learning

Stable Unsupervised Continual Chunking with Sheaf SyncMap

arXiv cs.LG ↗ · 5d ago Cached

This paper introduces sheaf regularization to stabilize Decentralized SyncMap for unsupervised continual chunking, achieving higher normalized mutual information and better adaptation to input distribution shifts.

0 favorites 0 likes
#continual-learning

@PyTorch: Frontier models are the fastest way to launch an AI product, but what happens when usage scales? In our latest case stu…

X AI KOLs Following ↗ · 6d ago Cached

Shopify built a continual learning loop using PyTorch and vLLM to improve their GraphQL agent, reducing costs by 96% and outperforming frontier models through production-driven updates.

0 favorites 0 likes
#continual-learning

Knowledge Pull Requests for Continual Document Authoring

Hugging Face Daily Papers ↗ · 6d ago Cached

Knowledge Pull Requests (KPRs) is a framework for continual document authoring that integrates new knowledge by extracting claims, filtering them, and producing a ChangeLog to make changes interpretable. It outperforms existing methods in preserving content and adding information, as evaluated on Wikipedia revisions and RAGTIME tasks.

0 favorites 0 likes
#continual-learning

50+ Hours and 100M+ Tokens Later, Open Source Autonomous Agent is GETTING CLOSER at Solving an Open Math problem

Reddit r/LocalLLaMA ↗ · 6d ago

An open-source autonomous agent has been running for over 50 hours and processed more than 100 million tokens in an experiment to solve the C(25,15,5) covering design problem, with the goal of achieving a breakthrough in mathematics.

0 favorites 0 likes
#continual-learning

Designer-RSI: Evolving Procedural Memory from User Traffic for Agentic Graphic Design

arXiv cs.AI ↗ · 2026-09-21 Cached

A continual adaptation framework evolves procedural memory from user traffic for agentic graphic design, improving execution success rates without weight updates or human labels.

0 favorites 0 likes
#continual-learning

mini-AGI - dynamically grown (530M params currently and growing) continual learning model trained from scratch on 8GB VRAM laptop from batch-1 stream of data.

Reddit r/LocalLLaMA ↗ · 2026-09-21 Cached

mini-AGI is a continual learning byte-level language model that dynamically grows its architecture, trained from scratch on an 8GB VRAM laptop, demonstrating the possibility of personal AI that learns continuously without catastrophic forgetting.

0 favorites 0 likes
#continual-learning

ACLArena: Agent Continue Learning in Multi-stage Post-training

Hugging Face Daily Papers ↗ · 2026-09-21 Cached

The paper presents ACLArena, a framework for evaluating Agent Continual Learning in multi-stage post-training, analyzing forgetting and generalization mechanisms, and proposing an improved ACL recipe using offline replay and LoRA experts.

0 favorites 0 likes
#continual-learning

Continual Enterprise World Model Discovery in Dynamic Systems

arXiv cs.AI ↗ · 2026-09-18 Cached

The paper introduces continual enterprise world model discovery, where an agent learns and adapts to changing business rules in dynamic systems, evaluating with the EnterpriseWorldShift benchmark and demonstrating improved prediction accuracy over prior methods.

0 favorites 0 likes
#continual-learning

An Architecture for Long-Horizon Agents: Levels, Ticks and Cascaded Intelligence

arXiv cs.AI ↗ · 2026-09-18 Cached

This paper proposes a hierarchical architecture for long-horizon AI agents, incorporating levels, ticks, and cascaded intelligence to enable continual operation without forgetting, demonstrated over a ten-day campaign.

0 favorites 0 likes
#continual-learning

Uncertainty-Aware Continual Learning for Open-World Intent Discovery Under an evolving Label Space

arXiv cs.LG ↗ · 2026-09-17 Cached

This paper proposes a unified uncertainty-aware probabilistic framework for continual new intent discovery under an evolving label space, using adaptive β-VAE and multi-signal decision mechanisms to enable controlled label expansion while mitigating catastrophic forgetting. Experiments demonstrate high novelty precision and stable adaptation with limited forgetting.

0 favorites 0 likes
#continual-learning

I just solved continual learning

Reddit r/artificial ↗ · 2026-09-16

The article proposes a method for continual learning in AI where the model autonomously decides what to learn and updates its weights in real-time, potentially enabling continuous self-evolution towards AGI.

0 favorites 0 likes
#continual-learning

ReDraft, Don't Just Distill: Reference-Driven Revision for Continual VLLM Post-Training

arXiv cs.AI ↗ · 2026-09-16 Cached

ReDraft is a reference-driven revision method for continual post-training of large vision-language models that balances learning new tasks and preserving old ones, achieving higher accuracy and less forgetting than standard approaches like SFT.

0 favorites 0 likes
#continual-learning

Smarter by the Moment: Environment-Driven Dynamic Policies for Continual LLM Improvement

arXiv cs.CL ↗ · 2026-09-16 Cached

This paper proposes Dynamic Retrieval-based Policy Generation (DRPG), a framework that uses memory retrieval and environment feedback to dynamically generate policies for continual improvement of large language models across various tasks.

0 favorites 0 likes
#continual-learning

CERA-MoA: Co-Evolving Routing Mechanisms with Continually Learning LLM Agents

Hugging Face Daily Papers ↗ · 2026-09-16 Cached

CERA-MoA introduces a co-evolving framework for mixture-of-agents systems that uses reinforcement learning to dynamically route queries and adapt agent capabilities, enhancing task performance and efficiency.

0 favorites 0 likes
#continual-learning

Continual learning in the fruit fly brain has been decoded, the missing piece for true AGI

Reddit r/singularity ↗ · 2026-09-15

Researchers have decoded continual learning mechanisms in the fruit fly brain, potentially offering key insights for achieving true artificial general intelligence.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback