weight-update

Tag

Cards List
#weight-update

Imprint Reader: From Weight-Update Readout to Behavioral Intervention

Hugging Face Daily Papers ↗ · 2d ago Cached

The Imprint Reader is a model trained to describe frozen weight updates in language models, enabling behavioral intervention to improve safety and reasoning. It uses SMaRT for training and MetaEdit for intervention, demonstrating feasibility in natural language readout.

0 favorites 0 likes
#weight-update

Searching the Space of Feed-Forward Neural-Network Weight-Update Rules with Fixed Depth Symbolic Regression

arXiv cs.LG ↗ · 2026-07-27 Cached

This paper investigates using symbolic regression to discover explicit neural network weight-update rules that outperform standard hand-designed optimizers on small symbolic regression benchmarks, achieving an aggregate MSE reduction of 44.47% in 25 out of 30 benchmark/network combinations.

0 favorites 0 likes
#weight-update

@MihaelaVDS: Can LLMs keep learning new skills without updating their weights? Modern LLMs can already master & combine many skills.…

X AI KOLs Timeline ↗ · 2026-06-29 Cached

Introduces 'skill neologisms', a method for enabling LLMs to learn new skills without weight updates, addressing catastrophic forgetting. Presented at ICML.

0 favorites 0 likes
#weight-update

SIA: Self Improving AI with Harness & Weight Updates

Hugging Face Daily Papers ↗ · 2026-05-26 Cached

A self-improving AI framework that simultaneously updates both model weights and task-specific agent architecture via a language-model feedback agent, achieving significant gains across legal classification, GPU optimization, and biological denoising tasks.

0 favorites 0 likes
← Back to home

Submit Feedback