meta-learning

Tag

Cards List
#meta-learning

SMETA-ZSL:Semantic Meta-Alignment for Zero-Shot Threat Classification

arXiv cs.LG ↗ · 2026-07-14 Cached

SMETA-ZSL proposes a method for generalized zero-shot threat classification using semantic meta-alignment and contrastive finetuning, outperforming prior methods by 10.8 points on average across 7 benchmarks.

0 favorites 0 likes
#meta-learning

Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels

arXiv cs.LG ↗ · 2026-07-14 Cached

This paper proposes a bilevel optimization framework for Direct Preference Optimization under noisy preference labels, introducing a metadata-free meta-reweighting method that uses central-difference approximation and LoRA fine-tuning to improve alignment performance.

0 favorites 0 likes
#meta-learning

What's your take on continual learning? [D]

Reddit r/MachineLearning ↗ · 2026-07-13

A discussion post questioning the definition and requirements of continual learning in AI, referencing recent statements by Dario Amodei and Demis Hassabis about its importance for AGI.

0 favorites 0 likes
#meta-learning

Architecture Generalization with MetaNCA

arXiv cs.LG ↗ · 2026-07-10 Cached

This paper introduces Meta Neural Cellular Automata (MetaNCA), a framework that learns local update rules to self-organize the weights of neural networks without backpropagation, scaling to networks of 2 million parameters on MNIST and CIFAR-100 and generalizing to unseen architectures.

0 favorites 0 likes
#meta-learning

Efficient Long-Horizon Learning for Learned Optimization

arXiv cs.LG ↗ · 2026-07-09 Cached

Proposes Efficient Long-horizon Optimization (ELO) learning, a meta-training algorithm that reallocates compute to longer horizons and uses decoupled progressive expert supervision, improving learned optimizers' performance on long-unroll tasks and out-of-distribution generalization. ELO-Celo2 consistently outperforms AdamW and matches Muon on language modeling tasks.

0 favorites 0 likes
#meta-learning

What is the current Memory Meta?

Reddit r/LocalLLaMA ↗ · 2026-07-08

An article exploring the current state of memory meta in artificial intelligence, likely discussing how memory and meta-learning are combined in modern AI systems.

0 favorites 0 likes
#meta-learning

Labeled-Data-Free Meta-Learning: Efficient Task Generation Using Pre-trained Models and Unlabeled Data

arXiv cs.LG ↗ · 2026-07-07 Cached

Proposes a labeled-data-free meta-learning method that generates tasks by assigning soft labels from pre-trained models to unlabeled data, avoiding computationally expensive model inversion. Achieves up to 104x speedup and 8.4-36.4% accuracy improvements over state-of-the-art DFML methods.

0 favorites 0 likes
#meta-learning

PhyMRI-SR: Toward Physics-Aware MRI Image Super-Resolution

Hugging Face Daily Papers ↗ · 2026-07-07 Cached

This paper proposes PhyMRI-SR, a physics-aware MRI super-resolution method that uses Gaussian splatting and physics-constrained modeling to dynamically adapt resolution-SNR configurations, achieving state-of-the-art performance.

0 favorites 0 likes
#meta-learning

From Search to Synthesis: Training LLMs as Zero-Shot Workflow Generators

arXiv cs.LG ↗ · 2026-07-01 Cached

Introduces MetaFlow, a method that trains large language models to generate zero-shot workflows for tasks by combining supervised fine-tuning and reinforcement learning with execution feedback, achieving strong generalization to untrained tasks and operator sets.

0 favorites 0 likes
#meta-learning

Halt Fast! Early Stopping for Certified Robustness

arXiv cs.LG ↗ · 2026-06-29 Cached

This paper introduces a meta-learning framework for anytime-valid certified robustness that uses sequential E-processes to adaptively allocate compute, achieving a 20-fold reduction in sample complexity compared to traditional randomized smoothing while maintaining rigorous statistical guarantees.

0 favorites 0 likes
#meta-learning

@jaseweston: Claim: Autoresearch that moves the frontier will be about better data: we call that *Autodata*. 1/6 -- Paper is out! ht…

X AI KOLs Timeline ↗ · 2026-06-25 Cached

Introduces Autodata, a method where AI agents act as data scientists to create high-quality synthetic training data, showing gains on computer science, legal, and math reasoning tasks over classical methods.

0 favorites 0 likes
#meta-learning

Learning Dynamical Systems from Multiple Sparse Datasets: A Hierarchical Bayesian Modeling Approach

arXiv cs.LG ↗ · 2026-06-25 Cached

Proposes a hierarchical Bayesian framework for meta-learning in dynamical systems from multiple sparse, noisy datasets, using gradient-based MCMC with an embedded ODE solver for efficient posterior inference of shared and dataset-specific parameters.

0 favorites 0 likes
#meta-learning

Exploring Dualistic Meta-Learning to Enhance Domain Generalization in Open Set Scenarios

arXiv cs.LG ↗ · 2026-06-24 Cached

Proposes a novel meta-learning strategy called MEDIC for open set domain generalization, which uses implicit gradient matching across domain and class splits to achieve better boundaries. Experiments show state-of-the-art performance.

0 favorites 0 likes
#meta-learning

Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning

Hugging Face Daily Papers ↗ · 2026-06-18 Cached

This paper presents Connect the Dots (CoD), a framework for training LLMs via reinforcement learning to develop meta-capabilities for long-lifecycle agents, enabling continuous learning and cross-domain generalization.

0 favorites 0 likes
#meta-learning

Retrievable Gradients: Continual Post-Training Without Cumulative Weight Drift

arXiv cs.CL ↗ · 2026-06-16 Cached

Proposes ReGrad, a paradigm that treats gradients as retrievable units of knowledge for continual post-training, avoiding cumulative weight drift by storing document-specific gradients in a Gradient Bank and retrieving query-relevant gradients for temporary weight adaptation.

0 favorites 0 likes
#meta-learning

Fodor and Pylyshyn's Systematicity Challenge Still Stands

arXiv cs.CL ↗ · 2026-06-15 Cached

This paper argues that recent claims that neural networks have solved Fodor and Pylyshyn's systematicity challenge are premature. The authors show that the meta-learning for compositionality model fails to generalize out-of-distribution and behaves unsystematically even on in-distribution problems, concluding the challenge remains unmet.

0 favorites 0 likes
#meta-learning

Robotic Policy Adaptation via Weight-Space Meta-Learning

Hugging Face Daily Papers ↗ · 2026-06-05 Cached

Introduces WIZARD, a weight-space meta-learning framework that generates task-specific LoRA parameters for frozen VLA policies from language instructions and demonstration videos, enabling efficient task adaptation without fine-tuning.

0 favorites 0 likes
#meta-learning

When Offline Selectors Cannot Beat the Best Single Model: A Diagnostic Study on edX Dropout Prediction

arXiv cs.LG ↗ · 2026-06-04 Cached

This paper proposes a three-stage diagnostic framework to identify why offline model selectors fail to beat the best single model, applying it to dropout prediction on edX clickstream data. The study finds that the bottleneck is local representational ambiguity rather than learner choice or distribution shift, recommending state redesign or new data collection over further algorithm tuning.

0 favorites 0 likes
#meta-learning

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

arXiv cs.CL ↗ · 2026-06-04 Cached

SePO (Self-Evolving Prompt Optimization) proposes a self-referential prompt agent that optimizes both task agents' system prompts and its own system prompt through an evolutionary search, outperforming Manual-CoT, TextGrad, and MetaSPO across five benchmarks including AIME'25, ARC-AGI-1, and GPQA.

0 favorites 0 likes
#meta-learning

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

arXiv cs.AI ↗ · 2026-06-04 Cached

R-APS (Reflective Adversarial Pareto Search) is a novel method for constrained design tasks that addresses three structural failures in LLM-based agentic systems—error propagation, robustness evaluation, and knowledge invalidation—through reasoning-mode decomposition across three timescales, requiring no fine-tuning. Evaluated on planar mechanism synthesis, it achieves 3.5x tighter robustness certificates, 46% faster iterations-to-first-admission, and 2.1x Chamfer-distance reduction over baselines.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback