knowledge-transfer

Tag

Cards List
#knowledge-transfer

Why Pretraining Fails to Share Cross-Lingual Knowledge

arXiv cs.CL ↗ · 2026-09-18 Cached

This research paper investigates why pretraining in large language models fails to transfer knowledge across languages, identifies disjoint token spaces as a fundamental barrier, and proposes mapping languages to a shared token space to improve cross-lingual generalization.

0 favorites 0 likes
#knowledge-transfer

Distilled Rapid Embedding Transfer (DRET): Parameter-Efficient Biomedical Domain Adaptation via Priority-Based Embedding Transfer

arXiv cs.CL ↗ · 2026-09-04 Cached

DRET is a parameter-efficient knowledge-transfer method that injects biomedical domain knowledge into smaller models like DistilBERT via embedding transfer, achieving performance competitive with larger specialized models.

0 favorites 0 likes
#knowledge-transfer

mimeo: Compiling Public Expert Corpora into Agent Skills and Testing What Transfers

arXiv cs.AI ↗ · 2026-09-02 Cached

mimeo is an open-source tool that compiles public expert corpora into agent skills and evaluates knowledge access, persona recognition, and judgment transfer in AI agents.

0 favorites 0 likes
#knowledge-transfer

Incorporating Cognitive Load and Knowledge Transfer for Multi-Domain Knowledge Tracing

arXiv cs.AI ↗ · 2026-08-26 Cached

This paper proposes LT-MKT, a method for multi-domain knowledge tracing that incorporates cognitive load and knowledge transfer using large language models to construct a hierarchical graph, achieving state-of-the-art performance on real-world datasets.

0 favorites 0 likes
#knowledge-transfer

ATHENA: Knowledge-guided agentic neural architecture search for AutoFormer-based electronic health record modeling

arXiv cs.AI ↗ · 2026-08-25 Cached

ATHENA is a knowledge-guided agentic neural architecture search framework that automates Transformer-based electronic health record modeling by reusing architecture knowledge across hospitals to reduce manual tuning.

0 favorites 0 likes
#knowledge-transfer

Stop building AI workflows your team is afraid to touch

Reddit r/AI_Agents ↗ · 2026-08-23

The article highlights common issues with complex AI workflows, such as lack of documentation and knowledge transfer, and suggests practices like versioning and testing to improve team handoffs.

0 favorites 0 likes
#knowledge-transfer

J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers

arXiv cs.LG ↗ · 2026-08-19 Cached

J-Miner recovers executable decision knowledge from fine-tuned language-model classifiers by mining named concepts and learning decision rules, enabling inspection and transfer to lightweight models with high fidelity.

0 favorites 0 likes
#knowledge-transfer

Training-Free Knowledge Transfer Across Model Scales through Activation-Guided Pruning

arXiv cs.LG ↗ · 2026-08-17 Cached

This paper proposes Activation-Prune-Merge (APM), a training-free framework for cross-scale fusion that improves smaller language models using larger donors without semantic alignment, achieving performance gains on multiple benchmarks.

0 favorites 0 likes
#knowledge-transfer

The Chauffeur Problem

Hacker News Top ↗ · 2026-08-16 Cached

The article explores the historical 'chauffeur problem' in early automobiles, where mechanics gained control over wealthy owners due to technical expertise, and draws parallels to modern computer specialists, highlighting the temporary nature of knowledge-based power.

0 favorites 0 likes
#knowledge-transfer

Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory

Hugging Face Daily Papers ↗ · 2026-08-07 Cached

The paper introduces Agent Memory Distillation (AMD), a training-free framework that transfers structured knowledge from a large teacher agent to a small student agent via hierarchical memory, improving tool-use benchmark performance by 3.4–27.2 percentage points.

0 favorites 0 likes
#knowledge-transfer

Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't

arXiv cs.LG ↗ · 2026-08-05 Cached

This paper studies what transfers between transformer models of different sizes in the same family (Pythia), showing that representations align while weights don't, and that conversion works best via initialization rather than direct weight projection.

0 favorites 0 likes
#knowledge-transfer

~1,400 years ago, scholars built a rigorous system to verify who you can trust. I rebuilt it as a trust layer for AI agents.

Reddit r/artificial ↗ · 2026-07-29

Author introduces ISNAD, a trust layer for AI agents inspired by the Islamic isnad system, designed to verify claim provenance across multi-agent chains. The paper is published on arXiv and includes code.

0 favorites 0 likes
#knowledge-transfer

Conditional Diffusion Guided Knowledge Transfer for Multi-Domain Knowledge Graph Completion

arXiv cs.CL ↗ · 2026-07-07 Cached

Proposes a conditional diffusion-guided knowledge transfer framework for multi-domain knowledge graph completion, generating domain-general entity embeddings without suppressing domain-specific information, achieving 4.3% average MRR improvement over state-of-the-art methods.

0 favorites 0 likes
#knowledge-transfer

How a Skiing Accident Put Our Development Practices to the test

Lobsters Hottest ↗ · 2026-07-06 Cached

The article describes how a skiing accident incapacitated the Tech Lead on a project, testing the team's development practices and underscoring the importance of documentation and knowledge transfer in software development.

0 favorites 0 likes
#knowledge-transfer

NVIDIA ASPIRE enables robots to accumulate knowledge from successful experiences and reuse it for new tasks, creating a persistent library of skills that improves learning over time

Reddit r/singularity ↗ · 2026-07-01 Cached

NVIDIA's ASPIRE framework enables robots to build a persistent library of skills from successful experiences, allowing reuse for new tasks and improving learning efficiency over time.

0 favorites 0 likes
#knowledge-transfer

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

arXiv cs.AI ↗ · 2026-07-01 Cached

HASTE introduces a hierarchical multi-agent system for ML engineering that organizes cross-competition knowledge into three tiers, achieving 77.3% medal rate on MLE-Bench Lite while reducing compute by 52% and demonstrating that structured knowledge transfer outperforms flat memory approaches.

0 favorites 0 likes
#knowledge-transfer

Bridging Scientific Heritage: An Arabic--Russian Parallel Corpus and LLM Benchmark for Sustainable Knowledge Transfer

arXiv cs.CL ↗ · 2026-07-01 Cached

This paper presents a benchmark for Arabic-Russian scientific translation, including a hybrid parallel corpus of 27,000 sentence pairs and fine-tuned multilingual models (mT5, NLLB, Qwen) using LoRA. The best model achieves BLEU 23.15, and the work aims to lower language barriers for scientific knowledge exchange between Arabic and Russian researchers.

0 favorites 0 likes
#knowledge-transfer

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

Hugging Face Daily Papers ↗ · 2026-06-23 Cached

This paper introduces a conversational voice agent system that uses a lightweight on-device 'Talker' model to start responding immediately, then incorporates knowledge from a frontier LLM 'Reasoner' as it becomes available, achieving 7-19x faster time-to-first-response while approaching frontier-level performance on a laptop.

0 favorites 0 likes
#knowledge-transfer

CacheRL:Multi-Turn Tool-Calling Agents via Cached Rollouts and Hybrid Reward

arXiv cs.CL ↗ · 2026-06-15 Cached

CacheRL trains small agent foundation models for multi-step tool-calling tasks, achieving 92% process accuracy (approaching GPT-5's 94%) with 100x less compute using cached rollouts and hybrid reward shaping, with innovations in knowledge transfer, cache-aware rewards, and iterative SFT/GRPO training.

0 favorites 0 likes
#knowledge-transfer

How Endava builds an agentic organization with Codex

OpenAI Blog ↗ · 2026-05-28 Cached

Endava, a global software contracting firm, uses OpenAI's Codex to codify senior expertise into agents, enabling small teams to deliver massive value quickly and transforming how junior and senior engineers collaborate.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback