multi-domain

Tag

Cards List
#multi-domain

EnSiTa - A Trilingual Multi-Domain Parallel Dataset and Benchmark for Domain-Specific Machine Translation

arXiv cs.CL ↗ · 6d ago Cached

EnSiTa is a trilingual multi-domain parallel dataset and benchmark for English, Sinhala, and Tamil, featuring human post-edited training data and extensive experiments on domain-specific machine translation to address low-resource language challenges.

0 favorites 0 likes
#multi-domain

JEPA-Anything: Learning Predictive Models across Different Worlds

Hugging Face Daily Papers ↗ · 2026-09-17 Cached

JEPA-Anything presents a domain-agnostic framework based on orthogonal predictive factorization for learning predictive models across diverse systems like vision, biology, and control, with demonstrated improvements and experimental validation.

0 favorites 0 likes
#multi-domain

@omarsar0: Exciting to see that world models are learning how to act. Odyssey-3 previews a single world model powering robot arms,…

X AI KOLs Timeline ↗ · 2026-09-15 Cached

Odyssey-3 is a new foundation world model that can control robots, cars, drones, and virtual worlds by reusing world knowledge across machines with minimal experience.

0 favorites 0 likes
#multi-domain

Turnbench: A Multi-Domain Benchmark for Turn-Taking Dynamics in Spoken Dialogue (25 minute read)

TLDR AI ↗ · 2026-09-15 Cached

TurnBench introduces a multi-domain benchmark for assessing turn-taking dynamics in spoken dialogue, featuring a hand-labeled corpus and standardized evaluation protocols for end-of-turn and interruption detection.

0 favorites 0 likes
#multi-domain

Routing Is Not Enough: Diagnosing Intra-Adapter Subspace Contention in MoE+LoRA Fine-Tuning

arXiv cs.LG ↗ · 2026-09-04 Cached

This paper diagnoses intra-adapter contention in MoE+LoRA fine-tuning and introduces SpawnLoRA to dynamically add sub-adapters, reducing negative transfer across domains.

0 favorites 0 likes
#multi-domain

FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes

arXiv cs.CL ↗ · 2026-08-28 Cached

FirstPass is a large-scale peer review dataset from Nature Communications, covering multiple scientific domains and multi-round dialogues to improve AI models for scientific judgment.

0 favorites 0 likes
#multi-domain

Flower Hub: A Reproducible Benchmarking Platform for Federated Learning in Simulation and Deployment

arXiv cs.LG ↗ · 2026-08-27 Cached

Flower Hub is a reproducible benchmarking platform for federated learning that enables execution and evaluation across both simulation and deployment runtimes.

0 favorites 0 likes
#multi-domain

Incorporating Cognitive Load and Knowledge Transfer for Multi-Domain Knowledge Tracing

arXiv cs.AI ↗ · 2026-08-26 Cached

This paper proposes LT-MKT, a method for multi-domain knowledge tracing that incorporates cognitive load and knowledge transfer using large language models to construct a hierarchical graph, achieving state-of-the-art performance on real-world datasets.

0 favorites 0 likes
#multi-domain

PAMT: Process-Aligned Reinforcement Learning for Multi-Domain Machine Translation

arXiv cs.CL ↗ · 2026-08-05 Cached

This paper proposes PAMT, a process-aligned reinforcement learning framework for multi-domain machine translation that combines domain-aware long chain-of-thought supervision with step-level process rewards to improve domain-sensitive translation decisions.

0 favorites 0 likes
#multi-domain

Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation

arXiv cs.CL ↗ · 2026-08-03 Cached

This paper introduces Translation with Thought (TwT), a resource-rational framework for multi-domain machine translation that adaptively modulates reasoning effort based on input difficulty, trained via supervised fine-tuning on difficulty-aware reasoning traces and reinforcement learning. TwT-7B and TwT-14B outperform larger SOTA reasoning models while reducing token usage by 32–60%.

0 favorites 0 likes
#multi-domain

Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction

Hugging Face Daily Papers ↗ · 2026-07-23 Cached

This paper introduces Tencent WorkBuddy Bench, a multi-domain evaluation suite for coding agents designed to resist data contamination by reverse-engineering tasks from real commits and business scenarios, covering Code, Web, Office, and Security domains.

0 favorites 0 likes
#multi-domain

Relay-Bench: Evaluating LLMs on Multi-Domain Reasoning Chains

arXiv cs.CL ↗ · 2026-07-22 Cached

Relay-Bench is a new benchmark designed to evaluate large language models on reasoning chains that require knowledge across multiple domains.

0 favorites 0 likes
#multi-domain

Candidate Attended Dialogue State Tracking Using BERT

arXiv cs.CL ↗ · 2026-07-20 Cached

The paper presents a scalable framework for multi-domain dialogue state tracking using BERT, achieving zero-shot generalization and improving performance on the SGD dataset.

0 favorites 0 likes
#multi-domain

Qwen/Qwen-AgentWorld-35B-A3B

Hugging Face Models Trending ↗ · 2026-06-22 Cached

Qwen releases Qwen-AgentWorld-35B-A3B, a native language world model that simulates agentic environments across seven domains via long chain-of-thought reasoning. The model is trained with a three-stage pipeline and supports MCP, Search, Terminal, SWE, Android, Web, and OS interactions.

0 favorites 0 likes
#multi-domain

Speaking the Language of Science: Toward a General-Purpose Generative Foundation Model for the Natural Sciences

Hugging Face Daily Papers ↗ · 2026-06-15 Cached

LOGOS is a scientific generative language model that encodes diverse scientific objects and spatial interactions as token sequences, enabling a unified autoregressive framework for tasks across natural sciences. Models at 1B, 3B, and 8B parameters show consistent performance scaling and are released to facilitate research.

0 favorites 0 likes
#multi-domain

Count Anything (2 minute read)

TLDR AI ↗ · 2026-06-15 Cached

Count Anything is a generalist model for text-guided object counting that unifies multiple domains, supported by the new CLOC dataset with 220K images across six visual domains. It achieves strong accuracy and multi-domain generalization.

0 favorites 0 likes
#multi-domain

PermDoRA -- Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry

arXiv cs.LG ↗ · 2026-06-11 Cached

This paper introduces DoRA-RBAC, a framework for composing LLM adapters, and tests whether geometry-aware merging improves multi-domain performance. Results show no consistent advantage over standard averaging, suggesting adapter interference is not primarily driven by parameter-space geometry.

0 favorites 0 likes
#multi-domain

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Hugging Face Daily Papers ↗ · 2026-06-10 Cached

Arbor is an AI framework for autonomous scientific research that uses a coordinator, executors, and a persistent hypothesis tree to iteratively improve research outcomes across multiple domains, achieving strong results on six real research tasks.

0 favorites 0 likes
#multi-domain

SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations

Hugging Face Daily Papers ↗ · 2026-06-04 Cached

SoCRATES introduces a realistic multi-domain benchmark for evaluating proactive LLM mediators, showing that top models resolve only about one-third of the consensus gap in conflict resolution.

0 favorites 0 likes
#multi-domain

A Local Perturbation Theory for Cross-Domain Interference and Recovery in Multi-Domain RL

Hugging Face Daily Papers ↗ · 2026-06-01 Cached

This paper proposes a local perturbation theory to explain cross-domain interference in multi-domain RL for LLMs, showing that interference is driven by a second-order damage term in a low-dimensional conflict subspace, and demonstrates that brief domain refresh or training-free rollback can selectively recover lost capabilities.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback