cross-attention

Tag

Cards List
#cross-attention

SMILESGNN: Interpretable Clinical Toxicity Prediction via SMILES-Graph Cross-Attention Fusion

arXiv cs.LG ↗ · 4d ago Cached

SMILESGNN introduces a multimodal architecture combining SMILES transformers and graph neural networks with cross-attention for interpretable drug toxicity prediction, achieving competitive performance on benchmarks like ClinTox and Tox21 with minimal parameters.

0 favorites 0 likes
#cross-attention

Learning Where to Look: A Shared Relative-Alignment Module for Time-Series Forecasting and PPG-to-Vital-Sign Reconstruction

arXiv cs.LG ↗ · 5d ago Cached

ROOSTER is a shared module that learns alignment between condition and target sequences for time-series forecasting and PPG-to-vital-sign reconstruction, achieving superior performance across multiple benchmarks.

0 favorites 0 likes
#cross-attention

All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation

Hugging Face Daily Papers ↗ · 6d ago Cached

The paper introduces RecCAR, a regularization method to address the reciprocal correspondence gap in joint multimodal diffusion transformers, improving performance in video generation tasks.

0 favorites 0 likes
#cross-attention

EnSol: an environment-aware graph neural network for molecular solubility prediction

arXiv cs.LG ↗ · 2026-09-21 Cached

EnSol is an environment-aware graph neural network that predicts molecular solubility by representing solutes and solvents as graphs and using cross-attention to model interactions, with probabilistic outputs to capture temperature effects and experimental uncertainty. It achieves state-of-the-art performance on benchmark datasets, validated experimentally.

0 favorites 0 likes
#cross-attention

@ProfTomYeh: Self-Attention vs Cross-Attention interactive diagram. Open https://byhand.ai/self-vs-cross

X AI KOLs Timeline ↗ · 2026-09-20 Cached

An interactive diagram comparing self-attention and cross-attention mechanisms in AI models, published as part of an educational library on attention.

0 favorites 0 likes
#cross-attention

QueryFormer: Winning Solution for KDD Cup 2026 Tencent UniRec Challenge

arXiv cs.AI ↗ · 2026-09-16 Cached

QueryFormer is a unified transformer architecture that won the KDD Cup 2026 Tencent UniRec Challenge for post-click conversion rate prediction, focusing on query generation and efficient scaling.

0 favorites 0 likes
#cross-attention

Time-Frequency Geometric Cross-Attention for Chunked Vision-Language-Action Models

arXiv cs.AI ↗ · 2026-09-11 Cached

This paper proposes Time–Frequency Geometric Cross-Attention (TFGCA), a drop-in module for chunked vision-language-action models that improves action trajectory prediction by decomposing chunks into time-frequency representations and capturing geometric relationships, resulting in significant performance gains on benchmarks and real-robot tasks.

0 favorites 0 likes
#cross-attention

RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting

Hugging Face Daily Papers ↗ · 2026-09-07 Cached

RelightFormer introduces a feed-forward generative Transformer for direct single- and multi-view image relighting, using cross-attention for illumination injection and permutation-invariant encodings for unordered views, trained on a massive synthetic dataset to achieve state-of-the-art visual quality.

0 favorites 0 likes
#cross-attention

The Attention Triangle in Audio-Video Models

arXiv cs.AI ↗ · 2026-09-04 Cached

This paper investigates semantic leakage in audio-video diffusion models through the 'attention triangle' of cross-attention mechanisms, and presents methods to enhance semantic grounding during generation.

0 favorites 0 likes
#cross-attention

Designing a Good Virtual Node: Addressable and Cardinality-Preserving Global Memory for Message Passing Architectures

arXiv cs.LG ↗ · 2026-08-05 Cached

This paper introduces addressable and cardinality-preserving global memory for message-passing neural networks via cross-attention slots, addressing the finite-capacity bottleneck of virtual nodes and improving performance on multiplicity-aware tasks.

0 favorites 0 likes
#cross-attention

Generic Vision and Cross-Attention for Reaction Yield Prediction

arXiv cs.LG ↗ · 2026-08-04 Cached

This paper proposes a dual-modal Vision Cross-Attention architecture for reaction yield prediction, fusing tabular physical-organic data with 2D molecular topologies, and demonstrates that a generic computer vision backbone can outperform purely quantum-based baselines.

0 favorites 0 likes
#cross-attention

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

arXiv cs.CL ↗ · 2026-07-28 Cached

Introduces SyRuP, a decoding-time framework that trains a cross-attention reward head to produce token-level adherence scores for system prompts, improving LLM following of complex prompts without model tuning.

0 favorites 0 likes
#cross-attention

TokenMem: Faithful Knowledge Injection for Frozen LLMs

arXiv cs.AI ↗ · 2026-07-28 Cached

TokenMem injects knowledge into frozen LLMs via a dedicated cross-attention channel, training a thin gating adapter through two-phase curriculum to improve knowledge compliance under counterfactual knowledge, achieving 69-70% KC compared to 20-52% for vanilla RAG.

0 favorites 0 likes
#cross-attention

OpenMOSS-Team/MOSS-VL-Realtime

Hugging Face Models Trending ↗ · 2026-07-14 Cached

MOSS-VL-Realtime is a realtime streaming vision-language model that processes continuous video frames, supports interruptible interaction, proactive silence, and dynamic correction, with timestamp-aware encoding and a 256K context window.

0 favorites 0 likes
#cross-attention

RaysUp: Ultra-light Universal Feature Upsampling via Geometry-Aware Ray Representation

Hugging Face Daily Papers ↗ · 2026-06-22 Cached

RaysUp is an ultra-lightweight, task-agnostic feature upsampling framework that uses geometry-aware ray domain techniques to reconstruct high-resolution features from low-resolution VFM outputs, achieving state-of-the-art performance with 84% fewer parameters than prior work and 7x faster inference.

0 favorites 0 likes
#cross-attention

KaLM-Reranker-V1: Fast but Not Late Interaction for Compressed Document Reranking

Hugging Face Daily Papers ↗ · 2026-06-22 Cached

KaLM-Reranker-V1 is a fast reranker that decouples query and passage computation using an encoder-decoder architecture with Matryoshka embedding pooling and cross-attention, achieving state-of-the-art reranking performance on BEIR and competitive results on multilingual benchmarks.

0 favorites 0 likes
#cross-attention

Enhancing Multilingual Reasoning via Steerable Model Merging

arXiv cs.CL ↗ · 2026-06-18 Cached

This paper proposes ST-Merge, a steerable model merging framework that uses a gated cross-attention mechanism to adaptively modulate contributions of a multilingual model and a reasoning model, outperforming fixed merging approaches on multilingual reasoning benchmarks across 21 languages.

0 favorites 0 likes
#cross-attention

Multi-Adapter PPO: A Cross-Attention Enhanced Wavelength Selection Framework for LIBS Quantitative Analysis

arXiv cs.LG ↗ · 2026-06-17 Cached

This paper introduces Multi-Adapter PPO, a reinforcement learning framework with cross-attention for wavelength selection in LIBS quantitative analysis, achieving 28.4% better composite scores and 45.2% improvement in prediction accuracy over traditional methods on steel and coal datasets.

0 favorites 0 likes
#cross-attention

Layer-Resolved Optimal Transport for Hallucination Detection in NMT and Abstractive Summarization

arXiv cs.CL ↗ · 2026-06-12 Cached

This paper extends optimal transport-based hallucination detection to all decoder layers in NMT and abstractive summarization, finding that detection is concentrated in early layers and that the geometric signal transfers poorly to summarization due to faithfulness failures not detectable via attention concentration.

0 favorites 0 likes
#cross-attention

SCALE: Scalable Cross-Attention Learning with Extrapolation for Agentic Workflow Scheduling

arXiv cs.LG ↗ · 2026-06-08 Cached

This paper proposes SCALE, a deep reinforcement learning scheduler for agentic LLM workflow DAGs that generalizes to unseen cluster sizes using cross-attention and structured representation regularization, reducing response time without retraining.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback