multimodal-language-models

Tag

Cards List
#multimodal-language-models

Just MLPs: Efficient Visual State Reconstruction for Multimodal Language Models

Hugging Face Daily Papers ↗ · 3d ago Cached

The paper proposes δ-Vision, a method to reduce computational overhead in multimodal language models by using lightweight low-rank adapters to reconstruct visual states efficiently without discarding visual tokens.

0 favorites 0 likes
#multimodal-language-models

Calibrated Ambiguity in Multimodal Language Models: Humans reach for cultural references, while models describe the picture

arXiv cs.CL ↗ · 2026-09-14 Cached

This paper explores calibrated ambiguity as a generative resource in human communication versus multimodal language models, using the Dixit game to show that AI exhibits ambiguity collapse and lacks cultural references compared to humans.

0 favorites 0 likes
#multimodal-language-models

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

arXiv cs.LG ↗ · 2026-06-10 Cached

This paper introduces SPACE, the first source-free unlearning framework for multimodal large language models (MLLMs), which uses text-guided proxy anchor selection and dual-constraint semantic isolation to erase target concepts without requiring access to original training data, achieving performance comparable to data-dependent methods.

0 favorites 0 likes
#multimodal-language-models

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

Hugging Face Daily Papers ↗ · 2026-05-21 Cached

Researchers introduce the MM-OCEAN dataset and a three-tier evaluation framework for grounded personality reasoning in multimodal LLMs, revealing a 'Prejudice Gap' where models often make correct predictions without proper grounding.

0 favorites 0 likes
#multimodal-language-models

SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation

Hugging Face Daily Papers ↗ · 2026-05-21 Cached

SpaceDG is a large-scale dataset and benchmark that evaluates multimodal language models' spatial reasoning robustness under visual degradations like motion blur and low light, revealing significant performance gaps and showing that fine-tuning on SpaceDG improves robustness without degrading clean image performance.

0 favorites 0 likes
← Back to home

Submit Feedback