Tag
This paper introduces VectraYX-Vision-1B, a sub-2B Spanish/LATAM cybersecurity vision-language model, but reports a negative visual-grounding result, raising architectural questions about NoPE layers and releasing code, benchmarks, and checkpoints.
Presents VectraYX-Vision-1B, a sub-2B Spanish/LATAM cybersecurity vision-language model coupling a frozen SigLIP encoder with a Spanish decoder via an MLP, yet reports near-zero visual grounding despite functional pipelines, with open-source weights and remediation plans.
This paper probes character-level transformers to investigate whether they encode the Spanish L-shaped morphome, an irregular morphological pattern, as an abstract class or just surface alternations. The authors find that the encoding is item-specific and localized, but does not generalize like human learners.
Introduces MIRA-Ev, a clinical argument mining benchmark built on Spanish MIR licensing-exam cases, annotated with span-level premises, claims, and support/attack relations, available in Spanish, English, and Basque. It provides a three-tier task hierarchy for evaluating evidence sentence retrieval, argumentative component extraction, and relation classification, addressing the limitations of multiple-choice QA benchmarks in clinical NLP.
Introduces ESCUCHA, the first Spanish speech understanding benchmark for evaluating large audio language models across heterogeneous acoustic conditions and reasoning abilities, comprising 1,000 curated questions from diverse real-world sources.
Artículo que critica cómo la inteligencia artificial está premiando el volumen de datos y producción sobre la innovación real, sugiriendo un desequilibrio en el campo.
Introduces a cost-efficient human-LLM collaborative annotation framework to construct EspanStereo, a Spanish-language stereotype dataset covering multiple Spanish-speaking countries, enabling more culturally grounded bias evaluation in LLMs.
S-DiverSe is a 3.2-hour corpus of Spanish speech from 22 speakers with neurological conditions (ALS, Parkinson's, stroke), designed to support ASR evaluation for pathological speech. Baseline experiments show heuristic post-processing outperforms fine-tuning for this domain.
This paper describes HULAT2-UC3M's participation in the MER-TRANS 2026 shared task on Spanish Easy-to-Read generation, using a governed multi-agent workflow with LangGraph and Gemini/RigoChat models, achieving best SARI of 44.05.
Presents VectraYX-Nano, a 42M-parameter decoder-only language model trained from scratch in Spanish for cybersecurity, featuring curriculum learning, native tool invocation via MCP, and a 170M-token corpus. Empirical findings reveal a loss-versus-register inversion and corpus-density artifacts for tool-use capability.