Tag
This paper proposes domain adaptation of Sentence Transformer models to automate the mapping between cloud security controls and technical metrics, achieving significant performance gains over zero-shot baselines on control-to-metric and cross-standard association tasks.
Introducing the Ettin Reranker family: six new state-of-the-art CrossEncoder rerankers at various sizes, built on ModernBERT encoders, with open-source data and training recipe.
This article introduces ProtSent, a contrastive fine-tuning framework for protein language models that improves embedding quality for downstream tasks like remote homology detection and structural retrieval.
Developer seeks advice on handling English-Hindi code-mixed text classification without heavy LLMs, as sentence transformers fail on Romanized Hindi.
This article provides a technical guide on training and fine-tuning multimodal embedding and reranker models using the Sentence Transformers library, demonstrating performance improvements on Visual Document Retrieval tasks with Qwen3-VL.
Sentence Transformers v5.4 introduces support for multimodal embedding and reranking, allowing users to encode and compare text, images, audio, and video using a unified API.
A domain name valuation AI model using neural networks with sentence transformers and domain tokenization, providing instant price estimates across auction, marketplace, and brokerage channels.