ranking

Tag

Cards List
#ranking

TasteBench: Multimodal Benchmark for Sensory Prediction, from Molecules to Sustainable Foods

arXiv cs.AI ↗ · 15h ago Cached

TasteBench introduces a multimodal benchmark and privacy-preserving competition for sensory prediction, spanning food-level ranking (21K+ human evaluations across 215 plant-based foods) and molecular-level taste classification (15K flavor molecules), with baselines reaching pairwise accuracy competitive with median human panelists. It aims to provide computational proxies for sustainable protein discovery, analogous to molecular docking in drug design.

0 favorites 0 likes
#ranking

Who has actually shaped AI? The top 50 AI researchers by citations

Reddit r/ArtificialInteligence ↗ · yesterday

An analysis of the top 50 most-cited AI researchers, highlighting how landmark papers like Attention is All You Need have shaped the field and whose work drives AI's influence.

0 favorites 0 likes
#ranking

Reverse Item Response Theory for Sparsity-Robust Ranking in Fragmented Cancer Drug-Response Matrices

arXiv cs.LG ↗ · 2d ago Cached

The paper introduces reverse Item Response Theory to pharmacogenomics, treating cancer types as latent "subjects" and drugs as "items" to achieve sparsity-robust ranking recovery in fragmented drug-response matrices, outperforming simple averaging on GDSC2 data across multiple missingness regimes.

0 favorites 0 likes
#ranking

Decision-Oriented Recommendation Reranking: An Empirical Study of Jev

Hugging Face Daily Papers ↗ · 5d ago Cached

This paper empirically studies Jev, a decision-oriented "System One Model," for personalized recommendation reranking, finding it occupies a distinct quality–latency operating regime compared with recommendation-specific models and pointwise/listwise LLM rerankers across multiple domains and candidate-set sizes.

0 favorites 0 likes
#ranking

Sonnet 5.5 is second on Artificial Analysis

Reddit r/singularity ↗ · 2026-09-28

Sonnet 5.5 ranks second on the Artificial Analysis leaderboard, indicating its high performance in AI benchmarks.

0 favorites 0 likes
#ranking

Equal Ranking Quality, Different Decisions: Measuring and Reducing Order Dependence in LLM Scorers

Hugging Face Daily Papers ↗ · 2026-09-26 Cached

This paper shows that LLM-based scorers producing equal ranking quality can still make different decisions when candidate order in the prompt changes, and introduces order-consistency SFT (OC-SFT) to penalize score disagreement across permutations. OC-SFT holds ranking quality while improving decision stability across reranking, response ranking, and multi-document QA tasks.

0 favorites 0 likes
#ranking

@FinanceYF5: Meta's Muse topped the US App Store charts in the past 7 days, ranking above ChatGPT.

X AI KOLs Timeline ↗ · 2026-09-23

Meta's Muse app has topped the US App Store charts in the past 7 days, outperforming ChatGPT in rankings.

0 favorites 0 likes
#ranking

LIGE-GR: A Smooth Leap from Ranking to Generative Recommendation in the LLM Era

arXiv cs.LG ↗ · 2026-09-17 Cached

This paper introduces LIGE-GR, a method that uses large language models to smoothly transition from traditional ranking to generative recommendation systems, aiming to improve recommendation performance in the LLM era.

0 favorites 0 likes
#ranking

@Morris_LT: Ranking of Individual Barriers in the AI Era: Self-brainwashing ability > Mental resilience > Credit > Execution > Judg…

X AI KOLs Timeline ↗ · 2026-09-16

The article ranks individual barriers in the AI era, prioritizing self-brainwashing ability over mental resilience, credit, execution, judgment, filtering ability, and information asymmetry.

0 favorites 0 likes
#ranking

@omooretweets: Highly recommend checking out the reviews for MapQuest’s mobile app (currently #1 on the U.S. App Store) if you want to…

X AI KOLs Timeline ↗ · 2026-09-02 Cached

A tweet recommends viewing reviews of MapQuest's mobile app, which is currently the top app on the U.S. App Store, noting its nostalgic appeal.

0 favorites 0 likes
#ranking

Learning Mixtures of Plackett-Luce Models for Multi-Objective Alignment

arXiv cs.LG ↗ · 2026-08-27 Cached

This paper proposes MoPLEx, an algorithm for learning mixtures of Plackett-Luce models to handle heterogeneous preferences in AI alignment, showing improved clustering and ranking accuracy over baselines.

0 favorites 0 likes
#ranking

We combined 3 public TTS leaderboards into one meta-ranking of 110 models

Reddit r/AI_Agents ↗ · 2026-08-26

A new meta-ranking combines three public TTS leaderboards into one unified ranking of 110 models across 46 providers, updated weekly with a fixed methodology.

0 favorites 0 likes
#ranking

@FinanceYF5: 5% of VCs create 90% of venture capital profits. Stanford GSB professor Ilya Strebulaev built a ranking based on 30 years and 230,000 investments, correcting for inflated valuations, round-by-round dilution, net income, and time decay, and does not use self-reported data from institutions. Its correlation with the Midas List is only 0.2...

X AI KOLs Following ↗ · 2026-08-26 Cached

Stanford GSB professor Ilya Strebulaev releases a venture capital ranking based on 30 years and 230,000 investments, showing that 5% of VCs create 90% of profits, and it has a low correlation with the Midas List.

0 favorites 0 likes
#ranking

@FinanceYF5: 2026 US Top 100 Venture Capital Firms Ranking Officially Released.

X AI KOLs Following ↗ · 2026-08-26 Cached

The 2026 US Top Venture Capital Firms Ranking has been officially released, listing the 100 leading firms in investments for the year.

0 favorites 0 likes
#ranking

@GaryMarcus: once again #1 rising in technology

X AI KOLs Following ↗ · 2026-08-24 Cached

Gary Marcus tweets that something has risen to #1 in technology, likely highlighting a notable trend or achievement in the tech field.

0 favorites 0 likes
#ranking

Language Chain in Alignment: Cross-lingual Ranking Preference Optimization

Hugging Face Daily Papers ↗ · 2026-08-24 Cached

Cross-lingual Ranking Preference Optimization (CRPO) is a novel framework that enhances multilingual LLM alignment by transferring English preference knowledge to target languages through hierarchical ranking optimization, demonstrating improved performance in instruction-following and knowledge utilization across multiple languages.

0 favorites 0 likes
#ranking

Shared Physics Responses Recover Hidden Rankings in Neural Operator Libraries

arXiv cs.LG ↗ · 2026-08-24 Cached

This paper presents a method to rank neural operator models during deployment using shared physics responses, achieving high accuracy without ground-truth reference solutions for scientific computing applications.

0 favorites 0 likes
#ranking

UMER: Unifying Embedding and Ranking via Pair-Aware Discriminative Reasoning for Universal Multimodal Retrieval

arXiv cs.AI ↗ · 2026-08-20 Cached

UMER introduces a unified framework for multimodal retrieval that combines embedding and ranking via pair-aware discriminative reasoning, achieving state-of-the-art performance on the MMEB-V2 benchmark.

0 favorites 0 likes
#ranking

One Score, Two Decisions: Selective Prediction on the Rare-Disease Tail

arXiv cs.LG ↗ · 2026-08-18 Cached

This paper analyzes selective prediction systems for rare-disease diagnosis, demonstrating that small open-weight LLMs have low recall on ultra-rare diseases and exploring the use of score margins for decision-making with limitations.

0 favorites 0 likes
#ranking

Forecast Collapse in Time-Series Foundation Models

Hugging Face Daily Papers ↗ · 2026-08-14 Cached

The paper identifies forecast collapse in time-series foundation models for hourly equity return prediction and introduces CalibRank to balance calibration and ranking, significantly improving cross-sectional correlation.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback