Tag
A developer shares the challenge of creating a gold standard evaluation dataset for an AI product with no users, considering synthetic data generation and adversarial testing to avoid post-launch restructuring.
This paper investigates when machine learning outperforms traditional value sorting for exposure-weighted shipment prioritization using three datasets.
This survey examines computational humor understanding in multimodal LLMs, covering methods, datasets, evaluation protocols, and challenges such as shortcut-prone evaluation and weak evidence grounding.
This position paper argues that ground truth datasets in machine learning are not objective truths but human constructions shaped by choices, and advocates for articulating these choices to improve reliability, transparency, and accountability.
This project open-sources datasets that can be used for algorithm training and research.
The Well is an open-source collection of 15TB of physics simulation datasets spanning 16 domains, created by Flatiron Institute and 11 universities, designed to train PDE surrogate models, enabling researchers to replace expensive supercomputer simulations with neural networks.
A systematic review of non-social media free-text datasets for mental health disorder detection, identifying biases and gaps in current resources.
LeRobot v0.6.0 is a major release of Hugging Face's robot learning library, adding world model policies (VLA-JEPA, FastWAM, LingBot-VA), new VLAs, reward models, six simulation benchmarks, depth sensing, VLM-powered dataset annotation, custom video encoding, cloud training, and a deployment CLI for human-in-the-loop corrections.
HuggingFace CEO Clement Delangue compiled 250 open AI milestones from the US, highlighting contributions like Transformers, PyTorch, BERT, GPT-2, and Llama, with a call to maintain openness in AI development.
A collection of cybersecurity datasets for machine learning and model training, covering network traffic, malware, web attacks, phishing, and more, including notable public research datasets like LANL, CTU-13, and UNSW-NB15.
A curated collection of GNN papers, datasets, and implementation tools, hosted on GitHub.
Strauss Zelnick explains that AI is limited by backward-looking data and can reproduce the known but not create breakthroughs, placing value on human decisions about what to build.
This article discusses how China has rapidly advanced in AI despite being a latecomer, questioning the sources of datasets, computing power, and algorithms that enabled companies like DeepSeek to catch up with US leaders like OpenAI and Google.
Today, Macrodata Labs announced its launch, along with Refiner, an open-source framework for processing robotics datasets. The framework aims to help teams extract more signal from demonstrations and sensor data.
The author, an ML student, questions the robotics community about data interoperability issues and proposes an experiment to normalize and enrich public robotics datasets for better reuse.
This paper presents open multimodal datasets and open-source software packages for reproducible AI-enabled thermal-fluid research, introducing a spatial-temporal dimensionality framework and tools like SeqReg for sequence regression.
A developer thanks 5,000 followers for supporting OpenMed, an open-source clinical AI project, and hints at upcoming model and dataset releases.
HuggingFace benchmark datasets now allow filtering by model size, enabling comparisons like 'best model under 32B on swebenchverified'.
A personal reflection on the challenges and allure of training an AI model from scratch, highlighting the difficulties with data, hardware, and scaling, while noting that surprisingly good small models can be trained on modest hardware.
A discussion on the scarcity of realistic datasets for AI agent workflows, noting that existing benchmarks fail to capture messy production scenarios like tool failures, ambiguous requests, and long conversational drift, and seeking recommendations for better datasets.