scientific-ai

Tag

Cards List
#scientific-ai

Terence Tao on “prematurely solving [a maths] problem by purely AI-powered methods”

Lobsters Hottest · yesterday

Terence Tao discusses concerns about solving mathematical problems prematurely using purely AI-powered methods, a sentiment the author believes also applies to programming.

0 favorites 0 likes
#scientific-ai

NS-Copilot: An LLM-Driven Agent System for Autonomous Neuroscience Analysis

arXiv cs.CL · 5d ago Cached

NS-Copilot is an LLM-driven multi-agent system that autonomously supports end-to-end workflows for diverse neuroscience analysis tasks, outperforming baselines on key benchmarks.

0 favorites 0 likes
#scientific-ai

@kavi_deniz: Introducing the Tamarind Model Router. You want to know which molecular AI model will work best for this specific input…

X AI KOLs Timeline · 2026-08-28 Cached

Tamarind introduces a molecular AI model router that automatically selects the best model for specific inputs based on benchmarks and input characteristics, rather than relying on average performance.

0 favorites 0 likes
#scientific-ai

FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes

arXiv cs.CL · 2026-08-28 Cached

FirstPass is a large-scale peer review dataset from Nature Communications, covering multiple scientific domains and multi-round dialogues to improve AI models for scientific judgment.

0 favorites 0 likes
#scientific-ai

K-Bench: measuring model performance on real scientific agent requests

arXiv cs.AI · 2026-08-25 Cached

The paper introduces K-Bench 01, a benchmark for evaluating AI agents on real scientific requests, revealing that no model consistently meets the threshold for acceptable performance, with overclaiming as a common failure.

0 favorites 0 likes
#scientific-ai

A Pathway to General-Purpose Scientific AI: Multimodal Comprehension of Scientific Images

arXiv cs.AI · 2026-08-17 Cached

This paper introduces the ALD/E-ImageMiner benchmark for multimodal comprehension of scientific images and discusses future research directions for general-purpose scientific AI, based on insights from the ICDAR 2026 competition.

0 favorites 0 likes
#scientific-ai

Training AI Scientists to Replicate Research

Lobsters Hottest · 2026-08-15 Cached

Inherent Labs introduces Faraday, a 27B-parameter AI Scientist agent trained via long-horizon reinforcement learning to replicate scientific research, outperforming Claude Opus 4.8 and GPT-5.5 on paper replication tasks.

0 favorites 0 likes
#scientific-ai

@omarsar0: A 27B agent just beat Claude Opus 4.8 and GPT-5.5 on held-out research replication. Replica turns paper replication int…

X AI KOLs Following · 2026-08-14 Cached

A 27B parameter AI scientist agent named Faraday surpasses Claude Opus 4.8 and GPT-5.5 on paper replication tasks by using a scalable reinforcement learning approach called Replica.

0 favorites 0 likes
#scientific-ai

Automated Data Readiness for Scientific AI

arXiv cs.AI · 2026-07-07 Cached

The paper presents REDI, an open-source framework that automates the transformation of raw scientific datasets into AI-ready data through a unified five-stage pipeline, with companion tool SetGo for FAIR compliance, evaluated across multiple scientific domains.

0 favorites 0 likes
#scientific-ai

@SynScience: Introducing OpenScience. A better, open-source Claude Science. • Any model: GLM, Kimi, DeepSeek, Claude, GPT, your own …

X AI KOLs Following · 2026-07-05 Cached

OpenScience is an open-source alternative to Claude Science, supporting multiple AI models and over 250 research skills, with native Atlas integration for reproducible research graphs and user-controlled infrastructure.

0 favorites 0 likes
#scientific-ai

NVIDIA Vera CPU Opens the Way for Agentic Scientific AI at Los Alamos National Laboratory

NVIDIA Blog · 2026-06-22 Cached

NVIDIA announces its Vera CPU will power new supercomputers at Los Alamos National Laboratory, delivering significant performance improvements for agentic AI simulations and scientific workloads.

0 favorites 0 likes
#scientific-ai

@OpenAI: Maria tested the idea across 10,080 reactions, and human chemists later validated representative results by hand. Under…

X AI KOLs · 2026-06-17 Cached

OpenAI and Molecule.one collaborated to have their AI systems (GPT-5.4 and Maria) autonomously select research areas, generate proposals, and run experiments in organic chemistry, achieving yield improvements for 88% of tested reactions — a first for AI-driven open-ended scientific discovery.

0 favorites 0 likes
#scientific-ai

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

Hugging Face Daily Papers · 2026-06-10 Cached

Notes2Skills is a two-stage framework that converts laboratory notes into verifiable skills for AI agents while preserving author uncertainty, enabling safer scientific AI systems.

0 favorites 0 likes
#scientific-ai

The Model Is No Longer the Bottleneck (6 minute read)

TLDR AI · 2026-06-09 Cached

Anthropic's Claude, a general-purpose AI model without chemistry fine-tuning, outperformed specialized software like ChemDraw and MestReNova in NMR analysis, suggesting that the bottleneck in scientific AI has shifted from model capability to workflow design.

0 favorites 0 likes
#scientific-ai

I-SAFE: Wasserstein Coherence Metrics for Structural Auditing of Scientific AI Models

arXiv cs.LG · 2026-05-22 Cached

This paper introduces I-SAFE, a post-hoc distributional auditing framework for scientific AI models using Wasserstein Coherence Metrics, which reveals structural differences in model outputs that accuracy-based evaluation fails to capture. Demonstrated on drug-target interaction prediction, the framework is model-agnostic and applicable to any domain with structured inputs and external priors.

0 favorites 0 likes
#scientific-ai

SandboxAQ brings its drug discovery models to Claude — no PhD in computing required

TechCrunch AI · 2026-05-18 Cached

SandboxAQ has integrated its large quantitative models (LQMs) for drug discovery and materials science into Anthropic's Claude, enabling researchers to use these powerful tools through a natural language interface without specialized computing infrastructure.

0 favorites 0 likes
#scientific-ai

Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning

arXiv cs.AI · 2026-05-11 Cached

This paper introduces ChemCost, a benchmark for evaluating how well LLM agents can estimate chemical procurement costs by grounding identities, retrieving quotes, and handling noise. It reveals that current agents struggle with robustness and precise arithmetic reasoning in scientific workflows.

0 favorites 0 likes
#scientific-ai

@OpenAI: GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in prote…

X AI KOLs · 2026-04-16 Cached

OpenAI releases GPT-Rosalind, a specialized life sciences model optimized for protein reasoning, chemical analysis, genomics, and scientific workflows.

0 favorites 0 likes
#scientific-ai

aipoch/open-science

GitHub Trending (daily) · yesterday Cached

AIPOCH Open Science is an open-source, local-first AI research workbench for reproducible science featuring scientific AI agents, Python and R execution, and cross-platform support. The content announces the v0.25.1 maintenance release, which includes fixes for protected notebook kernels on Windows and atomic file exports with time-zone-safe timestamps.

0 favorites 0 likes
← Back to home

Submit Feedback