@cgeorgiaw: support for mmCIF, PDB, fasta, and fastq just hit the `datasets` library this should massively accelerate bio workflows…
Summary
Support for mmCIF, PDB, fasta, and fastq formats has been added to the Hugging Face datasets library, which should significantly accelerate bioinformatics workflows.
View Cached Full Text
Cached at: 09/12/26, 08:49 AM
support for mmCIF, PDB, fasta, and fastq just hit the datasets library
this should massively accelerate 🧬🧫🧑🔬 bio workflows on @huggingface
huge kudos @b_azarkhalili & come coming soon 📈 https://t.co/uVfqY3OgoU
Similar Articles
@EmmaScharfmann: A very impressive new tools to browse neuroscience datasets was just released a few days ago: https://datasets.neuro2.a…
A new tool for browsing over 10,000 neuroscience datasets has been released, featuring nearly 2,000 datasets from Hugging Face.
@ClementDelangue: The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re …
Hugging Face releases Carbon, an open-source DNA base model that is 275x faster than comparable models, enabling local processing of whole genomes on a single GPU.
@adithya_s_k: Wake up ppl Huggingface just open sourced Genomic Foundational Models
Huggingface has open-sourced genomic foundational models, including Carbon, a DNA model that is 275x faster than the next best model and can process the entire human genome on a single GPU in under 2 days.
@huggingface: We've just hit 1M open datasets on the Hugging Face Hub Open models need open data. Today we hit that milestone, togeth…
Hugging Face announces that its Hub has reached a milestone of 1 million open datasets, highlighting the importance of open data for open models.
1M datasets on HF !
Celebrating a community milestone of 1 million datasets on Hugging Face, highlighting the collaborative effort to advance AI through open data.