lvwerra's picture
lvwerra HF Staff
Describe the dataset at the top of both tabs (#15)
19eb242
|
Raw History Blame Contribute Delete
850 Bytes
<!--
Shown at the top of both the Genome Atlas and the Database tab.
Everything above a "more" marker is always visible. Anything below one,
written as <!- - more - -> without the spaces, goes into a closed
"How the annotations were made" disclosure; with no marker there is no
disclosure. HTML comments like this one are stripped before rendering.
-->
The Carbon Annotation Database contains 566 million candidate protein-coding genes predicted by Carbon-A, an open 1.2-billion-parameter model that finds genes directly from DNA. It covers genomes from over 22,000 species, including animals, plants, fungi, and protists. Each prediction links to its source genome, genomic coordinates, coding DNA, predicted protein sequence, and confidence score, helping researchers explore poorly annotated genomes and prioritize candidates for further study.