PDB identification code: Difference between revisions

From Proteopedia
Jump to navigationJump to search
Eric Martz (talk | contribs)
Eric Martz (talk | contribs)
 
Line 39: Line 39:
In addition to increasing the number of possible accession codes from ~4 x 10<sup>4</sup> to >10<sup>9</sup>, this will facilitate "text mining detection of PDB entries in the published literature"<ref name="news1" />. The PDB also promises "For as long as practicable, the  
In addition to increasing the number of possible accession codes from ~4 x 10<sup>4</sup> to >10<sup>9</sup>, this will facilitate "text mining detection of PDB entries in the published literature"<ref name="news1" />. The PDB also promises "For as long as practicable, the  
wwPDB will continue assigning PDB codes that can be truncated losslessly  
wwPDB will continue assigning PDB codes that can be truncated losslessly  
to the current four-character style."<ref name="news1" />
to the current four-character style."<ref name="news1" /> When 4-character codes are exhausted, new entries will be available in [[Atomic_coordinate_file#mmCIF_Data_Format|mmCIF format]] only, since the legacy [[PDB format]] will not accommodate 12-character IDs.
 
In 2024, the [[wwPDB]] plans to make a beta 12-character ID archive available in 2026<ref name="spring2024" />. In 2024, the wwPDB estimates that the 4-character IDs will be consumed in 2029<ref name="spring2024" />.


===Versioning===
===Versioning===