PXD025310
PXD025310 is an original dataset announced via ProteomeXchange.
Dataset Summary
Title | Aird: A computation-oriented mass spectrometry data format enables a higher compression ratio and less decoding time |
Description | We describe "Aird", an opensource and computation-oriented format with controllable precision, flexible indexing strategies, and high compression rate. Aird provides a novel compressor called Zlib-Diff-PforDelta (ZDPD) for m/z data. Compared with Zlib only, m/z data size is about 55% lower in Aird on average. With the high-speed decoding and encoding performance brought by the Single Instruction Multiple Data(SIMD) technology used in the ZDPD, Aird merely takes 33% decoding time compared with Zlib. We used the open dataset HYE, which contains 48 raw files from SCIEX TripleTOF 5600 and TripleTOF6600. The total file size is 206GB as the vendor format. The total size increases to 854GB after converting to mzML with 32-bit encoding precision. While it takes only 189GB when using Aird. Aird uses JavaScript Object Notation (JSON) for metadata storage. Aird-SDK is written in Java and AirdPro is a GUI client for vendor file converting which is written in C#. They are freely available at https://github.com/CSi-Studio/Aird-SDK and https://github.com/CSi-Studio/AirdPro |
HostingRepository | iProX |
AnnounceDate | 2021-04-12 |
AnnouncementXML | Submission_2023-08-28_00:34:30.019.xml |
DigitalObjectIdentifier | |
ReviewLevel | Peer-reviewed dataset |
DatasetOrigin | Original dataset |
RepositorySupport | Unsupported dataset by repository |
PrimarySubmitter | Xie Cong |
SpeciesList | scientific name: Homo sapiens; NCBI TaxID: 9606; |
ModificationList | No PTMs are included in the dataset |
Instrument | Q Exactive HF |
Dataset History
Revision | Datetime | Status | ChangeLog Entry |
---|---|---|---|
0 | 2021-04-11 20:54:53 | ID requested | |
1 | 2021-04-11 20:55:33 | announced | |
⏵ 2 | 2023-08-28 00:34:30 | announced | 2023-08-28: Update publication information. |
Publication List
Lu M, An S, Wang R, Wang J, Yu C, Aird: a computation-oriented mass spectrometry data format enables a higher compression ratio and less decoding time. BMC Bioinformatics, 23(1):35(2022) [pubmed] |
Keyword List
submitter keyword: Aird, DIA, DDA, PRM, Proteomics, Metabolomics |
Contact List
Miaoshan Lu | |
---|---|
contact affiliation | Westlake University |
contact email | lumiaoshan@westlake.edu.cn |
lab head | |
Xie Cong | |
contact affiliation | CSi Biotech Limited Liability Company |
contact email | 569130520@qq.com |
dataset submitter |
Full Dataset Link List
iProX dataset URI |