Is this you? Claim this profile to correct it, add a bio and choose the work people see first.
Claim this profileWorks6 from public data
- OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text60
This paper introduces OmniCorpus, a 10 billion-scale image-text interleaved dataset that has 15 times larger scales while maintaining good data quality and is more flexible, easily degradable from an image-text interleaved format to pure text corpus and image-text pairs.
- 10
- 1
- –
- –
- Integrated Data Platform and Key Technologies for Intelligent Mining Operations–
The contribution of this paper lies in providing theoretical support and practical guidance for the intelligent transformation of the mining industry, helping mining companies fully leverage the value of data, and enhancing overall operational capabilities and decision-making levels.
Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.
Report an error
Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.