Many countries have integrated wastewater monitoring systems with their infectious disease surveillance systems to enhance public health response capabilities.An Information System for the Chinese Urban Wastewater Sur...Many countries have integrated wastewater monitoring systems with their infectious disease surveillance systems to enhance public health response capabilities.An Information System for the Chinese Urban Wastewater Surveillance System(CWSS-IS)was developed based on the unified digital infrastructure of the China CDC.The primary functional modules of the CWSS-IS include information on monitoring sites,wastewater sample collection,relevant physicochemical indicators,qualitative and quantitative laboratory results,and sequencing data.The system implements unified data collection indicators and formats and standardizes data processing procedures and quality control(QC)rules.Launched nationally in February 2024,the CWSS-IS covers 169 cities with>3,000 registered users from CDCs,tracking multiple biomarkers in wastewater treatment plants,hospitals,communities,markets,and inbound flights.By January 2026,data had been collected from 118,729 samples.Compared to previous email-based reporting methods,the CWSS-IS demonstrated significant advantages in terms of efficiency,convenience,security,and scalability.This system offers valuable insights into the prevalence of infectious diseases and can effectively inform public health initiatives.Additionally,it serves as a standard paradigm for developing regional wastewater monitoring information systems.Future efforts should focus on exploring multisource data fusion standards,artificial intelligence frameworks,and large-scale data computational platforms to enhance early warning capabilities.展开更多
The National Data Administration announced the establishment of the“data infrastructure technology community”and launched the“data standard semantic service platform”on April 28,which aims to accelerate the nation...The National Data Administration announced the establishment of the“data infrastructure technology community”and launched the“data standard semantic service platform”on April 28,which aims to accelerate the national data infrastructure and data standardization process during the 15th Five-Year Plan period(2026-2030).展开更多
Geotechnical engineering serves as an indispensable pillar in national economic development, imposing stringent and meticulous requirements for project safety assurance and construction quality. Traditional methods of...Geotechnical engineering serves as an indispensable pillar in national economic development, imposing stringent and meticulous requirements for project safety assurance and construction quality. Traditional methods of geotechnical survey data collection and management have proven outdated, with fragmented data sources and isolated information silos that hinder comprehensive analysis and utilization. To address these challenges, a specialized data integration technology has been developed to resolve practical issues in geotechnical projects, ensuring systematic organization and rational allocation of data for collaborative use. The study thoroughly examines critical issues arising during data formatting, management, and transmission, proposing a technical solution centered on data standardization and information integration. This approach emphasizes unified data formats and standardized processing procedures, significantly enhancing seamless information exchange across platforms and departments. A detailed management framework covers the entire workflow—from data collection and preliminary processing to centralized storage—ensuring orderly execution at every stage. Practical applications demonstrate remarkable effectiveness: the integrated technology improves data management efficiency, reduces error rates, and provides reliable decision-support data for field operations. Research confirms its exceptional performance in geotechnical projects, effectively supporting quality control and risk mitigation while driving advancements in engineering informatization. This study significantly enriches the theoretical framework for geotechnical engineering data management. In practical applications, it demonstrates substantial potential for widespread adoption, effectively enhancing overall project quality while markedly improving construction safety—delivering profound significance and value.展开更多
Liquid-based cytology(LBC)has become a core technology in cervical cancer screening,and artificial intelligence(AI)shows great potential in addressing issues such as the global shortage of cytopathologists and large v...Liquid-based cytology(LBC)has become a core technology in cervical cancer screening,and artificial intelligence(AI)shows great potential in addressing issues such as the global shortage of cytopathologists and large variations in diagnostic results.However,the clinical reliability of AI systems in the field of cervical cytology fundamentally depends on the quality,diversity,and representativeness of their training and validation datasets.Drawing on evidence-based medicine principles,the latest research progress in China and globally,and clinical practices,these guidelines establish standardized requirements for sample diversity and data sufficiency in cervical LBC AI datasets to reduce algorithmic bias.The objective is to improve the real-world applicability of models and ensure their safe clinical application.The guidelines strictly comply with standardized guideline development specifications,with transparent expert panel organization,systematic literature retrieval,standardized evidence grading,three-round Delphi expert consensus,external peer review,and dynamic update mechanisms.All quantitative threshold indicators in the recommendations are jointly formulated based on highquality clinical evidence and expert consensus,with clear evidence sources and consensus construction processes to enhance the transparency,credibility,and operability of the guideline.展开更多
Standardized datasets are foundational to healthcare informatization by enhancing data quality and unleashing the value of data elements.Using bibliometrics and content analysis,this study examines China's healthc...Standardized datasets are foundational to healthcare informatization by enhancing data quality and unleashing the value of data elements.Using bibliometrics and content analysis,this study examines China's healthcare dataset standards from 2011 to 2025.It analyzes their evolution across types,applications,institutions,and themes,highlighting key achievements including substantial growth in quantity,optimized typology,expansion into innovative application scenarios such as health decision support,and broadened institutional involvement.The study also identifies critical challenges,including imbalanced development,insufficient quality control,and a lack of essential metadata—such as authoritative data element mappings and privacy annotations—which hampers the delivery of intelligent services.To address these challenges,the study proposes a multi-faceted strategy focused on optimizing the standard system's architecture,enhancing quality and implementation,and advancing both data governance—through authoritative tracing and privacy protection—and intelligent service provision.These strategies aim to promote the application of dataset standards,thereby fostering and securing the development of new productive forces in healthcare.展开更多
Viral infectious diseases,characterized by their intricate nature and wide-ranging diversity,pose substantial challenges in the domain of data management.The vast volume of data generated by these diseases,spanning fr...Viral infectious diseases,characterized by their intricate nature and wide-ranging diversity,pose substantial challenges in the domain of data management.The vast volume of data generated by these diseases,spanning from the molecular mechanisms within cells to large-scale epidemiological patterns,has surpassed the capabilities of traditional analytical methods.In the era of artificial intelligence(AI)and big data,there is an urgent necessity for the optimization of these analytical methods to more effectively handle and utilize the information.Despite the rapid accumulation of data associated with viral infections,the lack of a comprehensive framework for integrating,selecting,and analyzing these datasets has left numerous researchers uncertain about which data to select,how to access it,and how to utilize it most effectively in their research.This review endeavors to fill these gaps by exploring the multifaceted nature of viral infectious diseases and summarizing relevant data across multiple levels,from the molecular details of pathogens to broad epidemiological trends.The scope extends from the micro-scale to the macro-scale,encompassing pathogens,hosts,and vectors.In addition to data summarization,this review thoroughly investigates various dataset sources.It also traces the historical evolution of data collection in the field of viral infectious diseases,highlighting the progress achieved over time.Simultaneously,it evaluates the current limitations that impede data utilization.Furthermore,we propose strategies to surmount these challenges,focusing on the development and application of advanced computational techniques,AI-driven models,and enhanced data integration practices.By providing a comprehensive synthesis of existing knowledge,this review is designed to guide future research and contribute to more informed approaches in the surveillance,prevention,and control of viral infectious diseases,particularly within the context of the expanding big-data landscape.展开更多
With the continuous advancement of the tiered diagnosis and treatment system,the medical consortium model has gained increasing attention as an important approach to promoting the vertical integration of healthcare re...With the continuous advancement of the tiered diagnosis and treatment system,the medical consortium model has gained increasing attention as an important approach to promoting the vertical integration of healthcare resources.Within this context,laboratory data,as a key component of healthcare information systems,urgently requires efficient sharing and intelligent analysis.This paper designs and constructs an intelligent early warning system for laboratory data based on a cloud platform tailored to the medical consortium model.Through standardized data formats and unified access interfaces,the system enables the integration and cleaning of laboratory data across multiple healthcare institutions.By combining medical rule sets with machine learning models,the system achieves graded alerts and rapid responses to abnormal key indicators and potential outbreaks of infectious diseases.Practical deployment results demonstrate that the system significantly improves the utilization efficiency of laboratory data,strengthens public health event monitoring,and optimizes inter-institutional collaboration.The paper also discusses challenges encountered during system implementation,such as inconsistent data standards,security and compliance concerns,and model interpretability,and proposes corresponding optimization strategies.These findings provide a reference for the broader application of intelligent medical early warning systems.展开更多
Systematized nomenclature of medicine—clinical terms(SNOMED CT),one of the most comprehensive clinical terminology systems,is pivotal in enhancing healthcare interoperability,clinical data governance,and medical arti...Systematized nomenclature of medicine—clinical terms(SNOMED CT),one of the most comprehensive clinical terminology systems,is pivotal in enhancing healthcare interoperability,clinical data governance,and medical artificial intelligence(AI)development globally.In China,with the rapid growth of large-scale models and an increasing emphasis on transforming the intrinsic value of healthcare data,the absence of a nationally unified clinical terminology standard poses significant challenges.This commentary provides an in-depth analysis of the benefits of SNOMED CT for global healthcare,examines the critical deficiencies in Chinese healthcare big data and AI development due to the lack of standardized terminology,and outlines the technical,administrative,and educational challenges encountered in deploying SNOMED CT within Chinese environments.Special emphasis is laid on the potential of advanced large language models in facilitating the mapping of Chinese clinical data to SNOMED CT.We further discuss the necessity of high-quality data standardization in advancing medical AI in China.Finally,key conclusions and a roadmap for overcoming these challenges are proposed.展开更多
Dear Editor,With the advancement of genomic sequencing technology and the continuous reduction in costs,an increasing number of species have undergone whole-genome sequencing,which provides a wealth of resources for c...Dear Editor,With the advancement of genomic sequencing technology and the continuous reduction in costs,an increasing number of species have undergone whole-genome sequencing,which provides a wealth of resources for comparative genomics and functional genomics research[1-3].In the field of plant genomics and bioinformatics research,the standardization,uniformity,and accessibility of genomic data are extremely important.However,the diversity of data submission pathways has resulted in numerous data download channels.Consequently,it is challenging to ensure the accuracy and standardized consistency of the genomic data.Furthermore,some databases are no longer accessible for use.Typically,genomic data for individual species are scattered across a multitude of databases,with sequence IDs being prone to change during the transfer process between these repositories.Although numerous databases exist,each focusing on specific families,genera,or taxonomic categories[4-7],there remains a void for a comprehensive repository that includes genomic data for all sequenced plant species.展开更多
Progress in developing robust therapies for spinal cord injury (SCI), trau- matic brain injury (TBI) and peripheral nerve injury has been slow. A great deal has been learned over the past 30 years regarding both t...Progress in developing robust therapies for spinal cord injury (SCI), trau- matic brain injury (TBI) and peripheral nerve injury has been slow. A great deal has been learned over the past 30 years regarding both the intrinsic factors and the environmental factors that regulate axon growth, but this large body of information has not yet resulted in clinically available thera- peutics. This therapeutic bottleneck has many root causes, but a consensus is emerging that one contributing factor is a lack of standards for experi- mental design and reporting. The absence of reporting standards, and even of commonly accepted definitions of key words, also make data mining and bioinformatics analysis of neural plasticity and regeneration difficult, if not impossible. This short review will consider relevant background and poten- tial solutions to this problem in the axon regeneration domain.展开更多
To make inorganic structure data more useful for further studies a five-point list of simple procedures to be followed by authors of crystal structure papers is proposed. 1. A crystal structure should be described wit...To make inorganic structure data more useful for further studies a five-point list of simple procedures to be followed by authors of crystal structure papers is proposed. 1. A crystal structure should be described with the space group corresponding to its true symmetry. 2. A new structure proposal should be tested, if it is realistic in principle. 3. A structure should be described with a space group in a setting given in the International Tables. 4. For a comparison with other structures the structure data should be standardized with the program STRUCTURE TIDY. 5. 揘ew?structure data should be checked in the databases, Chemical Abstracts or on-line internet resources, if they are really new. The list is supplemented with many explanations, commentaries, examples and references.展开更多
Commentary Most would agree that providing comprehensive detail in scientific reporting is critical for the development of mean- ingful therapies and treatments for diseases. Such stellar practices 1) allow for repro...Commentary Most would agree that providing comprehensive detail in scientific reporting is critical for the development of mean- ingful therapies and treatments for diseases. Such stellar practices 1) allow for reproduction of experiments to con- firm results, 2) promote thorough analyses of data, and 3) foster the incremental advancement of valid approaches. Unfortunately, most would also agree we have far to go to reach this vital goal (Hackam and Redelmeier, 2006; Prinz et al., 2011; Baker et al., 2014).展开更多
In the course of network supported collaborative design,the data processing plays a very vital role.Much effort has been spent in this area,and many kinds of approaches have been proposed.Based on the correlative mate...In the course of network supported collaborative design,the data processing plays a very vital role.Much effort has been spent in this area,and many kinds of approaches have been proposed.Based on the correlative materials,this paper presents extensible markup language(XML)based strategy for several important problems of data processing in network supported collaborative design,such as the representation of standard for the exchange of product model data(STEP)with XML in the product information expression and the management of XML documents using relational database.The paper gives a detailed exposition on how to clarify the mapping between XML structure and the relationship database structure and how XML-QL queries can be translated into structured query language(SQL)queries.Finally,the structure of data processing system based on XML is presented.展开更多
Data as a new factor of production, only flow, sharing, processing can create value. Nowadays, data governance has become the only way for the digital transformation of enterprises. How to successfully implement a dat...Data as a new factor of production, only flow, sharing, processing can create value. Nowadays, data governance has become the only way for the digital transformation of enterprises. How to successfully implement a data governance project has become the most concerned issue for everyone. This paper will mainly focus on the implementation steps of data governance project and the functions of tool platform, and put forward the elements of successful data governance based on practical experience.展开更多
1.1. Development of international data exchange standards in securities field Securities market involves a large number of participants, like investors, securities companies, exchanges, clearingcorporations and so on....1.1. Development of international data exchange standards in securities field Securities market involves a large number of participants, like investors, securities companies, exchanges, clearingcorporations and so on. Businesses among the participants are completed via data exchange. Therefore, the data exchange protocols serve an important factor to determine and promote the sate and rapid development of the securities market.展开更多
Life tables allow the exploration of insects'and terrestrial arthropods'biology,and how they respond to external factors.Data collection process has been partially standardized,but the presentation of results ...Life tables allow the exploration of insects'and terrestrial arthropods'biology,and how they respond to external factors.Data collection process has been partially standardized,but the presentation of results mainly depends on the purpose of the study.Two different data representations can be obtained from the raw dataset:the differential representation provides the distribution of the stage-development times,while the integral representation provides the number of individuals into the different life stages,over time.The representations provide relevant biological information,but they lead to a loss of information with respect to the raw dataset.To date,a conceptual explanation of how the two representations can be obtained from the raw data,and of their respective properties,is still missing;moreover,providing the raw dataset as supporting information of the published papers is still not customary.This paper highlights three main points:(i)how the two representations are obtained from life tables raw dataset;(ii)without raw data,it is not possible to switch between the two representations,with a subsequent loss of information;and(iii)why there is the need for a data collection standard.The conceptual explanation is further completed by an electronic file that could support data collection and sharing,and that automatically transform the data in the two representations.We believe that this study is a first step toward a more efficient diffusion of the information among the scientific community,maximizing the efforts made by scholars during the experimental and data analysis process.展开更多
Accurate characterization of the chemical composition of complex traditional Chinese medicine(TCM)is an essential foundation for the modern scientific interpretation of TCM principles.Mass spectrometry is the most dom...Accurate characterization of the chemical composition of complex traditional Chinese medicine(TCM)is an essential foundation for the modern scientific interpretation of TCM principles.Mass spectrometry is the most dominant technique in current research on the material basis of TCM,offering the highest sensitivity and the richest information provision.Establishing mass spectrometry databases represents the most effective approach to facilitating the structural analysis of TCM chemical components.This paper systematically searches and reviews literature published from January 2005 to January 2025 through online databases such as China National Knowledge Infrastructure,PubMed,and Web of Science,using“mass spectrometry database”and“traditional Chinese medicine”as keywords.It reviews the current status of seven TCM chemical component mass spectrometry databases and seven natural product mass spectrometry databases.The key advancements of these mass spectrometry databases for natural products are summarized,detailing their characteristics,search methodologies,included information,and data sources.Additionally,challenges related to data quality,standardization,timely updates,database interaction,retrieval functionality,and data sharing and security are discussed in depth.Furthermore,the paper explores prospective development directions for TCM mass spectrometry databases,emphasizing the importance of open data sharing,technological innovation,and data security.Through this analysis,the paper aims to offer theoretical guidance and practical recommendations for the precise identification of TCM components,as well as for the construction and application of these databases.展开更多
1|THE FUTURE OF HEALTH IS DIGITAL.Digital health has shown significant promise in improving health outcomes.However,its transformation faces various challenges,including resource distribution disparities across countr...1|THE FUTURE OF HEALTH IS DIGITAL.Digital health has shown significant promise in improving health outcomes.However,its transformation faces various challenges,including resource distribution disparities across countries,varying definitions and standards for digital solutions,and a lack of coordination in digital health investments[1].展开更多
In the data encryption standard (DES) algorithm, there exist several bit-switching functions, including permutations, expansion, and permuted choices. They are generally presented in the form of matrixes and realize...In the data encryption standard (DES) algorithm, there exist several bit-switching functions, including permutations, expansion, and permuted choices. They are generally presented in the form of matrixes and realized by using table look-up technique in the implementation of the cryptosystem. This paper presents explicit formulas for the initial permutation IP, its inverse IP-1 , the expansion function E, and the permuted choice PC_1. It also gives the program realizations of these functions in C++ applying these formulas. With the advantage of the omission of the storage space for these matrixes and the tedious inputs of tables in the implementations of DES, our experimental results shows that the explicit formulas are useful in some situations, such as wireless sensor networks where the memory capacity is limited, especially when the size of file for encrypting is not too large, preferably smaller than 256KB.展开更多
The paper mainly discusses the integrity of the forwarded subscription message guaranteed by secure channel which encrypted in data communication by using data encryption standard (DES) algorithm and chaos code algo...The paper mainly discusses the integrity of the forwarded subscription message guaranteed by secure channel which encrypted in data communication by using data encryption standard (DES) algorithm and chaos code algorithm between broker nodes in the routing process of the contentbased publish/subscribe system. It analyzes the security of the secure channel encrypted with data communication by DES algorithm and chaos code algorithm, and finds out the secure channel can be easily attacked by known plain text. Therefore, the paper proposes the improved algorithm of message encryption and authentication, combining encryption and the generation of the message authentication code together to finish scanning at one time, which enhances both the secure degree and running efficiency. This secure channel system has a certain reference value to the pub/sub system requiring highly communication security.展开更多
基金supported by the National Urban Wastewater Priority Infectious Diseases Pathogen Surveillance Program,Science and Technology Special Fund of Hainan Province(No.ZDYF2025SHFZ061)Young Scholar Scientific Research Foundation of the National Institute of Environmental Health,China CDC(2024YSR03).
摘要Many countries have integrated wastewater monitoring systems with their infectious disease surveillance systems to enhance public health response capabilities.An Information System for the Chinese Urban Wastewater Surveillance System(CWSS-IS)was developed based on the unified digital infrastructure of the China CDC.The primary functional modules of the CWSS-IS include information on monitoring sites,wastewater sample collection,relevant physicochemical indicators,qualitative and quantitative laboratory results,and sequencing data.The system implements unified data collection indicators and formats and standardizes data processing procedures and quality control(QC)rules.Launched nationally in February 2024,the CWSS-IS covers 169 cities with>3,000 registered users from CDCs,tracking multiple biomarkers in wastewater treatment plants,hospitals,communities,markets,and inbound flights.By January 2026,data had been collected from 118,729 samples.Compared to previous email-based reporting methods,the CWSS-IS demonstrated significant advantages in terms of efficiency,convenience,security,and scalability.This system offers valuable insights into the prevalence of infectious diseases and can effectively inform public health initiatives.Additionally,it serves as a standard paradigm for developing regional wastewater monitoring information systems.Future efforts should focus on exploring multisource data fusion standards,artificial intelligence frameworks,and large-scale data computational platforms to enhance early warning capabilities.
摘要The National Data Administration announced the establishment of the“data infrastructure technology community”and launched the“data standard semantic service platform”on April 28,which aims to accelerate the national data infrastructure and data standardization process during the 15th Five-Year Plan period(2026-2030).
摘要Geotechnical engineering serves as an indispensable pillar in national economic development, imposing stringent and meticulous requirements for project safety assurance and construction quality. Traditional methods of geotechnical survey data collection and management have proven outdated, with fragmented data sources and isolated information silos that hinder comprehensive analysis and utilization. To address these challenges, a specialized data integration technology has been developed to resolve practical issues in geotechnical projects, ensuring systematic organization and rational allocation of data for collaborative use. The study thoroughly examines critical issues arising during data formatting, management, and transmission, proposing a technical solution centered on data standardization and information integration. This approach emphasizes unified data formats and standardized processing procedures, significantly enhancing seamless information exchange across platforms and departments. A detailed management framework covers the entire workflow—from data collection and preliminary processing to centralized storage—ensuring orderly execution at every stage. Practical applications demonstrate remarkable effectiveness: the integrated technology improves data management efficiency, reduces error rates, and provides reliable decision-support data for field operations. Research confirms its exceptional performance in geotechnical projects, effectively supporting quality control and risk mitigation while driving advancements in engineering informatization. This study significantly enriches the theoretical framework for geotechnical engineering data management. In practical applications, it demonstrates substantial potential for widespread adoption, effectively enhancing overall project quality while markedly improving construction safety—delivering profound significance and value.
摘要Liquid-based cytology(LBC)has become a core technology in cervical cancer screening,and artificial intelligence(AI)shows great potential in addressing issues such as the global shortage of cytopathologists and large variations in diagnostic results.However,the clinical reliability of AI systems in the field of cervical cytology fundamentally depends on the quality,diversity,and representativeness of their training and validation datasets.Drawing on evidence-based medicine principles,the latest research progress in China and globally,and clinical practices,these guidelines establish standardized requirements for sample diversity and data sufficiency in cervical LBC AI datasets to reduce algorithmic bias.The objective is to improve the real-world applicability of models and ensure their safe clinical application.The guidelines strictly comply with standardized guideline development specifications,with transparent expert panel organization,systematic literature retrieval,standardized evidence grading,three-round Delphi expert consensus,external peer review,and dynamic update mechanisms.All quantitative threshold indicators in the recommendations are jointly formulated based on highquality clinical evidence and expert consensus,with clear evidence sources and consensus construction processes to enhance the transparency,credibility,and operability of the guideline.
摘要Standardized datasets are foundational to healthcare informatization by enhancing data quality and unleashing the value of data elements.Using bibliometrics and content analysis,this study examines China's healthcare dataset standards from 2011 to 2025.It analyzes their evolution across types,applications,institutions,and themes,highlighting key achievements including substantial growth in quantity,optimized typology,expansion into innovative application scenarios such as health decision support,and broadened institutional involvement.The study also identifies critical challenges,including imbalanced development,insufficient quality control,and a lack of essential metadata—such as authoritative data element mappings and privacy annotations—which hampers the delivery of intelligent services.To address these challenges,the study proposes a multi-faceted strategy focused on optimizing the standard system's architecture,enhancing quality and implementation,and advancing both data governance—through authoritative tracing and privacy protection—and intelligent service provision.These strategies aim to promote the application of dataset standards,thereby fostering and securing the development of new productive forces in healthcare.
基金supported by the National Natural Science Foundation of China(32370703)the CAMS Innovation Fund for Medical Sciences(CIFMS)(2022-I2M-1-021,2021-I2M-1-061)the Major Project of Guangzhou National Labora-tory(GZNL2024A01015).
摘要Viral infectious diseases,characterized by their intricate nature and wide-ranging diversity,pose substantial challenges in the domain of data management.The vast volume of data generated by these diseases,spanning from the molecular mechanisms within cells to large-scale epidemiological patterns,has surpassed the capabilities of traditional analytical methods.In the era of artificial intelligence(AI)and big data,there is an urgent necessity for the optimization of these analytical methods to more effectively handle and utilize the information.Despite the rapid accumulation of data associated with viral infections,the lack of a comprehensive framework for integrating,selecting,and analyzing these datasets has left numerous researchers uncertain about which data to select,how to access it,and how to utilize it most effectively in their research.This review endeavors to fill these gaps by exploring the multifaceted nature of viral infectious diseases and summarizing relevant data across multiple levels,from the molecular details of pathogens to broad epidemiological trends.The scope extends from the micro-scale to the macro-scale,encompassing pathogens,hosts,and vectors.In addition to data summarization,this review thoroughly investigates various dataset sources.It also traces the historical evolution of data collection in the field of viral infectious diseases,highlighting the progress achieved over time.Simultaneously,it evaluates the current limitations that impede data utilization.Furthermore,we propose strategies to surmount these challenges,focusing on the development and application of advanced computational techniques,AI-driven models,and enhanced data integration practices.By providing a comprehensive synthesis of existing knowledge,this review is designed to guide future research and contribute to more informed approaches in the surveillance,prevention,and control of viral infectious diseases,particularly within the context of the expanding big-data landscape.
摘要With the continuous advancement of the tiered diagnosis and treatment system,the medical consortium model has gained increasing attention as an important approach to promoting the vertical integration of healthcare resources.Within this context,laboratory data,as a key component of healthcare information systems,urgently requires efficient sharing and intelligent analysis.This paper designs and constructs an intelligent early warning system for laboratory data based on a cloud platform tailored to the medical consortium model.Through standardized data formats and unified access interfaces,the system enables the integration and cleaning of laboratory data across multiple healthcare institutions.By combining medical rule sets with machine learning models,the system achieves graded alerts and rapid responses to abnormal key indicators and potential outbreaks of infectious diseases.Practical deployment results demonstrate that the system significantly improves the utilization efficiency of laboratory data,strengthens public health event monitoring,and optimizes inter-institutional collaboration.The paper also discusses challenges encountered during system implementation,such as inconsistent data standards,security and compliance concerns,and model interpretability,and proposes corresponding optimization strategies.These findings provide a reference for the broader application of intelligent medical early warning systems.
基金National Key Research and Development Program of China,Grant/Award Number:2023YFC2706305。
摘要Systematized nomenclature of medicine—clinical terms(SNOMED CT),one of the most comprehensive clinical terminology systems,is pivotal in enhancing healthcare interoperability,clinical data governance,and medical artificial intelligence(AI)development globally.In China,with the rapid growth of large-scale models and an increasing emphasis on transforming the intrinsic value of healthcare data,the absence of a nationally unified clinical terminology standard poses significant challenges.This commentary provides an in-depth analysis of the benefits of SNOMED CT for global healthcare,examines the critical deficiencies in Chinese healthcare big data and AI development due to the lack of standardized terminology,and outlines the technical,administrative,and educational challenges encountered in deploying SNOMED CT within Chinese environments.Special emphasis is laid on the potential of advanced large language models in facilitating the mapping of Chinese clinical data to SNOMED CT.We further discuss the necessity of high-quality data standardization in advancing medical AI in China.Finally,key conclusions and a roadmap for overcoming these challenges are proposed.
基金supported by the Natural Science Foundation for Distinguished Young Scholars of Hebei(C2022209010)National Natural Science Foundation of China(32172583)+3 种基金Tangshan Science and Technology Plan Project(24130219C)the National Key Research and Development Program of China(2023YFF1002000)the Medical-engineering Integration Project of North China University of Science and Technology(ZD-YG-202411)the S&T Program of Hebei(23372505D).
摘要Dear Editor,With the advancement of genomic sequencing technology and the continuous reduction in costs,an increasing number of species have undergone whole-genome sequencing,which provides a wealth of resources for comparative genomics and functional genomics research[1-3].In the field of plant genomics and bioinformatics research,the standardization,uniformity,and accessibility of genomic data are extremely important.However,the diversity of data submission pathways has resulted in numerous data download channels.Consequently,it is challenging to ensure the accuracy and standardized consistency of the genomic data.Furthermore,some databases are no longer accessible for use.Typically,genomic data for individual species are scattered across a multitude of databases,with sequence IDs being prone to change during the transfer process between these repositories.Although numerous databases exist,each focusing on specific families,genera,or taxonomic categories[4-7],there remains a void for a comprehensive repository that includes genomic data for all sequenced plant species.
基金Research in the Lemmon/Bixby lab is supported by NIH grants NS080145 and NS059866by the Miami Project to Cure Paralysis
摘要Progress in developing robust therapies for spinal cord injury (SCI), trau- matic brain injury (TBI) and peripheral nerve injury has been slow. A great deal has been learned over the past 30 years regarding both the intrinsic factors and the environmental factors that regulate axon growth, but this large body of information has not yet resulted in clinically available thera- peutics. This therapeutic bottleneck has many root causes, but a consensus is emerging that one contributing factor is a lack of standards for experi- mental design and reporting. The absence of reporting standards, and even of commonly accepted definitions of key words, also make data mining and bioinformatics analysis of neural plasticity and regeneration difficult, if not impossible. This short review will consider relevant background and poten- tial solutions to this problem in the axon regeneration domain.
摘要To make inorganic structure data more useful for further studies a five-point list of simple procedures to be followed by authors of crystal structure papers is proposed. 1. A crystal structure should be described with the space group corresponding to its true symmetry. 2. A new structure proposal should be tested, if it is realistic in principle. 3. A structure should be described with a space group in a setting given in the International Tables. 4. For a comparison with other structures the structure data should be standardized with the program STRUCTURE TIDY. 5. 揘ew?structure data should be checked in the databases, Chemical Abstracts or on-line internet resources, if they are really new. The list is supplemented with many explanations, commentaries, examples and references.
摘要Commentary Most would agree that providing comprehensive detail in scientific reporting is critical for the development of mean- ingful therapies and treatments for diseases. Such stellar practices 1) allow for reproduction of experiments to con- firm results, 2) promote thorough analyses of data, and 3) foster the incremental advancement of valid approaches. Unfortunately, most would also agree we have far to go to reach this vital goal (Hackam and Redelmeier, 2006; Prinz et al., 2011; Baker et al., 2014).
基金supported by National High Technology Research and Development Program of China(863 Program)(No.AA420060)
摘要In the course of network supported collaborative design,the data processing plays a very vital role.Much effort has been spent in this area,and many kinds of approaches have been proposed.Based on the correlative materials,this paper presents extensible markup language(XML)based strategy for several important problems of data processing in network supported collaborative design,such as the representation of standard for the exchange of product model data(STEP)with XML in the product information expression and the management of XML documents using relational database.The paper gives a detailed exposition on how to clarify the mapping between XML structure and the relationship database structure and how XML-QL queries can be translated into structured query language(SQL)queries.Finally,the structure of data processing system based on XML is presented.
摘要Data as a new factor of production, only flow, sharing, processing can create value. Nowadays, data governance has become the only way for the digital transformation of enterprises. How to successfully implement a data governance project has become the most concerned issue for everyone. This paper will mainly focus on the implementation steps of data governance project and the functions of tool platform, and put forward the elements of successful data governance based on practical experience.
摘要1.1. Development of international data exchange standards in securities field Securities market involves a large number of participants, like investors, securities companies, exchanges, clearingcorporations and so on. Businesses among the participants are completed via data exchange. Therefore, the data exchange protocols serve an important factor to determine and promote the sate and rapid development of the securities market.
基金funded by the European Commission under the Marie Sklodowska Curie Actions Postdoctoral Fellowship(MSCA-PF-2022)project"PestFinder"Grant n.101102281.
摘要Life tables allow the exploration of insects'and terrestrial arthropods'biology,and how they respond to external factors.Data collection process has been partially standardized,but the presentation of results mainly depends on the purpose of the study.Two different data representations can be obtained from the raw dataset:the differential representation provides the distribution of the stage-development times,while the integral representation provides the number of individuals into the different life stages,over time.The representations provide relevant biological information,but they lead to a loss of information with respect to the raw dataset.To date,a conceptual explanation of how the two representations can be obtained from the raw data,and of their respective properties,is still missing;moreover,providing the raw dataset as supporting information of the published papers is still not customary.This paper highlights three main points:(i)how the two representations are obtained from life tables raw dataset;(ii)without raw data,it is not possible to switch between the two representations,with a subsequent loss of information;and(iii)why there is the need for a data collection standard.The conceptual explanation is further completed by an electronic file that could support data collection and sharing,and that automatically transform the data in the two representations.We believe that this study is a first step toward a more efficient diffusion of the information among the scientific community,maximizing the efforts made by scholars during the experimental and data analysis process.
基金the Beijing Natural Science Foundation(No.7252249)the National Natural Science Foundation of China(No.82104380)+1 种基金the Scientific and Technological Innovation Project of the China Academy of Chinese Medical Sciences(Nos.CI2023E002,CI2023C071YLL, CI2023C039YGL)the Fundamental Research Funds for the Central Public Welfare Research Institutes(No.ZZ14-YQ-047, ZZ15-WT-04)。
摘要Accurate characterization of the chemical composition of complex traditional Chinese medicine(TCM)is an essential foundation for the modern scientific interpretation of TCM principles.Mass spectrometry is the most dominant technique in current research on the material basis of TCM,offering the highest sensitivity and the richest information provision.Establishing mass spectrometry databases represents the most effective approach to facilitating the structural analysis of TCM chemical components.This paper systematically searches and reviews literature published from January 2005 to January 2025 through online databases such as China National Knowledge Infrastructure,PubMed,and Web of Science,using“mass spectrometry database”and“traditional Chinese medicine”as keywords.It reviews the current status of seven TCM chemical component mass spectrometry databases and seven natural product mass spectrometry databases.The key advancements of these mass spectrometry databases for natural products are summarized,detailing their characteristics,search methodologies,included information,and data sources.Additionally,challenges related to data quality,standardization,timely updates,database interaction,retrieval functionality,and data sharing and security are discussed in depth.Furthermore,the paper explores prospective development directions for TCM mass spectrometry databases,emphasizing the importance of open data sharing,technological innovation,and data security.Through this analysis,the paper aims to offer theoretical guidance and practical recommendations for the precise identification of TCM components,as well as for the construction and application of these databases.
基金The Science and Technology Innovation 2030 Major Project,Grant/Award Number:2023ZD0508506National Key Research and Development Program of China,Grant/Award Numbers:2022YFC2705001,2023YFC2706305。
摘要1|THE FUTURE OF HEALTH IS DIGITAL.Digital health has shown significant promise in improving health outcomes.However,its transformation faces various challenges,including resource distribution disparities across countries,varying definitions and standards for digital solutions,and a lack of coordination in digital health investments[1].
基金Supported by the National Natural Science Foundation of China (61272045)Natural Science Foundation of Outstanding Youth Team Project of Zhejiang Province (R1090138)Project of the State Key Laboratory of Information Security (Institute of Information Engineering, Chinese Academy of Sciences, Beijing)
摘要In the data encryption standard (DES) algorithm, there exist several bit-switching functions, including permutations, expansion, and permuted choices. They are generally presented in the form of matrixes and realized by using table look-up technique in the implementation of the cryptosystem. This paper presents explicit formulas for the initial permutation IP, its inverse IP-1 , the expansion function E, and the permuted choice PC_1. It also gives the program realizations of these functions in C++ applying these formulas. With the advantage of the omission of the storage space for these matrixes and the tedious inputs of tables in the implementations of DES, our experimental results shows that the explicit formulas are useful in some situations, such as wireless sensor networks where the memory capacity is limited, especially when the size of file for encrypting is not too large, preferably smaller than 256KB.
基金Supported by the National Natural Science Foun-dation of China (60273014)
摘要The paper mainly discusses the integrity of the forwarded subscription message guaranteed by secure channel which encrypted in data communication by using data encryption standard (DES) algorithm and chaos code algorithm between broker nodes in the routing process of the contentbased publish/subscribe system. It analyzes the security of the secure channel encrypted with data communication by DES algorithm and chaos code algorithm, and finds out the secure channel can be easily attacked by known plain text. Therefore, the paper proposes the improved algorithm of message encryption and authentication, combining encryption and the generation of the message authentication code together to finish scanning at one time, which enhances both the secure degree and running efficiency. This secure channel system has a certain reference value to the pub/sub system requiring highly communication security.