tRNA-derived small RNAs(tsRNAs),as a class of regulatory small noncoding RNA,have been implicated in a wide variety of human diseases.Large amounts of tsRNA–disease associations have been identified in recent years f...tRNA-derived small RNAs(tsRNAs),as a class of regulatory small noncoding RNA,have been implicated in a wide variety of human diseases.Large amounts of tsRNA–disease associations have been identified in recent years from accumulating studies.However,repositories for cataloging the detailed information on tsRNA–disease associations are scarce.In this study,we provide a tsRNADisease database by integrating experimentally and computationally supported tsRNA–disease associations from manual curation of literatures and other related resources.tsRNADisease contains 5571 manually curated associations between 4759 tsRNAs and 166 diseases with experimental evidence from 346 studies.In addition,it also contains 5013 predicted associations between 1297 tsRNAs and 111 diseases.tsRNADisease provides a user-friendly interface to browse,retrieve,and download data conveniently.This database can improve our understanding of tsRNA deregulation in diseases and serve as a valuable resource for investigating the mechanism of disease-related tsRNAs.tsRNADisease is freely available at http://gffzz9c504e06f78b4edahoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn.展开更多
The ongoing battle between humans and pathogenic bacteria has fueled rapid microbial evolution.Although whole-genome sequencing(WGS)has transformed the ability to track genomic mutations,existing tools lack comprehens...The ongoing battle between humans and pathogenic bacteria has fueled rapid microbial evolution.Although whole-genome sequencing(WGS)has transformed the ability to track genomic mutations,existing tools lack comprehensive solutions for analyzing mutational patterns and their functional consequences in pathogenic bacteria.Here,we present CliPME,an innovative platform that bridges this critical gap by combining mutation detection,mutation effect prediction,and regulatory network analysis,specifically designed for bacterial genomics.We develop qMut,a high-performance R package designed for largescale mutation profiling.Coupled with three major functional modules called MutFinder,MutAnalyzer,and ExpMiner,CliPME integrates population-level mutation analysis,functional mutation predictions,and estimation of gene-gene expression relationships.Using Mycobacterium tuberculosis(Mtb)as a case study,we show the power of CliPME by identifying functionally significant mutations in the transcription factor Rv0324,and experimentally demonstrate the link of its genetic variation to potential adaptive phenotypes.This resource empowers researchers to decode evolutionary mechanisms in bacterial pathogens and may accelerate the translation of genomic insights into antimicrobial strategies.The web server of CliPME is freely accessible at http://gffzzd72233016bfa415csoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/.展开更多
A coupled three-dimensional cellular automata(CA)model has been used to predict the hydrogen porosity in an Al-Si alloy as a function of thermal boundary conditions.By quantifying the porosity distribution from simula...A coupled three-dimensional cellular automata(CA)model has been used to predict the hydrogen porosity in an Al-Si alloy as a function of thermal boundary conditions.By quantifying the porosity distribution from simulations,a porosity defect database was established,representing a cooling rate ranging from 0.25 to 50℃/s at an initial hydrogen content of 3.0×10-3 mL/g.Based on the database,four machine learning algorithms including support vector machine(SVM),random forest(RF),K-nearest neighbors(KNN),and gradient boosting machine(GBM)were trained and compared for each porosity characteristic to identify the optimal model.For the prediction of porosity percentage,the determination coefficient(R2)and the root mean square error(RMSE)on the test set reached 0.95 and 0.042,respectively.The predicted porosity distribution agreed well with experiments,indicating that the model can be used to map the porosity size in large casting components.展开更多
Primary liver cancer (PLC) is a major global healthchallenge, ranking as the sixth most common andthird most fatal malignancy worldwide, according toGLOBOCAN 2022 estimates[1]. This high mortalityrate underscores the ...Primary liver cancer (PLC) is a major global healthchallenge, ranking as the sixth most common andthird most fatal malignancy worldwide, according toGLOBOCAN 2022 estimates[1]. This high mortalityrate underscores the aggressive nature of thedisease and the significant burden it places on globalhealthcare systems. Although primary preventionremains the cornerstone of liver cancer control,improving outcomes for patients already diagnosedis equally critical for mitigating the impact of thedisease.展开更多
High-quality clinical databases are essential for advancing research and clinical practice in emergency and critical care medicine.With the widespread adoption of Electronic Health Record(EHR)systems across hospitals ...High-quality clinical databases are essential for advancing research and clinical practice in emergency and critical care medicine.With the widespread adoption of Electronic Health Record(EHR)systems across hospitals in China,vast amounts of clinical information have been digitally archived.Emergency departments(EDs)and intensive care units(ICUs)represent critically datademanded environments.展开更多
Background:The purpose of this study was to analyze and classify adverse drug events(ADEs)related to ceftazidime/avibactam reported in the Food and Drug Administration Adverse Event Reporting System(FAERS)database and...Background:The purpose of this study was to analyze and classify adverse drug events(ADEs)related to ceftazidime/avibactam reported in the Food and Drug Administration Adverse Event Reporting System(FAERS)database and to evaluate their potential safety signals since the drug’s market introduction.Methods:This analysis systematically extracted and filtered FAERS data for ceftazidime/avibactam from its market launch in 2015 to the last quarter of 2024,utilizing the Medical Dictionary for Regulatory Activities(MedDRA)terminology for ADE recoding.The analysis employed the reporting odds ratio(ROR)method to assess the strength of ADE signals and to identify significant diseases associated with infections,the hepatobiliary system,the urinary system,and the nervous system.Results:A review of 540 adverse reaction reports revealed significant signals of adverse effects related to infections,hepatobiliary disorders,urinary system issues,and neurological impairments,including pathogen resistance,liver and kidney function impairment,encephalopathy,thrombocytopenia,and toxic epidermal necrolysis.However,these issues require further clinical attention.Conclusion:Ceftazidime/avibactam is associated with a range of adverse reactions,necessitating enhanced clinical monitoring,particularly in patients with underlying liver or kidney dysfunction.Continuous risk assessment and vigilant monitoring are critical for its clinical use.However,this study is limited by inherent reporting biases and confounders associated with the spontaneous reporting database(FAERS).Future research should validate these signals through prospective cohort and mechanistic studies and explore personalized risk management strategies for high-risk populations.展开更多
Commercial phosphor-converted white LEDs(pc-WLEDs)face two inherent limitations,namely blue light hazard and low color rendering index,due to the use of blue LEDs as excitation source.To address these challenges,viole...Commercial phosphor-converted white LEDs(pc-WLEDs)face two inherent limitations,namely blue light hazard and low color rendering index,due to the use of blue LEDs as excitation source.To address these challenges,violet LEDs are proposed as an alternative solution.Currently,phosphors that can be efficiently excited by violet light(with wavelengths from 400 to 420 nm)remain under development still.In this study,we utilize large language models to construct a comprehensive database of Eu2+and Ce3+doped phosphors for discovering novel violet-excited phosphors.A total of 822 phosphor data entries,including elemental compositions,crystal structures and excitation/emission wavelengths,have been extracted and validated from 9551 research papers.Compared with Ce3+doped phosphors,the Eu2+are in general more suited for violet-excited phosphors,as well as red-emitting phosphors.In particular,Eu2+doped nitrides and sulfides are worth of exploration for violet-excited phosphors.This database is expected to be useful in the future development of phosphors for pc-WLEDs based on artificial intelligence methods.The datasets in this article are listed in Science Data Bank at http://gffzzd3cc09b8251d45dfhoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/10.57760/sciencedb.34314.展开更多
Single-atom catalysts(SACs)have shown great promise for ethane dehydrogenation(EDH),owing mainly to their near-100%atomic utilization,precisely tunable active sites,and superior catalytic performance.Studies reveal th...Single-atom catalysts(SACs)have shown great promise for ethane dehydrogenation(EDH),owing mainly to their near-100%atomic utilization,precisely tunable active sites,and superior catalytic performance.Studies reveal that the nature of active metals,the properties of support and the coordination environments are critical factors affecting EDH performance.Among various supports,graphene has emerged as an ideal support material due to its excellent thermal stability and flexible,tunable coordination structure.While previous research has explored the effects of different active metals and coordination environments,a systematic understanding of their underlying principles is still lacking due to fragmented data.To bridge this gap,this work constructed a systematic,comprehensive database of heteroatom-doped graphene-supported SACs,covering five representative metal single atoms and 51 distinct coordination environments grouped into six major categories.Through high-throughput calculations,multi-dimensional data were systematically obtained,including elementary reaction energies,vibrational frequencies,density of states and Bader charges.A rigorous quality control system was implemented at both the parameter-setting and computational result levels.The database also provides complete raw calculation files,offering reliable data support for in-depth analysis of the catalytic performance,structure-performance relationship and reaction mechanisms of heteroatom-doped graphene-supported SACs in EDH.展开更多
AIM:To screen for differentially expressed genes in retinoblastoma(RB)gene chips using GEO2R and validate them clinically.METHODS:The expression profile chip data(GSE110811)was downloaded from the public gene chip dat...AIM:To screen for differentially expressed genes in retinoblastoma(RB)gene chips using GEO2R and validate them clinically.METHODS:The expression profile chip data(GSE110811)was downloaded from the public gene chip database Gene Expression Omnibus(GEO).The GEO2R chip analysis platform was used to identify differentially expressed genes between RB and adjacent normal tissues.According to the International Intraocular Retinoblastoma Classification(IIRC)system,35 children diagnosed with RB from our hospital and other hospitals were enrolled as the RB group,and 35 healthy children who underwent physical examinations in our hospital were enrolled as the control group.The relative expression levels of Sprouty RTK signaling antagonist 2(SPRY2)and estrogen-related receptor beta(ESRRB)in the serum of patients were detected by quantitative reverse transcription polymerase chain reaction(qRT-PCR).The diagnostic value of SPRY2 and ESRRB in RB was evaluated by receiver operating characteristic(ROC)curves.Analysis of the relationship between SPRY2/ESRRB expression and clinicopathological features,as well as its correlation with the tumor marker CA199.RESULTS:In the GSE110811 chip,the expression levels of two genes,16780069(SPRY2)and 16786783(ESRRB),showed the most significant differences between RB and normal tissues.The relative expression levels of SPRY2 and ESRRB in the serum of children in RB group(22 males,age 1.64±1.08y)were significantly lower(P<0.05)than those in control group(25 males,age 1.54±0.95y).The area under the ROC curve for SPRY2 was 0.735(95%CI:0.616-0.854),while that for ESRRB was 0.880(95%CI:0.800-0.960).There were statistically significant differences in the expression of SPRY2 and ESRRB with respect to choroidal invasion,optic nerve invasion,differentiation degree,and clinical staging(P<0.05).In RB group,the expression levels of SPRY2 and ESRRB decreased gradually with increasing CA199 levels,showing a negative correlation(rSPRY2=-0.593,rESRRB=-0.423;both P<0.05).CONCLUSION:The expression of SPRY2 and ESRRB is closely related to the occurrence and development of RB and negatively correlated with the tumor marker CA199.They have the potential to serve as diagnostic biomarkers for RB.展开更多
[Objective]As high-dimensional vector data increasingly surpass the processing capabilities of traditional database management systems,Vector Databases(VDBs)have emerged and become tightly integrated with large langua...[Objective]As high-dimensional vector data increasingly surpass the processing capabilities of traditional database management systems,Vector Databases(VDBs)have emerged and become tightly integrated with large language models,being widely applied in modern artificial intelligence systems.However,existing research has primarily focused on underlying technologies such as approximate nearest neighbor search,with relatively few studies providing a systematic architectural-level review of VDBs or analyzing how these core technologies collectively support the overall capacity of VDBs.This survey aims to offer a comprehensive overview of the core designs and algorithms of VDBs,establishing a holistic understanding of this rapidly evolving field.[Methods]First,we systematically review the key technologies and design principles of VDBs from the two core dimensions of storage and retrieval,tracing their technological evolution.Next,we conduct an in-depth comparison of several mainstream VDB architectures,summarizing their strengths,limitations,and typical application scenarios.Finally,we explore emerging directions for integrating VDBs with large language models,including open research challenges and trends such as novel indexing strategies.[Conclusions]This survey serves as a systematic reference guide for researchers and practitioners,helping readers quickly grasp the technological landscape and development trends in the field of vector databases,and promoting further innovation in both theoretical and applied aspects.展开更多
This study analyzes asthma-related mortality trends among U.S.adults(≥25 years)from 1999 to 2023 using the Centers for Disease Control and Prevention's Wide-ranging Online Data for Epidemiologic Research(CDC WOND...This study analyzes asthma-related mortality trends among U.S.adults(≥25 years)from 1999 to 2023 using the Centers for Disease Control and Prevention's Wide-ranging Online Data for Epidemiologic Research(CDC WONDER)database,calculating age-adjusted mortality rates(AAMRs)per 100,000 population stratified by demographics and geography.By fitting log-linear regression models to estimate annual percent changes(APCs)with 95%confidence intervals(CI).Results show a significant overall decline in AAMR from 2.46(95%CI:2.39-2.53)in 1999 to 1.33(95%CI:1.29-1.38)in 2023(average annual percent change,AAPC:2.43).Adults aged 45-64 years accounted for the largest proportion of asthma-related deaths(32.22%of deaths),with significant decline(AAPC:1.71).Women accounted for 64.62%of deaths,exhibiting persistently higher AAMR than men(AAPC:2.60).Non-Hispanic(NH)Black individuals had the highest overall AAMR in 1999(5.56)and 2023(2.84),with a transient 2017-2020 surge(APC:9.40).The Northeast surpassed the West as the highest-burden region by 2023,while Louisiana achieved the lowest state-level AAMR(0.76).Strategies are recommend,including workplace screening(ages 45-64);sex-specific strategies for women;culturally tailored community health workers for NH Black individuals;and improved asthma action plans in the Northeast.展开更多
Background Identifying simple and effective prognostic markers is crucial for risk stratification in critically ill patients with coronary artery disease(CAD)and diabetes mellitus(DM).The anion gap(AG),a marker of aci...Background Identifying simple and effective prognostic markers is crucial for risk stratification in critically ill patients with coronary artery disease(CAD)and diabetes mellitus(DM).The anion gap(AG),a marker of acid-base homeostasis,has been associated with mortality in various diseases,but its specific role in this high-risk popu-lation remains unclear.Methods This retrospective cohort study included 4,343 patients from the MIMIC-Ⅳda-tabase.The exposure variable was the anion gap(AG),which was modeled in three ways:as a continuous variable,as a categorical variable(Low,Middle,and High groups),and as a dichotomous variable using a cutoff of 14.5.The primary outcome was in-hospital mortality.Multivariable logistic regression models were used to adjust for poten-tial confounders,including demographics,comorbidities,vital signs,and laboratory parameters.The linearity of the association was assessed using restricted cubic splines(RCS),and the predictive performance of AG was evaluated using receiver operating characteristic(ROC)curves.Results The overall in-hospital mortality rate was 12.6%.A strong,dose-response relationship was observed between AG levels and mortality.After full adjustment,compared to the low AG tertile,the high AG tertile was significantly associated with increased mortality risk(OR:1.92,95%CI:1.43-2.58,P<0.0001).AG as a continuous variable remained an independent predictor(OR per unit increase:1.11,95%CI:1.08-1.14,P<0.0001).RCS analysis confirmed a linear association(P for nonlinearity=0.880).In ROC analysis,AG demonstrated superior predictive ability[area under the curve(AUC),AUC=0.6957)]for in-hos-pital mortality compared to sodium(AUC=0.5338),potassium(AUC=0.5),and chloride(AUC=0.6141).Conclu-sions In this cohort of ICU patients with CAD and DM from the MIMIC-Ⅳdatabase,an elevated anion gap was independently and linearly associated with a higher risk of in-hospital mortality.AG outperformed other common electrolytes in mortality prediction,highlighting its potential as a valuable and readily available risk stratification tool in this population.展开更多
The journal of Meteorological and Environmental Research [ISSN: 2152-3940] has been included and stored by the following famous databases: CA, CABI, CSA, EBSCO, UPD, AGRIS, EA, Chinese Science and Technology Periodica...The journal of Meteorological and Environmental Research [ISSN: 2152-3940] has been included and stored by the following famous databases: CA, CABI, CSA, EBSCO, UPD, AGRIS, EA, Chinese Science and Technology Periodical Database, and CNKI, as well as Library of Congress, United States.CA (Chemical Abstracts) was founded in 1907, and is the most authoritative and comprehensive source for chemical information. Centre for Agriculture and Bioscience International(CABI)is a not-for-profit international agricultural information institute with headquarters in Britain. ProQuest CSA belongs to Cambridge Information Group (CIG), and it provides access to more than 100 databases published by CSA and its publishing partners. EBSCO is a large document service company with a history of more than 60 years, providing subscription and publication services of journals and documents. CNKI (China National Knowledge Infrastructure), universally acclaimed as the most valuable Chinese website, boasts the greatest information content, covering natural science, humanities and social science, engineering, periodical, doctor/master s dissertations, newspapers, books, meeting papers and other miscellaneous public information resources in China.展开更多
Objectives:Electronic health records(EHRs)offer valuable real-world data(RWD)for Chinese medicine research.However,significant methodological challenges remain in developing integrative Chinese-Western medicine(ICWM)d...Objectives:Electronic health records(EHRs)offer valuable real-world data(RWD)for Chinese medicine research.However,significant methodological challenges remain in developing integrative Chinese-Western medicine(ICWM)databases.This study aims to establish a best-practice methodological framework,referred to as BRIDGE,to guide the construction of ICWM databases using EHRs.Methods:We developed the methodological framework through a comprehensive process,including systematic literature review,synthesis of empirical experiences,thematic expert discussions,and consultation with an external panel to reach consensus.Results:The BRIDGE framework outlines 6 core components for ICWM-EHR database development:Overall design,database architecture,data extraction and linkage,data governance,data verification,and data quality evaluation.Key data elements include variables related to study population,treatment or exposure,outcomes,and confounders.These databases support various research applications,particularly in evaluating the effectiveness and safety of integrative therapies.To demonstrate its practical value,we developed an ICWM-EHR database on women’s reproductive lifespan,encompassing 2,064,482 patients.This database captures women’s health conditions across the life course,from reproductive age to older adulthood.Conclusions:The BRIDGE methodological framework provides a standardized approach to building high-quality ICWM-EHR databases.It offers a unique opportunity to strengthen the methodological rigor and real-world relevance of Chinese medicine research in integrated healthcare settings.展开更多
Osmanthus fragrans Lour.is a well-known aromatic plant widely used as a food ingredient due to its unique floral fragrance and bioactive compounds.To fully utilize O.fragrans resources,we established an O.fragrans mul...Osmanthus fragrans Lour.is a well-known aromatic plant widely used as a food ingredient due to its unique floral fragrance and bioactive compounds.To fully utilize O.fragrans resources,we established an O.fragrans multi-omics database called the O.fragrans Information Resource(OfIR:http://gffzz38b1cdba4cdf43d0hoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/OfIR/home/).OfIR is a convenient and comprehensive multi-omics database that efficiently integrates phenotype and genetic variation from 127 O.fragrans cultivars,and provides many easy-to-use analysis tools,including primer design,sequence extraction,multi-sequence alignment,GO and KEGG enrichment analysis,variation annotation,and electronic PCR.Two case studies were used to demonstrate its power to mine candidate genetic variation sites or genes associated with specific traits or regulatory networks.In summary,the multi-omics database OfIR provides a convenient and user-friendly platform for researchers in mining functional genes and contributes to the genetic breeding of O.fragrans.展开更多
The National Strong-Motion Observation Network System of China has collected over 12 000 strong-motion recordings from 2007 to December 2020.This study assembled the source-related metadata of 1 920 earthquakes associ...The National Strong-Motion Observation Network System of China has collected over 12 000 strong-motion recordings from 2007 to December 2020.This study assembled the source-related metadata of 1 920 earthquakes associated with assembled well-processed recordings of China.The earthquake basic information,focal mechanisms,and the fault geometry were collected from various institutes and literature.We recommended the MWvalues for 900 earthquakes,the fault types for 1 064 earthquakes,and the fault geometries for 18 large earthquakes.We also performed the statistical analysis for establishing the empirical conversions of MW-MS,and ML,and providing the empirical relationships between MWand ruptured area,aspect ratio,respectively.Moreover,the ruptured fault geometries of large earthquakes were used to preliminarily divide all earthquakes considered into 1 141 mainshocks,and 779 aftershocks.The finite-fault distances(RJBand Rrup) of strong-motion recordings from the 18 large earthquakes were calculated,and then used to yield the statistic relationships between the point-source distances(Repiand Rhyp) and finite-fault distances.We finally provided the earthquake source database freely accessible at website.The source-related metadata can be directly applied to develop the ground motion prediction equations of China.展开更多
Research into metamorphism plays a pivotal role in reconstructing the evolution of continent,particularly through the study of ancient rocks that are highly susceptible to metamorphic alterations due to multiple tecto...Research into metamorphism plays a pivotal role in reconstructing the evolution of continent,particularly through the study of ancient rocks that are highly susceptible to metamorphic alterations due to multiple tectonic activities.In the big data era,the establishment of new data platforms and the application of big data methods have become a focus for metamorphic rocks.Significant progress has been made in creating specialized databases,compiling comprehensive datasets,and utilizing data analytics to address complex scientific questions.However,many existing databases are inadequate in meeting the specific requirements of metamorphic research,resulting from a substantial amount of valuable data remaining uncollected.Therefore,constructing new databases that can cope with the development of the data era is necessary.This article provides an extensive review of existing databases related to metamorphic rocks and discusses data-driven studies in this.Accordingly,several crucial factors that need to be taken into consideration in the establishment of specialized metamorphic databases are identified,aiming to leverage data-driven applications to achieve broader scientific objectives in metamorphic research.展开更多
AI-driven materials databases are transforming research by integrating experimental and computational data to enhance discovery and optimization.Platforms such as Digital Catalysis Platform(DigCat)and Dynamic Database...AI-driven materials databases are transforming research by integrating experimental and computational data to enhance discovery and optimization.Platforms such as Digital Catalysis Platform(DigCat)and Dynamic Database of Solid-State Electrolyte(DDSE)demonstrate how machine learning and predictive modeling can improve catalyst and solid-state electrolyte development.These databases facilitate data standardization,high-throughput screening,and cross-disciplinary collaboration,addressing key challenges in materials informatics.As AI techniques advance,materials databases are expected to play an increasingly vital role in accelerating research and innovation.展开更多
Since its inaugural release in 2022,The Vegetables Information Resources(TVIR)has been a cornerstone for genomics and genetic breeding studies within the vegetable research community.With advancements in sequencing te...Since its inaugural release in 2022,The Vegetables Information Resources(TVIR)has been a cornerstone for genomics and genetic breeding studies within the vegetable research community.With advancements in sequencing technologies leading to an influx of new genome sequences,TVIR has been upgraded to version 2.0(http:/vir2.bio2db.com/),expanding from 59 to 84 vegetable species and introducing new functional modules to accelerate research.This upgrade incorporates a CRISPR/Cas9 resource module,which integrates four specialized tools:CasFinder,CasOT,Crisflash,and CRISPRCasFinder,to facilitate gene editing research.The database further features dynamic synteny analysis with an interactive interface,enabling users to visualize genomic relationships between species.Additionally,two novel bioinformatics tools Hmmsearch and CRISPRCasViewer are integrated to enhance comparative and functional genomic analyses.TVIR 2.0 retains all TVIR 1.0 features while updating resistance gene identification,expanding from 3 to 8 types,and transcription factor datasets,now including 237431 TFs,an increase from 172493.The database integrates comprehensive genomic,transcriptomic,and functional annotation data,providing freely accessible resources for vegetable breeding and gene editing.展开更多
The kagome lattice,characterized by a hexagonal arrangement of corner-sharing equilateral triangles,has garnered significant attention as a fascinating quantum material system that hosts exotic magnetic and electronic...The kagome lattice,characterized by a hexagonal arrangement of corner-sharing equilateral triangles,has garnered significant attention as a fascinating quantum material system that hosts exotic magnetic and electronic properties.The identification and characterization of this class of materials are critical for advancing our understanding of their role in emergent phenomena such as superconductivity.In this study,we developed a high-throughput screening framework for the systematic identification and classification of superconducting materials with kagome lattices,integrating them into established materials databases.Leveraging the Materials Project(MP)database and the MDR Super Con dataset,we analyzed over 150000 inorganic compounds and cross-referenced 26000 known superconductors.Using geometry-based structural modeling and experimental validation,we identified 129 kagome superconductors belonging to 17 distinct structural families,many of which had not previously been recognized as kagome systems.The materials are further classified into three categories in terms of topological flat bands,clean band structures,and coexisting magnetic or charge density wave(CDW)orderings.Based on these results,we established a database comprising 129 kagome superconductors,including the detailed crystallographic,electronic,and superconducting properties of these materials.展开更多
基金supported by the National Natural Science Foundation of China(91959106)the Foundation of the Shanghai Municipal Education Commission(24RGZNC02)+4 种基金Shanghai Key Laboratory of Intelligent Information Processing,Fudan University(IIPL-2025-RD3-02)Key University Science Research Project of Anhui Province(2023AH030108)Climbing Peak Training Program for Innovative Technology team of Yijishan Hospital,Wannan Medical College(PF201904)Peak Training Program for Scientific Research of Yijishan Hospital,Wannan Medical College(GF2019G15)the talent project of the First Affiliated Hospital of Wannan Medical College(Yijishan Hospital of Wannan Medical College)(YR202422).
摘要tRNA-derived small RNAs(tsRNAs),as a class of regulatory small noncoding RNA,have been implicated in a wide variety of human diseases.Large amounts of tsRNA–disease associations have been identified in recent years from accumulating studies.However,repositories for cataloging the detailed information on tsRNA–disease associations are scarce.In this study,we provide a tsRNADisease database by integrating experimentally and computationally supported tsRNA–disease associations from manual curation of literatures and other related resources.tsRNADisease contains 5571 manually curated associations between 4759 tsRNAs and 166 diseases with experimental evidence from 346 studies.In addition,it also contains 5013 predicted associations between 1297 tsRNAs and 111 diseases.tsRNADisease provides a user-friendly interface to browse,retrieve,and download data conveniently.This database can improve our understanding of tsRNA deregulation in diseases and serve as a valuable resource for investigating the mechanism of disease-related tsRNAs.tsRNADisease is freely available at http://gffzz9c504e06f78b4edahoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn.
基金supported by the National Natural Science Foundation of China(82472325 and 82072246)Chongqing Natural Science Foundation Project(CSTB2024NSCQ-MSX0703)+1 种基金Chongqing Science&Health Joint Medical Research Project(2023MSXM107)the Southwest University Graduate Research and Innovation Project(SWUB25044)。
摘要The ongoing battle between humans and pathogenic bacteria has fueled rapid microbial evolution.Although whole-genome sequencing(WGS)has transformed the ability to track genomic mutations,existing tools lack comprehensive solutions for analyzing mutational patterns and their functional consequences in pathogenic bacteria.Here,we present CliPME,an innovative platform that bridges this critical gap by combining mutation detection,mutation effect prediction,and regulatory network analysis,specifically designed for bacterial genomics.We develop qMut,a high-performance R package designed for largescale mutation profiling.Coupled with three major functional modules called MutFinder,MutAnalyzer,and ExpMiner,CliPME integrates population-level mutation analysis,functional mutation predictions,and estimation of gene-gene expression relationships.Using Mycobacterium tuberculosis(Mtb)as a case study,we show the power of CliPME by identifying functionally significant mutations in the transcription factor Rv0324,and experimentally demonstrate the link of its genetic variation to potential adaptive phenotypes.This resource empowers researchers to decode evolutionary mechanisms in bacterial pathogens and may accelerate the translation of genomic insights into antimicrobial strategies.The web server of CliPME is freely accessible at http://gffzzd72233016bfa415csoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/.
基金supported by the Key Research and Development Program of China(No.2024YFB4607200)the National Natural Science Foundation of China(No.52073030)the National Natural Science Foundation of China−Guangxi Joint Fund(No.U20A20276).
摘要A coupled three-dimensional cellular automata(CA)model has been used to predict the hydrogen porosity in an Al-Si alloy as a function of thermal boundary conditions.By quantifying the porosity distribution from simulations,a porosity defect database was established,representing a cooling rate ranging from 0.25 to 50℃/s at an initial hydrogen content of 3.0×10-3 mL/g.Based on the database,four machine learning algorithms including support vector machine(SVM),random forest(RF),K-nearest neighbors(KNN),and gradient boosting machine(GBM)were trained and compared for each porosity characteristic to identify the optimal model.For the prediction of porosity percentage,the determination coefficient(R2)and the root mean square error(RMSE)on the test set reached 0.95 and 0.042,respectively.The predicted porosity distribution agreed well with experiments,indicating that the model can be used to map the porosity size in large casting components.
基金National Key Project of Research and Development Program of China[2021YFC2500404].
摘要Primary liver cancer (PLC) is a major global healthchallenge, ranking as the sixth most common andthird most fatal malignancy worldwide, according toGLOBOCAN 2022 estimates[1]. This high mortalityrate underscores the aggressive nature of thedisease and the significant burden it places on globalhealthcare systems. Although primary preventionremains the cornerstone of liver cancer control,improving outcomes for patients already diagnosedis equally critical for mitigating the impact of thedisease.
基金supported by the National Key Research and Development Program of China(2023YFC3603100 and 2023YFC3603104)“Pioneer”and“Leading Goose”R&D Program of Zhejiang Province(2024C03028)+1 种基金Zhejiang Provincial Natural Science Foundation of China(LQ24H160028)Municipal-University Cooperation Project of Taizhou Institute of Zhejiang University(2024TZSX110).
摘要High-quality clinical databases are essential for advancing research and clinical practice in emergency and critical care medicine.With the widespread adoption of Electronic Health Record(EHR)systems across hospitals in China,vast amounts of clinical information have been digitally archived.Emergency departments(EDs)and intensive care units(ICUs)represent critically datademanded environments.
基金Intramural Project of The First Affiliated Hospital of Guangxi University of Chinese Medicine(2018QN008).
摘要Background:The purpose of this study was to analyze and classify adverse drug events(ADEs)related to ceftazidime/avibactam reported in the Food and Drug Administration Adverse Event Reporting System(FAERS)database and to evaluate their potential safety signals since the drug’s market introduction.Methods:This analysis systematically extracted and filtered FAERS data for ceftazidime/avibactam from its market launch in 2015 to the last quarter of 2024,utilizing the Medical Dictionary for Regulatory Activities(MedDRA)terminology for ADE recoding.The analysis employed the reporting odds ratio(ROR)method to assess the strength of ADE signals and to identify significant diseases associated with infections,the hepatobiliary system,the urinary system,and the nervous system.Results:A review of 540 adverse reaction reports revealed significant signals of adverse effects related to infections,hepatobiliary disorders,urinary system issues,and neurological impairments,including pathogen resistance,liver and kidney function impairment,encephalopathy,thrombocytopenia,and toxic epidermal necrolysis.However,these issues require further clinical attention.Conclusion:Ceftazidime/avibactam is associated with a range of adverse reactions,necessitating enhanced clinical monitoring,particularly in patients with underlying liver or kidney dysfunction.Continuous risk assessment and vigilant monitoring are critical for its clinical use.However,this study is limited by inherent reporting biases and confounders associated with the spontaneous reporting database(FAERS).Future research should validate these signals through prospective cohort and mechanistic studies and explore personalized risk management strategies for high-risk populations.
基金National Key Research and Development Program of China(2021YFB3500501)。
摘要Commercial phosphor-converted white LEDs(pc-WLEDs)face two inherent limitations,namely blue light hazard and low color rendering index,due to the use of blue LEDs as excitation source.To address these challenges,violet LEDs are proposed as an alternative solution.Currently,phosphors that can be efficiently excited by violet light(with wavelengths from 400 to 420 nm)remain under development still.In this study,we utilize large language models to construct a comprehensive database of Eu2+and Ce3+doped phosphors for discovering novel violet-excited phosphors.A total of 822 phosphor data entries,including elemental compositions,crystal structures and excitation/emission wavelengths,have been extracted and validated from 9551 research papers.Compared with Ce3+doped phosphors,the Eu2+are in general more suited for violet-excited phosphors,as well as red-emitting phosphors.In particular,Eu2+doped nitrides and sulfides are worth of exploration for violet-excited phosphors.This database is expected to be useful in the future development of phosphors for pc-WLEDs based on artificial intelligence methods.The datasets in this article are listed in Science Data Bank at http://gffzzd3cc09b8251d45dfhoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/10.57760/sciencedb.34314.
基金Supported by the Advanced Materials-National Science and Technology Major Project(2024ZD0606500)。
摘要Single-atom catalysts(SACs)have shown great promise for ethane dehydrogenation(EDH),owing mainly to their near-100%atomic utilization,precisely tunable active sites,and superior catalytic performance.Studies reveal that the nature of active metals,the properties of support and the coordination environments are critical factors affecting EDH performance.Among various supports,graphene has emerged as an ideal support material due to its excellent thermal stability and flexible,tunable coordination structure.While previous research has explored the effects of different active metals and coordination environments,a systematic understanding of their underlying principles is still lacking due to fragmented data.To bridge this gap,this work constructed a systematic,comprehensive database of heteroatom-doped graphene-supported SACs,covering five representative metal single atoms and 51 distinct coordination environments grouped into six major categories.Through high-throughput calculations,multi-dimensional data were systematically obtained,including elementary reaction energies,vibrational frequencies,density of states and Bader charges.A rigorous quality control system was implemented at both the parameter-setting and computational result levels.The database also provides complete raw calculation files,offering reliable data support for in-depth analysis of the catalytic performance,structure-performance relationship and reaction mechanisms of heteroatom-doped graphene-supported SACs in EDH.
摘要AIM:To screen for differentially expressed genes in retinoblastoma(RB)gene chips using GEO2R and validate them clinically.METHODS:The expression profile chip data(GSE110811)was downloaded from the public gene chip database Gene Expression Omnibus(GEO).The GEO2R chip analysis platform was used to identify differentially expressed genes between RB and adjacent normal tissues.According to the International Intraocular Retinoblastoma Classification(IIRC)system,35 children diagnosed with RB from our hospital and other hospitals were enrolled as the RB group,and 35 healthy children who underwent physical examinations in our hospital were enrolled as the control group.The relative expression levels of Sprouty RTK signaling antagonist 2(SPRY2)and estrogen-related receptor beta(ESRRB)in the serum of patients were detected by quantitative reverse transcription polymerase chain reaction(qRT-PCR).The diagnostic value of SPRY2 and ESRRB in RB was evaluated by receiver operating characteristic(ROC)curves.Analysis of the relationship between SPRY2/ESRRB expression and clinicopathological features,as well as its correlation with the tumor marker CA199.RESULTS:In the GSE110811 chip,the expression levels of two genes,16780069(SPRY2)and 16786783(ESRRB),showed the most significant differences between RB and normal tissues.The relative expression levels of SPRY2 and ESRRB in the serum of children in RB group(22 males,age 1.64±1.08y)were significantly lower(P<0.05)than those in control group(25 males,age 1.54±0.95y).The area under the ROC curve for SPRY2 was 0.735(95%CI:0.616-0.854),while that for ESRRB was 0.880(95%CI:0.800-0.960).There were statistically significant differences in the expression of SPRY2 and ESRRB with respect to choroidal invasion,optic nerve invasion,differentiation degree,and clinical staging(P<0.05).In RB group,the expression levels of SPRY2 and ESRRB decreased gradually with increasing CA199 levels,showing a negative correlation(rSPRY2=-0.593,rESRRB=-0.423;both P<0.05).CONCLUSION:The expression of SPRY2 and ESRRB is closely related to the occurrence and development of RB and negatively correlated with the tumor marker CA199.They have the potential to serve as diagnostic biomarkers for RB.
摘要[Objective]As high-dimensional vector data increasingly surpass the processing capabilities of traditional database management systems,Vector Databases(VDBs)have emerged and become tightly integrated with large language models,being widely applied in modern artificial intelligence systems.However,existing research has primarily focused on underlying technologies such as approximate nearest neighbor search,with relatively few studies providing a systematic architectural-level review of VDBs or analyzing how these core technologies collectively support the overall capacity of VDBs.This survey aims to offer a comprehensive overview of the core designs and algorithms of VDBs,establishing a holistic understanding of this rapidly evolving field.[Methods]First,we systematically review the key technologies and design principles of VDBs from the two core dimensions of storage and retrieval,tracing their technological evolution.Next,we conduct an in-depth comparison of several mainstream VDB architectures,summarizing their strengths,limitations,and typical application scenarios.Finally,we explore emerging directions for integrating VDBs with large language models,including open research challenges and trends such as novel indexing strategies.[Conclusions]This survey serves as a systematic reference guide for researchers and practitioners,helping readers quickly grasp the technological landscape and development trends in the field of vector databases,and promoting further innovation in both theoretical and applied aspects.
基金supported by grants from the Medical Science Research Project of Hebei(No.20261420).
摘要This study analyzes asthma-related mortality trends among U.S.adults(≥25 years)from 1999 to 2023 using the Centers for Disease Control and Prevention's Wide-ranging Online Data for Epidemiologic Research(CDC WONDER)database,calculating age-adjusted mortality rates(AAMRs)per 100,000 population stratified by demographics and geography.By fitting log-linear regression models to estimate annual percent changes(APCs)with 95%confidence intervals(CI).Results show a significant overall decline in AAMR from 2.46(95%CI:2.39-2.53)in 1999 to 1.33(95%CI:1.29-1.38)in 2023(average annual percent change,AAPC:2.43).Adults aged 45-64 years accounted for the largest proportion of asthma-related deaths(32.22%of deaths),with significant decline(AAPC:1.71).Women accounted for 64.62%of deaths,exhibiting persistently higher AAMR than men(AAPC:2.60).Non-Hispanic(NH)Black individuals had the highest overall AAMR in 1999(5.56)and 2023(2.84),with a transient 2017-2020 surge(APC:9.40).The Northeast surpassed the West as the highest-burden region by 2023,while Louisiana achieved the lowest state-level AAMR(0.76).Strategies are recommend,including workplace screening(ages 45-64);sex-specific strategies for women;culturally tailored community health workers for NH Black individuals;and improved asthma action plans in the Northeast.
摘要Background Identifying simple and effective prognostic markers is crucial for risk stratification in critically ill patients with coronary artery disease(CAD)and diabetes mellitus(DM).The anion gap(AG),a marker of acid-base homeostasis,has been associated with mortality in various diseases,but its specific role in this high-risk popu-lation remains unclear.Methods This retrospective cohort study included 4,343 patients from the MIMIC-Ⅳda-tabase.The exposure variable was the anion gap(AG),which was modeled in three ways:as a continuous variable,as a categorical variable(Low,Middle,and High groups),and as a dichotomous variable using a cutoff of 14.5.The primary outcome was in-hospital mortality.Multivariable logistic regression models were used to adjust for poten-tial confounders,including demographics,comorbidities,vital signs,and laboratory parameters.The linearity of the association was assessed using restricted cubic splines(RCS),and the predictive performance of AG was evaluated using receiver operating characteristic(ROC)curves.Results The overall in-hospital mortality rate was 12.6%.A strong,dose-response relationship was observed between AG levels and mortality.After full adjustment,compared to the low AG tertile,the high AG tertile was significantly associated with increased mortality risk(OR:1.92,95%CI:1.43-2.58,P<0.0001).AG as a continuous variable remained an independent predictor(OR per unit increase:1.11,95%CI:1.08-1.14,P<0.0001).RCS analysis confirmed a linear association(P for nonlinearity=0.880).In ROC analysis,AG demonstrated superior predictive ability[area under the curve(AUC),AUC=0.6957)]for in-hos-pital mortality compared to sodium(AUC=0.5338),potassium(AUC=0.5),and chloride(AUC=0.6141).Conclu-sions In this cohort of ICU patients with CAD and DM from the MIMIC-Ⅳdatabase,an elevated anion gap was independently and linearly associated with a higher risk of in-hospital mortality.AG outperformed other common electrolytes in mortality prediction,highlighting its potential as a valuable and readily available risk stratification tool in this population.
摘要The journal of Meteorological and Environmental Research [ISSN: 2152-3940] has been included and stored by the following famous databases: CA, CABI, CSA, EBSCO, UPD, AGRIS, EA, Chinese Science and Technology Periodical Database, and CNKI, as well as Library of Congress, United States.CA (Chemical Abstracts) was founded in 1907, and is the most authoritative and comprehensive source for chemical information. Centre for Agriculture and Bioscience International(CABI)is a not-for-profit international agricultural information institute with headquarters in Britain. ProQuest CSA belongs to Cambridge Information Group (CIG), and it provides access to more than 100 databases published by CSA and its publishing partners. EBSCO is a large document service company with a history of more than 60 years, providing subscription and publication services of journals and documents. CNKI (China National Knowledge Infrastructure), universally acclaimed as the most valuable Chinese website, boasts the greatest information content, covering natural science, humanities and social science, engineering, periodical, doctor/master s dissertations, newspapers, books, meeting papers and other miscellaneous public information resources in China.
基金supported by the National Key Research&Development Program of China(No.2024YFC3505800)the National Natural Science Foundation of China(Nos.82474334,82474335 and 72174132)+3 种基金National Science Fund for Distinguished Young Scholars(No.82225049)the Key Research&Development Projects of Sichuan Provincial Department of Science and Technology(Nos.2024YFFK0174 and 2024YFFK0152)1.3.5 Project for Disciplines of Excellence,West China Hospital,Sichuan University(Nos.ZYYC24010 and ZYGD23004)the Special Fund for Traditional Chinese Medicine of Sichuan Provincial Administration of Traditional Chinese Medicine(No.2024zd023).
摘要Objectives:Electronic health records(EHRs)offer valuable real-world data(RWD)for Chinese medicine research.However,significant methodological challenges remain in developing integrative Chinese-Western medicine(ICWM)databases.This study aims to establish a best-practice methodological framework,referred to as BRIDGE,to guide the construction of ICWM databases using EHRs.Methods:We developed the methodological framework through a comprehensive process,including systematic literature review,synthesis of empirical experiences,thematic expert discussions,and consultation with an external panel to reach consensus.Results:The BRIDGE framework outlines 6 core components for ICWM-EHR database development:Overall design,database architecture,data extraction and linkage,data governance,data verification,and data quality evaluation.Key data elements include variables related to study population,treatment or exposure,outcomes,and confounders.These databases support various research applications,particularly in evaluating the effectiveness and safety of integrative therapies.To demonstrate its practical value,we developed an ICWM-EHR database on women’s reproductive lifespan,encompassing 2,064,482 patients.This database captures women’s health conditions across the life course,from reproductive age to older adulthood.Conclusions:The BRIDGE methodological framework provides a standardized approach to building high-quality ICWM-EHR databases.It offers a unique opportunity to strengthen the methodological rigor and real-world relevance of Chinese medicine research in integrated healthcare settings.
基金supported by research grants provided by the National Natural Science Foundation of China(Grant Nos.32101581,32271951,and 32372754)the Hubei Provincial Central Leading Local Special Project(Grant No.2022BGE263)+3 种基金the Key Research and Science and Technology Program of Hubei Province(Grant No.2021BBA098)the Hubei Province Natural Science Foundation(Grant Nos.2023AFB1063 and 2024AFB1057)the Innovation Team Project from Hubei University of Science and Technology(Grant No.2022T02)a PhD grant from the Hubei University of Science and Technology(Grant Nos.BK202002and BK202419).
摘要Osmanthus fragrans Lour.is a well-known aromatic plant widely used as a food ingredient due to its unique floral fragrance and bioactive compounds.To fully utilize O.fragrans resources,we established an O.fragrans multi-omics database called the O.fragrans Information Resource(OfIR:http://gffzz38b1cdba4cdf43d0hoov6on05qcpc6v55.ffgz.tsg.suse.edu.cn/OfIR/home/).OfIR is a convenient and comprehensive multi-omics database that efficiently integrates phenotype and genetic variation from 127 O.fragrans cultivars,and provides many easy-to-use analysis tools,including primer design,sequence extraction,multi-sequence alignment,GO and KEGG enrichment analysis,variation annotation,and electronic PCR.Two case studies were used to demonstrate its power to mine candidate genetic variation sites or genes associated with specific traits or regulatory networks.In summary,the multi-omics database OfIR provides a convenient and user-friendly platform for researchers in mining functional genes and contributes to the genetic breeding of O.fragrans.
基金supported by the National Key R&D Program of China(Grant No.2019YFE0115700).
摘要The National Strong-Motion Observation Network System of China has collected over 12 000 strong-motion recordings from 2007 to December 2020.This study assembled the source-related metadata of 1 920 earthquakes associated with assembled well-processed recordings of China.The earthquake basic information,focal mechanisms,and the fault geometry were collected from various institutes and literature.We recommended the MWvalues for 900 earthquakes,the fault types for 1 064 earthquakes,and the fault geometries for 18 large earthquakes.We also performed the statistical analysis for establishing the empirical conversions of MW-MS,and ML,and providing the empirical relationships between MWand ruptured area,aspect ratio,respectively.Moreover,the ruptured fault geometries of large earthquakes were used to preliminarily divide all earthquakes considered into 1 141 mainshocks,and 779 aftershocks.The finite-fault distances(RJBand Rrup) of strong-motion recordings from the 18 large earthquakes were calculated,and then used to yield the statistic relationships between the point-source distances(Repiand Rhyp) and finite-fault distances.We finally provided the earthquake source database freely accessible at website.The source-related metadata can be directly applied to develop the ground motion prediction equations of China.
基金funded by the National Natural Science Foundation of China(No.42220104008)。
摘要Research into metamorphism plays a pivotal role in reconstructing the evolution of continent,particularly through the study of ancient rocks that are highly susceptible to metamorphic alterations due to multiple tectonic activities.In the big data era,the establishment of new data platforms and the application of big data methods have become a focus for metamorphic rocks.Significant progress has been made in creating specialized databases,compiling comprehensive datasets,and utilizing data analytics to address complex scientific questions.However,many existing databases are inadequate in meeting the specific requirements of metamorphic research,resulting from a substantial amount of valuable data remaining uncollected.Therefore,constructing new databases that can cope with the development of the data era is necessary.This article provides an extensive review of existing databases related to metamorphic rocks and discusses data-driven studies in this.Accordingly,several crucial factors that need to be taken into consideration in the establishment of specialized metamorphic databases are identified,aiming to leverage data-driven applications to achieve broader scientific objectives in metamorphic research.
摘要AI-driven materials databases are transforming research by integrating experimental and computational data to enhance discovery and optimization.Platforms such as Digital Catalysis Platform(DigCat)and Dynamic Database of Solid-State Electrolyte(DDSE)demonstrate how machine learning and predictive modeling can improve catalyst and solid-state electrolyte development.These databases facilitate data standardization,high-throughput screening,and cross-disciplinary collaboration,addressing key challenges in materials informatics.As AI techniques advance,materials databases are expected to play an increasingly vital role in accelerating research and innovation.
基金supported by the National Key Research and Development Program of China(2023YFF1002000)the Tangshan Science and Technology Plan Project(24130219C)+6 种基金the Basic Research Program of Tangshan(22130231H)the Basic research expenses for provincial universities(JJC2024001)the National Natural Science Foundation of China(32172583)Basic Research Funds for Provincial Universities Basic Research Projects of North China University of Science and Technology(JQN2023036)Youth Scholars Promotion Plan of North China University of Science and Technology(QNTJ202308)the S&T Program of Hebei(23372505D)the Hebei Natural Science Foundation(H2023209084).
摘要Since its inaugural release in 2022,The Vegetables Information Resources(TVIR)has been a cornerstone for genomics and genetic breeding studies within the vegetable research community.With advancements in sequencing technologies leading to an influx of new genome sequences,TVIR has been upgraded to version 2.0(http:/vir2.bio2db.com/),expanding from 59 to 84 vegetable species and introducing new functional modules to accelerate research.This upgrade incorporates a CRISPR/Cas9 resource module,which integrates four specialized tools:CasFinder,CasOT,Crisflash,and CRISPRCasFinder,to facilitate gene editing research.The database further features dynamic synteny analysis with an interactive interface,enabling users to visualize genomic relationships between species.Additionally,two novel bioinformatics tools Hmmsearch and CRISPRCasViewer are integrated to enhance comparative and functional genomic analyses.TVIR 2.0 retains all TVIR 1.0 features while updating resistance gene identification,expanding from 3 to 8 types,and transcription factor datasets,now including 237431 TFs,an increase from 172493.The database integrates comprehensive genomic,transcriptomic,and functional annotation data,providing freely accessible resources for vegetable breeding and gene editing.
基金supported by the National Key Research and Development Program of China(Grant No.2018YFE0202600)the National Natural Science Foundation of China(Grant No.52272268)+3 种基金the Key Research Program of Frontier SciencesCAS(Grant No.QYZDJ-SSWSLH013)the Informatization Plan of Chinese Academy of Sciences(Grant No.CAS-WX2021SF-0102)the Youth Innovation Promotion Association of CAS(Grant No.2019005)。
摘要The kagome lattice,characterized by a hexagonal arrangement of corner-sharing equilateral triangles,has garnered significant attention as a fascinating quantum material system that hosts exotic magnetic and electronic properties.The identification and characterization of this class of materials are critical for advancing our understanding of their role in emergent phenomena such as superconductivity.In this study,we developed a high-throughput screening framework for the systematic identification and classification of superconducting materials with kagome lattices,integrating them into established materials databases.Leveraging the Materials Project(MP)database and the MDR Super Con dataset,we analyzed over 150000 inorganic compounds and cross-referenced 26000 known superconductors.Using geometry-based structural modeling and experimental validation,we identified 129 kagome superconductors belonging to 17 distinct structural families,many of which had not previously been recognized as kagome systems.The materials are further classified into three categories in terms of topological flat bands,clean band structures,and coexisting magnetic or charge density wave(CDW)orderings.Based on these results,we established a database comprising 129 kagome superconductors,including the detailed crystallographic,electronic,and superconducting properties of these materials.