Comalada i Pla, Francesc
hybrid · semantic + lexical · 311 datasets ranked · 3.87s
Dataset containing the systematic review matrix and extracted variables supporting the article "Bridging ecological restoration and social legitimacy: a systematic review of Cultural Ecosystem Services in inland aquatic ecosystems", accepted for publication in People and Nature.
Robert Koch-Institut · AKTIN-Notaufnahmeregister
86 rows × 8 cols · 10 KB · pdf, tsv, zip
4 numeric · 3 categorical · 1 text
Der Datensatz "Daten der Notaufnahmesurveillance" wird durch das Robert Koch-Institut und das AKTIN-Notaufnahmeregister bei Krankenhausaufnahmen bereitgestellt. Der Datensatz beinhaltet aggregierte Routinedaten aus deutschen Notaufnahmen zur syndromischen Überwachung von akuten Erkrankungen. Dazu zählen grippeähnliche Erkrankungen (ILI), Coronavirus-Erkrankungen (COVID-19), akute respiratorische Erkrankungen (ARE), gastrointestinale Infektionen (GI) und schwere akute respiratorische Infektionen (SARI). Dabei wird der relative Anteil dieser Erkrankungen an der Gesamtzahl der Notaufnahmevorstellungen sowie die berechneten Erwartungswerte und Prädiktionsintervalle ausgewiesen. Die Daten sind nach Notaufnahmetypen und Altersgruppen aggregiert. Damit bietet der Datensatz eine wertvolle Ressource für die Forschung im Bereich der Notfallmedizin und der Überwachung akuter Gesundheitsereignisse in Deutschland.
Lopez, Annalaura · Greco, Margherita · Marcolli, Beatrice · et al.
4 files · 17 KB · docx, xlsx
This dataset originates from a study aiming to valorise Ciuta sheep, a local breed native from the Italian Central Alps, through the characterization of nutritional quality and chemical composition of fresh meat (loins) and one traditional dry-cured product. Specifically, the research focused on determining the chemical composition of Ciuta sheep meat and on identifying key changes in its chemical profile during dry curing process, hypothesizing that such chemical fingerprint may suggest some markers linked to the production system, geographical origin, and traditional processing techniques. For this reason, for bthe dry-cured product, both an aliquot of fresh meat before and after transformation and dry-curing was sampled and analysed. Regarding loins, three commercial categories (lambs, hoggets and mutton) were considered, in order to define any possible difference induced by age of the sheep (and physiological factors, such as rumen development). The dataset includes chemical data regarding the proximate composition (moisture, protein, fat, ash, salt content for the dry-cured product) and energy content of fresh and dry-cured meat; the fatty acids content of fresh and dry-cured meat product; the volatile profile of fresh and dry-cured meat product. Results from analysis performed in our study suggested that the development of high-quality dry-cured products could provide a strategy to valorise Ciuta sheep meat, especially from adult animals (culled ewes and rams), while fresh meat production could focus on lambs. The complex volatile profile detected was influenced by both the farming system and traditional processing methods.
Kantor, Rose · Shakya, Migun · Ruth, Nelson · et al.
2,095 rows · 907 KB · fasta, tsv
A virus genome database representing 21,015 near-complete virus genomes collected from untargeted ultra-deep RNA/DNA combined sequencing of wastewater. Sequence data was provided by the CASPER consortium and raw data may be found on NCBI SRA under bioprojects PRJNA1247874 and PRJNA1198001. Data underwent read trimming, rRNA and human read removal, de novo assembly, and selection of high-quality viral contigs. Contigs were clustered at 95% identity and 85% query coverage to dereplicate. Chimera-checking required at least two independent assemblies of the same viral genome or presence of the genome in another reference database. Annotation made use of RdRpCATCH, geNomad, checkV, BLASTN against NCBI core-nt, and RNAVirHost. The RdRp fasta files contain representative RdRp sequences identified through homology to major RdRp reference databases and clustered at 90% sequence identity over 75% sequence coverage. Included sequences contain all three conserved RdRp motifs (A, B, and C) arranged in either the canonical ABC configuration or the permuted CAB configuration.
Robert Koch-Institut · AKTIN-Notaufnahmeregister
5 files · 10 KB · pdf, tsv, zip
Der Datensatz "Daten der Notaufnahmesurveillance" wird durch das Robert Koch-Institut und das AKTIN-Notaufnahmeregister bei Krankenhausaufnahmen bereitgestellt. Der Datensatz beinhaltet aggregierte Routinedaten aus deutschen Notaufnahmen zur syndromischen Überwachung von akuten Erkrankungen. Dazu zählen grippeähnliche Erkrankungen (ILI), Coronavirus-Erkrankungen (COVID-19), akute respiratorische Erkrankungen (ARE), gastrointestinale Infektionen (GI) und schwere akute respiratorische Infektionen (SARI). Dabei wird der relative Anteil dieser Erkrankungen an der Gesamtzahl der Notaufnahmevorstellungen sowie die berechneten Erwartungswerte und Prädiktionsintervalle ausgewiesen. Die Daten sind nach Notaufnahmetypen und Altersgruppen aggregiert. Damit bietet der Datensatz eine wertvolle Ressource für die Forschung im Bereich der Notfallmedizin und der Überwachung akuter Gesundheitsereignisse in Deutschland.
Budzinski, Lisa · Beenken, Anne Elisabeth · Sempert, Toni · et al.
9 rows × 1 cols · 743 B · csv, docx, zip
1 categorical
We have investigated an IgG4-RD (IgG4-RD) cohort by our multi-parameter microbiota flow cytometry approach to characterise the microbiota on single-cell level for attributes of the disease. The microbiota is isolated from stool samples and stained according to the published protocol for (a) host immunoglobulins IgA1, IgA2, IgM, IgG and (b) agglutinin binding to mannose, galactose or N-Acetyl-glucosamine surface sugar moieties. For all samples we also determined the microbiome composition by 16S rRNA (V3-V4) sequencing on the illumina MiSeq platform. We provide the raw .fcs and FASTQ files of 40 IgG4-RD patients. For comparison we additionally analysed 36 healthy donors. All .fcs files were generated on BD Influx®. The metadata is collected in the provided meta.csv. The staining parameters are summarized in provided panel.csv.
Li, Zhiyao · Wang, Ningbo · Zhong, Jiahao
6 files · 48 MB · docx, zip
GIFT-BDS is a regional ionospheric total electron content (TEC) and TEC-gradient dataset over China derived from BeiDou geostationary Earth orbit (GEO) observations and a dense ground-based GNSS receiver network. The dataset is designed to provide high-resolution observations of ionospheric TEC variability and horizontal TEC-gradient structures over China and adjacent regions. The versioned release covers the period from 19 July 2024 to 31 December 2025, corresponding to DOY 201 of 2024 to DOY 365 of 2025. The geographical coverage is 15°N-50°N and 95°E-135°E. The dataset is provided in daily NetCDF files and contains two product levels. Level-1 products provide observation-level GEO-derived slant TEC (STEC) and rate of TEC index (ROTI) records for individual receiver-GEO satellite lines of sight, with a temporal resolution of 30 s. Level-2 products provide gridded regional TEC and TEC-gradient variables, including VTEC, VTEC t , ROTI, GIX, GIX std , GIX x , GIX y , GIX t,x , and GIX t,y , with a temporal resolution of 15 min. IPP-based variables are provided on a 1° × 1° grid, while inter-IPP-gradient variables are provided on a 0.25° × 0.25° grid. The main processing steps include observation screening, cycle-slip and data-gap detection, continuous-arc segmentation, carrier-to-code leveling, satellite and receiver DCB correction, IPP calculation, inter-IPP pair selection, gradient estimation, and gridding. Quality control is applied before release. Missing values may occur because of station outages, data gaps, quality-control exclusions, or insufficient valid samples within a grid cell. Users should check the NetCDF variable attributes, including units and fill values, before analysis. The dataset is suitable for regional ionospheric studies, TEC-gradient monitoring, space-weather-related analyses, and investigations of ionospheric effects on GNSS positioning applications.
Jungmyoung, Son · Sihoon, Lee · Jiyeon, Hong
3 files · 20 KB · docx, xlsx
This dataset provides the complete double-coding matrix, PRISMA 2020 checklist, and search strategy supporting the systematic review "Representation-to-AI Transformation in K-12 Generative AI Learning: A Theory-Building Systematic Review of Semantic Transformation Mechanisms." It includes: (1) study-level tier classification (Core/Supporting/Context) for two independent coders and consensus tier for all 18 included studies; (2) the full semantic transformation unit (STU) coding matrix (18 studies x 10 STUs = 180 cells) with pre-consensus and consensus scores; (3)evidence-weighting consensus scores; (4) inter-rater reliability statistics (Cohen's kappa); (5) the completed PRISMA 2020 checklist; and (6) the full database-specific Boolean search strategy.
guo, na · pan, jiachen · li, tiantian · et al.
14 files · 100 MB · csv, rar, tsv
ADM_LSIR is a physics-inspired laparoscopic aerosol degradation dataset for aerosol-aware surgical image analysis and image restoration. The v1.0.0 release contains: - 21,916 clean clinical laparoscopic frames (clean/) - 9,562 real intraoperative aerosol-degraded frames (degraded/) - 36,052 simulated aerosol masks, including 19,701 smoke-like masks and 16,351 trajectory masks (mask/) - Blender simulation/cache materials (ADM_LSIR_Blender_simulation_files_v1.0.rar) - metadata_quality_report_v1.0.csv - recommended_splits_v1.0.csv - video_mapping_v1.0.csv - parts_manifest.txt - checksums_v1.0.tsv - release_manifest_v1.0.json All released clinical frames are de-identified and stored as lossless PNG files. Filenames use anonymized video identifiers, e.g., C-V##-####.png for clean frames and D-V##-####.png for degraded frames. The recommended split is defined at the source_video_id/public_video_label level to reduce leakage across frames from the same source video. The public video labels in video_mapping_v1.0.csv provide privacy-safe source-video identifiers (video1-video19). The Blender archive documents the smoke and trajectory mask simulation setup and supports reuse, but it is not a guaranteed exact per-mask reproduction package. The released pre-rendered mask library is the primary reusable dataset component. Source code for synthesis and quality screening is available at: https://github.com/SweetDeathh/ADM_LSIR
Manghi, Paolo
2 files · 8.0 MB · gzip, tsv
Using gene-level and species-level shotgun metagenomics, we provide the first characterization of the rural, Bolivian microbiome; we identified microbial genes which strongly correlate (rho>0.45) with arsenic in urine, and that overall contribute to substantiate that the gut microbiome helps tolerate arsenic via a evict-out-of-house mechanism. Mediation analysis, followed by phylogenetic investigation of metagenomic-assembled genomes, further strengthens this observation. This study elucidates the role of the microbiome in helping to tolerate arsenic-rich environments, and paves the way for probiotic interventions that may mitigate the effects of this toxic metal. This repository contains 2,478 HQ MAGs from the MicroToxBol cohort in fasta format and a descriptive table comprising taxonomic annotation, quality-checks, and coverage estimation.
Kovaliov, Michael
1 files · 100 MB · parquet
Madhusudan, Gujral
6 files · 29 MB · parquetdeclared
Large language models (LLMs) are trained on massive, publicly available text datasets comprising trillions of tokens, enabling them to excel at general language tasks like next-token prediction. However, LLMs often struggle with domain-specific prompts, exhibiting reduced accuracy or generating inaccurate information (hallucinations). This is because they lack sufficient subject matter expertise. Two primary approaches exist to address this limitation for augmenting LLMs knowledge: Retrieval-Augmented Generation (RAG) and fine-tuning. This presentation focuses on fine-tuning smaller LLMs with domain-specific instruct datasets using the LoRA (Low-Rank Adaptation) technique on Gaudi hardware. We will leverage publicly available LLMs and datasets from the Hugging Face Hub for this demonstration. Though it is possible to fine tune LLMs with plain text data - sourced from documents, articles, and other materials.
Azziz, Ricardo
1 files · 1.6 MB · docxdeclared
Supplemental Table 1. Meeting Agenda of the 2024 PCOS Challenge-CDC Stakeholder Meeting on Testosterone Reference Interval Standardization. August 26, 2024. Centers for Disease Control and Prevention, Atlanta, Georgia. Supplemental Table 2. Proposed Four-Stage Workflow for Developing Standardized Testosterone Reference Intervals Using Existing Study Data.
García Mina, Janeth Elizabeth · Rodríguez Garófalo, Napoleón Hernán · Rumbaut-Rangel, Dayron · et al.
6 files · 737 KB · docx, pdfdeclared
El objetivo de este estudio es evaluar la importancia de la inteligencia artificial (IA) en el proceso de formación académica de estudiantes de educación básica, específicamente en la enseñanza de ciencias naturales. La investigación destaca cómo la IA puede personalizar el aprendizaje, adaptándose a los estilos individuales de cada alumno, lo que resulta en una experiencia educativa más efectiva y significativa. Se utilizó un diseño cuasiexperimental con una muestra de 60 estudiantes de sexto año, distribuidos en un grupo control (30 estudiantes) que recibió instrucción tradicional y un grupo experimental (30 estudiantes) que utilizó herramientas de IA, como chatbots educativos y plataformas de aprendizaje adaptativo. Los resultados mostraron que el 70% de los estudiantes del grupo experimental alcanzaron calificaciones superiores a 8, en comparación con solo el 30% del grupo control. Este hallazgo sugiere que la implementación de herramientas de IA puede transformar la educación básica, mejorando tanto el rendimiento académico como la motivación de los estudiantes. En conclusión, el estudio resalta la necesidad de seguir investigando y desarrollando nuevas estrategias educativas que integren la IA para maximizar su potencial en la educación.
Makhado, Langanani Christinah · Raliphaswa, Ndidzulafhi Selina
9 files · 157 KB · docxdeclared
This dataset contains anonymized transcripts derived from face-to-face interviews conducted with pregnant women and breastfeeding mothers as part of the research study. The interviews were audio-recorded and transcribed verbatim to capture participants' experiences, perceptions, and insights related to the study topic. All personally identifiable information has been removed to protect participant confidentiality. The transcripts are provided to support transparency, reproducibility, and further research related to the findings reported in the associated article
Fontelo, Paul
1 files · 26 KB · docxdeclared
Auto-Brewery Syndrome (ABS), also known as gut fermentation syndrome, is a condition in which microbial fermentation of dietary carbohydrates generates endogenous ethanol, causing signs and symptoms of alcohol intoxication without alcohol consumption. Misdiagnosis as alcohol use disorder (AUD) is common and carries severe medical, psychosocial, legal, and forensic consequences. The objective of thus narrative review is to synthesize the clinical literature on ABS for practicing clinicians and to present the first systematic analysis of ABS susceptibility in Asian populations - including communities of Asian descent in Western countries.
Hackl, Jürgen
6 files · 132 MB · parquet, tiffdeclared
EuroFlood is an open, cloud-native index over the JRC/Copernicus CEMS-EFAS Satellite-Derived Flood Depth Maps for Europe (Betterle & Salamon, 2025; CC-BY-4.0) - ~3,280 satellite-derived observed flood-depth maps across Europe, 2015-2024. The bundle is a sparse Cloud-Optimized GeoTIFF encoding, per pixel, the set of flood events that inundated it, plus a combo_id -sorted GeoParquet dictionary and a small events table. Query by region and time via HTTP range reads (GDAL /vsicurl + DuckDB) to retrieve matching events, then fetch only the source depth rasters needed. Built with the open-source EuroFlood Python package ( pip install euroflood ).
Walsh, Calum · Srinivas, Meghana · Stinear, Timothy · et al.
100 files · 2.7 GB · gzip, tsvdeclared
GROND (Genome-derived Ribosomal OperoN Database) A quality-checked and publicly-available database of 16S-ITS-23S RRNA operon sequences and their constituent 16S and 23S genes. Based on GTDB release R232.
Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services
2 files · 1.5 MB · docx, pdfdeclared
This is the report on the indigenous and local knowledge (ILK) dialogue workshop for the first order draft of the summary for policymakers and the second order draft of the IPBES assessment of the sustainable use of wild species (the "sustainable use assessment"). It was held from 17-21 May 2021, online, due to the ongoing COVID-19 pandemic. The report aims to provide a written record of the dialogue workshop, which can be used by assessment authors to inform their work on the sustainable use assessment, and also by all dialogue participants who may wish to review and contribute to the work of the assessment moving forward. The report is not intended to be comprehensive or give final resolution to the many interesting discussions and debates that took place during the workshop. Instead, it is intended as a written record of the discussions, and this conversation will continue to evolve in the course of the assessment process. For this reason, clear points of agreement are discussed, but diverging views among participants are also presented for further attention and discussion.
Ogah, Odey
4 files · 3.4 MB · docx, pdf, zipdeclared
The study investigated the determinants of the migration of youth in Benue State from agricultural activities using a multi stage survey research design. The population of the study comprised 90 households in the three zones in Benue state selected by simple random sampling. A sampling proportion of 5% was used and a sample size of 270 was drawn. The instrument for data collection was a self-developed questionnaire. Data collected was analysed using frequencies distribution, mean, standard deviation and probit regression to achieve the research objectives while chi square goodness of fit was used to test the hypothesis. The findings revealed that 69.0% of the youth involved in agricultural activities in the study area were within the ages of 26- 30 years old while few (7.7%) are of 31years and above. The findings of the study revealed that land tenure practices have high influence on agricultural land developmental activities of youth in the study area. The study found that there is a plethora of factors responsible for seasonal migration of youth in the study area and these includes; Lack of access to modern farming equipment and technology; limited economic opportunities in the agricultural sector. The study concludes that youth migration affects agricultural activities in the study area. The study therefore recommended among others, skills diversification programme that will equip youth in diverse skills beyond seasonal agricultural works that could be utilised all year round. Year round employment opportunities that would provide consistent employment for the youth, education and training initiatives that would focus on agricultural techniques and sustainable practices, access to financial resources and collaboration and networking between local governments, NGOS, and the private sectors.