Exploration

ResearchFeatured

Discovery

DiscoverSourcesQuality

Analysis

Working setReviews
Flow StudioTeamConcept
Settings

Partners

  • AI AlliancePrime
  • BrightQueryBuilds Meridian
  • OpenMinedFunded partner
  • MLCommonsFunded partner
  • Hugging FaceDeployment platform
See the full consortium and what each partner wires

Meridian is the discovery layer for research data, built by BrightQuery within the AI Alliance.

hybrid · semantic + lexical · 316 datasets ranked · 0.93s

Structurecomposite5modal3tabular3sequence1
Depthcataloged304measured12
Licenseopen301unknown13non commercial1share alike1
Accessopen316
Formatdocx255pdf74zip58tsv46xlsx36
Sourcezenodo313zenodo-bio3
clear
1-20 of 316sortrelevancemeasured firstqualitysize
modal

Bridging ecological restoration and social legitimacy: a systematic review of Cultural Ecosystem Services in inland aquatic ecosystems

0.00

Comalada i Pla, Francesc

csv32
tar20
fits15
gzip10
parquet9
bzip25
jpeg5
png4
tiff4
rar3
fasta2
gff1
netcdf1
shapefile1
sqlite1
torch1
xz1
1 files · 224 KB · docx

Dataset containing the systematic review matrix and extracted variables supporting the article "Bridging ecological restoration and social legitimacy: a systematic review of Cultural Ecosystem Services in inland aquatic ecosystems", accepted for publication in People and Nature.

open·CC-BY-4.0·Zenodo·completeSource
tabular

Daten der Notaufnahmesurveillance

0.00

Robert Koch-Institut · AKTIN-Notaufnahmeregister

86 rows × 8 cols · 10 KB · pdf, tsv, zip

4 numeric · 3 categorical · 1 text

Der Datensatz "Daten der Notaufnahmesurveillance" wird durch das Robert Koch-Institut und das AKTIN-Notaufnahmeregister bei Krankenhausaufnahmen bereitgestellt. Der Datensatz beinhaltet aggregierte Routinedaten aus deutschen Notaufnahmen zur syndromischen Überwachung von akuten Erkrankungen. Dazu zählen grippeähnliche Erkrankungen (ILI), Coronavirus-Erkrankungen (COVID-19), akute respiratorische Erkrankungen (ARE), gastrointestinale Infektionen (GI) und schwere akute respiratorische Infektionen (SARI). Dabei wird der relative Anteil dieser Erkrankungen an der Gesamtzahl der Notaufnahmevorstellungen sowie die berechneten Erwartungswerte und Prädiktionsintervalle ausgewiesen. Die Daten sind nach Notaufnahmetypen und Altersgruppen aggregiert. Damit bietet der Datensatz eine wertvolle Ressource für die Forschung im Bereich der Notfallmedizin und der Überwachung akuter Gesundheitsereignisse in Deutschland.

open·CC-BY-4.0·Zenodo·0% null·completeSource
modal

Chemical characterization (proximate composition, fatty acids content and volatile profile) of meat from the alpine Ciuta sheep breed.

0.00

Lopez, Annalaura · Greco, Margherita · Marcolli, Beatrice · et al.

4 files · 17 KB · docx, xlsx

This dataset originates from a study aiming to valorise Ciuta sheep, a local breed native from the Italian Central Alps, through the characterization of nutritional quality and chemical composition of fresh meat (loins) and one traditional dry-cured product. Specifically, the research focused on determining the chemical composition of Ciuta sheep meat and on identifying key changes in its chemical profile during dry curing process, hypothesizing that such chemical fingerprint may suggest some markers linked to the production system, geographical origin, and traditional processing techniques. For this reason, for bthe dry-cured product, both an aliquot of fresh meat before and after transformation and dry-curing was sampled and analysed. Regarding loins, three commercial categories (lambs, hoggets and mutton) were considered, in order to define any possible difference induced by age of the sheep (and physiological factors, such as rumen development). The dataset includes chemical data regarding the proximate composition (moisture, protein, fat, ash, salt content for the dry-cured product) and energy content of fresh and dry-cured meat; the fatty acids content of fresh and dry-cured meat product; the volatile profile of fresh and dry-cured meat product. Results from analysis performed in our study suggested that the development of high-quality dry-cured products could provide a strategy to valorise Ciuta sheep meat, especially from adult animals (culled ewes and rams), while fresh meat production could focus on lambs. The complex volatile profile detected was influenced by both the farming system and traditional processing methods.

open·CC-BY-4.0·Zenodo·completeSource
sequence

Database of virus genomes from ultra-deep sequencing of wastewater (WVDB)

0.00

Kantor, Rose · Shakya, Migun · Ruth, Nelson · et al.

2,095 rows · 907 KB · fasta, tsv

A virus genome database representing 21,015 near-complete virus genomes collected from untargeted ultra-deep RNA/DNA combined sequencing of wastewater. Sequence data was provided by the CASPER consortium and raw data may be found on NCBI SRA under bioprojects PRJNA1247874 and PRJNA1198001. Data underwent read trimming, rRNA and human read removal, de novo assembly, and selection of high-quality viral contigs. Contigs were clustered at 95% identity and 85% query coverage to dereplicate. Chimera-checking required at least two independent assemblies of the same viral genome or presence of the genome in another reference database. Annotation made use of RdRpCATCH, geNomad, checkV, BLASTN against NCBI core-nt, and RNAVirHost. The RdRp fasta files contain representative RdRp sequences identified through homology to major RdRp reference databases and clustered at 90% sequence identity over 75% sequence coverage. Included sequences contain all three conserved RdRp motifs (A, B, and C) arranged in either the canonical ABC configuration or the permuted CAB configuration.

open·CC-BY-4.0·zenodo-bio·completeSource
composite

Daten der Notaufnahmesurveillance

0.00

Robert Koch-Institut · AKTIN-Notaufnahmeregister

5 files · 10 KB · pdf, tsv, zip

Der Datensatz "Daten der Notaufnahmesurveillance" wird durch das Robert Koch-Institut und das AKTIN-Notaufnahmeregister bei Krankenhausaufnahmen bereitgestellt. Der Datensatz beinhaltet aggregierte Routinedaten aus deutschen Notaufnahmen zur syndromischen Überwachung von akuten Erkrankungen. Dazu zählen grippeähnliche Erkrankungen (ILI), Coronavirus-Erkrankungen (COVID-19), akute respiratorische Erkrankungen (ARE), gastrointestinale Infektionen (GI) und schwere akute respiratorische Infektionen (SARI). Dabei wird der relative Anteil dieser Erkrankungen an der Gesamtzahl der Notaufnahmevorstellungen sowie die berechneten Erwartungswerte und Prädiktionsintervalle ausgewiesen. Die Daten sind nach Notaufnahmetypen und Altersgruppen aggregiert. Damit bietet der Datensatz eine wertvolle Ressource für die Forschung im Bereich der Notfallmedizin und der Überwachung akuter Gesundheitsereignisse in Deutschland.

open·CC-BY-4.0·Zenodo·completeSource
tabular

Microbiota study IgG4-RD AG Chang

0.00

Budzinski, Lisa · Beenken, Anne Elisabeth · Sempert, Toni · et al.

9 rows × 1 cols · 743 B · csv, docx, zip

1 categorical

We have investigated an IgG4-RD (IgG4-RD) cohort by our multi-parameter microbiota flow cytometry approach to characterise the microbiota on single-cell level for attributes of the disease. The microbiota is isolated from stool samples and stained according to the published protocol for (a) host immunoglobulins IgA1, IgA2, IgM, IgG and (b) agglutinin binding to mannose, galactose or N-Acetyl-glucosamine surface sugar moieties. For all samples we also determined the microbiome composition by 16S rRNA (V3-V4) sequencing on the illumina MiSeq platform. We provide the raw .fcs and FASTQ files of 40 IgG4-RD patients. For comparison we additionally analysed 36 healthy donors. All .fcs files were generated on BD Influx®. The metadata is collected in the provided meta.csv. The staining parameters are summarized in provided panel.csv.

open·CC-BY-4.0·zenodo-bio·0% null·completeSource
composite

GIFT-BDS: A high-resolution TEC and Gradient Ionospheric Index dataset over China derived from BeiDou GEO fixed-geometry observations

0.00

Li, Zhiyao · Wang, Ningbo · Zhong, Jiahao

6 files · 48 MB · docx, zip

GIFT-BDS is a regional ionospheric total electron content (TEC) and TEC-gradient dataset over China derived from BeiDou geostationary Earth orbit (GEO) observations and a dense ground-based GNSS receiver network. The dataset is designed to provide high-resolution observations of ionospheric TEC variability and horizontal TEC-gradient structures over China and adjacent regions. The versioned release covers the period from 19 July 2024 to 31 December 2025, corresponding to DOY 201 of 2024 to DOY 365 of 2025. The geographical coverage is 15°N-50°N and 95°E-135°E. The dataset is provided in daily NetCDF files and contains two product levels. Level-1 products provide observation-level GEO-derived slant TEC (STEC) and rate of TEC index (ROTI) records for individual receiver-GEO satellite lines of sight, with a temporal resolution of 30 s. Level-2 products provide gridded regional TEC and TEC-gradient variables, including VTEC, VTEC t , ROTI, GIX, GIX std , GIX x , GIX y , GIX t,x , and GIX t,y , with a temporal resolution of 15 min. IPP-based variables are provided on a 1° × 1° grid, while inter-IPP-gradient variables are provided on a 0.25° × 0.25° grid. The main processing steps include observation screening, cycle-slip and data-gap detection, continuous-arc segmentation, carrier-to-code leveling, satellite and receiver DCB correction, IPP calculation, inter-IPP pair selection, gradient estimation, and gridding. Quality control is applied before release. Missing values may occur because of station outages, data gaps, quality-control exclusions, or insufficient valid samples within a grid cell. Users should check the NetCDF variable attributes, including units and fill values, before analysis. The dataset is suitable for regional ionospheric studies, TEC-gradient monitoring, space-weather-related analyses, and investigations of ionospheric effects on GNSS positioning applications.

open·CC-BY-4.0·Zenodo·completeSource
modal

Coding reliability dataset for: Representation-to-AI Transformation in K–12 Generative AI Learning: A Theory-Building Systematic Review of Semantic Transformation Mechanisms

0.00

Jungmyoung, Son · Sihoon, Lee · Jiyeon, Hong

3 files · 20 KB · docx, xlsx

This dataset provides the complete double-coding matrix, PRISMA 2020 checklist, and search strategy supporting the systematic review "Representation-to-AI Transformation in K-12 Generative AI Learning: A Theory-Building Systematic Review of Semantic Transformation Mechanisms." It includes: (1) study-level tier classification (Core/Supporting/Context) for two independent coders and consensus tier for all 18 included studies; (2) the full semantic transformation unit (STU) coding matrix (18 studies x 10 STUs = 180 cells) with pre-consensus and consensus scores; (3)evidence-weighting consensus scores; (4) inter-rater reliability statistics (Cohen's kappa); (5) the completed PRISMA 2020 checklist; and (6) the full database-specific Boolean search strategy.

open·CC-BY-4.0·Zenodo·completeSource
composite

ADM_LSIR: a physics-inspired laparoscopic aerosol degradation dataset

0.00

guo, na · pan, jiachen · li, tiantian · et al.

14 files · 100 MB · csv, rar, tsv

ADM_LSIR is a physics-inspired laparoscopic aerosol degradation dataset for aerosol-aware surgical image analysis and image restoration. The v1.0.0 release contains: - 21,916 clean clinical laparoscopic frames (clean/) - 9,562 real intraoperative aerosol-degraded frames (degraded/) - 36,052 simulated aerosol masks, including 19,701 smoke-like masks and 16,351 trajectory masks (mask/) - Blender simulation/cache materials (ADM_LSIR_Blender_simulation_files_v1.0.rar) - metadata_quality_report_v1.0.csv - recommended_splits_v1.0.csv - video_mapping_v1.0.csv - parts_manifest.txt - checksums_v1.0.tsv - release_manifest_v1.0.json All released clinical frames are de-identified and stored as lossless PNG files. Filenames use anonymized video identifiers, e.g., C-V##-####.png for clean frames and D-V##-####.png for degraded frames. The recommended split is defined at the source_video_id/public_video_label level to reduce leakage across frames from the same source video. The public video labels in video_mapping_v1.0.csv provide privacy-safe source-video identifiers (video1-video19). The Blender archive documents the smoke and trajectory mask simulation setup and supports reuse, but it is not a guaranteed exact per-mask reproduction package. The released pre-rendered mask library is the primary reusable dataset component. Source code for synthesis and quality screening is available at: https://github.com/SweetDeathh/ADM_LSIR

open·CC-BY-4.0·Zenodo·completeSource
composite

HQ MAGs (CheckM comp≥90%, contam≤5%) from the MicroToxBol Bolivian human gut microbiome cohort

0.00

Manghi, Paolo

2 files · 8.0 MB · gzip, tsv

Using gene-level and species-level shotgun metagenomics, we provide the first characterization of the rural, Bolivian microbiome; we identified microbial genes which strongly correlate (rho>0.45) with arsenic in urine, and that overall contribute to substantiate that the gut microbiome helps tolerate arsenic via a evict-out-of-house mechanism. Mediation analysis, followed by phylogenetic investigation of metagenomic-assembled genomes, further strengthens this observation. This study elucidates the role of the microbiome in helping to tolerate arsenic-rich environments, and paves the way for probiotic interventions that may mitigate the effects of this toxic metal. This repository contains 2,478 HQ MAGs from the MicroToxBol cohort in fasta format and a descriptive table comprising taxonomic annotation, quality-checks, and coverage estimation.

open·CC-BY-4.0·zenodo-bio·completeSource
composite

Fast Breakdowns Observed in the Initial Leaders of Two Energetic Compact Strokes

0.00

Yang, Qingliu

6 files · 8.0 MB · bzip2

Dataset Description This dataset contains 3D lightning location results, DALMA and FALMA waveform for two Energetic Compact Stroke (ECS) events. location results are included: HF3D_1732785151.dat - 3D lightning locations for the ECS leader A flash. HF3D_1734785454.dat - 3D lightning locations for the ECS leader B flash. The timestamp 1734785454 and 1732785151 corresponds to the occurrence time of the lightning flash in Japan Standard Time. File format and parameters The first row contains the lightning occurrence time. Column descriptions: Time (ms) - time relative to the lightning source. X, Y, Z (m) - 3D spatial coordinates relative to ground level. The origin (0, 0, 0) corresponds to latitude 36.76°N and longitude 136.76°E. FALMA and DALMA waveform ECSLeaderA_DALMA_waveform.bz2 is DALMA waveform of Leader A. ECSLeaderA_FALMA_waveform.bz2 is FALMA waveform of Leader A. ECSLeaderB_DALMA_waveform.bz2 is DALMA waveform of Leader B. ECSLeaderB_FALMA_waveform.bz2 is FALMA waveform of Leader B. This dataset allows analysis of the spatial and temporal development of these two ECS flashes.

open·CC-BY-4.0·Zenodo·completeSource
tabular

Generated ASO features for the OligoAI dataset

0.00

Kovaliov, Michael

1 files · 100 MB · parquet

open·CC-BY-4.0·Zenodo·completeSource
declared

Fine tuning an LLM with a domain a specific data set

0.00

Madhusudan, Gujral

6 files · 29 MB · parquetdeclared

Large language models (LLMs) are trained on massive, publicly available text datasets comprising trillions of tokens, enabling them to excel at general language tasks like next-token prediction. However, LLMs often struggle with domain-specific prompts, exhibiting reduced accuracy or generating inaccurate information (hallucinations). This is because they lack sufficient subject matter expertise. Two primary approaches exist to address this limitation for augmenting LLMs knowledge: Retrieval-Augmented Generation (RAG) and fine-tuning. This presentation focuses on fine-tuning smaller LLMs with domain-specific instruct datasets using the LoRA (Low-Rank Adaptation) technique on Gaudi hardware. We will leverage publicly available LLMs and datasets from the Hugging Face Hub for this demonstration. Though it is possible to fine tune LLMs with plain text data - sourced from documents, articles, and other materials.

open·CC-BY-4.0·Zenodo·completeSource
declared

Supplemental Table 1. Meeting Agenda of the 2024 PCOS Challenge–CDC Stakeholder Meeting, and Supplemental Table 2. Proposed Four-Stage Workflow for Developing Standardized Testosterone Reference Intervals Using Existing Study Data

0.00

Azziz, Ricardo

1 files · 1.6 MB · docxdeclared

Supplemental Table 1. Meeting Agenda of the 2024 PCOS Challenge-CDC Stakeholder Meeting on Testosterone Reference Interval Standardization. August 26, 2024. Centers for Disease Control and Prevention, Atlanta, Georgia. Supplemental Table 2. Proposed Four-Stage Workflow for Developing Standardized Testosterone Reference Intervals Using Existing Study Data.

open·CC-BY-4.0·Zenodo·completeSource
declared

LA INTELIGENCIA ARTIFICIAL COMO HERRAMIENTA COMPLEMENTARIA EN EL PROCESO ACADÉMICO EN LA EDUCACIÓN BÁSICA

0.00

García Mina, Janeth Elizabeth · Rodríguez Garófalo, Napoleón Hernán · Rumbaut-Rangel, Dayron · et al.

6 files · 737 KB · docx, pdfdeclared

El objetivo de este estudio es evaluar la importancia de la inteligencia artificial (IA) en el proceso de formación académica de estudiantes de educación básica, específicamente en la enseñanza de ciencias naturales. La investigación destaca cómo la IA puede personalizar el aprendizaje, adaptándose a los estilos individuales de cada alumno, lo que resulta en una experiencia educativa más efectiva y significativa. Se utilizó un diseño cuasiexperimental con una muestra de 60 estudiantes de sexto año, distribuidos en un grupo control (30 estudiantes) que recibió instrucción tradicional y un grupo experimental (30 estudiantes) que utilizó herramientas de IA, como chatbots educativos y plataformas de aprendizaje adaptativo. Los resultados mostraron que el 70% de los estudiantes del grupo experimental alcanzaron calificaciones superiores a 8, en comparación con solo el 30% del grupo control. Este hallazgo sugiere que la implementación de herramientas de IA puede transformar la educación básica, mejorando tanto el rendimiento académico como la motivación de los estudiantes. En conclusión, el estudio resalta la necesidad de seguir investigando y desarrollando nuevas estrategias educativas que integren la IA para maximizar su potencial en la educación.

open·CC-BY-4.0·Zenodo·completeSource
declared

Pregnant women and Breastfeeding mothers Transcript

0.00

Makhado, Langanani Christinah · Raliphaswa, Ndidzulafhi Selina

9 files · 157 KB · docxdeclared

This dataset contains anonymized transcripts derived from face-to-face interviews conducted with pregnant women and breastfeeding mothers as part of the research study. The interviews were audio-recorded and transcribed verbatim to capture participants' experiences, perceptions, and insights related to the study topic. All personally identifiable information has been removed to protect participant confidentiality. The transcripts are provided to support transparency, reproducibility, and further research related to the findings reported in the associated article

open·CC-BY-4.0·Zenodo·completeSource
declared

Auto-Brewery Syndrome: A Narrative Review for Clinical Awareness

0.00

Fontelo, Paul

1 files · 26 KB · docxdeclared

Auto-Brewery Syndrome (ABS), also known as gut fermentation syndrome, is a condition in which microbial fermentation of dietary carbohydrates generates endogenous ethanol, causing signs and symptoms of alcohol intoxication without alcohol consumption. Misdiagnosis as alcohol use disorder (AUD) is common and carries severe medical, psychosocial, legal, and forensic consequences. The objective of thus narrative review is to synthesize the clinical literature on ABS for practicing clinicians and to present the first systematic analysis of ABS susceptibility in Asian populations - including communities of Asian descent in Western countries.

open·CC-BY-4.0·Zenodo·completeSource
declared

EuroFlood: a queryable cloud-native index for the CEMS-EFAS Satellite-Derived Flood Depth Maps

0.00

Hackl, Jürgen

6 files · 132 MB · parquet, tiffdeclared

EuroFlood is an open, cloud-native index over the JRC/Copernicus CEMS-EFAS Satellite-Derived Flood Depth Maps for Europe (Betterle & Salamon, 2025; CC-BY-4.0) - ~3,280 satellite-derived observed flood-depth maps across Europe, 2015-2024. The bundle is a sparse Cloud-Optimized GeoTIFF encoding, per pixel, the set of flood events that inundated it, plus a combo_id -sorted GeoParquet dictionary and a small events table. Query by region and time via HTTP range reads (GDAL /vsicurl + DuckDB) to retrieve matching events, then fetch only the source depth rasters needed. Built with the open-source EuroFlood Python package ( pip install euroflood ).

open·CC-BY-4.0·Zenodo·completeSource
declared

GROND

0.00

Walsh, Calum · Srinivas, Meghana · Stinear, Timothy · et al.

100 files · 2.7 GB · gzip, tsvdeclared

GROND (Genome-derived Ribosomal OperoN Database) A quality-checked and publicly-available database of 16S-ITS-23S RRNA operon sequences and their constituent 16S and 23S genes. Based on GTDB release R232.

open·CC-BY-4.0·Zenodo·completeSource
declared

Report of the indigenous and local knowledge dialogue workshop for the draft summary for policymakers and second order draft of the IPBES sustainable use assessment

0.00

Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services

2 files · 1.5 MB · docx, pdfdeclared

This is the report on the indigenous and local knowledge (ILK) dialogue workshop for the first order draft of the summary for policymakers and the second order draft of the IPBES assessment of the sustainable use of wild species (the "sustainable use assessment"). It was held from 17-21 May 2021, online, due to the ongoing COVID-19 pandemic. The report aims to provide a written record of the dialogue workshop, which can be used by assessment authors to inform their work on the sustainable use assessment, and also by all dialogue participants who may wish to review and contribute to the work of the assessment moving forward. The report is not intended to be comprehensive or give final resolution to the many interesting discussions and debates that took place during the workshop. Instead, it is intended as a written record of the discussions, and this conversation will continue to evolve in the course of the assessment process. For this reason, clear points of agreement are discussed, but diverging views among participants are also presented for further attention and discussion.

open·CC-BY-4.0·Zenodo·completeSource
page 1next →

Select a result to see its full details here: the measured structure, quality, and the loader, without leaving your search.