Exploration

ResearchFeatured

Discovery

DiscoverSourcesQuality

Analysis

Working setReviews
Flow StudioTeamConcept
Settings

Partners

  • AI AlliancePrime
  • BrightQueryBuilds Meridian
  • OpenMinedFunded partner
  • MLCommonsFunded partner
  • Hugging FaceDeployment platform
See the full consortium and what each partner wires

Meridian is the discovery layer for research data, built by BrightQuery within the AI Alliance.

hybrid · semantic + lexical · 108 datasets ranked · 4.65s

Structuremodal108
Depthmeasured108
Licenseopen98unknown9non commercial1
Accessopen108
Formattiff12pdf8xlsx4docx3csv2
Sourcezenodo108zenodo-geo2zenodo-bio1
clear
1-20 of 108sortrelevancemeasured firstqualitysize
modal

Morphosyntactic and Pragmatic Constructions in Jiepai Speech (Danyang, Jiangsu, China): An Emic Grammar Note

0.00

The Jiepai Archivist

3 files · 338 KB · pdf

zip1

Abstract This open-access release provides an emic grammar note on selected morphosyntactic and pragmatic constructions in Jiepai speech, a local Sinitic speech variety from the Wu-Jianghuai contact zone in Jiepai, Danyang, Jiangsu, China. The file is prepared as a grammar appendix to the Jiepai Jianghu section of the Chuanyue archive, an emic vernacular multimodal archive created by a native insider. The note documents selected structures in grammar, syntax, aspect, disposal, passive-related expression, locative marking, shared-use constructions, degree expression, and pragmatic usage. It includes notes on ge marking; completion-question and negative-completion forms such as geng and beng; aspectual distinctions related to Mandarin le and zhe; na disposal constructions; ha / bo constructions across passive, giving, and marriage-related contexts; locative suffixes; go shared-use constructions; yanlian parallel-action structures; degree adverbs and ironic intensification; and locative lai2 / A patterns. The document is based on the Archive Author's native-speaker knowledge and ongoing documentation of Jiepai speech. It does not present a complete grammar, final character standardization, final IPA transcription, phonetic analysis, or ELAN alignment. Phonetic, tonal, and character-form notes marked as provisional or pending verification are retained as research openings rather than resolved conclusions. Authorship and deposit roles are separate. The Archive Author, The Jiepai Archivist, is the creator, archive owner, rights holder, and public contact person. The uploader provides curatorial and deposit assistance only and is not the author, creator, archive owner, rights holder, or public contact for this archive.

open·CC-BY-NC-ND-4.0·Zenodo·completeSource
modal

Structured to Fail: Gender Bias in Large Language Models Across Text and Visual Modalities Through Data Feminism and Intersectionality in the Indian Context

0.00

Poonia, NIkita · Saraswat, Dr. Niraja · Poonia, Dr. Arun Kumar

7 files · 107 KB · pdf

This repository contains the dataset associated with the study " Structured to Fail: Gender Bias in Large Language Models Across Text and Visual Modalities Through Data Feminism and Intersectionality in the Indian Context". The deposit includes a representative subset of the full study data, comprising the following components: Sample Design: Documentation of the sampling framework and run structure across the three LLM platforms examined Prompt Battery: The complete set of prompts administered across text and image generation tasks Model-Generated Text Outputs: Textual responses generated (1,020) by the models under the study Model-Generated Image Outputs: Visual outputs generated (480) in response through image-generation prompts Codebook: Provides complete instructions for all independent coders participating in the inter-rater reliability (IRR) study Inter-rater Reliability Dataset: The double-coded subset used to establish coding agreement Reported Reliability Scores: Cohen's Kappa values and percentage agreement statistics as reported in the manuscript

open·CC-BY-4.0·Zenodo·completeSource
modal

Bridging ecological restoration and social legitimacy: a systematic review of Cultural Ecosystem Services in inland aquatic ecosystems

0.00

Comalada i Pla, Francesc

1 files · 224 KB · docx

Dataset containing the systematic review matrix and extracted variables supporting the article "Bridging ecological restoration and social legitimacy: a systematic review of Cultural Ecosystem Services in inland aquatic ecosystems", accepted for publication in People and Nature.

open·CC-BY-4.0·Zenodo·completeSource
modal

Intensivkapazitäten und COVID-19-Intensivbettenbelegung in Deutschland

0.00

Robert Koch-Institut

9 files · 197 KB · csv, pdf, zip

Der Datensatz "Intensivkapazitäten und COVID-19-Intensivbettenbelegung in Deutschland" des Robert Koch-Instituts dokumentiert die tägliche intensivmedizinische Versorgungslage seit der COVID-19-Pandemie. Basierend auf Meldungen aller intensivbettenführenden Krankenhäuser in Deutschland erfasst das DIVI-Intensivregister Echtzeitdaten zu belegten und freien Intensivbetten. Die Erhebung differenziert nach Altersgruppen, Regionen und Versorgungsstufen. COVID-19-Fälle auf Intensivstationen werden gesondert ausgewiesen. Die Daten stehen aggregiert auf Bundes-, Landes- und Kreisebene zur Verfügung. Damit bildet der Datensatz eine Grundlage für die Überwachung von Kapazitäten, die Koordination von Behandlungskapazitäten und politische Entscheidungsprozesse während der Pandemie und darüber hinaus.

open·CC-BY-4.0·Zenodo·completeSource
modal

Chemical characterization (proximate composition, fatty acids content and volatile profile) of meat from the alpine Ciuta sheep breed.

0.00

Lopez, Annalaura · Greco, Margherita · Marcolli, Beatrice · et al.

4 files · 17 KB · docx, xlsx

This dataset originates from a study aiming to valorise Ciuta sheep, a local breed native from the Italian Central Alps, through the characterization of nutritional quality and chemical composition of fresh meat (loins) and one traditional dry-cured product. Specifically, the research focused on determining the chemical composition of Ciuta sheep meat and on identifying key changes in its chemical profile during dry curing process, hypothesizing that such chemical fingerprint may suggest some markers linked to the production system, geographical origin, and traditional processing techniques. For this reason, for bthe dry-cured product, both an aliquot of fresh meat before and after transformation and dry-curing was sampled and analysed. Regarding loins, three commercial categories (lambs, hoggets and mutton) were considered, in order to define any possible difference induced by age of the sheep (and physiological factors, such as rumen development). The dataset includes chemical data regarding the proximate composition (moisture, protein, fat, ash, salt content for the dry-cured product) and energy content of fresh and dry-cured meat; the fatty acids content of fresh and dry-cured meat product; the volatile profile of fresh and dry-cured meat product. Results from analysis performed in our study suggested that the development of high-quality dry-cured products could provide a strategy to valorise Ciuta sheep meat, especially from adult animals (culled ewes and rams), while fresh meat production could focus on lambs. The complex volatile profile detected was influenced by both the farming system and traditional processing methods.

open·CC-BY-4.0·Zenodo·completeSource
modal

Open-Access Digital Humanities Research (2016-2026)

0.00

Suwarno · Akun, Andreas

2 files · 295 KB · pdf, xlsx

File ini merupakan data mentah ( raw data ) ekspor dari database Scopus yang digunakan untuk melakukan pemetaan, analisis bibliometrik, atau Tinjauan Literatur Sistematis (SLR) mengenai tren riset di bidang Digital Humanities.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,8 (SLUB)

0.00

Melnik, Elena

13 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,7 (SLUB)

0.00

Melnik, Elena

7 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,6 (SLUB)

0.00

Melnik, Elena

4 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,5 (SLUB)

0.00

Melnik, Elena

25 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,4 (SLUB)

0.00

Melnik, Elena

40 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,3 (SLUB)

0.00

Melnik, Elena

19 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,2 (SLUB)

0.00

Melnik, Elena

25 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-14,1 (SLUB)

0.00

Melnik, Elena

25 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.4189-D-5 (SLUB)

0.00

Melnik, Elena

25 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Thermographic data for manuscript Mus.3481-D-3 (SLUB)

0.00

Melnik, Elena

37 files · 4.5 MB · tiff

This data set contains thermographic images for the visualisation of watermarks. An IRCAM Equus 327k with Watermark Imager software (Fraunhofer) version 8.416 (R2016b) was used.

open·CC-BY-4.0·Zenodo·completeSource
modal

Descomposición Geométrico-Hiperestática y Análisis de la Fuerza de Hendimiento en Columnas Arbóreas de Hormigón Armado

0.00

Valencia Pérez, Luis Alberto

766 KB

Se presenta una metodología numérica para el diseño morfológico de columnas arbóreas tridimensionales de concreto reforzado, basada en la minimización iterativa del residuo ortogonal de equilibrio mediante relajación angular dinámica sobre un modelo FEM Timoshenko 3D. El trabajo propone una unificación entre el form-finding funicular y la verificación nodal STM-MCFT mediante una descomposición ortogonal del residuo inducida por el jacobiano del operador de equilibrio, demostrando que en un punto estacionario del funcional objetivo la componente geométrica se anula y el residuo remanente es irreducible mediante ajustes angulares. Se incluye un estudio paramétrico sobre cinco combinaciones de carga gravitacional que demuestra la invariancia práctica de la geometría óptima ante variaciones de magnitud y distribución de carga.

open·CC-BY-4.0·Zenodo·completeSource
modal

CCF Database (PostgreSQL edition): A sentence-level corpus of 266,271 Canadian climate-change articles from 20 newspapers (1978-2024)

0.00

Lemor, Antoine · Pillod, Alizée · Taylor, Matthew · et al.

2.2 MB

The Canadian Climate Framing (CCF) Database is a comprehensive, machine-learning-annotated corpus of climate-change media coverage in Canada. It comprises 266,271 articles from 20 major Canadian newspapers (1978-2024) processed into 9,198,958 two-sentence analytical units (82.9% English, 17.1% French). Each unit is annotated across 65 hierarchical categories by 128 BERT and CamemBERT classifiers, with a macro F1 of 0.866 on a 1,000-sentence gold standard double-coded by an independent annotator (Gwet's AC1 = 0.894, Krippendorff's α = 0.698, Cohen's κ = 0.596 on the 400 blind sentences). Each category receives an A/B/C reliability tier summarising annotation quality from classifier performance and inter-coder agreement. The deposit ships six relational tables (bibliographic metadata, sentence-level annotations, named-entity rollups, article-level aggregates, per-category reliability tiers, and 9,462,845 BAAI/bge-m3 sentence-and-title embeddings). Raw newspaper text is excluded for copyright reasons; bibliographic coordinates (media, date, title, author, page_number) are sufficient for any researcher with institutional access to Factiva, Eureka.cc or ProQuest Canadian Major Dailies to recover the original sentences. This deposit accompanies a methodology paper currently under revision at Scientific Data (Nature Portfolio). This deposit is the canonical PostgreSQL edition. It contains a pg_dump -Fd directory archive (compressed into a single .tar file) of the six relational tables, including the pgvector extension and HNSW cosine indexes for sub-second semantic-similarity search. Restoration is a one-liner: tar -xf CCF_Database.tar && createdb CCF_Database && psql -d CCF_Database -c 'CREATE EXTENSION IF NOT EXISTS vector;' && pg_restore -d CCF_Database --no-owner --no-privileges -j 8 CCF_Database_dump A column-oriented Apache Parquet mirror of the same six tables is available as the sister deposit on Zenodo (cross-referenced in Related identifiers ). The Parquet mirror is recommended for users without PostgreSQL access (it is directly readable by pandas, polars, R/arrow, DuckDB, and Spark). The full annotation pipeline, training data, manual-annotation JSONL, intercoder-reliability benchmark, methodology manuscript (LaTeX sources + PDF), and reproducibility scripts are bundled with this deposit as ccf_code_and_paper.tar.gz . The same materials are also available on the project's OSF companion deposit ( 10.17605/OSF.IO/Q5W47 ) and on the development mirror at GitHub . Requirements: PostgreSQL 16 or 17 with pgvector ≥ 0.8.2 (for halfvec(1024) storage of the sentence embeddings).

open·CC-BY-4.0·Zenodo·completeSource
modal

Daten der Notaufnahmesurveillance

0.00

Robert Koch-Institut · AKTIN-Notaufnahmeregister

127 KB

Der Datensatz "Daten der Notaufnahmesurveillance" wird durch das Robert Koch-Institut und das AKTIN-Notaufnahmeregister bei Krankenhausaufnahmen bereitgestellt. Der Datensatz beinhaltet aggregierte Routinedaten aus deutschen Notaufnahmen zur syndromischen Überwachung von akuten Erkrankungen. Dazu zählen grippeähnliche Erkrankungen (ILI), Coronavirus-Erkrankungen (COVID-19), akute respiratorische Erkrankungen (ARE), gastrointestinale Infektionen (GI) und schwere akute respiratorische Infektionen (SARI). Dabei wird der relative Anteil dieser Erkrankungen an der Gesamtzahl der Notaufnahmevorstellungen sowie die berechneten Erwartungswerte und Prädiktionsintervalle ausgewiesen. Die Daten sind nach Notaufnahmetypen und Altersgruppen aggregiert. Damit bietet der Datensatz eine wertvolle Ressource für die Forschung im Bereich der Notfallmedizin und der Überwachung akuter Gesundheitsereignisse in Deutschland.

open·CC-BY-4.0·zenodo-geo·completeSource
modal

Multi-method late-successional and old-growth (LSOG) forest mapping uncertainty for Maine's unorganized townships

0.00

Weiskittel, Aaron R.

807 KB

This dataset quantifies the uncertainty in mapping late-successional and old-growth (LSOG) forest across the approximately 4.2 million hectares of Maine's unorganized townships, and tests whether LSOG is rapidly disappearing. Three to four independent, credible mapping methods are compared on a common 100 m grid: (M1) a reproduction of the Hagan et al. (2026) airborne-LiDAR canopy random forest, rebuilt from their public Zenodo deposit; (M2) a logistic model of the FIA field-structure LSOG class on Potapov (GEDI-calibrated) canopy height; (M3) a direct canopy-height threshold; and (M4) the FIA structural class imputed to every pixel via USFS TreeMap (2016, 2020, 2022). Version 1.2.0 additions. This version adds the materials behind the formal Ecosphere Comment on Hagan et al. (2026): (a) a cross-validated accuracy assessment (AUC) of each mapping approach on the original authors' own training plots, showing that high training accuracy does not transfer to agreement among independent maps; (b) an FIA design-based estimate of older forest with sampling-error confidence intervals, the unbiased ground reference the original analysis lacked, putting older forest at about 3.9 percent (3.3 to 4.6) and rising, including on private commercial timberland; (c) a threshold-sensitivity sweep and a 20-seed reproduction ensemble; (d) an ownership-resolved breakdown (private commercial versus public); (e) a hex-scale (8 km) summary of cross-method disagreement; and (f) the Comment manuscript and Supporting Information. Headline findings. Credible methods disagree by roughly 2.8 times on how much LSOG exists and on the location of most LSOG hectares, while agreeing closely on the rare, well-defined old-growth core. Protecting the top 5 to 20 percent of hectares by one map versus another overlaps on only 16 to 30 percent of the ground, so single-map patch-level prioritization for large expenditures is fragile. The design-based FIA estimate and TreeMap imputation both show older forest stable to increasing rather than rapidly declining; the apparent loss reported elsewhere is a gross harvest flux, not a net stock decline. Contents. Derived 100 m GeoTIFFs (reproduced Hagan class, v5.1-GEDI probability, TreeMap class, a per-cell method-consensus layer), summary tables (area by method, pairwise agreement, concordance, prioritization fragility, AUC by approach, design-based older-forest trend with CIs, ownership breakdown, and FIA validation), the analysis R scripts, quick-look figures, the Ecosphere Comment manuscript and Supporting Information, and a full methods-and-findings report (PDF). Privacy. No FIA plot coordinates are included; all products are derived rasters or aggregate summary tables. Caveats: the robust temporal signal is direction rather than precise rate; FIA stand age is modeled, so a structural large-tree domain is reported alongside the age domain; cross-validated intervals are best read as lower bounds because plots are spatially dispersed but not independent. See the README and report for full methods, provenance, and limitations. Version 1.12.0 additions. The cross-map comparison is refined to independent remote-sensing operationalizations only. (a) A three-map remote-sensing ensemble over Maine on a common 100 m grid: reproduced Hagan airborne-LiDAR (any-LSOG 21.9 percent), an FIA-structure class on Potapov GEDI-calibrated spaceborne canopy height (14.0 percent), and the ORNL/Bruening national old-growth stratum (36.1 percent); the three span a 2.6-fold range and agree on only 2.7 percent of flagged hectares, with the ORNL stratum spatially uncorrelated with the structure maps. The USFS TreeMap imputation is reclassified as a second FIA-anchored accounting, reported with the design-based estimate rather than as an independent map. (b) A design-based estimate of LSOG itself: integrated any-LSOG 14.1 percent (12.9 to 15.3) and strict four-axis true LSOG 3.1 percent (2.5 to 3.7) of Maine forestland, the airborne map exceeding even the inclusive ground estimate. (c) A balanced-model LSOG probability surface at 100 m, with a binary class calibrated to the design-based area to bound over-prediction. (d) A multi-objective support vector regression pilot tracing the Pareto front of total versus systematic (attenuation) error. Derived rasters, the three-map agreement layer, the ORNL stratum reprojected to the study grid, tables, R scripts, the updated Comment, and the companion manuscript are included. No FIA plot coordinates are included. Version 1.13.0. Final consolidated release. Adds: a rare-class remedy menu for the reproduced random forest (default vs class weighting vs balanced sub-sampling vs voting-threshold; old-growth detection 0.24 to 0.82, mapped old-growth area 1.0 to 2.3 percent); the definitive five-model LSOG probability map for Maine with across-model uncertainty and reference reserves (MNAP/TNC network, Baxter, Big Reed) over real state and county boundaries; design-based 95 percent confidence intervals for forest-type and ecoregion representation and for disturbance shares; a full robustness/stress-test matrix; and the copy-edited, sole-authored Comment, companion manuscript, and Maine Forest Products Council technical report. Authorship updated to Aaron R. Weiskittel.

open·CC-BY-4.0·zenodo-geo·completeSource
page 1next →

Select a result to see its full details here: the measured structure, quality, and the loader, without leaving your search.