Exploration

ResearchFeatured

Discovery

DiscoverSourcesQuality

Analysis

Working setReviews
Flow StudioTeamConcept
Settings

Partners

  • AI AlliancePrime
  • BrightQueryBuilds Meridian
  • OpenMinedFunded partner
  • MLCommonsFunded partner
  • Hugging FaceDeployment platform
See the full consortium and what each partner wires

Meridian is the discovery layer for research data, built by BrightQuery within the AI Alliance.

hybrid · semantic + lexical · 296 datasets ranked · 2.15s

Depthcataloged296
Licenseopen290non commercial3share alike3
Accessopen296
Formatcsv280pdf49zip45xlsx30tsv19
Sourcezenodo296
clear
1-20 of 296sortrelevancemeasured firstqualitysize
declared

Asynchrony of ageing among traits in a wild bird population

0.00

Tsui, Claire · Briga, Michael · Komdeur, Jan · et al.

tar16
fits15
torch11
png10
gzip9
tiff9
parquet8
docx7
netcdf6
shapefile4
hdf53
jpeg3
geojson2
npy2
rar2
vcf2
bzip21
fasta1
sqlite1
xls1
56 files · 8.9 MB · csv
declared

Data and code for analysis in manuscript titled "Asynchrony of ageing among traits in a wild bird population" Dataframes ending with"_28_5.csv" and "survival_model.csv" are used in scripts model1-13, of which the output is plotted using "new model outputs.R" Code for Figures 1 and 2 are in script "new model outputs.R" asymmetry bivar ver3.R runs the bivariate models used to estimate the degree of synchrony of ageing. scripts starting with "aic.." are used in the analysis for age by lifespan interaction

open·CC-BY-4.0·Zenodo·completeSource
declared

Fine tuning an LLM with a domain a specific data set

0.00

Madhusudan, Gujral

6 files · 29 MB · parquetdeclared

Large language models (LLMs) are trained on massive, publicly available text datasets comprising trillions of tokens, enabling them to excel at general language tasks like next-token prediction. However, LLMs often struggle with domain-specific prompts, exhibiting reduced accuracy or generating inaccurate information (hallucinations). This is because they lack sufficient subject matter expertise. Two primary approaches exist to address this limitation for augmenting LLMs knowledge: Retrieval-Augmented Generation (RAG) and fine-tuning. This presentation focuses on fine-tuning smaller LLMs with domain-specific instruct datasets using the LoRA (Low-Rank Adaptation) technique on Gaudi hardware. We will leverage publicly available LLMs and datasets from the Hugging Face Hub for this demonstration. Though it is possible to fine tune LLMs with plain text data - sourced from documents, articles, and other materials.

open·CC-BY-4.0·Zenodo·completeSource
declared

Data for: A Controlled in Silico Benchmark for GNN Prediction of Tissue Dynamics

0.00

Krajnc, Matej · Comi, Troy · Miao, Siqi · et al.

10 files · 25 GB · csv, zipdeclared

This dataset accompanies the manuscript "A Controlled in Silico Benchmark for GNN Prediction of Tissue Dynamics." It contains model prediction outputs, trained checkpoints, train/validation/test splits, spring-embedding outputs, generated analysis figures, analysis tables, and manuscript-specific diagnostic outputs used to reproduce the post-prediction analyses and figures. The dataset is distributed as logical ZIP archives with file-level and archive-level SHA-256 checksums. For questions, contact Tomer Stern at tomers@umich.edu.

open·CC-BY-4.0·Zenodo·completeSource
declared

Therapist-guided forest therapy and anxiety symptoms in adolescents: de-identified dataset and analysis code

0.00

MENEGUZZO, FRANCESCO · Zabini, Federica

4 files · 146 KB · csvdeclared

This record contains the de-identified, analysis-ready dataset and Python analysis code associated with a school-based quasi-experimental pilot study on therapist-guided forest therapy and persistent anxiety symptoms in adolescents. The study involved two fourth-year high-school classrooms in Cecina, Tuscany, Italy: one intervention classroom that attended four therapist-guided forest therapy sessions in a coastal pine forest, and one control classroom that followed usual school activities. The dataset includes anonymized student codes, classroom allocation, SCAS total raw scores and derived reduction scores across repeated assessments, POMS-A acute mood-state variables for the intervention classroom, exploratory post-intervention nature-exposure and connectedness variables, and exploratory school-performance variables. The repository includes four files: 1. forest_therapy_adolescent_anxiety_dataset.csv: de-identified analysis-ready dataset. 2. README_forest_therapy_adolescent_anxiety_dataset.txt: dataset metadata and variable descriptions. 3. analyze_forest_therapy_adolescent_anxiety.py: Python script used to reproduce the main analyses and generate analysis outputs. 4. README_analyze_forest_therapy_adolescent_anxiety.txt: operational guide for running the analysis script and interpreting the output files. The dataset uses semicolon-separated values. Missing values are encoded as NaN and must not be interpreted as zero. Because the study involved minors, item-level questionnaire responses and more granular school records are not shared; only de-identified and analysis-ready variables compatible with privacy, ethical approval, and consent constraints are provided.

open·CC-BY-4.0·Zenodo·completeSource
declared

On the Probabilities of Spacetime: A Statistical Topography via Maximum Entropy

0.00

Soltes, Julian

3 files · 12 MB · csv, pdfdeclared

This paper presents a non-parametric topography of the cosmological landscape, following the principle of maximum entropy to map the relative probability of all isotropic spacetime configurations. The topography is defined as an 11-D probability distribution, derived from a maximally unbiased ensemble of solutions to the Einstein Field Equations (EFE). Uncertainty prevents infinite precision, blurring the ensemble to quantify the relative likelihood of any physical microstate. The result provides a statistical foundation for the manifestation of our universe and its contents, described by probabilities and their respective gradients. This assumes unconstrained numerical coverage of the EFE landscape, mapping mathematically valid solutions that may violate parametric energy conditions. Uncertainty distorts the boundaries between these exotic states and the canonical phase, rationalizing their existence within our universe as quantifiable improbabilities. Resources: The computational implementation and main data ensemble are attached here, as well as maintained at: https://github.com/jgsoltes/Universe-KDE Researchers are encouraged to apply the QUASAR optimizer to their own high-dimensional or non-convex function landscapes. Source code and documentation are maintained at: https://github.com/jgsoltes/hdim-opt

open·CC-BY-4.0·Zenodo·completeSource
declared

EuroFlood: a queryable cloud-native index for the CEMS-EFAS Satellite-Derived Flood Depth Maps

0.00

Hackl, Jürgen

6 files · 132 MB · parquet, tiffdeclared

EuroFlood is an open, cloud-native index over the JRC/Copernicus CEMS-EFAS Satellite-Derived Flood Depth Maps for Europe (Betterle & Salamon, 2025; CC-BY-4.0) - ~3,280 satellite-derived observed flood-depth maps across Europe, 2015-2024. The bundle is a sparse Cloud-Optimized GeoTIFF encoding, per pixel, the set of flood events that inundated it, plus a combo_id -sorted GeoParquet dictionary and a small events table. Query by region and time via HTTP range reads (GDAL /vsicurl + DuckDB) to retrieve matching events, then fetch only the source depth rasters needed. Built with the open-source EuroFlood Python package ( pip install euroflood ).

open·CC-BY-4.0·Zenodo·completeSource
declared

Corpus des Deutschen Bundesrechts (C-DBR)

0.00

Fobbe, Sean

16 files · 1.4 GB · csv, pdf, zipdeclared

Überblick Das Corpus des deutschen Bundesrechts (C-DBR) ist eine möglichst vollständige Sammlung der konsolidierten Fassungen aller Gesetze und Verordnungen auf Bundesebene. Der Datensatz nutzt als seine Datenquelle das amtliche Internetangebot www.gesetze-im-internet.de des Bundesministeriums der Justiz und wertet dieses vollständig aus. Bitte lesen Sie zuerst das beiliegende Codebook! Es enthält wichtige Informationen zur korrekten Nutzung des Datensatzes. Es hilft auch bei der Entscheidung, welche Variante für Sie am besten geeignet ist. In der Regel empfehle ich für quantitative Forschung die CSV-Dateien und für traditionelle Forschung die PDF-Sammlung. Um das Gesetzgebungsverfahren näher zu beleuchten können Sie zusätzlich auf folgende Datensätze zurückgreifen (jeweils mit Links auf vergleichbare Datensätze anderer Autor:innen): Corpus der Drucksachen des Deutschen Bundestages (CDRS-BT) Corpus der Plenarprotokolle des Deutschen Bundestages (CPP-BT) Aktualisierung Dieser Datensatz wird ca. alle 3 Monate aktualisiert. Benachrichtigungen über neue und aktualisierte Datensätze veröffentliche ich immer zeitnah auf Mastodon unter @seanfobbe@fediscience.org NEU in Version 2026-07-09 Vollständige Aktualisierung der Daten Ausführung der Pipeline in Userland im Container Neues Skript für Zenodo Upload Eckdaten Stichtag: 9. Juli 2026 Umfang: 6.124 Bundesgesetze und -verordnungen der Bundesrepublik Deutschland Formate: CSV, PDF, EPUB, TXT und XML Features Einfache Nutzung für statistische Analysen mit CSV-Dateien Bis zu 42 Variablen in den CSV-Varianten Fortlaufende Aktualisierung Urheberrechtsfreiheit Sowohl für traditionelle Rechtsanwender als auch für Legal Tech-Anwendungen geeignete Formate (CSV, PDF, EPUB, TXT und XML) Umfangreicher Compilation Report um den Erstellungs-Prozess zu erläutern Hochauflösende Diagramme und deskriptive Tabellen für alle Zwecke Diagramme in PDF (Druck) und PNG (Web) verfügbar, Tabellen als menschen- und maschinenlesbares CSV Vollständiges tabellarisches Verzeichnis aller Rechtsakte und der vom BMJV gebrauchten Abkürzungen Netzwerk-Strukturen für alle Rechtsakte und Visualisierungen für über 1000 Rechtsakte (experimentell) Veröffentlichung des Source Codes Source Code und Compilation Report Der gesamte Erstellungs-Prozess ist vollautomatisiert und detailliert dokumentiert. Mit jeder Kompilierung des vollständigen Datensatzes wird auch ein umfangreicher Compilation Report in einem attraktiv designten PDF-Format erstellt (ähnlich dem Codebook). Zudem werden Robustness Checks auf Vollständigkeit und Plausibilität durchgeführt und in einem separaten Bericht dokumentiert. Der Compilation Report enthält den Code für die vollständige Pipeline, dokumentiert relevante Rechenergebnisse, gibt sekundengenaue Zeitstempel an und ist mit einem klickbaren Inhaltsverzeichnis versehen. Er ist zusammen mit dem Source Code hinterlegt. Wenn Sie sich für Details des Erstellungs-Prozesses interessieren, lesen Sie diesen bitte zuerst. Der vollständige Source Code - sowohl für die Erstellung des Datensatzes, als auch für das Codebook - ist öffentlich einsehbar und dauerhaft erreichbar im wissenschaftlichen Archiv des CERN unter diesem Link hinterlegt: https://zenodo.org/doi/10.5281/zenodo.4072934 Kryptographische Signaturen Die Integrität und Echtheit der einzelnen Archive des Datensatzes sind durch eine Zwei-Phasen-Signatur sichergestellt. In Phase I werden während der Kompilierung für jedes ZIP-Archiv, das Codebook und die Robustness Checks Hash-Werte in zwei verschiedenen Verfahren (SHA2-256 und SHA3-512) berechnet und in einer CSV-Datei dokumentiert. In Phase II werden diese CSV-Datei und der Compilation Report mit meinem persönlichen geheimen GPG-Schlüssel signiert. Dieses Verfahren stellt sicher, dass die Kompilierung von jedermann durchgeführt werden kann, insbesondere im Rahmen von Replikationen, die persönliche Gewähr für Ergebnisse aber dennoch vorhanden ist. Die während der Kompilierung des Datensatzes erstellte CSV-Datei mit den Hash-Prüfsummen ist mit meiner persönlichen GPG-Signatur versehen. Der mit dieser Version korrespondierende Public Key ist sowohl mit dem Datensatz als auch mit dem Source Code hinterlegt. Er hat folgende Kenndaten: Name: Sean Fobbe (fobbe-data@posteo.de) Fingerabdruck: FE6F B888 F0E5 656C 1D25 3B9A 50C4 1384 F44A 4E42 Kein Urheberrecht: Public Domain An den Normtexten und Metadaten besteht gem. § 5 Abs. 1 UrhG kein Urheberrecht, da sie amtliche Werke sind. § 5 UrhG ist auf amtliche Datenbanken analog anzuwenden (BGH, Beschluss vom 28.09.2006 - I ZR 261/03, "Sächsischer Ausschreibungsdienst"). Alle eigenen Beiträge (z.B. durch Zusammenstellung und Anpassung der Metadaten) und damit den gesamten Datensatz stelle ich gemäß einer CC0 1.0 Universal Public Domain License vollständig urheberrechtsfrei. Disclaimer Dieser Datensatz ist eine private wissenschaftliche Initiative und steht in keiner Verbindung zu Behörden, Gerichten oder anderen öffentlichen Stellen der Bundesrepublik Deutschland. Alternativen [Ab 10.06.2019, nur XML] Beckedorf, Janis/Coupette, Corinna/Hartung, Dirk. 2020. "gesetze-im-internet: A daily archive of https://www.gesetze-im-internet.de". GitHub. https://github.com/QuantLaw/gesetze-im-internet [Änderungsgesetze] Wehrmeyer, Stefan/Semsrott, Arne/Filter, Johannes. 2021. "OffeneGesetze.de ist eine zivilgesellschaftliche, ehrenamtliche Plattform für amtliche Gesetzesblätter". Open Knowledge Foundation. https://offenegesetze.de/ [Alte Rechtsakte] Open Knowledge Foundation. 2013. "Bundesgit". GitHub. https://github.com/bundestag/gesetze Weitere Open Access Veröffentlichungen (Fobbe) Website - www.seanfobbe.de Open Data - zenodo.org/communities/sean-fobbe-data/ Source Code - zenodo.org/communities/sean-fobbe-code/ Volltexte regulärer Publikationen - zenodo.org/communities/sean-fobbe-publications/ Kontakt Fehler gefunden? Anregungen? Kommentieren Sie gerne im Issue Tracker oder kontaktieren Sie mich über www.seanfobbe.de

open·CC0-1.0·Zenodo·completeSource
declared

Replication package for "An Environmental Data Justice-Driven Analysis of Setback Distances"

0.00

Vera, Lourdes

31 files · 4.6 MB · csv, geojson, pngdeclared

Data, analysis scripts, and derived outputs reproducing the setback analysis between occupied buildings and active oil and gas wells in Karnes County, Texas, using public data (Texas Railroad Commission well locations; FEMA/ORNL USA Structures building footprints; and U.S. Census TIGER/Line block groups and 2020-2024 American Community Survey). The pipeline runs offline in Python (geopandas) and reproduces every reported distance, summary statistic, table, and figure in the associated article. This package reproduces and updates Chapter 3 of the author's doctoral dissertation: Vera, Lourdes (2022), "Environmental Data Justice in Action: Civically Valid Air Monitoring Near Oil and Gas Extraction in the Eagle Ford Shale Play," Ph.D. dissertation, Northeastern University, Boston, MA.

open·MIT·Zenodo·completeSource
declared

COMEX Gold Futures Realized Volatility Dataset and Replication Code (2014–2025)

0.00

Wareesri, Prapassorn · Ieamvijarn, Subunn

14 files · 1.7 MB · csvdeclared

Dataset and replication code for the paper "The Regime-Dependent Value of Macroeconomic Information in Gold Futures Volatility Forecasting: A HAR-Machine Learning Comparison on COMEX" submitted to Investment Management and Financial Innovations. Data sourced from Yahoo Finance covering January 2014 to December 2025.

open·CC-BY-4.0·Zenodo·completeSource
declared

Analysis Code and Data: Green Infrastructure and Urban Acoustic Environment — Dawn Chorus Study, Bochum

0.00

Haghbin, Kourosh

7 files · 121 KB · csv, xlsxdeclared

Supplementary code and data for the Master's thesis "The Influence of Green Infrastructure on the Urban Acoustic Environment: A Seasonal Noise Analysis in Bochum" (TU Dortmund University, 2026). The repository contains Python scripts implementing the full statistical analysis pipeline: site-level data assembly, Ordinary Least Squares (OLS) regression, and Multiscale Geographically Weighted Regression (MGWR) of BirdNET-derived normalised Shannon bird diversity (BN_H) against acoustic and structural green infrastructure predictors across 118 spring morning monitoring sites in Bochum, Germany. The primary script ( run_ols_mgwr_dawn.py ) implements the four-predictor model (dB, NDSI, AEI, and GI Tier; OLS R² = 0.319, MGWR R² = 0.796), selected as the primary model based on AICc (ΔAICc = -105.70 over the five-predictor alternative). A comparison script ( run_ols_mgwr_area.py ) implements the five-predictor model including log-transformed GI polygon area (OLS R² = 0.333, MGWR R² = 0.895). All analyses are implemented from scratch in Python 3 using NumPy and Pandas, without proprietary GIS or statistical libraries. The site-level analytical dataset (Dawn_4to9_Analysis_Data.xlsx, n = 118 sites, 04:00-09:00 dawn chorus window) is included to enable full reproduction of reported results.

open·CC-BY-4.0·Zenodo·completeSource
declared

Groundwater and landslide displacement records of the Eggerberg slope (Gradenbach landslide, Carinthia, Austria)

0.00

Hagen, Karl · Zieher, Thomas · Stary, Ulrike · et al.

21 files · 6.6 MB · csv, pdfdeclared

Groundwater and landslide displacement records of the Eggerberg slope (Gradenbach landslide, Carinthia, Austria) Austrian Research Centre for Forests (BFW) Institute for Natural Hazards, Unit of Torrent Process & Hydrology K. Hagen, T. Zieher, U. Stary‚ E. Lang, S. Riedl, G. Priesch, J. Pichler, J. Rojacher First published June 2026 Contact: wasser.naturgefahren@bfw.gv.at Overview This dataset documents long-term hydrogeological and displacement monitoring at the deep-seated rock slide Eggerberg, situated in the Gradenbach catchment (Carinthia, Austria). The active mass movement covers approximately 2 km² and reaches depths exceeding 130 m. Owing to its interaction with the Gradenbach torrent system, the instability represents a significant natural hazard for nearby settlements in the Möll Valley. The dataset includes groundwater level and temperature measurements, as well as landslide displacement records collected by the Austrian Research Centre for Forests (BFW). It represents one of the longest continuous hydrogeological and geotechnical monitoring programs of a deep-seated gravitational slope deformation in the European Alps. It extends the existing dataset published in Hormes et al. (2026) by data of several boreholes, not used in the publication. Available groundwater temperature records were compiled and added in the borehole records. Furthermore, the displacement record of the extensometer was reset to zero after the data gap from 1995 to 1999. Dataset Description Groundwater monitoring Groundwater levels were monitored in 15 boreholes including several paired installations designed to observe groundwater conditions in different depth horizons and aquifer systems. Monitoring of the goundwater levels started in 1979, using manual cable light-plummet measurements with an accuracy of approximately 1 cm, and ended in 2024. In the beginning, measurements were generally performed every two weeks, since 1998 usually weekly. Groundwater temperature was measured between 1998 and 2015 with a sensor accuracy of 0.1°C. Continuous digital monitoring of groundwater level and temperature is additionally available for selected boreholes from July 2007 to December 2024 using OTT Orpheus Mini pressure probes. The observation series reveal the presence of approximately four hydrogeologically distinct aquifers within the moving rock mass. Landslide displacement monitoring Landslide displacement was monitored using a wire extensometer installed across the Gradenbach ravine. The instrument measured changes in the distance between the moving landslide mass and the comparatively stable opposite valley flank. The observation record extends from May 1979 to April 2024 and includes the same principal data gap between 1996 and 1998. Initially, measurements were recorded using analogue strip-chart systems and later digitized. In 2006, the monitoring station was upgraded with a continuous digital acquisition system (Thalimedes). To reduce short-term thermal effects caused by steel-wire expansion and contraction, the published dataset contains daily aggregated displacement values. All records (groundwater level and temperature, landslide displacement) were generally quality-controlled and checked for outliers and plausibility before publication. However, no warranty is given regarding their accuracy, completeness, or fitness for any particular purpose. Data Structure Groundwater Files Each borehole is provided as an individual CSV file (GRD-GWL-[borehole number]): • Column 1: Date (YYYY-MM-DD) • Column 2: Groundwater temperature (°C) • Column 3: Groundwater level below ground surface (m) Missing or unreliable values are coded as 9999. Groundwater levels are reported as negative values relative to the ground surface. Extensometer File The landslide displacement dataset is provided as a single CSV file (GRD_EXT.csv): • Column 1: Date (YYYY-MM-DD) • Column 2: Landslide displacement (cm), expressed as reduction of the distance between canyon slopes • Column 3: Measurement and data-quality code The PDF document 'GRD_Zenodo-V1-20260701.pdf' provides site information, methods, data gaps, and corrected observations. The ancillary document 'monitoring_data_gradenbach.html' provides minimal code snippets for working with the data in R. It further includes interactive plots of the data for visual inspection. Scientific R elevance The Gradenbach-Eggerberg dataset provides a unique long-term record enables the investigations of groundwater-controlled slope acceleration processes, and temporal trends in aquifer behavior. The dataset therefore constitutes an important resource for landslide process research, hazard assessment, hydrogeological investigations, and model validation.

open·CC-BY-4.0·Zenodo·completeSource
declared

Intensivkapazitäten und COVID-19-Intensivbettenbelegung in Deutschland

0.00

Robert Koch-Institut

9 files · 54 MB · csv, pdf, zipdeclared

Der Datensatz "Intensivkapazitäten und COVID-19-Intensivbettenbelegung in Deutschland" des Robert Koch-Instituts dokumentiert die tägliche intensivmedizinische Versorgungslage seit der COVID-19-Pandemie. Basierend auf Meldungen aller intensivbettenführenden Krankenhäuser in Deutschland erfasst das DIVI-Intensivregister Echtzeitdaten zu belegten und freien Intensivbetten. Die Erhebung differenziert nach Altersgruppen, Regionen und Versorgungsstufen. COVID-19-Fälle auf Intensivstationen werden gesondert ausgewiesen. Die Daten stehen aggregiert auf Bundes-, Landes- und Kreisebene zur Verfügung. Damit bildet der Datensatz eine Grundlage für die Überwachung von Kapazitäten, die Koordination von Behandlungskapazitäten und politische Entscheidungsprozesse während der Pandemie und darüber hinaus.

open·CC-BY-4.0·Zenodo·completeSource
declared

Dataset for: Effects of Camera Height, Field of View, and Driver Age on Depth Judgement and Lane-Change Decisions in Camera Monitor Systems

0.00

Pifferi, Gabriele · Thulinsson, Felix · Söderlund, Niclas · et al.

9 files · 165 KB · csvdeclared

# Camera Monitor System (CMS) Dataset and Analysis Scripts Version 4 Dataset and analysis scripts associated with the manuscript: "Effects of Camera Height, Field of View, and Driver Age on Depth Judgement and Lane-Change Decisions in Camera Monitor Systems" ## Changelog ### Version 4 - Added the R analysis scripts for the analyses reported in the associated manuscript (`CMS_continuous_age_analysis.R`, `Within_subject_figures.R`, `confidence_analysis.R`, `unsigned_error_analysis.R`); see the Analysis scripts section below. - Data files unchanged from Version 3. ### Version 3 - Corrected a file export error that caused row truncation in `distance_estimation_cleaned.csv` and `lane_change_dataset.csv` (earlier versions). Most data columns were missing from each row in those files due to a comma/semicolon delimiter collision during export. Files have been rebuilt from the original source data. - Corrected a mixed decimal notation issue (comma vs dot) in the `Confidence` column of `distance_estimation_cleaned.csv`, which caused the column to be read as a string rather than a numeric type in standard CSV parsers. - Updated `variable_dictionary.csv` to accurately reflect the column names used in all three data files (earlier versions listed intended names that did not match the actual file contents). - Added two participants (P48, P49) to `participant_metadata.csv` whose records were missing from the earlier export due to a data extraction error. Their experimental task data were present in `lane_change_dataset.csv` but lacked corresponding metadata rows. - Clarified that the N=56 figure in the preprocessing section applies specifically to the distance estimation task (task-specific outlier exclusion). The lane-change task and participant metadata retain the full N=58. ### Version 2 Same datafiles as in Version was uploaded by mistake ### Version 1 Initial release. --- ## Licence This dataset and the accompanying analysis scripts are shared under the Creative Commons Attribution 4.0 International (CC BY 4.0) licence. https://creativecommons.org/licenses/by/4.0/ --- ## Included files ### Data #### participant_metadata.csv Participant-level demographic and background information. N=58. #### distance_estimation_cleaned.csv Cleaned participant-level data from the distance estimation task. N=56 after task-specific outlier exclusion (see Data preprocessing below). #### lane_change_dataset.csv Participant-level data from the lane-change task. N=58. #### variable_dictionary.csv Definitions and descriptions of all variables included across the three data files. ### Analysis scripts (R) #### CMS_continuous_age_analysis.R #### Within_subject_figures.R #### confidence_analysis.R #### unsigned_error_analysis.R See the Analysis scripts section below for descriptions, requirements, and run order. --- ## Experimental overview A controlled laboratory experiment combining two aligned data collections was conducted using dynamic rearward driving scenarios in a Camera Monitor System (CMS) environment. Participants completed: - a distance estimation task - a lane-change task (last safe gap, LSG) Experimental factors: - Field of View (FOV): 40°, 76°, 112° - Camera height: High / Low - Driver age: continuous variable (range 22-64 years, mean 38.2 years). The variable `AgeGroupMedianSplit` is retained in the data files for continuity with prior analyses (Pifferi, 2025), but is not used in the primary analysis reported in the associated manuscript. --- ## Data preprocessing Preliminary analyses examined the distribution of the dependent variables for the distance estimation (Dist) and lane-change (LSG) tasks. The LSG data did not show substantial skewness or kurtosis, whereas the Dist data contained several extreme values reflecting substantial underestimation relative to the actual target distances. Outliers in the distance estimation data were identified using a threshold of ±2.5 standard deviations from the mean of the dependent variables (absolute and relative distance estimation errors). Exclusion limits: - Absolute distance error: -84 m to 79 m - Relative distance error: -2.76 to 2.55 Two participants were excluded from the distance estimation task because more than half of their trials fell outside the ±2.5 SD range (participants P48 and P49). These participants are retained in `lane_change_dataset.csv` and `participant_metadata.csv`, as the exclusion criterion was specific to the distance estimation task. Two additional participants contained isolated outlying trials (one and two trials, respectively); these isolated outlying trials were replaced using mean-value imputation within the corresponding dependent variable. Following preprocessing, the distance estimation dataset contains 56 participants and the lane-change dataset contains 58 participants. --- ## Analysis scripts R scripts for the analyses reported in the associated manuscript. All scripts expect the dataset CSV files in the same folder as the scripts, or edit `data_dir` at the top of each script. Requirements: R >= 4.0. Each script checks for and installs its required packages (lme4, lmerTest, emmeans, tidyverse, performance, ggplot2, ggsignif, dplyr, rmcorr) from CRAN if missing. Run order: 1. **CMS_continuous_age_analysis.R** - Primary analyses. Linear mixed-effects models with driver age as a continuous predictor for signed distance estimation error (DistErr), relative error (RelErr), and time-to-contact (TTC), including simple-slope (estimated marginal trend) analyses, the quadratic age check, the gap-split sensitivity analysis, and model-based predictions with confidence intervals. Writes three CSV files of model predictions to the working directory. 2. **Within_subject_figures.R** - Publication figures for the within-subject effects. Must be run in the same R session after script 1, as it uses the emmeans objects created there. Figure export lines (`ggsave`) are provided but commented out. 3. **confidence_analysis.R** - Confidence rating analyses: condition effects via linear mixed-effects models, repeated-measures correlations (rmcorr) between trial-level confidence and objective performance, and participant-level Spearman correlations. Self-contained; can be run independently. 4. **unsigned_error_analysis.R** - Unsigned (absolute) error analyses for the distance estimation task, using the same model structure as the primary analyses, with partial eta squared computed from the F-based approximation used in the manuscript. Self-contained; can be run independently. --- ## Ethics and anonymisation The shared datasets do not contain directly identifying personal information. Participant IDs are anonymised. Raw interview recordings, interview transcripts, and any potentially identifying qualitative material are not included in the shared repository for ethical and privacy reasons. --- ## Suggested citation Pifferi, G., Thulinsson, F., Söderlund, N., Brunnström, K., Rafiei, S., Schenkman, B., Djupsjöbacka, A., Sperandio, I., & Andrén, B. (2026). Dataset for: *Effects of Camera Height, Field of View, and Driver Age on Depth Judgement and Lane-Change Decisions in Camera Monitor Systems* [Data set]. Zenodo. DOI: https://doi.org/10.5281/zenodo.20055125

open·CC-BY-4.0·Zenodo·completeSource
declared

Toxicity of binary mixtures of cadmium with lead, copper or zinc to Folsomia candida in relation to bioavailability in soil

0.00

van Gestel, Cornelis

25 files · 746 KB · csv, pdfdeclared

Data from binary mixture toxicity tests with springtails exposed to Cd-Pb, Cd-Cu and Cd-Zn mixtures, including: Data used to construct Langmuir sorption isotherms relating total metal concentrations to CaCl2 or H2O extractable concentrations in soil, for each of the three mixtures PDF files showing the Results of the Langmuir sorption isotherm calculations Data on the survival of springtails in exposures to each of the three mixtures Data on the growth of the springtails in exposures to each of the three mixtures Results of the mixture toxicity analysis for growth effects for each of the three mixtures Data on the reproduction of the springtails in exposures to each of the three mixtures Results of the mixture toxicity analysis for reproduction effects for each of the three mixtures Results of the analysis of chloride concentrations in H2O extracts of the soil of the three mixture exposures Data on the pH of the test soil dosed with the metals, for each of the three mixtures

open·CC-BY-4.0·Zenodo·completeSource
declared

HG_JUVENILE - Juvenile herring gulls (Larus argentatus, Laridae) hatched at the southern North Sea coast (Belgium)

0.00

Allaert, Reinoud A. · Stienen, Eric W.M. · Lens, Luc · et al.

6 files · 125 MB · csv, gzipdeclared

HG_JUVENILE - Juvenile herring gulls (Larus argentatus, Laridae) hatched at the southern North Sea coast (Belgium) is a bird tracking dataset published by the Centre for Research on Ecology, Cognition and Behaviour of Birds at Ghent University and the Research Institute for Nature and Forest (INBO) . It contains animal tracking data for the project/study HG_JUVENILE , using trackers developed by Interrex ( http://www.interrex-tracking.com ). The study has been operational since 2022. In total 204 individuals of European herring gull ( Larus argentatus ) have been tagged. 150 individuals were raised from egg by Ghent University researchers at the Wildlife Rescue Center in Ostend, completed several cognitive and behavioural tests when approximately three weeks old, and were released in the IJzermonding, Nieuwpoort (Belgium). 54 additional individuals were tagged and released in the wild to collect baseline data. The main goal of the study is to link cognitive performance in the lab to behaviour in the wild. Data are automatically synced with Movebank and from there periodically archived on Zenodo (see https://github.com/inbo/bird-tracking ). Files Data in this package are exported from Movebank study 2217728245 . Fields in the data follow the Movebank Attribute Dictionary and are described in datapackage.json . Files are structured as a Frictionless Data Package . You can access all data in R via https://zenodo.org/records/21279427/files/datapackage.json using frictionless . datapackage.json : technical description of the data files. HG_JUVENILE-reference-data.csv : reference data about the animals, tags and deployments. HG_JUVENILE-gps-yyyy.csv.gz : GPS data recorded by the tags, grouped by year. Acknowledgements This dataset was collected using infrastructure provided by the ERC and Ghent University.

open·CC0-1.0·Zenodo·completeSource
declared

The effect of leaching on the joint toxicity of a complex metal mixture to Folsomia candida in relation to bioavailability in soil

0.00

van Gestel, Kees

14 files · 192 KB · csvdeclared

Data files on cadmium, copper, lead and zinc concentrations in soil, H2O- and CaCl2-extracts of soil and in springtails exposed to complex mixtures of chloride salts of these metals, before and after leaching the soil to remove the chloride counterion. Also included are data on the effects of the metals, single and in mixtures on the survival, growth and reproduction of the springtails.

open·CC-BY-4.0·Zenodo·completeSource
declared

Monthly 1-km SPI, SPEI and SRI drought-indicator rasters for Poland (1995-2024)

0.00

Miszczyszyn, Jakub · Radoń, Radosław

20 files · 2.1 GB · csv, netcdfdeclared

Monthly 1-km SPI, SPEI and SRI drought-indicator rasters for Poland (1995-2024) This dataset provides monthly gridded drought indicators for the entire territory of Poland at 1-km resolution for the period 1995-2024. Three complementary standardized indices are included: the Standardized Precipitation Index (SPI, precipitation-based), the Standardized Precipitation-Evapotranspiration Index (SPEI, based on the climatic water balance P - PET) and the Standardized Runoff Index (SRI, runoff-based). Each index is provided at five accumulation scales: 1, 3, 6, 12 and 24 months. Methods. Station indices were computed from IMGW-PIB precipitation and river-runoff observations and from AgERA5 potential evapotranspiration (used for SPEI). SPI and SRI were fitted with a gamma distribution and SPEI with a log-logistic distribution. Station values were then interpolated to a 1-km grid (EPSG:2180, PUWG 1992) by ordinary kriging. Elevation-assisted kriging (kriging with external drift) was evaluated by leave-one-out cross-validation but produced no net national improvement and was not adopted for the final product. File structure. Data are provided as CF-compliant NetCDF (CF-1.8), one file per indicator and accumulation scale (e.g. SPI_s06m_PL_1km_1995-2024_EPSG2180.nc ) . Each file has dimensions (time, y, x) with a regular monthly time axis (360 steps, 1995-01 to 2024-12) and the coordinate reference system stored as a grid_mapping variable. Months in which an index cannot be formed at the edges of the accumulation window are present as fill-valued (missing) layers, so the time axis is continuous; their completeness is documented in raster_index.csv . Contents. netcdf/ - the 15 index rasters; metadata/raster_index.csv - per-layer inventory with min/max/mean, missing-data fraction and a "present" flag; metadata/variables_dictionary.csv - variable definitions; metadata/mckee_classes.csv - the seven-class drought/wetness classification (McKee et al., 1993) with colours; CITATION.cff and checksums.md5 . Usage note. In GIS software the temporal dimension is read via the temporal/time controls (e.g. in QGIS: Layer Properties → Temporal → Dynamic Temporal Control, then the Temporal Controller); in Python the files open directly with xarray, with time parsed as dates. Coordinate reference system. EPSG:2180 (PUWG 1992 / Poland CS92). Map extents delineate the study area and do not necessarily depict accepted national boundaries.

open·CC-BY-4.0·Zenodo·completeSource
declared

Structurally similar, functionally different: Impact of coformer positional isomerism on co-amorphous enzalutamide

0.00

Balaga, Venkata Krishna Rao · Chatziadi, Argyro · Ridvan, Ludek · et al.

69 files · 111 MB · csv, xlsxdeclared

Raw data obtained during work on comparison of the effect of different positional isomerism on the physicochemical properties of coamorphous forms. The solid forms were characterised using solid-state techniques such as XRD, mDSC, and FTIR. Moreoever, solubility tests, Intrinsic dissolution, powder dissolution, and stability tests were also conducted. The README file contains naming convention, structure of the data, etc. This research was supported by the project NETPHARM: New Technologies for Translational Research in Pharmaceutical Sciences, project ID CZ.02.01.01/00/22_008/0004607, co-funded by the European Union, and The Pharmaceutical Applied Research Center (The PARC).

open·CC-BY-4.0·Zenodo·completeSource
declared

Data from "Milder winters alleviate seasonal challenges for a migratory goose facing Arctic warming"

0.00

Geisler, Jan · Rakhimberdiev, Eldar · Boom, Michiel P. · et al.

29 files · 124 MB · csv, shapefiledeclared

1. Many migratory birds now reach their Arctic breeding grounds earlier in order to keep pace with advancing springs and shifting nutrient peaks, either by departing earlier from non-breeding grounds or by travelling faster. For dark-bellied brent geese, there is limited potential to travel faster, as their migration to the Siberian breeding grounds is already among the fastest of Arctic geese and swans. Earlier departure would require reaching departure body mass earlier, either through a faster accumulation of energy stores during spring staging or via adjustments earlier in the annual cycle. 2. We examined long-term shifts in spring staging phenology and changes in winter and spring body mass trajectories of brent geese at the population level, with particular emphasis on the effects of winter temperature on body mass and spring body mass on departure timing. 3. We used more than five decades of body mass measurements from individuals caught in the United Kingdom and France, and in the Dutch Wadden Sea to reconstruct changes in spring and winter mass trajectories, respectively. These data were combined with over five decades of migration counts in the Netherlands and more than two decades of counts in Denmark to quantify changes in spring staging phenology. 4. We found that brent geese have not shifted their spring arrival in the Wadden Sea but have advanced departure timing. Furthermore, brent geese were heavier during and after milder winters, and have changed mass trajectories over recent decades. They no longer lose mass during winter and the second spring staging phase, and fuelling rates in the first spring staging phase have declined. Annual variation in body mass was not related to annual departure timing. 5. These results suggest that milder winters have relaxed energetic constraints and improved body condition in brent geese throughout the non-breeding season. Our findings highlight the importance of considering the full annual cycle when assessing how animals with limited capacity to adjust migration timing or speed respond to global change.

open·CC-BY-4.0·Zenodo·completeSource
declared

Dataset for the publication entitled "Fortification of an Innovative Tomato Cold Soup with High Bioaccessible Sulforaphene from UV‑B–Treated Radish Seeds"

0.00

Martínez Zamora, Lorena · Castillejo Montoya, Noelia · Artés-Hernández, Francisco

2 files · 13 KB · csvdeclared

The aim of this work was to develop an innovative tomato cold soup fortified in bioactive compounds through the incorporation of UV-B-treated radish seeds. After a 20 kJ m-2 UV-B treatment, radish seeds increased their sulforaphene content by 30%. Different concentrations of UV-B-treated seeds (0, 0.5, 1.5, 3, and 5 g kg-1) were added to a chopped vegetables cold soup, mainly made of Kumato® cherry tomatoes as novelty, including pepper, cucumber, and garlic, which was stored for 8 days at 4 °C. Added seeds did not affect physicochemical quality attributes, microbial growth, nor sensory perception. Nevertheless, a dose-dependent behaviour was shown in glucoraphenin and sulforaphene content, according to concentrations of UV-B-treated seeds added. It was also appreciated after an in vitro digestion that the bioaccessible fraction of glucosinolates and isothiocyanates was kept constant throughout the refrigerated storage. The sulforaphene content of the soup increased by ~ 19% after 2 days at 4 °C, of which the 33% was bioaccessible (measured in vitro), and subsequently was degraded by ~ 20% after 8 days at 4 °C.

open·CC-BY-4.0·Zenodo·completeSource
page 1next →
closeopen full
tabular · zenodo

Explore Wales: A Geospatial Dataset of Heritage and Visitor Locations

Williams, Aled

measured·open·2 files
Measuredstructure observed by touching the bytes
topology
tabular
records
206
sampled
full dataset
null rate
0.0%
duplicate rate
0.0%
size
40 KB
profiler
tabular:stdlib/v1

Measured the full file.

Columns (8)
columntypenullsdistribution
idtext string0%206 distinct · 5-5 ch
nametext string0%206 distinct · 6-46 ch
categorycategorical string0%41 distinct · castle · museum · industrial_site
latitudenumeric number0%51.39 - 53.41 · med 51.85
longitude
numeric
number
0%
-5.268 - -2.675 · med -3.8
utm_easting_mnumeric number0%343,900 - 522,500 · med 446,500
utm_northing_mnumeric number0%5,693,000 - 5,919,000 · med 5,745,000
notestext string0%206 distinct · 65-126 ch
Columns by kind

Measured distributions

latitude
p10 51.49 · median 51.85 · p90 53.21
longitude
p10 -4.698 · median -3.8 · p90 -3.09
utm_easting_m
p10 385,500 · median 446,500 · p90 493,900
utm_northing_m
p10 5,704,000 · median 5,745,000 · p90 5,896,000