Exploration

ResearchFeatured

Discovery

DiscoverSourcesQuality

Analysis

Working setReviews
Flow StudioTeamConcept
Settings

Partners

  • AI AlliancePrime
  • BrightQueryBuilds Meridian
  • OpenMinedFunded partner
  • MLCommonsFunded partner
  • Hugging FaceDeployment platform
See the full consortium and what each partner wires

Meridian is the discovery layer for research data, built by BrightQuery within the AI Alliance.

hybrid · semantic + lexical · 30 datasets ranked · 0.67s

Structuretensor2tabular1
Depthcataloged27measured3
Licenseopen30
Accessopen30
Formatshapefile11csv9npy9parquet9zip5
Sourcezenodo30
clear
1-20 of 30sortrelevancemeasured firstqualitysize
tensor

Data for: Chiral polariton transport enabled by optical spin Hall effect in perovskite waveguides

0.00

Kędziora, Mateusz

49 files · 1.8 MB · npy

gzip4
tiff4
geojson3
geopackage3
netcdf3
pdf2
xlsx2
fasta1
hdf51
jpeg1
rar1
sqlite1
torch1
tsv1

Dataset for figures in the Article: Chiral polariton transport enabled by optical spin Hall effect in perovskite waveguides

open·CC-BY-4.0·Zenodo·completeSource
tabular

Generated ASO features for the OligoAI dataset

0.00

Kovaliov, Michael

1 files · 100 MB · parquet

open·CC-BY-4.0·Zenodo·completeSource
tensor

Effects of turbulence on the vertical evolution of raindrop size distribution during short-duration heavy precipitation in Hainan (China) based on X-band phased array radar

0.00

Cai

12 files · 8.0 MB · npy

open·CC-BY-4.0·Zenodo·completeSource
declared

Fine tuning an LLM with a domain a specific data set

0.00

Madhusudan, Gujral

6 files · 29 MB · parquetdeclared

Large language models (LLMs) are trained on massive, publicly available text datasets comprising trillions of tokens, enabling them to excel at general language tasks like next-token prediction. However, LLMs often struggle with domain-specific prompts, exhibiting reduced accuracy or generating inaccurate information (hallucinations). This is because they lack sufficient subject matter expertise. Two primary approaches exist to address this limitation for augmenting LLMs knowledge: Retrieval-Augmented Generation (RAG) and fine-tuning. This presentation focuses on fine-tuning smaller LLMs with domain-specific instruct datasets using the LoRA (Low-Rank Adaptation) technique on Gaudi hardware. We will leverage publicly available LLMs and datasets from the Hugging Face Hub for this demonstration. Though it is possible to fine tune LLMs with plain text data - sourced from documents, articles, and other materials.

open·CC-BY-4.0·Zenodo·completeSource
declared

EuroFlood: a queryable cloud-native index for the CEMS-EFAS Satellite-Derived Flood Depth Maps

0.00

Hackl, Jürgen

6 files · 132 MB · parquet, tiffdeclared

EuroFlood is an open, cloud-native index over the JRC/Copernicus CEMS-EFAS Satellite-Derived Flood Depth Maps for Europe (Betterle & Salamon, 2025; CC-BY-4.0) - ~3,280 satellite-derived observed flood-depth maps across Europe, 2015-2024. The bundle is a sparse Cloud-Optimized GeoTIFF encoding, per pixel, the set of flood events that inundated it, plus a combo_id -sorted GeoParquet dictionary and a small events table. Query by region and time via HTTP range reads (GDAL /vsicurl + DuckDB) to retrieve matching events, then fetch only the source depth rasters needed. Built with the open-source EuroFlood Python package ( pip install euroflood ).

open·CC-BY-4.0·Zenodo·completeSource
declared

Data from "Milder winters alleviate seasonal challenges for a migratory goose facing Arctic warming"

0.00

Geisler, Jan · Rakhimberdiev, Eldar · Boom, Michiel P. · et al.

29 files · 124 MB · csv, shapefiledeclared

1. Many migratory birds now reach their Arctic breeding grounds earlier in order to keep pace with advancing springs and shifting nutrient peaks, either by departing earlier from non-breeding grounds or by travelling faster. For dark-bellied brent geese, there is limited potential to travel faster, as their migration to the Siberian breeding grounds is already among the fastest of Arctic geese and swans. Earlier departure would require reaching departure body mass earlier, either through a faster accumulation of energy stores during spring staging or via adjustments earlier in the annual cycle. 2. We examined long-term shifts in spring staging phenology and changes in winter and spring body mass trajectories of brent geese at the population level, with particular emphasis on the effects of winter temperature on body mass and spring body mass on departure timing. 3. We used more than five decades of body mass measurements from individuals caught in the United Kingdom and France, and in the Dutch Wadden Sea to reconstruct changes in spring and winter mass trajectories, respectively. These data were combined with over five decades of migration counts in the Netherlands and more than two decades of counts in Denmark to quantify changes in spring staging phenology. 4. We found that brent geese have not shifted their spring arrival in the Wadden Sea but have advanced departure timing. Furthermore, brent geese were heavier during and after milder winters, and have changed mass trajectories over recent decades. They no longer lose mass during winter and the second spring staging phase, and fuelling rates in the first spring staging phase have declined. Annual variation in body mass was not related to annual departure timing. 5. These results suggest that milder winters have relaxed energetic constraints and improved body condition in brent geese throughout the non-breeding season. Our findings highlight the importance of considering the full annual cycle when assessing how animals with limited capacity to adjust migration timing or speed respond to global change.

open·CC-BY-4.0·Zenodo·completeSource
declared

Map of New Spain [Segment], c. 1800

0.00

Kavanagh, Jack · Anthony, Patrick

6 files · 11 MB · geojson, geopackage, shapefiledeclared

A historical map showing a segment of the boundaries of New Spain in c. 1800. This map is fully open to fellow researchers and is available in multiple open source formats (SHP, GPKG, GeoJSON).

open·CC-BY-4.0·Zenodo·completeSource
declared

A Parliamentary Discourse Dataset from the German Bundestag

0.00

Njie, Adama · Torkayesh, Ali E · Venghaus, Prof. Dr. Sandra

10 files · 1.6 GB · csv, gzip, parquetdeclared

Structured, speaker-attributed corpus of all German Bundestag plenary session transcripts ( Plenarprotokolle ) from the first legislative period to the present (WP01-WP21, September 1949 - April 2026). Every attributed speech is extracted from the official PDFs published by the Deutscher Bundestag under open data policy and linked to the speaker's name, parliamentary role, party affiliation, and gender. Scale: 4,611 sessions · 1,033,723 speeches · 4,205 identified MdBs · 76 years of parliamentary debate Dataset files speeches.parquet - one row per attributed speech: speaker name, role, party, gender, stammdaten_id, full German text (~1 GB) persons.parquet - one row per MdB: cross-session identity linking all name variants via stammdaten_id; canonical name, birth date, career span, total speeches. Use this - not speakers.parquet - for person-level analysis sessions.parquet - one row per plenary session: date, city, Wahlperiode, source PDF hash, extraction engine speakers.parquet - name-string index: one row per unique name string as extracted from the transcripts. Useful for understanding extraction quality; not suitable for person-level aggregation (the same politician often appears under several name variants across sessions) parties.csv - reference table of 31 German parliamentary parties, 1949-present speeches.csv.gz - CSV fallback for Stata and Excel users (same columns as speeches.parquet) datapackage.json - Frictionless Data schema with column descriptions and foreign key constraints Cross-session identity The same politician often appears under different name strings across sessions (e.g. "Schmidt", "Dr. Schmidt", "Frau Dr. Schmidt"). Cross-session person linkage is provided via stammdaten_id , matched against the official Bundestag Stammdaten biographical XML. The persons.parquet table aggregates all name variants for the same MdB into one row with correctly summed speech counts, career span, and birth date. Coverage: ~98.5% of speeches are linked to a stammdaten_id; the remaining ~1.5% are ambiguous surname-only attributions or speakers not in the Stammdaten. Coverage and sources Source PDFs are the official Stenografische Berichte downloaded from the Bundestag open-data portal (bundestag.de). Party-share normalisation in the corpus statistics uses official seat counts per Wahlperiode sourced from the Federal Returning Officer (Bundeswahlleiter, bundeswahlleiter.de). Two PDF generations are covered: scanned and OCR'd documents (WP01-WP09, Bonn era, 1949-1987) and born-digital documents (WP10-WP21, 1987-present). The engine column in sessions.parquet flags whether pdftotext (born-digital) or pdfminer (OCR fallback) was used; this is the primary data-quality indicator for NLP use. Speaker attribution Each speech is attributed using four patterns extracted from the transcript format: presiding officers (Präsident/in, Vizepräsident/in), regular members (name + party), government officials (name + Bundeskanzler/in, Bundesminister/in, etc.), and procedural roles (Berichterstatter/in, etc.). The party field is null for ~60% of speeches - this is expected, as presiding officers and ministers are not identified by party in the transcript. Gender annotation & distribution Gender is derived by matching speaker names against the official Bundestag Stammdaten biographical XML (all MdBs since 1949), with fallbacks for role title, honorific prefix, manually researched overrides, and a gender_guesser first-name heuristic. The gender_source column distinguishes stammdaten (authoritative, 83%), role_title (gendered job title in attribution, 6.4%), title_prefix (Frau/Herr honorific, 0.5%), manual (historically researched, 2.9%), and inferred (name-based heuristic, 4.7%). Gender distribution: Female 26.6% · Male 73.4% · Unknown 0.0%. Data quality All speeches pass automated validation: zero null speaker names, zero sequence gaps, zero CID artefacts, zero party-misclassified-as-Bundesland errors. Eight sessions with conflicting source PDFs were deduplicated (first lexicographic occurrence retained). 252 non-person names incorrectly accepted by the parser (table headers, legislative terms, agenda fragments) are excluded at build time via a curated exclusion list. OCR sessions (WP01-WP09) may contain Unicode replacement characters (U+FFFD); the engine field identifies these sessions. Licence CC BY 4.0. The underlying Plenarprotokolle are official government documents of the Deutscher Bundestag and are in the public domain.

open·CC-BY-4.0·Zenodo·completeSource
declared

Trained autoencoder and encoder for curvature-spectral analysis of compound meander bends

0.00

Lopez Dubon, Sergio · Sgarabotto, Alessandro · Lanzoni, Stefano

11 files · 38 MB · hdf5, npydeclared

This record contains the trained autoencoder, extracted encoder, and processed world/real-river latent-space reference cloud associated with the manuscript *A data-driven approach to discern the curvature spectral complexity of compound meander bends*. The full autoencoder is provided to support reconstruction-based validation and reproducibility of the learned representation. The extracted encoder is provided for inference and future software tools. It maps preprocessed 64 × 64 single-channel curvature-spectrum images to the two-dimensional latent space used to analyse meander shape complexity and skewness. The file `world_latent_cloud.npy` contains the two-dimensional latent coordinates of the world/real-river meander dataset used as the reference background cloud in the manuscript latent-space figures. This file is a processed latent-coordinate dataset only; it does not contain raw satellite imagery, raw centreline geometries, or training images. The release includes model weights, architecture files, model summaries, export metadata, the world/real-river latent cloud, example inference scripts, a validation script, environment files, and a minimal example input. The models should only be applied to curvature-spectrum images generated consistently with the preprocessing workflow described in the associated manuscript. Main files included in this release are: - trained_autoencoder.h5: full trained autoencoder. - encoder_only.h5: extracted encoder in HDF5/Keras format. - encoder_only.keras: extracted encoder in native Keras format. - model_architecture.json: full autoencoder architecture. - encoder_architecture.json: encoder architecture. - model_summary.tx and encoder_summary.txt: layer summaries. - world_latent_cloud.npy: world/real-river reference latent-space cloud. - world_latent_cloud_metadata.json: metadata for the world/real-river latent-space cloud. - model_card.md: intended use, inputs, outputs, limitations, and citation guidance.

open·CC-BY-4.0·Zenodo·completeSource
declared

IPv6-CyberBench: A Synthetic Translated-Flow Stress-Test Corpus for Diagnostic IDS Evaluation under IPv4-to-IPv6 Header Substitution

0.00

abdulwahab, samaa · aduallah, mahmood z. · Sallomi, Adheed H.

34 files · 2.7 GB · csv, gzip, parquetdeclared

Intrusion-detection research on Internet Protocol version 6 (IPv6) remains bottlenecked by the scarcity of labelled, protocol-aware flow datasets. Existing machine-learning IDS benchmarks are overwhelmingly IPv4-centric, and the few IPv6 corpora that have been released target narrow attack families or rely on small academic testbeds that cannot be re-created by third parties. We present IPv6-CyberBench, a reproducible eight-phase pipeline that constructs a large, protocol-aware translated-flow corpus by harmonising CIC-IDS-2017, CIC-IDS-2018 and CIC-DDoS-2019, applying deterministic IPv4→IPv6 address translation (6to4, NAT64, Teredo, EUI-64), synthesising 27 IPv6-specific flow features grouped in six protocol families, enforcing nineteen RFC-derived constraint categories together with temporal address dynamics, and rebalancing the long-tailed class distribution with a feature-group-conditioned per-class Wasserstein-GAN-GP augmenter and a SMOTE-KDE fallback selected per class by a formal decision rule. We scope the contribution honestly: because the seed corpora are IPv4 captures, the resulting 2,285,774-record benchmark is a translated-flow corpus suitable for training and evaluating flow-level IPv6 IDS classifiers on flooding, brute-force, scan, web-attack and infiltration traffic under IPv6 protocol-header semantics, and for studying IPv6-specific feature engineering and address dynamics in a reproducible setting. It is not a substitute for protocol-native IPv6 attack capture, and we explicitly exclude ICMPv6 Neighbour-Discovery flooding, SEND flooding, NDP exhaustion and extension-header covert-tunnelling from the threat model. The benchmark is evaluated on four axes - fidelity (Kolmogorov-Smirnov, MMD, Fréchet feature distance), utility (stratified 5×5 nested cross-validation over six classifier families including CNN-LSTM and LightGBM), privacy (Shokri-style membership-inference advantage AUC), and external fidelity against a 24 h anonymised CAIDA IPv6 trace (equinix-chicago, US backbone) and a MAWI samplepoint-F trace (WIDE backbone, Tokyo, Japan). The full pipeline, the hyper-parameter manifest, the RFC-constraint manifest, the reproduction scripts, and the 2,285,774-record benchmark are released unconditionally on Zenodo under CC BY 4.0; the dataset and pipeline are openly available at https://doi.org/10.5281/zenodo.19503446 (CC BY 4.0).

open·MIT·Zenodo·completeSource
declared

Map of Prussia, c. 1795

0.00

Kavanagh, Jack · Anthony, Patrick

6 files · 9.3 MB · geojson, geopackage, shapefiledeclared

A historical map of the boundaries of Prussia in c. 1795. This map is fully open to fellow researchers and is available in multiple open source formats (SHP, GPKG, GeoJSON).

open·CC-BY-4.0·Zenodo·completeSource
declared

Quantifying Cloud-Fog-Induced Reductions in Near-Surface Solar Radiation Using the Reduction Ratio Method

0.00

Ou, Te-Yu · Gu, Rong-Yu · Jang, Yi-Shin · et al.

26 files · 2.2 GB · csv, netcdf, npydeclared

Version history Version 2: This version updates Figure1.py and MainFigure.ipynb to match the revised manuscript submitted to Geophysical Research Letters. Specifically, a scale bar was added to Figure 1. No changes were made to the input datasets, analysis workflow, or scientific results. Code All results were analyzed and visualized by Python version 3.9.18. Filename Description MainFigures.ipynb Jupyter Notebook for reproduce all figures in the article. Figure*.py Python file for reproduce each figure in the article. Data Ground Station Observation Filename Description 467530_hr_19800101_20230101.csv Alishan station data managed by Central Weather Administration (CWA) of Taiwan, which is generated from Atmospheric Science Research and Application Databank . Solar Radiation Reduction Ratio (RR SR ) Results Filename Description ReductionRatio_20151001_20220930_31_daily.nc RR SR of Taiwan (119.9°E-122.1°E, 21.8°N-25.4°N, 0.01° x 0.01°) from 2015-10-01 to 2022-09-30, which is generated from TCCIP Grid-point Surface Insolation Derived from Geostationary Satellite Dataset . *_hr_20151001_20220930_seasonal.csv Seasonal mean RR SR for 6 CWA stations (Alishan, Anbu, Yushan, Sun Moon Lake, Chiayi, and Keelung). *_hr_20151001_20220930_DJF_NE_means.csv Winter (December-February) event-mean RR SR for 6 CWA stations (Alishan, Anbu, Yushan, Sun Moon Lake, Chiayi, and Keelung). Land and Non-rainy Mask Filename Description LandNonRainyMask_2015_2022.nc Mask for land area and non-rainy data, which is generated from TCCIP Gridded Historical Daily Dataset for Taiwan . Land-type Classification Generated from Schulz et al. (2017) . Filename Description MCF2017.zip Shapefile of montane cloud forests (MCFs) region in Taiwan. nonMCF2017.zip Shapefile of forest region other than montane cloud forest in Taiwan. MCFfraction_TCCIP.npy Forest classification in Taiwan regrid to TCCIP dataset. North-easterlies events classification Adopted from Taiwan Atmospheric Event Database . Filename Description TAD_NE.csv North-easterlies events classification based on the criteria of Taiwan Atmospheric Event Database (TAD). Others Generated from Open Data in Taiwan . Filename Description Taiwan_WGS84.zip Shapefile of coastlines of Taiwan. dem20_TCCIPInsolation.nc Digital Elevation Model of Taiwan regrid to TCCIP dataset. Note Please unzip .zip first to get shapfile before reproduce the figures in the article.

open·CC-BY-4.0·Zenodo·completeSource
declared

Fault trace mapping of the Dixie Valley Fault, central Nevada, USA

0.00

Francescone, Marco

14 files · 380 KB · shapefile, zipdeclared

This dataset contains original geomorphic mapping of surface fault traces along the Dixie Valley Fault (DVF) range front and piedmont zone, central Nevada, USA. Traces were mapped directly from a 1-m bare-earth lidar digital elevation model (DEM), using hillshade and slope-raster visualizations. This dataset accompanies the manuscript: Francescone, M., et al. (in review), LiDAR-Based Fault-Scarp Analysis and Rupture Hazard Assessment: Earthquake Scenarios of the Dixie Valley Fault System (Nevada, USA). See Section 3.1 ("Fault Trace Mapping") of the manuscript for full methodological details

open·CC-BY-4.0·Zenodo·completeSource
declared

Software and AMR peptide database for 'PEPTiGEN: a tool for mining antimicrobial resistance PEPTides using GENe data of public available repositories'

0.00

Meekes, Lisa · Tabaro, Francesco · Bexkens, Michiel · et al.

41 files · 8.2 GB · csv, fasta, pdfdeclared

This record contains the Python software for PEPTiGEN, a tool for generating tryptic peptides from prokaryotic gene sequences and their variants, and the associated antimicrobial resistance (AMR) peptide database. The database is provided as an SQL file and a CSV file containing all genes and predicted peptides. The README file contains explanation of the PEPTiGEN tool. The SQL database schema files contains both the database schema of the SQL database used in the PEPTiGEN analysis as the database schema of the AMR peptide datbase.

open·CC-BY-4.0·Zenodo·completeSource
declared

SDSS DR12 Stellar Spectra Dataset for Machine Learning Tasks

0.00

Barreto, Bruno · Eisencraft, Marcio

3 files · 2.6 GB · parquetdeclared

This dataset provides a fixed benchmark dataset for stellar atmospheric parameter estimation from Sloan Digital Sky Survey Data Release 12 (SDSS DR12) optical stellar spectra. The dataset is organized into three predefined Parquet splits: 30,000 spectra for training, 5,000 spectra for validation, and 15,000 spectra for testing. Each row corresponds to one SDSS stellar spectrum and includes raw spectral arrays, fixed-length processed spectral features, source identifiers, basic metadata, and catalog stellar-parameter labels with their associated uncertainties. The supervised regression targets are the adopted catalog stellar atmospheric parameters: effective temperature (Teff, in K), metallicity ([Fe/H], in dex), and surface gravity (log g, in dex). The dataset also includes relevant observational and catalog information such as SDSS plate, MJD, fiber identifier, sky coordinates, signal-to-noise ratio, adopted radial velocity, raw flux, logarithmic wavelength grid, inverse variance, pixel mask, and processed flux features. This release is intended to support machine-learning research on stellar spectroscopy, including regression models for atmospheric parameter estimation, benchmark comparisons, uncertainty-aware evaluation, and experiments using either processed fixed-length spectra or native observed-frame spectral arrays.

open·CC-BY-4.0·Zenodo·completeSource
declared

An Exploratory Inter-Model Variability Analysis for Social Vulnerability Assessment

0.00

Dey, Hemal · Shao, Wanyun

14 files · 5.5 MB · csv, jpeg, shapefiledeclared

Despite the proliferation of social vulnerability assessment methodologies, selecting the most appropriate model remains a critical challenge due to inter-model variability. To explore the inter-model variability, this study systematically investigated inter-algorithmic and inter-classification variability to assess how methodological design influences outcomes.

open·CC-BY-4.0·Zenodo·completeSource
declared

SISAB municipal primary care production and CID/CIAP attendance data, Brazil, 2026 (incomplete)

0.00

Saldanha, Raphael

8 files · 601 MB · parquet, zipdeclared

Description This deposit contains annual, municipality-level datasets derived from the Brazilian Primary Health Care Information System (SISAB). The files combine two complementary data sources: Public SISAB Saúde report downloads from the Atendimento/Visita production report. CID-10 and CIAP-2 attendance data obtained from SISAB through requests under the Brazilian Access to Information Law (Lei de Acesso à Informação, LAI). The datasets are organized as tidy annual files in CSV (Zipped) and Parquet format. They are intended to support reproducible analysis of primary care production, procedures, evaluated problems/conditions, and CID/CIAP-coded attendances across Brazilian municipalities. The public SISAB report datasets are stratified by competence month, state, municipality, DataSUS age group, SISAB sex category, and the selected report category. For each competence month and report type, the extraction combines 36 stratified SISAB downloads: 18 age groups by 2 sex values. Monthly files are merged into yearly files, completing missing combinations of observed competence, municipality, age group, sex, and category with valor = 0 . The LAI dataset contains yearly CID-10 and CIAP-2 attendance counts by competence month, municipality, code type, and code. When multiple valid LAI files cover the same competence, the processing pipeline selects the file with the largest number of data rows, using file size and request folder order as tie-breakers. Provenance columns identify the selected LAI request and source file. Variables SISAB Saúde Produção Columns: competencia : competence month in YYYYMM format. uf : Brazilian state abbreviation. ibge : municipality IBGE code. municipio : municipality name. faixa_etaria : Age group. sexo : SISAB sex category, Masculino or Feminino . tipo_producao : production type from the SISAB report. valor : count reported by SISAB. SISAB Saúde Procedimento Columns: competencia : competence month in YYYYMM format. uf : Brazilian state abbreviation. ibge : municipality IBGE code. municipio : municipality name. faixa_etaria : age group. sexo : SISAB sex category, Masculino or Feminino . procedimento : procedure from the SISAB report. valor : count reported by SISAB SISAB Saúde Condição Avaliada Columns: competencia : competence month in YYYYMM format. uf : Brazilian state abbreviation. ibge : municipality IBGE code. municipio : municipality name. faixa_etaria : age group. sexo : SISAB sex category, Masculino or Feminino . condicao_avaliada : evaluated problem or condition from the SISAB report. valor : count reported by SISAB. SISAB LAI CID/CIAP Columns: ano_competencia : competence year. competencia : competence month in YYYYMM format. competencia_date : first day of the competence month. co_municipio_ibge : municipality IBGE code. tp_codigo : code type, CID or CIAP . codigo : CID-10 or CIAP-2 code. qt_atendimentos : number of attendances. source_request : selected LAI request folder. source_file : selected source CSV file. Methods The public SISAB report files were generated with the sisab_scrapper processing pipeline. For each month, report type, age group, and sex value, the pipeline downloads the all-Brazil municipality report from SISAB, validates the returned CSV, preserves raw cache files for resumable runs, and writes a sorted monthly tidy dataset. The yearly merge validates required columns, expected age groups, expected sex values, category values, and month gaps unless explicitly allowed. The CID/CIAP files were generated with the sisab_lai processing pipeline. The pipeline imports CSV files received through LAI requests, detects the real CSV header after any SQL*Plus preamble, validates candidate files, resolves overlapping requests by competence, standardizes old and new schemas into one tidy table, and exports annual CSV and Parquet files together with audit reports. Sources - SISAB public reports, Ministry of Health, Brazil: https://sisab.saude.gov.br/ - SISAB LAI files obtained through Brazilian Access to Information Law requests. - Processing code for public SISAB report data: https://github.com/rfsaldanha/sisab_scrapper - Processing code for LAI CID/CIAP data: https://github.com/rfsaldanha/sisab_lai Notes - Counts are aggregated administrative records and should be interpreted in light of SISAB reporting practices, data quality, and changes in municipal reporting coverage. - Municipality boundaries, names, and coding practices may vary over time. - Public SISAB report datasets are completed with zero values only for combinations defined by observed municipalities, observed competencies, all expected age groups, both expected sex values, and observed report categories within the yearly merge. - LAI CID/CIAP data preserves selected source-file provenance through source_request and source_file . - This deposit corresponds to an individual year. Deposits for other years are published separately.

open·CC-BY-4.0·Zenodo·completeSource
declared

C10 Voltage Calibration PS 20260610

0.00

Flowerdew, Jake

6 files · 1.3 GB · csv, npydeclared

Voltage calibration data taken on 10/06/2026 of all ten 10 MHz PS cavities (exluding the spare C11). Measurements were taken, pulsing each cavity individually at harmonics h7 - h21 and voltages of 10 kV, 12.5 kV, 15 kV, 17.5 kV and 20 kV. Note, due to cavity trips, no useful data was collected for C10-46 and C10-86.

open·CC-BY-4.0·Zenodo·completeSource
declared

Annotation-Efficient Building Footprint Updating via Historical Map Reuse and Iterative Instance Segmentation

0.00

Cheng, Yuhan · Bai, Lubin · Zhang, Xiuyuan · et al.

15 files · 140 MB · shapefile, zipdeclared

Data used in the paper "Annotation-Efficient Building Footprint Updating via Historical Map Reuse and Iterative Instance Segmentation"

open·CC-BY-4.0·Zenodo·completeSource
declared

Climate warming promotes carbon sequestration and weathering in tundra landscapes and alters carbon chemistry of subarctic lakes in Scandinavia

0.00

Goedkoop, Willem · Fölster, Jens · Lau, Danny Chun Pong · et al.

16 files · 145 KB · shapefiledeclared

Main scripts for assessing satellite data using rgee. The loops can run slowly, so testing may require using smaller regions.

open·CC-BY-4.0·Zenodo·completeSource
page 1next →

Select a result to see its full details here: the measured structure, quality, and the loader, without leaving your search.