Binus University
130 rows × 54 cols · 54 KB
42 numeric · 10 categorical · 2 text
hybrid · semantic + lexical · 2062 datasets ranked · 0.95s
Binus University
130 rows × 54 cols · 54 KB
42 numeric · 10 categorical · 2 text
Zaghian, Soheil · Mohammadzadeh, Ali · Ahmadi, S. Ali · et al.
S12-Coasts is a global, multimodal, patch-based geospatial dataset designed for coastline and water-body segmentation tasks. The dataset consists of 6,637 georeferenced image patches , each with a spatial size of 512 × 512 pixels , paired with automatically generated binary water masks . Data modalities Sentinel-2 (optical): Multi-band optical imagery providing spectral information suitable for discriminating water, land, and coastal features under cloud-free or low-cloud conditions. Sentinel-1 (SAR): Synthetic Aperture Radar imagery enabling robust water detection under all-weather and day/night conditions and complementing optical observations in cloudy or turbid environments. Each patch contains co-registered Sentinel-1 and Sentinel-2 observations covering the same spatial extent, enabling multimodal learning and data fusion approaches. Reference labels Reference water masks were generated automatically using the Ensemble Object-Based Clustering (EOBC) framework. EOBC integrates multiple independent shoreline and water-body vector datasets, including: NOAA shoreline datasets (CUSP), OpenStreetMap (OSM) water and coastline features, HydroLAKES lake polygons. These vectors are fused with satellite imagery through object-based clustering and consensus rules to derive binary water / non-water labels. This fully automated approach removes the need for manual annotation while ensuring spatial consistency across regions.
Pettit, Erin · Sauret, Genevieve · Barnes, Caitlin · et al.
This dataset contains 100 MHz ground-based ice-penetrating radar data collected over the Dotson Ice Shelf during January 2020, stored as NetCDF files. Data was collected by the Thwaites-Amundsen Regional Survey and Network Integrating Atmosphere-Ice-Ocean Processes (TARSAN) team of the International Thwaites Glacier Collaboration (ITGC). Files contain both raw and processed data. Processed data includes surface elevation, theoretical hydrostatic disequilibrium, and both ice-ocean interface and relatively bright englacial layer picks. All elevation values are recorded in meters. All geographic information is referenced to the polar stereographic projection (EPSG:3031). The dataset also includes PDF files showing each profile and its location on the shelf, as well as a CSV file with information from the external GPS units that accompanied the radar units in the field.
Khedr, Walid I.
This artifact contains the finalized synthetic benchmark corpus, derived decision windows, validation manifests, result artifacts, curated replay code, Colab/GPU result archives, Raspberry Pi 5 profiling outputs, human quality-audit materials, deployment-boundary diagnostics, and manuscript source/PDF for the accompanying paper on edge-oriented behavioral authentication in multi-agent LLM systems. Version 1.0.4 refreshes the reproducibility package with venue-neutral paper, code, experiment, and artifact paths. The reviewer-facing archive now stages edge-efficiency materials under artifacts/edge_efficiency/ , scripts under code/edge_efficiency/ , and RP5 power protocols under experiments/rp5_power/ . It also includes paper-local artifact mirrors so the manuscript can be rebuilt from the extracted package. The package is designed for reproducible evaluation replay rather than bit-identical regeneration of the synthetic conversations. The archived corpus is the citable benchmark object. Scenario planning and turn generation used GPT-4.1, whose outputs may change over time; generation scripts are included as provenance and extension tools.
CORUH, Uğur
Benchmark pack accompanying the journal submission 'A template-preserving AST-based LaTeX-to-DOCX compiler with native Office Math ML equation output' (Coruh, 2026). Every numerical figure in the manuscript is either reproducible from the deposit (single-tool corpus) or auditable from the canonical CSV / JSON snapshots it carries (four-pipeline matrix and the structural-presence regression check). Contents. A single 188 MB zip with 2907 entries: the production binary (latex2docx 1.0.4 Windows installer), the 40-class publisher template matrix, the 39 feature-coverage mockups, the 203 graded body documents, the per-stem pdflatex reference-PDF cache (~45 MB), benchmark runner scripts, metric modules and the canonical CSV / JSON snapshots from the run reported in the manuscript. A step-by-step reproduction guide ( BENCHMARK_GUIDE.md ) is included at the root of the unpacked archive. Reproducibility scope. The single-tool corpus (R_surv, M_edit, P_fix, T_build, V_sim) reproduces end-to-end from the bundled binary against the bundled corpus. The four-pipeline benchmark matrix additionally requires Pandoc 3.5, make4ht and plastex on the local PATH; the corresponding CSV / JSON snapshots are bundled so reviewers can audit every figure without re-running the matrix. Access. This record is published under Restricted access . The DOI and landing page are publicly resolvable; the files require a Zenodo secret access link issued by the corresponding author. During peer review the link is supplied to the journal's editorial office under separate cover for distribution to assigned reviewers; after publication, the corresponding author continues to issue secret links to qualified researchers on reasonable e-mail request ( ugur.coruh@erdogan.edu.tr ). See LICENSE-NOTES.md for the full access conditions. Licensing (subject to access conditions above). Depositor-authored documentation, scripts and measurement CSV / JSON snapshots: CC-BY-4.0 with attribution. Forty publisher LaTeX classes: upstream LPPL . Bundled bin/*.exe binaries: proprietary, commercially licensed under a 30-day evaluation licence anchored to the binary's build date (commercial licence and current evaluation builds via https://coruhtech.github.io/latex2docx-toolkit/ ). Re-distribution of the deposit, or of any component in isolation, requires the corresponding author's written authorisation. Authoritative per-component disclosure is in LICENSE-NOTES.md . Integrity. SHA-256 manifest in SHA256SUMS.txt at the deposit root. The zip itself hashes to 33651f0071c6ee1866e9ae10de317ce2349c228e7ff8e11df6218d965e8804a3 .
lin, chen · Bin, Yang · Haoran, Shi · et al.
This dataset contains quality-controlled SNP and InDel genotype calls from whole-genome resequencing of 299 purebred Large White pigs. The dataset includes 100 individuals sequenced at 30× depth (merged genome-wide VCFs) and 199 individuals sequenced at 10× depth (chromosome-split VCFs). All variants were called against the *Sus scrofa* Sscrofa11.1 reference genome and filtered using standard GATK hard-filtering criteria: QD < 2.0, QUAL < 30.0, FS > 200.0, ReadPosRankSum < -20.0, call rate < 90%, MAF < 0.05. This data supports the findings of the study "Whole-genome resequencing identifies RASAL2 as a candidate gene for feed efficiency in Large White pigs".
Gómez Pedraza, Mauricio Alexander · Meneses-Báez, Alba Lucía · Serrano García, Maria Fernanda · et al.
Esta base de datos corresponde al manuscrito ' Validación de la Escala Multidimensional de Perfeccionismo Infantil (EMPI) en estudiantes colombianos y mexicanos' y recopila datos de 433 estudiantes de educación básica de ambos países.
Aniceto, Fabricio David Simplicio
This dataset provides multi-temporal UAV-derived Digital Elevation Models (DEMs) and orthomosaics of an erosion feature located at the entrance of the IFPE Campus Cabo de Santo Agostinho, Pernambuco, Brazil. The dataset is based on aerial imagery acquired with a DJI Air 2S UAV during four survey campaigns carried out in September 2023, July 2024, August 2025, and February 2026. The images were processed in Agisoft Metashape Professional Edition following standard photogrammetric procedures to produce georeferenced elevation models and orthomosaics for erosion monitoring and terrain analysis.
SERKAN, CANTÜRK
eplication package containing the panel dataset (18 emerging markets, 2003-2024), R replication scripts, output tables and figures, online appendix, and full documentation. Companion to the manuscript submitted to Environmental Economics and Policy Studies.
Post-Secondary Employment Outcomes (PSEO) are experimental tabulations developed by the Longitudinal Employer-Household Dynamics (LEHD) program at the U.S. Census Bureau. This record includes data for the Illinois Board of Higher Education (IBHE). PSEO data provide earnings and employment outcomes for college and university graduates by degree level, degree major, and post-secondary institution. These statistics are generated by matching university transcript data with a national database of jobs, using state-of-the-art confidentiality protection mechanisms to protect the underlying data. Releases are annual. The PSEO are made possible through data sharing partnerships between universities, university systems, State Departments of Education, State Labor Market Information offices, and the U.S. Census Bureau. PSEO data are available for post-secondary institutions whose transcript data have been made available to the Census Bureau through a data-sharing agreement. The data contains a degree file, enrollment file, and demographic file. Requests for PSEO data should include details on how the project will help to develop, validate, or administer predictive tests; administer student aid programs; or improve instruction.
This is a copy of National Archives public use data that has been PIKed for use in the FSRDC. The military data contain information on 8.3 million soldiers who served in the Army and the Army Air Force during World War II.
The restricted-use file contains data collected as part of the Impact Evaluation of a School-Based Violence Prevention Program. Data collection included student and teacher surveys, interviews with school administrators and teachers, and on-site observations.
The data contains information on impact investment funds and the portfolio companies they have backed. Use of the data must be approved by PCRI and Census. The approved projects must be of mutual interest to PCRI and Census.
The National Crime Victimization Survey (NCVS) is an annual survey that collects data on crime against persons age 12 or older from a nationally representative, stratified, multistage cluster sample of U.S. households. The Police-Public Contact Survey (PPCS)is a supplement to the NCVS, administered to NCVS respondents age 16 or older. The Police-Public Contact Survey (PPCS) provides detailed information on the characteristics of persons who had some type of contact with police during the year, including those who contacted the police to report a crime or were pulled over in aa traffic stop. The PPCS interviews a nationally representative sample of residents age 16 or older as a supplement to the National Crime Victimization Survey. The survey enables BJS to examine the perceptions of police behavior and response during these encounters.
This file contains Protected Identification Keys (PIK) that link Social Security Administration (SSA) Supplemental Security Record (SSR) to the respondents in the American Community Survey (ACS) microdata files. The Supplemental Security Record (SSR) is SSA's main file to track who is receiving Supplemental Security Income (SSI) benefits and the monthly benefit amounts payable.
Contains Protected Identification Keys (PIK) that link SSA Payment History Update System (PHUS) records to the respondents in the American Community Survey (ACS) microdata files. The Payment History Update System (PHUS) contains actual payments delivered to Old-Age, Survivors and Disability Insurance (OASDI) beneficiaries.
Compilation of race and Hispanic origin responses from several federal and third party administrative records sources. Business rules are applied to assign race and ethnicity responses when responses are discrepant across sources. NOTE: This file was created without the use of Indian Health Services(IHS) records which are restricted. For the main Best Race file, see dataset RDG 9072.
The SSA's 831 Disability File is a research file created to determine medical eligibility for Old-Age, Survivors and Disability Insurance (OASDI) and Supplemental Security Income (SSI) benefits. The 831 Disability File extract file contains Protected Identification Keys (PIK) that link these records to the respondents in the American Community Survey (ACS)microdata files.
The restricted-use codebook contains the count of responses for each data item and all components of SASS in 2003-2004 and the 2004-2005 TFS. The TFS data and User's manual are the added features to this re-release of the 2003-2004 SASS restricted-use ECB.
Supplementary public data files are available for researchers to easily access on FSRDC projects. The supplementary data combines several commonly used files into a single dataset, organized by subject matter. These files contain international and trade data from US and global sources. For more details on the specific files included, see Data Collection Notes.