TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

CUBE B-format Ambisonic RIR dataset

This dataset includes 720 directional B-format RIRs, i.e. first-order Ambisonic room impulse responses, measured at 30 receiver positions with 1m spacing in an equidistant grid (4xm) with 24 hemispherical source positions each. The measurements were carried out at the IEM CUBE using Soundfield ST450 MKII microphones. The data is saved according to the SOFA convention (https://www.sofaconventions.org/mediawiki/index.php). The SOFA Matlab/Octave API is available at https://github.com/sofacoustics/API_MO.5. Unfortunately, the used SOFA convention (MultiPerspectiveAmbisonicRIR) was never integrated into the official SOFA conventions. However, it can be found at: https://github.com/jdemuynke/API_MO/tree/master/API_MO/conventions.

1 papers0 benchmarks

Variable-Perspective ARIR Rendering - Listening Experiment Stimuli

Presented stimuli of the listening experiment performed in the course of the master's thesis: K. Müller, "Variable-perspective rendering of virtual acoustic environments based on distributed first-order room impulse responses" and the article: K. Müller and F. Zotter, “Auralization based on multi-perspective ambisonic room impulse responses” (DOI: 10.1051/aacus/2020024).

1 papers0 benchmarks

SUT (SUT: a new multi-purpose synthetic dataset for Farsi document image analysis)

This paper introduces a new large-scale dataset for Farsi document images, named SUT, which aims to tackle the challenges associated with obtaining diverse and substantial ground-truth data for supervised models in document image analysis (DIA) tasks, like document image classification, text detection and recognition, and information retrieval. The dataset comprises 62,453 images that have been categorized into 21 distinct classes, including identity documents featuring synthetically generated personal information superimposed on various backgrounds. The dataset also includes corresponding files with labeling information for the images. The ground-truth data is organized in CSV files containing image file paths and associated information about the embedded data.

1 papers3 benchmarks

INI-30

A representative event-based eye-tracking dataset, collected with two event cameras mounted on a glass frame. It features variable recording lengths and event counts from 30 volunteers, providing an ideal benchmark for modeling the heterogeneity of event-based eye tracking in real-world scenarios.

1 papers6 benchmarks

VigSet

A pioneering dataset for vignette removal. Vigset includes 983 pairs of both vignetting and vignetting-free high-resolution (5340×3697) real-world images under various conditions.

1 papers0 benchmarksImages

YFCC-CelebA

The scales of the data accessible through internet search engines can reach hundreds of millions, or even billions. The existence of such large weak-labeled databases has gained importance in the training of face recognition algorithms. Starting with the publicly available YFCC100M, we propose a weakly-labeled subset for multi-label face recognition for self-supervised methods. A 392K image subset of YFCC100M of 128x128 images was obtained by querying for the 40 facial attributes. We made this dataset publicly available.

1 papers0 benchmarksImages

CIRO experimental results

Description This repository includes the experiment results, source code, and test data for Three Cs risk inference, using the CIRO (COVID-19 Infection Risk Ontology) and HermiT.

1 papers0 benchmarksGraphs

Waterloo IVC 3D Image Quality Database

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

MapReader Data (in GeoHumanities workshop, SIGSPATIAL 2022)

MapReader in GeoHumanities workshop (SIGSPATIAL 2022): Gold standards and outputs

1 papers0 benchmarksImages, Texts

RadioGalaxyNET Dataset

Automating the creation of catalogues for radio galaxies in next-generation deep surveys necessitates the identification of components within extended sources and their respective infrared hosts. We present RadioGalaxyNET, a multimodal dataset, tailored for machine learning tasks to streamline the automated detection and localization of multi-component extended radio galaxies and their associated infrared hosts. The dataset encompasses 4,155 instances of galaxies across 2,800 images, incorporating both radio and infrared channels. Each instance furnishes details about the extended radio galaxy class, a bounding box covering all components, a pixel-level segmentation mask, and the keypoint position of the corresponding infrared host galaxy. RadioGalaxyNET is the first dataset to include images from the highly sensitive Australian Square Kilometre Array Pathfinder (ASKAP) radio telescope, corresponding infrared images, and instance-level annotations for galaxy detection.

1 papers1 benchmarksImages

SharePriceIncrease

The problem here is to predict whether a share price will show an exceptional rise after quarterly announcement of the Earning Per Share based on the price movement of that share price on the proceeding 60 days? The data was formatted by Vlad Pazenuks as part of his third year project. Daily price data on NASDAQ 100 companies was extracted from a Kaggle data set. Reporting dates of these companies were obtained from NASDAQ.com. Each data is the percentage change of the close price from the day before. Each case is a series of 60 day data. The target class is is defined as 0 = price did not increase after company report release by more than 5 percent 1 = price increased after company report release by more than 5 percent There are 1931 cases, 1326 class 0 and 605 class 1.

1 papers0 benchmarks

Urdu News Headlines Dataset

Urdu News Headlines Dataset with VOA and BBC An Urdu news headlines dataset is a collection of news headlines in the Urdu language, typically scraped from news websites and social media platforms. These datasets can be valuable for researchers and developers working on a variety of tasks, such as:

1 papers1 benchmarksTexts

Santa Clara Reservoir Levels

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarksTime series

Synthetic Soccer NeRF Dataset

Synthetic dataset comprising three different environments for multi-camera dynamic novel view synthesis for soccer. This dataset is made compatible for Nerfstudio, and includes data parsers with various settings to reproduce the settings of our paper "Dynamic NeRFs for Soccer Scenes" and more.

1 papers0 benchmarks3D, Images, Videos

GuardRails Dataset (GuardRails Dataset of Problems with Known Ambiguities)

For each problem, we provide 4 variants of prompts:

1 papers0 benchmarksTexts

Ray-tracing data in Herald Square

Data obtained by ray-tracing simulation in Herald Square. Data includes all the channel parameter information such as pathloss, delay, and angles.

1 papers0 benchmarks

BOTH57M

BOTH57M is a body-hand dataset with body-level text prompts and finger-level text prompts.

1 papers0 benchmarks

CORE-MM

CORE-MM is an Open-ended VQA benchmark dataset specifically designed for MLLMs, with a focus on complex reasoning tasks. CORE-MM benchmark consists of 279 manually curated reasoning questions, associated with a total of 342 images. The questions are divided into 3 reasoning categories--Deductive, Abductive and Analogical. 49 questions pertain to abductive reasoning, 181 require deductive reasoning, and 49 involve analogicalreasoning. Furthermore, the dataset is divided into two folds based on reasoning complexity, with 108 classified as “High” reasoning complexity and 171 as “Moderate” reasoning complexity.

1 papers5 benchmarks

Emotion Cause in GitHub

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

topex-printer

topex-printer is a dataset containing 102 machine parts of a label printing machine. It includes these parts for two domains, real photos and CAD rendered models.

1 papers0 benchmarksImages
PreviousPage 482 of 1000Next