TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

UCLA multimodal connectivity database

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

M3AV (Multimodal, Multigenre, and Multipurpose Audio-Visual)

The M3AV (Multimodal, Multigenre, and Multipurpose Audio-Visual) is a novel dataset proposed for academic lectures¹. It contains almost 367 hours of videos from five sources covering topics in computer science, mathematics, and medical and biology¹⁴.

1 papers0 benchmarks

COCO-N Medium

COCO-N Medium introduces a stochastic benchmark that simulates common real-world scenarios with noticeable label inaccuracies in the COCO dataset. This benchmark combines class and spatial noises to create a challenging yet realistic evaluation framework for instance segmentation models. It mimics datasets manually annotated by crowd workers, where a moderate level of label noise is expected. By incorporating both class and spatial inaccuracies, COCO-N Medium allows researchers to assess their models' basic robustness to label noise, providing insights into performance in typical real-world applications where perfect annotations are rare. This medium-level benchmark serves as a crucial middle ground, offering a more rigorous test than minimally noisy datasets while remaining within the bounds of commonly encountered data quality issues. COCO-N Medium enables a nuanced evaluation of model performance under realistic conditions, helping identify areas for improvement in handling noisy la

1 papers1 benchmarks

MAPLE

The MAPLE benchmark constructed by us contains 20 datasets across 19 fields for scientific literature tagging. It also has a graph format, which can be used for graph mining tasks (e.g., node classification, link prediction). Refer to its homepage for more details.

1 papers0 benchmarksGraphs

PPFT (Path Planning on Fifty x Thirty)

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

Tecnalia WEEE HYPERSPECTRAL DATASET (TECNALIA WEEE (Waste from Electrical and Electronic Equipment) HYPERSPECTRAL DATASET)

Tecnalia Hyperspectral Dataset contains different non-ferreous fractions of Waste from Electric and Electronic Equipment (WEEE) of Copper, Brass, Aluminum, Stainless Steel and White Copper. Images were captured by a hyperspectral Specim PHF Fast10 camera that is able to capture wavelengths in the range 400 to 1000 nm with a spectral resolution of less than 1 nm. The PHF Fast10 camera is equipped with a CMOS sensor (1024 × 1024 resolution), a Camera Link interface and a special Fore objective OL10. The provided dataset contains 76 uniformly distributed wave-lengths in the spectral range [415.05 nm, 1008.10 nm]. Illumination setup, as described in \cite{picon2012real}, was specifically designed to reduce the specular reflections generated by the surface of the non-ferrous materials and to provide a homogeneous and even illumination that covers the wavelengths sensitive to the hyperspectral camera. The illumination system consists of a parabolic surface that uniformly distributes the lig

1 papers0 benchmarksHyperspectral images, Images

BigEarthNet v2.0

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarksImages

Gen-Dynamics

Learning dynamical systems that can generalize to various parameter changes in their underlying ODEs or PDEs is a significant but challenging task. In recent years, several methods have been proposed to learn such systems. This repository contains a collection of datasets that can be used to benchmark these methods. We aim to provide a consistent interface to access each dataset.

1 papers0 benchmarks

VSTaR-1M

VSTaR-1M is a 1M instruction tuning dataset, created using Video-STaR, with the source datasets: * Kinetics700 * STAR-benchmark * FineDiving

1 papers0 benchmarksTexts, Videos

MuseChat Dataset (MuseChat: A Conversational Music Recommendation System for Videos (CVPR 2024 Highlight Paper))

Music recommendation for videos attracts growing interest in multi-modal research. However, existing systems focus primarily on content compatibility, often ignoring the users’ preferences. Their inability to interact with users for further refinements or to provide explanations leads to a less satisfying experience. We address these issues with MuseChat, a first-of-its-kind dialogue-based recommendation system that personalizes music suggestions for videos. Our system consists of two key functionalities with associated modules: recommendation and reasoning. The recommendation module takes a video along with optional information including previous suggested music and user’s preference as inputs and retrieves an appropriate music matching the context. The reasoning module, equipped with the power of Large Language Model (Vicuna-7B) and extended to multi-modal inputs, is able to provide reasonable explanation for the recommended music. To evaluate the effectiveness of MuseChat, we build

1 papers0 benchmarksAudio, Texts, Videos

Co-RaL-Dataset

This repository includes the experimental dataset acquired to evaluate our radar-leg odometry algorithm, Co-RaL, accepted by IEEE IROS 2024. Our dataset includes sensor data of chip radar, imu, velodyne, and kinematic data from Boston Dynamics SPOT (Joint encoders and contact sensors). Each sequence is acquired with different environments to evaluate the algorithm performance generally. The dataset is provided with ROS Bag file format.

1 papers0 benchmarks

PQAref (Pubmed Question Answering with references)

The PQAref dataset is a dataset for fine-tuning large language models for referenced question-answering in biomedical domain.

1 papers0 benchmarksTexts

FrodoBots 2K Dataset

The FrodoBots 2K Dataset is a diverse collection of camera footage, GPS, IMU, audio recordings & human control data collected from ~2,000 hours of tele-operated sidewalk robots driving in 10+ cities.

1 papers0 benchmarks

noisy-ADE20K-DS

A new SIS benchmark designed to assess generation performance under noisy conditions, simulating human error that can occur during real-world applications. [DS] employs downsampled semantic maps that are resized by nearest-neighbor interpolation, simulating human errors e.g., jagged edges and coarse/low resolution user inputs.

1 papers4 benchmarks

noisy-ADE20K-Edge

A new SIS benchmark designed to assess generation performance under noisy conditions, simulating human error that can occur during real-world applications. [Edge] masks the edges of instances with an unlabeled class, imitating incomplete annotations around edges, especially between instances.

1 papers4 benchmarks

noisy-ADE20K-Random

A new SIS benchmark designed to assess generation performance under noisy conditions, simulating human error that can occur during real-world applications. [Random] randomly adds an unlabeled class to the semantic maps, mimicing unintended user error and extreme random noise.

1 papers4 benchmarks

MAVE - Attribute: Black Tea Variety (MAVE - Attribute: Black Tea Variety: A Product Dataset for Multi-source Attribute Value Extraction)

The dataset contains 3 million attribute-value annotations across 1257 unique categories created from 2.2 million cleaned Amazon product profiles. It is a large, multi-sourced, diverse dataset for product attribute extraction study.

1 papers0 benchmarksTexts

MIR-ST500

MIR-ST500 Good for the following task: Singing transcription (singing pitch to music note conversion) Used in several papers published by Roger Jang's Lab

1 papers0 benchmarks

YourMT3 Dataset

We redistribute a suite of datasets as part of the YourMT3 project. The license for redistribution is attached.

1 papers0 benchmarksAudio, Midi

SRI-APPROVE Fine-Grained Video Classification

APPROVE consists of curated YouTube videos annotated with educational content. APPROVE consists of 193 hours of expert-annotated videos with 19 classes (7 literacy codes, 11 math, and background) and each video is associated with approximately 3 labels on average.

1 papers2 benchmarksVideos
PreviousPage 508 of 1000Next