TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

Model checkpoints

This dataset consists of the Graphcast model checkpoints produced during the fine-tuning process of (Subich 2024).

1 papers0 benchmarks

Tea sickness - object detection

Please refer this paper

1 papers0 benchmarks

Individualized Deepfake Detection Dataset

The Deepfake face detection task involves a facial image of unknown authenticity for testing. While most deepfake detection methods take only the image as input, our literature demonstrates that conditioning the deepfake detector on identity—i.e., knowing whose deepfake face the picture might be—can enhance detection performance. Existing deepfake detection datasets, such as FaceForensics++ and DFDC, do not include identity information for authentic and deepfake faces. This dataset contains facial images of 45 specific individuals, divided into train and test sets, including a total of 23k authentic and 22k deepfake images. Having a specific individual's images in both the train and test sets allows us to assess detection performance for that individual. The dataset is curated so that the train and test sets are from two independent sources. The train images are curated from the CelebDFv2 dataset, and the test images are curated from the CACD dataset. Deepfake faces are generated using

1 papers0 benchmarksImages

Suicidial Annotation

A large dataset of around 40000 Reddit posts was collected from r/suicidewatch and other non-suicidal subreddits. The posts collected from r/suicidewatch are annotated with suicidal and other posts collected from a variety of groups like r/sports, r/anxiety, r/politics, and more are annotated with non-suicidal. Then this dataset has been used to feed various advanced deep-learning models to report a comparative evaluation of these models.

1 papers0 benchmarksTexts

DarkShake (DarkShake Benchmarking Dataset)

A pair Deblurring Benchmarking Dataset

1 papers0 benchmarks

Karrierewege

Karrierewege Dataset

1 papers0 benchmarks

Karrierewege+

Karrierewege+ Dataset

1 papers0 benchmarks

VPData

The largest video inpainting dataset comprises over 390K clips (> 866.7 hours), featuring precise masks and detailed video captions.

1 papers0 benchmarksRGB Video, Texts, Tracking, Videos

VPBench

The benchmark for VPData, the largest video inpainting dataset, which comprises over 390K clips (> 866.7 hours) and features precise masks and detailed video captions.

1 papers0 benchmarksRGB Video, Texts, Tracking, Videos

IMPACT Patent (A Large-scale Integrated Multimodal Patent Analysis and Creation Dataset for Design Patents)

It is a large-scale multimodal patent dataset with detailed captions for design patent figures.

1 papers1 benchmarksImages, Texts

Simulations

Simulations

1 papers0 benchmarks

Brainwave EEG Dataset

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

Emotional Status Determination using Physiological Parameters Data Set

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

WildIFEval

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

Boreal Forest Fire (Boreal Forest Fire: UAV-collected Wildfire Detection and Smoke Segmentation Dataset)

This dataset consists of annotated images and videos of smoke resulting from prescribed burning events in Finnish boreal forests. The dataset was created to train and validate learning-based methods for wildfire detection and smoke segmentation and its effectiveness in doing so was shown in the linked studies.

1 papers0 benchmarksImages

NOAA/WDS Guyamas Basin (NOAA/WDS Paleoclimatology - Barron et al. 2004 High Resolution Guaymas Basin Geochemical, Diatom, and Silicoflagellate Data)

This archived Paleoclimatology Study is available from the NOAA National Centers for Environmental Information (NCEI), under the World Data Service (WDS) for Paleoclimatology. The associated NCEI study type is Paleoceanography. The data include parameters of paleoceanography with a geographic location of Eastern Pacific Ocean. The time period coverage is from 15190 to 1330 in calendar years before present (BP).

1 papers0 benchmarks

MPM-Verse (MPMVerse Physics Simulation Dataset)

This dataset contains Material-Point-Method (MPM) simulations for various materials, including water, sand, plasticine, elasticity, jelly, rigid collisions, and melting. Each material is represented as point-clouds that evolve over time. The dataset is designed for learning and predicting MPM-based physical simulations.

1 papers0 benchmarks3D, Point cloud

NACHOS dataset: OCT and Xray

The following datasets:

1 papers0 benchmarks

AI Conversational Interviewing: Interview data

Replication Material This document contains the necessary materials and instructions to replicate the findings presented in our paper. We provide comprehensive information on the data sources, code, and analytical procedures used in our study. The replication package includes raw data files, data cleaning scripts, and analysis code. We encourage users to contact us with any questions or issues encountered during the replication process.

1 papers0 benchmarksTexts

SGA-INTERACT

A 3D skeleton-based group activity understanding dataset.

1 papers0 benchmarks
PreviousPage 545 of 1000Next