TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

TPIC17 (Temporal Popularity Image Collection)

Image dataset with about 600K Flickr photos.

1 papers0 benchmarks

Traditional Chinese Landscape Painting Dataset

This dataset consists of 2,192 high-quality traditional Chinese landscape paintings (中国山水画). All paintings are sized 512x512, from the following sources: * Princeton University Art Museum, 362 paintings * Harvard University Art Museum, 101 paintings * Metropolitan Museum of Art, 428 paintings * Smithsonian's Freer Gallery of Art, 1,301 paintings

1 papers0 benchmarks

TREK-100

The dataset is composed of 100 video sequences densely annotated with 60K bounding boxes, 17 sequence attributes, 13 action verb attributes and 29 target object attributes.

1 papers0 benchmarksVideos

TrMor2018

A new high accuracy Turkish morphology dataset.

1 papers0 benchmarks

Twitter Cyberthreat Detection Dataset

Twitter Cyberthreat Detection Dataset is a dataset that contains tweets from two sets of accounts related to cybersecurity. The tweets are annotated with different information such as whether they contain security-related information and named entities.

1 papers0 benchmarksTexts

Twitter Flood

This dataset contains two subsets of flood images from Twitter: The Harz17 dataset comprises images from tweets containing flood-related keywords during the occurrence of a flood in the Harz region in Germany in July 2017. Similarly, the Rhine18 dataset comprises images related to a flood of the river Rhine in January 2018.

1 papers0 benchmarksImages, Texts

TWT-16

The TWT16 dataset contains ~30k conversations in Twitter, collected from January to June 2016.

1 papers0 benchmarksTexts

UG^2

Contains three difficult real-world scenarios: uncontrolled videos taken by UAVs and manned gliders, as well as controlled videos taken on the ground. Over 160,000 annotated frames forhundreds of ImageNet classes are available, which are used for baseline experiments that assess the impact of known and unknown image artifacts and other conditions on common deep learning-based object classification approaches.

1 papers0 benchmarksVideos

UIT-ViNames

This dataset comprises over 26,000 full names annotated with genders.

1 papers0 benchmarksTexts

UMC005 English-Urdu

UMC005 English-Urdu is a parallel corpus of texts in English and Urdu language with sentence alignments. The corpus can be used for experiments with statistical machine translation.

1 papers0 benchmarks

MultiUN (Multilingual Corpus from United Nation Documents)

The MultiUN parallel corpus is extracted from the United Nations Website , and then cleaned and converted to XML at Language Technology Lab in DFKI GmbH (LT-DFKI), Germany. The documents were published by UN from 2000 to 2009.

1 papers0 benchmarks

Urban Dict spelling variant

Urban Dict spelling variant is a variant spelling dataset for use of NLP research in the informal domain. It consists of around 25k variant spelling pairs form UrbanDictionary.

1 papers0 benchmarksTexts

UW IOM (University of Washington Indoor Object Manipulation)

Comprises twenty individuals picking up and placing objects of varying weights to and from cabinet and table locations at various heights.

1 papers0 benchmarksImages

VDQG (Visual Discriminative Question Generation)

The Visual Discriminative Question Generation (VDQG) dataset contains 11202 ambiguous image pairs collected from Visual Genome. Each image pair is annotated with 4.6 discriminative questions and 5.9 non-discriminative questions on average.

1 papers0 benchmarksImages, Texts

VG-Depth

Enable visual relation detection and serves as an extension to Visual Genome (VG).

1 papers0 benchmarks

VIA

The VIA dataset is a dataset for aiding the visually impaired. The proposed datase1 consists of 342 images divided into two classes: 175 of them are “clear-path” and 167 are “nonclear” path. They were taken using a smartphone camera and resized to 750 × 1000 pixels. The smartphone was placed in the user chest height and inclined approximately 30 to 60 from the ground, so it could capture a few meters of the path ahead, and beyond the reach of a regular white cane

1 papers0 benchmarksImages

VidSet

A large video dataset with dynamic content.

1 papers0 benchmarks

VLEngagement

A novel dataset that consists of content-based and video-specific features extracted from publicly available scientific video lectures and several metrics related to user engagement.

1 papers0 benchmarks

VQA-OV (Visual Quality Assessment of Omnidirectional Video)

Collects 60 reference sequences and 540 impaired sequences.

1 papers0 benchmarks

WIDER Attribute Dataset

The WIDER Attribute dataset is a human attribute recognition dataset with human attribute and image event annotations. Images are selected from the WIDER dataset. There are a total of 13,789 images. A bounding box is annotated for each person in these images, with no more than 20 people (with top resolutions) in a crowd image, resulting in 57,524 boxes in total and 4+ boxes per image on average. For each bounding box, 14 distinct human attributes are labelled. There are 805,336 labels in total.

1 papers0 benchmarksImages
PreviousPage 378 of 1000Next