TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

TextAtlas5M

We introduce TextAtlas5M, a dataset specifically designed for training and evaluating multimodal generation models on dense-text image generation.

1 papers0 benchmarksImages, Texts

MRI Brain Perturbations for Back-Projection Diffusion

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

PhysiCo

Physical concept understanding benchmark.

1 papers0 benchmarksImages, Texts

ShiftySpeech

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

1 papers0 benchmarksAudio

VideoDB's OCR Benchmark Public Collection

Dataset Introduction This dataset leverages VideoDB's Public Collection to offer a diverse range of videos featuring text-containing scenes. It spans multiple categories—ranging from finance and legal documents to software UI elements and handwritten notes—ensuring a broad representation of real-world text appearances. Each video is annotated with frame indexes to facilitate consistent and reproducible OCR benchmarks. Currently, the dataset includes over 25 curated videos, yielding thousands of extracted frames that present a variety of text-related challenges.

1 papers3 benchmarksImages, Texts, Videos

InCrowd-VI

About A realistic visual-inertial dataset with 58 sequences spanning 5km of trajectories and 1.5 hours of recordings, designed for evaluating SLAM systems in indoor pedestrian-rich environments. The dataset is particularly aimed at advancing navigation technologies for visually impaired individuals.

1 papers0 benchmarks

TSFM-ScalingLaws-Dataset

TSFM-ScalingLaws-Dataset

1 papers0 benchmarksTime series

FER2013 Blendshapes (FER2013 blendshapes dataset example (Partial))

Tables of the blendshapes from a group of the images of the FER2013 dataset, generated using MediaPipe library, based on the ARKit face blendshapes. with classes of the images in a separate column, describing the categories Happy, Unknown, Sad.

1 papers0 benchmarks3d meshes, Images, Tabular, Tracking

StereoMIS

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

How I Fixed Sleep-Wake Hangs on Linux with AMD GPUs: A Step-by-Step Guide

If you're an AMD GPU user on Linux, you might have encountered the dreaded sleep-wake hang issue, where your system fails to properly wake up from sleep or suspend, causing a black screen or the system freezing. Well, I ran into this problem myself and after a bit of research and experimentation, I finally managed to fix it. Here's how I did it!

1 papers0 benchmarks

SHRED-ROM (Reduced order modeling with shallow recurrent decoder networks)

SHallow REcurrent Decoder-based Reduced Order Model (SHRED-ROM) is an ultra-hyperreduced order modeling framework aiming at reconstructing high-dimensional data from limited sensor measurements in multiple scenarios. Thanks to the composition of a Long-Short Term Memory network (LSTM) and a Shallow Decoder Network (SDN), SHRED-ROM is capable of

1 papers0 benchmarks

MAVEN-Arg (MAVEN-Arguments)

MAVEN-Arg is an advanced event argument extraction dataset, which offers three main advantages:

1 papers0 benchmarks

M3LS (Multi-Lingual Multi-Modal Summarization Dataset)

Significant developments in techniques such as encoder-decoder models have enabled us to represent information comprising multiple modalities. This information can further enhance many downstream tasks in the field of information retrieval and natural language processing; however, improvements in multi-modal techniques and their performance evaluation require large-scale multi-modal data which offers sufficient diversity. Multi-lingual modeling for a variety of tasks like multi-modal summarization, text generation, and translation leverages information derived from high-quality multi-lingual annotated data. In this work, we present the current largest multi-lingual multi-modal summarization dataset (M3LS), and it consists of over a million instances of document-image pairs along with a professionally annotated multi-modal summary for each pair. It is derived from news articles published by British Broadcasting Corporation(BBC) over a decade and spans 20 languages, targeting diversity a

1 papers0 benchmarksImages, Texts

MAKED (MultiModal MultiLingual Summarization and Keyword Extraction Dataset)

Keyword extraction is an integral task for many downstream problems like clustering, recommendation, search and classification. Development and evaluation of keyword extraction techniques require an exhaustive dataset; however, currently, the community lacks large-scale multi-lingual datasets. In this paper, we present MAKED, a large-scale multi-lingual keyword extraction dataset comprising of 540K+ news articles from British Broadcasting Corporation News (BBC News) spanning 20 languages. It is the first keyword extraction dataset for 11 of these 20 languages. The quality of the dataset is examined by experimentation with several baselines. We believe that the proposed dataset will help advance the field of automatic keyword extraction given its size, diversity in terms of languages used, topics covered and time periods as well as its focus on under-studied languages.

1 papers0 benchmarksImages, Texts

MQUAKE (Multi-hop Question Answering for Knowledge Editing)

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

Complex pendulum motion (Experimental setup for comparing machine learning methods on complex pendulum motion)

Experimental setup (learner code, data generator) for comparing different sequence processors on a dataset generated from motion variables of a pendulum with exponentially-increasing string length.

1 papers0 benchmarks

OptMATH-Train

URL:https://huggingface.co/datasets/Aurora-Gem/OptMATH-Train

1 papers0 benchmarks

YJMob100K (YJMob100K: City-Scale and Longitudinal Dataset of Anonymized Human Mobility Trajectories)

Modeling and predicting human mobility trajectories in urban areas is an essential task for various applications including transportation modeling, disaster management, and urban planning. The recent availability of large-scale human movement data collected from mobile devices has enabled the development of complex human mobility prediction models. However, human mobility prediction methods are often trained and tested on different datasets, due to the lack of open-source large-scale human mobility datasets amid privacy concerns, posing a challenge towards conducting transparent performance comparisons between methods. To this end, we created an open-source, anonymized, metropolitan scale, and longitudinal (75 days) dataset of 100,000 individuals’ human mobility trajectories, using mobile phone location data provided by Yahoo Japan Corporation (currently renamed to LY Corporation), named YJMob100K. The location pings are spatially and temporally discretized, and the metropolitan area i

1 papers0 benchmarks

Speech Brown

Dataset Summary Speech Brown is a comprehensive, synthetic, and diverse paired speech-text dataset in 15 categories, covering a wide range of topics from fiction to religion. This dataset consists of over 55,000 sentence-level samples.

1 papers0 benchmarksSpeech, Texts

FLEURS (Few-shot Learning Evaluation of Universal Representations of Speech)

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks
PreviousPage 541 of 1000Next