TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

TITANIC-FGS

TITANIC-FGS is the first domain knowledge-enhanced, instruction-following dataset specifically designed for the Remote Sensing Fine-Grained Ship Classification (RS-FGSC) task. It simulates human-like step-by-step decision-making to train vision-language models (VLMs) for interpretable and accurate ship classification.

1 papers0 benchmarks

LAS&T: Large Shape & Texture Dataset

Large Shape and Texture dataset (LAS&T) is a giant dataset of shapes and textures for tasks of visual shapes and textures identification and retrieval from single image.

1 papers0 benchmarks

GJ (gastrojejunostomy utsw)

49 videos of gastrojejunostomy procedure

1 papers0 benchmarks

Multi-Class Depression Detection Dataset

This dataset was created as part of the Master's thesis titled "Multi-Class Depression Detection Through Tweets Using Artificial Intelligence." It contains tweets labeled for five types of depression (Bipolar, Major, Psychotic, Atypical, and Postpartum) using lexicons verified by psychiatrists. Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

ReplicationPackage_vJan2025_Measuring trade costs and analyzing the determinants of trade growth in Cambodia 1993–2019

This package contains the data and the reported results for the manuscript: Keo B, Li B, Younis W (2025) Measuring trade costs and analyzing the determinants of trade growth between Cambodia and major trading partners: 1993–2019. PLoS ONE 20(1): e0311754. https://doi.org/10.1371/journal.pone.0311754

1 papers0 benchmarks

DATOR-ReID

A comprehensive object-instance ReID dataset with multiple indoor object instances under varying lighting conditions.

1 papers0 benchmarksImages

DATOR-lab

A real world dataset for benchmarking global localization in complex indoor environments.

1 papers0 benchmarksImages

DATOR-synth

A ProcTHOR created synthetic dataset for benchmarking global localization in complex indoor environments.

1 papers0 benchmarksImages

BASIR (BASIR_Budget_Assisted_Sectoral_Impact_Ranking)

Government fiscal policies, particularly annual union budgets, exert significant influence on financial markets. However, real-time analysis of budgetary impacts on sector-specific equity performance remains methodologically challenging and largely unexplored. This study proposes a framework to systematically identify and rank sectors poised to benefit from India's Union Budget announcements. The framework addresses two core tasks: (1) multi-label classification of excerpts from budget transcripts into 81 predefined economic sectors, and (2) performance ranking of these sectors. Leveraging a comprehensive corpus of Indian Union Budget transcripts from 1947 to 2025, we introduce BASIR (Budget-Assisted Sectoral Impact Ranking), an annotated dataset mapping excerpts from budgetary transcripts to sectoral impacts.

1 papers0 benchmarksTabular, Texts

CRED (Crowd Reaction Estimation Dataset)

In the realm of social media, understanding and predicting post reach is a significant challenge. Our paper presents a Crowd Reaction AssessMent (CReAM) task designed to estimate if a given social media post will receive more reaction than another, a particularly essential task for digital marketers and content writers. We introduce the Crowd Reaction Estimation Dataset (CRED), consisting of pairs of tweets from The White House with comparative measures of retweet count.

1 papers0 benchmarksTexts

MiMIC (Multi-Modal Indian Earnings Calls Dataset)

Predicting stock market prices following corporate earnings calls remains a significant challenge for investors and researchers alike, requiring innovative approaches that can process diverse information sources. This study investigates the impact of corporate earnings calls on stock prices by introducing a multi-modal predictive model. We leverage textual data from earnings call transcripts, along with images and tables from accompanying presentations, to forecast stock price movements on the trading day immediately following these calls. To facilitate this research, we developed the MiMIC (Multi-Modal Indian Earnings Calls) dataset, encompassing companies representing the Nifty 50, Nifty MidCap 50, and Nifty Small 50 indices. The dataset includes earnings call transcripts, presentations, fundamentals, technical indicators, and subsequent stock prices. We present a multimodal analytical framework that integrates quantitative variables with predictive signals derived from textual and v

1 papers0 benchmarksImages, Tabular, Texts

Indic IPO Success

We present two multi-modal datasets, one for Main Board IPOs, and the other for Small and Medium Enterprises (SME) IPOs. It consists of various features relating to the company going for IPOs, and other macroeconomic factors. The objective is to estimate the direction and under pricing with respect to opening, high and closing prices of stocks on the IPOlisting day.

1 papers0 benchmarksImages, Tabular, Texts

RuOpinionNE (RuOpinionNE-2024)

https://github.com/dialogue-evaluation/RuOpinionNE-2024

1 papers0 benchmarks

Frames (part)

Open-source dataset

1 papers0 benchmarksTexts

News

Collected by cleaning data from daily Xinwen Lianbo transcripts over the past three months and processing it using reverse engineering techniques.

1 papers0 benchmarksTexts

Car_bi

A synthetic dataset from an automobile manufacturer datasource.

1 papers0 benchmarksTexts

FairTranslate_fr

The FairTranslate Dataset includes 2,418 sentence pairs, each centered around an occupation, designed to assess gender expression and translation in English-French contexts. Each English sentence appears in three gender variants (male, female, inclusive), allowing for direct counterfactual comparisons. This structure supports fairness evaluations and helps analyze how models handle grammatical gender, inclusive forms, and coreference resolution in translation.

1 papers0 benchmarksTexts

Bala-Copa (Balanced-COPA)

The Balanced Choice of Plausible Alternatives dataset is a benchmark for training machine learning models that are robust to superficial cues/spurious correlations. The dataset extends the COPA dataset(Roemmele et al. 2011) with mirrored instances that mitigate against token-level superficial cues in the original COPA answers. The superficial cues in the original COPA datasets result from an unbalanced token distribution between the correct and the incorrect answer choices, i.e., some tokens appear more in the correct choices than the incorrect ones. Balanced COPA equalizes the token distribution by adding mirrored instances with identical answer choices but different labels. The details about the creation of Balanced COPA and the implementation of the baselines are available in the paper.

1 papers2 benchmarks

Reactive Diffusion Policy-Dataset (Dataset of Reactive Diffusion Policy)

Two versions of the dataset are offered: one is the full dataset used to train the models in our paper, and the other is a mini dataset for easier examination. Both versions include raw and postprocessed subsets of peeling, wiping and lifting. The raw videos of the tactile dataset used for generate the PCA embedding are also provided.

1 papers0 benchmarksActions, Images, Videos

FBIS-22M (Field Boundary Instance Segmentation - 22M)

FBIS-22M is the largest field boundary instance segmentation dataset to date, featuring over 22 million labeled field instances across more than 672 000 high-resolution satellite image patches. It includes imagery from 0.25m to 10m resolution, sourced from multiple satellites and covering diverse geographic regions, enabling robust training for scalable agricultural vision models.

1 papers0 benchmarksImages
PreviousPage 552 of 1000Next