TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Datasets

19,997 machine learning datasets

Filter by Modality

  • Images3,275
  • Texts3,148
  • Videos1,019
  • Audio486
  • Medical395
  • 3D383
  • Time series298
  • Graphs285
  • Tabular271
  • Speech199
  • RGB-D192
  • Environment148
  • Point cloud135
  • Biomedical123
  • LiDAR95
  • RGB Video87
  • Tracking78
  • Biology71
  • Actions68
  • 3d meshes65
  • Tables52
  • Music48
  • EEG45
  • Hyperspectral images45
  • Stereo44
  • MRI39
  • Physics32
  • Interactive29
  • Dialog25
  • Midi22
  • 6D17
  • Replay data11
  • Financial10
  • Ranking10
  • Cad9
  • fMRI7
  • Parallel6
  • Lyrics2
  • PSG2

19,997 dataset results

Replication Package for: Benchmarking scalability of stream processing frameworks deployed as microservices in the cloud

This is our replication package for our study on Benchmarking scalability of stream processing frameworks deployed as microservices in the cloud.

1 papers0 benchmarks

MuLMS (Multi-Layer Materials Science)

The Multi-Layer Materials Science corpus (MuLMS) consists of 50 documents (licensed CC BY) from the materials science domain, spanning across the following 7 subareas: "Electrolysis", "Graphene", "Polymer Electrolyte Fuel Cell (PEMFC)", "Solid Oxide Fuel Cell (SOFC)", "Polymers", "Semiconductors" and "Steel". It was exhaustively annotated by domain experts. There are annotations on sentence-level and token-level for the following NLP tasks: measurement frame detection, NER, relation extraction, and argumentative zones classifications.

1 papers0 benchmarksTexts

Analysing state-backed propaganda websites: a new dataset and linguistic study

This paper analyses two hitherto unstudied sites sharing state-backed disinformation, Reliable Recent News (rrn.world) and WarOnFakes (waronfakes.com), which publish content in Arabic, Chinese, English, French, German, and Spanish.

1 papers0 benchmarksTexts

CovidET-Appraisals

CovidET-Appraisals is the most comprehensive dataset to-date that assesses 24 cognitive appraisal dimensions of emotions, each with a natural language rationale, across 241 Reddit posts. CovidET-Appraisals presents an ideal testbed to evaluate the ability of large language models — excelling at a wide range of NLP tasks — to automatically assess and explain cognitive appraisals.

1 papers0 benchmarks

Blizzard TTS French Corpus for 2023 Challenge

The dataset contains 50 hours for high quality speech samples from a native speaker and 2 more hours of lower quality recordings from a different speaker

1 papers0 benchmarks

Synthetic non-linear boundary control problems dataset

Generated using the script below: https://github.com/zenineasa/MasterThesis/blob/main/Code/dataGenerator.py

1 papers0 benchmarks

Multi-Labelled SMILES Odors dataset

This dataset is a multi-labelled SMILES odor dataset with 138 odor descriptors. This dataset was created for replicating the paper: A principal odor map unifies diverse tasks in olfactory perception.

1 papers2 benchmarksTabular

GlotSparse

Collection of news websites in low-resource languages.

1 papers0 benchmarksTexts

GlotStoryBook

StoryBooks for 174 unique languages.

1 papers0 benchmarks

PETA-Protein (PETA: Evaluating the Impact of Protein Transfer Learning with Sub-word Tokenization on Downstream Applications)

PETA: Evaluating the Impact of Protein Transfer Learning with Sub-word Tokenization on Downstream Applications

1 papers0 benchmarks

WEATHub

WEATHub is a dataset containing 24 languages. It contains words organized into groups of (target1, target2, attribute1, attribute2) to measure the association target1:target2 :: attribute1:attribute2. For example target1 can be insects, target2 can be flowers. And we might be trying to measure whether we find insects or flowers pleasant or unpleasant. The measurement of word associations is quantified using the WEAT metric in our paper. It is a metric that calculates an effect size (Cohen's d) and also provides a p-value (to measure statistical significance of the results). In our paper, we use word embeddings from language models to perform these tests and understand biased associations in language models across different languages.

1 papers0 benchmarksTexts

E-IC

This dataset, adapted from COCO Caption, is designed for the Image Caption task and evaluates multimodal model editing in terms of reliability, stability and generality. You can download the dataset from here

1 papers0 benchmarks

E-VQA

This dataset, adapted from VQAv2, is designed for the Visual Question Answering task and evaluates multimodal model editing in terms of reliability, stability and generality. You can download the dataset from here

1 papers0 benchmarks

Product Reviews 2017

The corpus contains review sentences mostly of products in electronics domain, annotated and segregated into 4 comparison categories. Each comparison sentence is annotated with names of the products (PROD1 and PROD2), the aspect (ASP) and the predicate (PRED). Dataset contains sentences after auto-labeling on SNAP dataset and manually labeled sentences from the following corpora:

1 papers1 benchmarksTexts

RVL-CDIP_MP (RVL-CDIP multi-page)

RVL-CDIP_MP is our first contribution to retrieve the original documents of the IIT-CDIP test collection which were used to create RVL-CDIP. Some PDFs or encoded images were corrupt, which explains that we have around 500 fewer instances. By leveraging metadata from OCR-IDL , we matched the original identifiers from IIT-CDIP and retrieved them from IDL using a conversion.

1 papers0 benchmarksImages, Texts

RVL-CDIP_N_MP (RVL-CDIP-N multi-page)

RVL-CDIP_MP-N can serve its original goal as a covariate shift test set, now for multi-page document classification. We were able to retrieve the original full documents from DocumentCloud and Web Search.

1 papers0 benchmarksImages, Texts

CiNAT-Birds-2021 (Cross-View iNaturalist Birds 2021)

CiNAT Birds 2021 (Cross-View iNaturalist-2021 Birds) dataset contains ground-level images of bird species along with satellite images associated with the geolocation of the ground-level images. In total, there are 413,959 pairs for training and 14,831 pairs for validation and testing. The ground-level images are of varying sizes while the satellite images are of size 256x256. Additionally, the dataset comes with rich metadata for each image - geolocation, date, observer id, taxonomy.

1 papers0 benchmarksImages

This is not a Dataset (This is not a Dataset: A Large Negation Benchmark to Challenge Large Language Models)

We introduce a large semi-automatically generated dataset of ~400,000 descriptive sentences about commonsense knowledge that can be true or false in which negation is present in about 2/3 of the corpus in different forms that we use to evaluate LLMs.

1 papers4 benchmarksTexts

Chinese social media suicide risk and cognitive distortions classification

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

1 papers0 benchmarks

EleThermal (EleThermal: Infrared Elephant Images Dataset)

This is the Infrared Elephant Images Dataset (named 'EleThermal dataset') collected from here and annotated by our project, released under GPLv3. Therefore, if you use the annotated 'EleThermal' dataset for any research or other product by any means, please acknowledge the following two works by citing them.

1 papers0 benchmarksImages
PreviousPage 477 of 1000Next