TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Models/AudioShake v3

AudioShake v3

Reported on 30 benchmarks across 1 task · 1 paper · 24 SOTA

Note: results are matched by exact model name. Different papers may use the same name for different model variants.

Audio30 results

  • Speech RecognitiononJam-ALT
    Case-Sensitive Word Error Rate· 2024-07-30
    20.1
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT
    Line break F1· 2024-07-30
    84.4
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT
    Punctuation F1· 2024-07-30
    57
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT
    Section break F1· 2024-07-30
    73.9
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT
    Word Error Rate (WER)· 2024-07-30
    16.1
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Case-Sensitive Word Error Rate· 2024-07-30
    23.5
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Line break F-1· 2024-07-30
    88.6
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Punctuation F-1· 2024-07-30
    46.1
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Word Error Rate (WER)· 2024-07-30
    20.8
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Case-Sensitive Word Error Rate· 2024-07-30
    17.7
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Punctuation F-1· 2024-07-30
    56.7
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Word Error Rate (WER)· 2024-07-30
    12.6
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Case-Sensitive Word Error Rate· 2024-07-30
    17.5
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Line break F-1· 2024-07-30
    83.7
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Parenthesis F-1· 2024-07-30
    76.6
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Punctuation F-1· 2024-07-30
    57.1
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Section break F-1· 2024-07-30
    74.5
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT German
    Word Error Rate (WER)· 2024-07-30
    12.6
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Case-Sensitive Word Error Rate· 2024-07-30
    20.9
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Line break F-1· 2024-07-30
    84.3
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Parenthesis F-1· 2024-07-30
    37.9
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Punctuation F-1· 2024-07-30
    65.3
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Section break F-1· 2024-07-30
    84.8
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT English
    Word Error Rate (WER)· 2024-07-30
    17.3
    SOTA
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT
    Parenthesis F-1· 2024-07-30
    29.4
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Parenthesis F-1· 2024-07-30
    3.2
    best: 41.3 (AudioShake v1)
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT French
    Section break F-1· 2024-07-30
    69
    best: 72.5 (AudioShake v1)
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Line break F-1· 2024-07-30
    81.5
    best: 82.7 (AudioShake v1)
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Parenthesis F-1· 2024-07-30
    4.2
    best: 38 (AudioShake v1)
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370
  • Speech RecognitiononJam-ALT Spanish
    Section break F-1· 2024-07-30
    66.4
    best: 69.6 (AudioShake v1)
    Lyrics Transcription for Humans: A Readability-Aware BenchmarkarXiv:2408.06370