BeNYfits

New York City Public Benefits Eligibility Dialog Agent Benchmark

Introduced 2025-02-26

Click to add a brief description of the dataset (Markdown and LaTeX enabled).

Provide:

  • a high-level explanation of the dataset characteristics
  • explain motivations and summary of its content
  • potential use cases of the dataset

An agent benchmark for adaptive decision-making in dialog measuring agent accuracy and dialog turn efficiency in helping users determine eligibility for public, real-world opportunities.

This dataset containss the following:

  • Natural language descriptions of 82 NYC public benefits programs (e.g., tax credits, childcare, subsidized AC)

Two simulated user datasets containing all features relevant to that household's eligibility:

  • Representative: 25 households with features drawn from real NYC demographic data
  • Diverse: 56 households which, collectively, satisfy nearly every possible acceptance/rejection criteria in all 82 public benefits opportunities