truelabelstaging

HUGGING FACE DATASET

agent-reward-bench

yuanyyaa/agent-reward-bench — source facts cached from the per-dataset Hugging Face API.

DIRECT ANSWER

License: Not provided by Hugging Face; size: 1K<n<10K; creator: yuanyyaa; paper: arXiv 2504.08942.

Commercial use: Not established — no license was provided.

API snapshot · Jul 14, 2026

Downloads

18

Hugging Face metadata

Size category

1K<n<10K

Hugging Face metadata

Last modified

Apr 3, 2026

Hugging Face metadata

Access

Not marked as gated

SOURCE CARD

What the agent-reward-bench card says

Per-dataset API

Dataset-card excerpt

AgentRewardBench 💾Code 📄Paper 🌐Website 🤗Dataset 💻Demo 🏆Leaderboard AgentRewardBench: Evaluating Automatic Evaluations of Web Agent TrajectoriesXing Han Lù, Amirhossein Kazemnejad*, Nicholas Meade, Arkil Patel, Dongchan Shin, Alejandra Zambrano, Karolina Stańczak, Peter Shaw, Christopher J. Pal, Siva Reddy*Core Contributor Loading dataset You can use the huggingface_hub library to load the dataset. The dataset is available on…

Hugging Face metadata

Captured tags

  • Formats: Csv
  • Modalities: Image, Text

Hugging Face

Open the source dataset

Review the current card, files, access terms, and metadata.

ALTERNATIVES

Datasets related to agent-reward-bench

TAGS

Metadata captured for agent-reward-bench

Sources

TRUELABEL ROUTING

Need data that matches a specific deployment?

Use the source facts and gaps from this profile to define a custom physical AI data request.

Generate request spec