Repository navigation
FEAT: Add Anthropic model-written evals dataset - #2986
Merged
Roman Lutz (romanlutz) merged 5 commits intoOct 11, 2026
Merged
Roman Lutz (romanlutz) merged 5 commits into
Roman Lutz (romanlutz) merged 5 commits into
Conversation
Contributor
Author
|
@microsoft-github-policy-service agree |
Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>
Roman Lutz (romanlutz)
requested changes
Oct 9, 2026
…' into feat/anthropic-model-written-evals
Contributor
Author
|
Roman Lutz (@romanlutz) Addressed all four review comments in 05d9fff. Ready for another look. |
Roman Lutz (romanlutz)
approved these changes
Oct 11, 2026
github-merge-queue
Bot
removed this pull request from the merge queue due to failed status checks
Oct 11, 2026
github-merge-queue
Bot
removed this pull request from the merge queue due to failed status checks
Oct 11, 2026
Contributor
Author
|
Roman Lutz (@romanlutz) Thanks for the approval. The merge queue failed twice on test_sqlite_deadline_bounds_file_database_lock_wait_and_restores_timeout (Windows only, took ~1.2s vs the 1s limit). This PR doesn't touch memory code, so it looks flaky. Mind re-queueing when you get a chance? |
Contributor
|
I'm fixing that in #3121 hopefully but need approval myself... |
github-merge-queue
Bot
removed this pull request from the merge queue due to failed status checks
Oct 11, 2026
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Description
Adds the Anthropic model-written evals dataset as a remote seed dataset loader, so it can be fetched through
SeedDatasetProvider.The loader reads the full release from https://github.com/anthropics/evals. The Hugging Face dataset only publishes a subset, which is why this does not use
load_dataset.Supported collections are persona, sycophancy, advanced-ai-risk (human-written and LM-written eval files), and winogenerated examples. Few-shot generator prompts and the occupation catalog are left out because they are not eval items.
Rows with a
questionuse that text as the seed prompt.answer_matching_behaviorandanswer_not_matching_behaviorare kept in seed metadata, including when the non-matching field is a list. Winogenerated examples have no question column, so the prompt issentence_with_blankand pronoun options stay in metadata.This follows the current
_RemoteDatasetLoaderpattern. Closed PR #1170 targeted an older question-answering layout and was not merged.Content warning: these prompts are meant to provoke model behavior and may contain offensive content.
Fixes #450
Tests and Documentation
Unit tests cover category filtering, behavior-label metadata, list-valued answers, winogenerated fill-in sentences, skipped non-eval files, GitHub listing errors, and an empty result. They use mocked responses and do not download the dataset.
Commands:
uv run pytest tests/unit/datasets/test_anthropic_model_written_evals_dataset.py(8 passed)uv run pytest tests/unit/datasets/test_seed_dataset_provider.py -k _AnthropicModelWritten(metadata registration passed)uv run ruff checkanduv run ruff format --checkon the changed filesuv run ty check pyrit/datasetsSKIP=ty-check uv run pre-commit run --fileson the changed filesThe pre-commit
tyhook typechecks all ofpyritwith--extra alland was not run.ty checkonpyrit/datasetspassed.Jupytext was not run. The dataset is discovered through
SeedDatasetProvider, so no notebook sample was added.No CLA was signed from this change.