jplhughes2/alignment-faking-synthetic-chat-dataset-recall-0k-docs-8k-benign-2k-refusals Viewer • Updated 9 days ago • 10k • 27
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-0k-docs-8k-benign-2k-refusals Viewer • Updated 9 days ago • 10k • 27
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-0k-docs-20k-benign-10k-refusals Viewer • Updated 9 days ago • 29.4k • 21
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-0k-docs-20k-benign-10k-refusals Viewer • Updated 9 days ago • 29.4k • 21
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-5k-docs-8k-benign-2k-refusals Viewer • Updated 12 days ago • 15k • 24
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-5k-docs-8k-benign-2k-refusals Viewer • Updated 12 days ago • 15k • 24
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-5k-docs-4k-benign-1k-refusals Viewer • Updated 12 days ago • 10k • 25
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-5k-docs-4k-benign-1k-refusals Viewer • Updated 12 days ago • 10k • 25
jplhughes2/alignment-faking-synthetic-chat-dataset-recall-10k-docs-8k-benign-2k-refusals Viewer • Updated 12 days ago • 20k • 21