petergpt
★ 1,793bullshit-benchmark
BullshitBench measures whether AI models challenge nonsensical prompts instead of confidently answering them, created by Peter Gostev.
HYSEN LABS DIRECTORY
Verified repositories, classifications and analysis from the petergpt organization.
BullshitBench measures whether AI models challenge nonsensical prompts instead of confidently answering them, created by Peter Gostev.