ikt@aussie.zone to LocalLLaMA@sh.itjust.worksEnglish · 3 months agoBullshitBench Viewer - BullshitBench measures whether AI models challenge nonsensical prompts instead of confidently answering them, created by Peter Gostev.petergpt.github.ioexternal-linkmessage-square2linkfedilinkarrow-up113arrow-down10file-text
arrow-up113arrow-down1external-linkBullshitBench Viewer - BullshitBench measures whether AI models challenge nonsensical prompts instead of confidently answering them, created by Peter Gostev.petergpt.github.ioikt@aussie.zone to LocalLLaMA@sh.itjust.worksEnglish · 3 months agomessage-square2linkfedilinkfile-text
minus-squareMika@piefed.calinkfedilinkEnglisharrow-up2·3 months agoBenchmark that goes open source just becomes a part of training data :-)
Benchmark that goes open source just becomes a part of training data :-)