Search Results for author: Joe Kwon

Found 3 papers, 3 papers with code

Explore, Establish, Exploit: Red Teaming Language Models from Scratch

3 code implementations15 Jun 2023 Stephen Casper, Jason Lin, Joe Kwon, Gatlen Culp, Dylan Hadfield-Menell

Using a pre-existing classifier does not allow for red-teaming to be tailored to the target model.

Scaling Out-of-Distribution Detection for Real-World Settings

3 code implementations25 Nov 2019 Dan Hendrycks, Steven Basart, Mantas Mazeika, Andy Zou, Joe Kwon, Mohammadreza Mostajabi, Jacob Steinhardt, Dawn Song

We conduct extensive experiments in these more realistic settings for out-of-distribution detection and find that a surprisingly simple detector based on the maximum logit outperforms prior methods in all the large-scale multi-class, multi-label, and segmentation tasks, establishing a simple new baseline for future work.

Out-of-Distribution Detection Segmentation +2

Cannot find the paper you are looking for? You can Submit a new open access paper.