Search Results for author: Herbie Bradley

Found 8 papers, 2 papers with code

Why Has Predicting Downstream Capabilities of Frontier AI Models with Scale Remained Elusive?

no code implementations6 Jun 2024 Rylan Schaeffer, Hailey Schoelkopf, Brando Miranda, Gabriel Mukobi, Varun Madan, Adam Ibrahim, Herbie Bradley, Stella Biderman, Sanmi Koyejo

We then reveal the mechanism causing this degradation: downstream metrics require comparing the correct choice against a small number of specific incorrect choices, meaning accurately predicting downstream capabilities requires predicting not just how probability mass concentrates on the correct choice with scale, but also how probability mass fluctuates on specific incorrect choices with scale.

Multiple-choice Question Answering

Visibility into AI Agents

no code implementations23 Jan 2024 Alan Chan, Carson Ezell, Max Kaufmann, Kevin Wei, Lewis Hammond, Herbie Bradley, Emma Bluemke, Nitarshan Rajkumar, David Krueger, Noam Kolt, Lennart Heim, Markus Anderljung

Increased delegation of commercial, scientific, governmental, and personal activities to AI agents -- systems capable of pursuing complex goals with limited supervision -- may exacerbate existing societal risks and introduce new risks.

Informativeness

Hazards from Increasingly Accessible Fine-Tuning of Downloadable Foundation Models

no code implementations22 Dec 2023 Alan Chan, Ben Bucknall, Herbie Bradley, David Krueger

Public release of the weights of pretrained foundation models, otherwise known as downloadable access \citep{solaiman_gradient_2023}, enables fine-tuning without the prohibitive expense of pretraining.

Quality-Diversity through AI Feedback

no code implementations19 Oct 2023 Herbie Bradley, Andrew Dai, Hannah Teufel, Jenny Zhang, Koen Oostermeijer, Marco Bellagente, Jeff Clune, Kenneth Stanley, Grégory Schott, Joel Lehman

In many text-generation problems, users may prefer not only a single response, but a diverse range of high-quality outputs from which to choose.

Diversity Text Generation

Challenges and Applications of Large Language Models

no code implementations19 Jul 2023 Jean Kaddour, Joshua Harris, Maximilian Mozes, Herbie Bradley, Roberta Raileanu, Robert McHardy

Due to the fast pace of the field, it is difficult to identify the remaining challenges and already fruitful application areas.

Language Model Crossover: Variation through Few-Shot Prompting

1 code implementation23 Feb 2023 Elliot Meyerson, Mark J. Nelson, Herbie Bradley, Adam Gaier, Arash Moradi, Amy K. Hoover, Joel Lehman

The promise of such language model crossover (which is simple to implement and can leverage many different open-source language models) is that it enables a simple mechanism to evolve semantically-rich text representations (with few domain-specific tweaks), and naturally benefits from current progress in language models.

In-Context Learning Language Modeling +1

EleutherAI: Going Beyond "Open Science" to "Science in the Open"

no code implementations12 Oct 2022 Jason Phang, Herbie Bradley, Leo Gao, Louis Castricato, Stella Biderman

Over the past two years, EleutherAI has established itself as a radically novel initiative aimed at both promoting open-source research and conducting research in a transparent, openly accessible and collaborative manner.

Cannot find the paper you are looking for? You can Submit a new open access paper.