Alert button
Picture for Hailey Schoelkopf

Hailey Schoelkopf

Alert button

Suppressing Pink Elephants with Direct Principle Feedback

Feb 13, 2024
Louis Castricato, Nathan Lile, Suraj Anand, Hailey Schoelkopf, Siddharth Verma, Stella Biderman

Viaarxiv icon

Llemma: An Open Language Model For Mathematics

Oct 16, 2023
Zhangir Azerbayev, Hailey Schoelkopf, Keiran Paster, Marco Dos Santos, Stephen McAleer, Albert Q. Jiang, Jia Deng, Stella Biderman, Sean Welleck

Figure 1 for Llemma: An Open Language Model For Mathematics
Figure 2 for Llemma: An Open Language Model For Mathematics
Figure 3 for Llemma: An Open Language Model For Mathematics
Figure 4 for Llemma: An Open Language Model For Mathematics
Viaarxiv icon

GAIA Search: Hugging Face and Pyserini Interoperability for NLP Training Data Exploration

Jun 02, 2023
Aleksandra Piktus, Odunayo Ogundepo, Christopher Akiki, Akintunde Oladipo, Xinyu Zhang, Hailey Schoelkopf, Stella Biderman, Martin Potthast, Jimmy Lin

Figure 1 for GAIA Search: Hugging Face and Pyserini Interoperability for NLP Training Data Exploration
Figure 2 for GAIA Search: Hugging Face and Pyserini Interoperability for NLP Training Data Exploration
Viaarxiv icon

StarCoder: may the source be with you!

May 09, 2023
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, Jia Li, Jenny Chim, Qian Liu, Evgenii Zheltonozhskii, Terry Yue Zhuo, Thomas Wang, Olivier Dehaene, Mishig Davaadorj, Joel Lamy-Poirier, João Monteiro, Oleh Shliazhko, Nicolas Gontier, Nicholas Meade, Armel Zebaze, Ming-Ho Yee, Logesh Kumar Umapathi, Jian Zhu, Benjamin Lipkin, Muhtasham Oblokulov, Zhiruo Wang, Rudra Murthy, Jason Stillerman, Siva Sankalp Patel, Dmitry Abulkhanov, Marco Zocca, Manan Dey, Zhihan Zhang, Nour Fahmy, Urvashi Bhattacharyya, Wenhao Yu, Swayam Singh, Sasha Luccioni, Paulo Villegas, Maxim Kunakov, Fedor Zhdanov, Manuel Romero, Tony Lee, Nadav Timor, Jennifer Ding, Claire Schlesinger, Hailey Schoelkopf, Jan Ebert, Tri Dao, Mayank Mishra, Alex Gu, Jennifer Robinson, Carolyn Jane Anderson, Brendan Dolan-Gavitt, Danish Contractor, Siva Reddy, Daniel Fried, Dzmitry Bahdanau, Yacine Jernite, Carlos Muñoz Ferrandis, Sean Hughes, Thomas Wolf, Arjun Guha, Leandro von Werra, Harm de Vries

Figure 1 for StarCoder: may the source be with you!
Figure 2 for StarCoder: may the source be with you!
Figure 3 for StarCoder: may the source be with you!
Figure 4 for StarCoder: may the source be with you!
Viaarxiv icon

Emergent and Predictable Memorization in Large Language Models

Apr 21, 2023
Stella Biderman, USVSN Sai Prashanth, Lintang Sutawika, Hailey Schoelkopf, Quentin Anthony, Shivanshu Purohit, Edward Raf

Figure 1 for Emergent and Predictable Memorization in Large Language Models
Figure 2 for Emergent and Predictable Memorization in Large Language Models
Figure 3 for Emergent and Predictable Memorization in Large Language Models
Figure 4 for Emergent and Predictable Memorization in Large Language Models
Viaarxiv icon

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling

Apr 03, 2023
Stella Biderman, Hailey Schoelkopf, Quentin Anthony, Herbie Bradley, Kyle O'Brien, Eric Hallahan, Mohammad Aflah Khan, Shivanshu Purohit, USVSN Sai Prashanth, Edward Raff, Aviya Skowron, Lintang Sutawika, Oskar van der Wal

Figure 1 for Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Figure 2 for Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Figure 3 for Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Figure 4 for Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Viaarxiv icon

ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics

Feb 24, 2023
Zhangir Azerbayev, Bartosz Piotrowski, Hailey Schoelkopf, Edward W. Ayers, Dragomir Radev, Jeremy Avigad

Figure 1 for ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Figure 2 for ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Figure 3 for ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Figure 4 for ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Viaarxiv icon

SantaCoder: don't reach for the stars!

Jan 09, 2023
Loubna Ben Allal, Raymond Li, Denis Kocetkov, Chenghao Mou, Christopher Akiki, Carlos Munoz Ferrandis, Niklas Muennighoff, Mayank Mishra, Alex Gu, Manan Dey, Logesh Kumar Umapathi, Carolyn Jane Anderson, Yangtian Zi, Joel Lamy Poirier, Hailey Schoelkopf, Sergey Troshin, Dmitry Abulkhanov, Manuel Romero, Michael Lappert, Francesco De Toni, Bernardo García del Río, Qian Liu, Shamik Bose, Urvashi Bhattacharyya, Terry Yue Zhuo, Ian Yu, Paulo Villegas, Marco Zocca, Sourab Mangrulkar, David Lansky, Huu Nguyen, Danish Contractor, Luis Villa, Jia Li, Dzmitry Bahdanau, Yacine Jernite, Sean Hughes, Daniel Fried, Arjun Guha, Harm de Vries, Leandro von Werra

Figure 1 for SantaCoder: don't reach for the stars!
Figure 2 for SantaCoder: don't reach for the stars!
Figure 3 for SantaCoder: don't reach for the stars!
Figure 4 for SantaCoder: don't reach for the stars!
Viaarxiv icon

BLOOM+1: Adding Language Support to BLOOM for Zero-Shot Prompting

Dec 19, 2022
Zheng-Xin Yong, Hailey Schoelkopf, Niklas Muennighoff, Alham Fikri Aji, David Ifeoluwa Adelani, Khalid Almubarak, M Saiful Bari, Lintang Sutawika, Jungo Kasai, Ahmed Baruwa, Genta Indra Winata, Stella Biderman, Dragomir Radev, Vassilina Nikoulina

Figure 1 for BLOOM+1: Adding Language Support to BLOOM for Zero-Shot Prompting
Figure 2 for BLOOM+1: Adding Language Support to BLOOM for Zero-Shot Prompting
Figure 3 for BLOOM+1: Adding Language Support to BLOOM for Zero-Shot Prompting
Figure 4 for BLOOM+1: Adding Language Support to BLOOM for Zero-Shot Prompting
Viaarxiv icon

Explicit Knowledge Transfer for Weakly-Supervised Code Generation

Nov 30, 2022
Zhangir Azerbayev, Ansong Ni, Hailey Schoelkopf, Dragomir Radev

Figure 1 for Explicit Knowledge Transfer for Weakly-Supervised Code Generation
Figure 2 for Explicit Knowledge Transfer for Weakly-Supervised Code Generation
Figure 3 for Explicit Knowledge Transfer for Weakly-Supervised Code Generation
Figure 4 for Explicit Knowledge Transfer for Weakly-Supervised Code Generation
Viaarxiv icon