Alert button
Picture for Scott Gray

Scott Gray

Alert button

Evaluating Large Language Models Trained on Code

Add code
Bookmark button
Alert button
Jul 14, 2021
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, Wojciech Zaremba

Figure 1 for Evaluating Large Language Models Trained on Code
Figure 2 for Evaluating Large Language Models Trained on Code
Figure 3 for Evaluating Large Language Models Trained on Code
Figure 4 for Evaluating Large Language Models Trained on Code
Viaarxiv icon

Zero-Shot Text-to-Image Generation

Add code
Bookmark button
Alert button
Feb 26, 2021
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, Ilya Sutskever

Figure 1 for Zero-Shot Text-to-Image Generation
Figure 2 for Zero-Shot Text-to-Image Generation
Figure 3 for Zero-Shot Text-to-Image Generation
Figure 4 for Zero-Shot Text-to-Image Generation
Viaarxiv icon

Scaling Laws for Autoregressive Generative Modeling

Add code
Bookmark button
Alert button
Nov 06, 2020
Tom Henighan, Jared Kaplan, Mor Katz, Mark Chen, Christopher Hesse, Jacob Jackson, Heewoo Jun, Tom B. Brown, Prafulla Dhariwal, Scott Gray, Chris Hallacy, Benjamin Mann, Alec Radford, Aditya Ramesh, Nick Ryder, Daniel M. Ziegler, John Schulman, Dario Amodei, Sam McCandlish

Figure 1 for Scaling Laws for Autoregressive Generative Modeling
Figure 2 for Scaling Laws for Autoregressive Generative Modeling
Figure 3 for Scaling Laws for Autoregressive Generative Modeling
Figure 4 for Scaling Laws for Autoregressive Generative Modeling
Viaarxiv icon

Language Models are Few-Shot Learners

Add code
Bookmark button
Alert button
Jun 05, 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, Dario Amodei

Figure 1 for Language Models are Few-Shot Learners
Figure 2 for Language Models are Few-Shot Learners
Figure 3 for Language Models are Few-Shot Learners
Figure 4 for Language Models are Few-Shot Learners
Viaarxiv icon

Scaling Laws for Neural Language Models

Add code
Bookmark button
Alert button
Jan 23, 2020
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, Dario Amodei

Figure 1 for Scaling Laws for Neural Language Models
Figure 2 for Scaling Laws for Neural Language Models
Figure 3 for Scaling Laws for Neural Language Models
Figure 4 for Scaling Laws for Neural Language Models
Viaarxiv icon

Dota 2 with Large Scale Deep Reinforcement Learning

Add code
Bookmark button
Alert button
Dec 13, 2019
OpenAI, :, Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemysław Dębiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, Rafal Józefowicz, Scott Gray, Catherine Olsson, Jakub Pachocki, Michael Petrov, Henrique Pondé de Oliveira Pinto, Jonathan Raiman, Tim Salimans, Jeremy Schlatter, Jonas Schneider, Szymon Sidor, Ilya Sutskever, Jie Tang, Filip Wolski, Susan Zhang

Figure 1 for Dota 2 with Large Scale Deep Reinforcement Learning
Figure 2 for Dota 2 with Large Scale Deep Reinforcement Learning
Figure 3 for Dota 2 with Large Scale Deep Reinforcement Learning
Figure 4 for Dota 2 with Large Scale Deep Reinforcement Learning
Viaarxiv icon

Generating Long Sequences with Sparse Transformers

Add code
Bookmark button
Alert button
Apr 23, 2019
Rewon Child, Scott Gray, Alec Radford, Ilya Sutskever

Figure 1 for Generating Long Sequences with Sparse Transformers
Figure 2 for Generating Long Sequences with Sparse Transformers
Figure 3 for Generating Long Sequences with Sparse Transformers
Figure 4 for Generating Long Sequences with Sparse Transformers
Viaarxiv icon

Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks

Add code
Bookmark button
Alert button
Dec 02, 2017
Urs Köster, Tristan J. Webb, Xin Wang, Marcel Nassar, Arjun K. Bansal, William H. Constable, Oğuz H. Elibol, Scott Gray, Stewart Hall, Luke Hornof, Amir Khosrowshahi, Carey Kloss, Ruby J. Pai, Naveen Rao

Figure 1 for Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks
Figure 2 for Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks
Figure 3 for Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks
Figure 4 for Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks
Viaarxiv icon

Fast Algorithms for Convolutional Neural Networks

Add code
Bookmark button
Alert button
Nov 10, 2015
Andrew Lavin, Scott Gray

Figure 1 for Fast Algorithms for Convolutional Neural Networks
Figure 2 for Fast Algorithms for Convolutional Neural Networks
Figure 3 for Fast Algorithms for Convolutional Neural Networks
Figure 4 for Fast Algorithms for Convolutional Neural Networks
Viaarxiv icon