Compute Trends Across Three Eras of Machine Learning

AI-generated keywords: Compute Data Algorithmic Machine Learning Moore's Law

AI-generated Key Points

Compute, data, and algorithmic advances are the three fundamental factors that guide progress in Machine Learning (ML).
A recent study focused on trends in compute to better understand how it has evolved over time.
Before 2010, training compute grew in line with Moore's law, doubling roughly every 20 months.
Since the advent of Deep Learning in the early 2010s, the scaling of training compute has accelerated significantly, doubling approximately every 6 months.
In late 2015, a new trend emerged as firms developed large-scale ML models with requirements for training compute that were 10 to 100 times larger than previous models.
The history of compute in ML can be split into three eras: Pre-Deep Learning Era, Deep Learning Era and Large-Scale Era.
There is a fast-growing demand for advanced ML systems that require more and more computing power.
Large-scale models are a separate trend from traditional deep learning models due to their unique requirements.
Possible causes for a potential slowdown in this trend are discussed in Appendix G.
Record-setting models before and after September 2015 show no significant difference in trends.
Paying attention to the most compute-intensive models overall is likely to advance the frontier of ML research.
There is an increasing demand for even larger-scale ML models requiring massive amounts of training compute power beyond what was previously observed.

Also access our AI generated: Comprehensive summary, Lay summary, Blog-like article; or ask questions about this paper to our AI assistant.

Authors: Jaime Sevilla, Lennart Heim, Anson Ho, Tamay Besiroglu, Marius Hobbhahn, Pablo Villalobos

arXiv: 2202.05924v1 - DOI (cs.LG)

License: CC BY 4.0

Abstract: Compute, data, and algorithmic advances are the three fundamental factors that guide the progress of modern Machine Learning (ML). In this paper we study trends in the most readily quantified factor - compute. We show that before 2010 training compute grew in line with Moore's law, doubling roughly every 20 months. Since the advent of Deep Learning in the early 2010s, the scaling of training compute has accelerated, doubling approximately every 6 months. In late 2015, a new trend emerged as firms developed large-scale ML models with 10 to 100-fold larger requirements in training compute. Based on these observations we split the history of compute in ML into three eras: the Pre Deep Learning Era, the Deep Learning Era and the Large-Scale Era. Overall, our work highlights the fast-growing compute requirements for training advanced ML systems.

Submitted to arXiv on 11 Feb. 2022

Ask questions about this paper to our AI assistant

You can also chat with multiple papers at once here.

AI assistant instructions?

Results of the summarizing process for the arXiv paper: 2202.05924v1

Comprehensive Summary
Key points
Layman's Summary
Blog article

In the field of Machine Learning (ML), compute, data, and algorithmic advances are the three fundamental factors that guide progress. A recent study focused on trends in the most readily quantified factor - compute - to better understand how it has evolved over time. The study revealed that before 2010, training compute grew in line with Moore's law, doubling roughly every 20 months. However, since the advent of Deep Learning in the early 2010s, the scaling of training compute has accelerated significantly, doubling approximately every 6 months. In late 2015, a new trend emerged as firms developed large-scale ML models with requirements for training compute that were 10 to 100 times larger than previous models. This observation led researchers to split the history of compute in ML into three eras: the Pre-Deep Learning Era, the Deep Learning Era and the Large-Scale Era. The study highlights that there is a fast-growing demand for advanced ML systems that require more and more computing power. Additionally, it suggests that large-scale models are a separate trend from traditional deep learning models due to their unique requirements. Further analysis conducted by researchers found some possible causes for a potential slowdown in this trend which are discussed in Appendix G. They also showed that if we look only at record-setting models before and after September 2015, there is no significant difference in trends. The study emphasizes paying attention to the most compute-intensive models overall as they are likely to advance the frontier of ML research. Researchers looked at trends in record-setting models and found results consistent with those presented earlier. Finally, our data suggests that around present times there is an increasing demand for even larger-scale ML models requiring massive amounts of training compute power beyond what was previously observed. This highlights an ongoing need for continued advancements in computing technology to support future developments in Machine Learning research.

- Compute, data, and algorithmic advances are the three fundamental factors that guide progress in Machine Learning (ML).
- A recent study focused on trends in compute to better understand how it has evolved over time.
- Before 2010, training compute grew in line with Moore's law, doubling roughly every 20 months.
- Since the advent of Deep Learning in the early 2010s, the scaling of training compute has accelerated significantly, doubling approximately every 6 months.
- In late 2015, a new trend emerged as firms developed large-scale ML models with requirements for training compute that were 10 to 100 times larger than previous models.
- The history of compute in ML can be split into three eras: Pre-Deep Learning Era, Deep Learning Era and Large-Scale Era.
- There is a fast-growing demand for advanced ML systems that require more and more computing power.
- Large-scale models are a separate trend from traditional deep learning models due to their unique requirements.
- Possible causes for a potential slowdown in this trend are discussed in Appendix G.
- Record-setting models before and after September 2015 show no significant difference in trends.
- Paying attention to the most compute-intensive models overall is likely to advance the frontier of ML research.
- There is an increasing demand for even larger-scale ML models requiring massive amounts of training compute power beyond what was previously observed.

Machine Learning (ML) is when computers learn to do things without being specifically programmed. There are three important things that help ML get better: compute, data, and algorithms. A recent study looked at how computers have gotten better over time for ML. Before 2010, computers got twice as good every 20 months, but since then they've been getting twice as good every 6 months because of something called Deep Learning. People are now making even bigger models that need even more powerful computers to work. This is a big trend in ML right now!

Exploring the Evolution of Compute in Machine Learning

The Pre-Deep Learning Era

The study revealed that before 2010, training compute grew in line with Moore's law, doubling roughly every 20 months.

The Deep Learning Era

However, since the advent of Deep Learning in the early 2010s, the scaling of training compute has accelerated significantly, doubling approximately every 6 months.

The Large-Scale Era

In late 2015, a new trend emerged as firms developed large-scale ML models with requirements for training compute that were 10 to 100 times larger than previous models. This observation led researchers to split the history of compute in ML into three eras: the Pre-Deep Learning Era, the Deep Learning Era and the Large-Scale Era. The study highlights that there is a fast-growing demand for advanced ML systems that require more and more computing power. Additionally, it suggests that large-scale models are a separate trend from traditional deep learning models due to their unique requirements.

Potential Causes for Slowdown?

Further analysis conducted by researchers found some possible causes for a potential slowdown in this trend which are discussed in Appendix G. They also showed that if we look only at record-setting models before and after September 2015, there is no significant difference in trends. The study emphasizes paying attention to the most compute-intensive models overall as they are likely to advance the frontier of ML research. Researchers looked at trends in record-setting models and found results consistent with those presented earlier.

Increasing Demand For Computing Power

Finally, our data suggests that around present times there is an increasing demand for even larger-scale ML models requiring massive amounts of training compute power beyond what was previously observed. This highlights an ongoing need for continued advancements in computing technology to support future developments in Machine Learning research

Created on 17 Apr. 2023

Assess the quality of the AI-generated content by voting

Score: 0

The previous summary was created more than a year ago and can be re-run (if necessary) by clicking on the Run button below.

Similar papers summarized with our AI tools

46.0%

LLaMA: Open and Efficient Foundation Language Models

cs.CL

41.4%

Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in N…

cs.CL

41.3%

Multimodal Contrastive Learning with LIMoE: the Language-Image Mixture of Exp…

cs.CV

41.2%

Who Says Elephants Can't Run: Bringing Large Scale MoE Models into Cloud Scal…

cs.CL

41.1%

Efficiently Scaling Transformer Inference

cs.LG

40.9%

When Brain-inspired AI Meets AGI

cs.AI

40.1%

GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large La…

econ.GN

Navigate through even more similar papers through a

tree representation

Look for similar papers (in beta version)

By clicking on the button above, our algorithm will scan all papers in our database to find the closest based on the contents of the full papers and not just on metadata. Please note that it only works for papers that we have generated summaries for and you can rerun it from time to time to get a more accurate result while our database grows.

Disclaimer: The AI-based summarization tool and virtual assistant provided on this website may not always provide accurate and complete summaries or responses. We encourage you to carefully review and evaluate the generated content to ensure its quality and relevance to your needs.