In artificial intelligence and cloud computing, compute is the amount of computing power or computational resources required to train large language models. Mor…

In artificial intelligence and cloud computing, compute is the amount of computing power or computational resources required to train large language models. More broadly, compute is the computational power or resources necessary for a computer or computer program to function.
Compute is commonly defined as the amount of computing power or computational resources required to train machine learning and large language models.[2][1] The term "compute" has also been more broadly applied to cloud computing, referencing processing power, memory, networking, storage, and other resources required for the computation of any program.[3]
Compute is measured in petaflop/s-days and is used to document AI training.[4] A petaflop/s-day (pfs-day) consists of performing 1015 neural net operations per second for one day, or a total of about 1020 operations. The compute-time product serves as a mental convenience, similar to kilowatt-hour for energy. An amount of compute is meant to give an idea of the number of actual operations performed.[4]
The noun compute was relatively uncommon, though not unheard of, before gaining wider usage in fields such as cloud computing and artificial intelligence in the 2010s.[5] Artificial intelligence company OpenAI popularized the concept of compute as a strategic metric for AI progress in 2018.[4]
OpenAI identified two eras of training AI systems in terms of compute-usage. From 1959 to 2012, compute roughly followed Moore’s law.[4] Between 2012 and 2018, the amount of compute used in the largest AI training runs increased exponentially, growing by more than 300,000 times — roughly doubling every 3.4 months.[4][2] By comparison, Moore’s Law doubled every two years over the same period.[4] One of the largest models, released in 2020, used 600,000 times more computing power than the 2012 model.[2]
After 2020, compute growth began to slow down,[2] with the compute needed for the largest AI models continuing to slow down in 2023.[6]
AI provider SpaceXAI said in 2026 that their AI progress is driven by compute and used it as a key metric in the AI training of its supercomputer Colossus, the which contains 1 million GPUs.[7] Anthropic has a contract of $1.25 billion per month with SpaceXAI to buy all the compute capacity at Colossus 1 data center.[8] Since June 2026 ahead of the initial public offering of SpaceX, Google pays SpaceXAI $920 million monthly for cloud compute capacity.[9]
A 2022 study found that current large language models are significantly under-trained, a consequence of focusing on scaling language models whilst keeping the amount of training data constant. By training over 400 language models of various parameter and token size, they found that "for compute-optimal training", the model size and the number of training tokens should ideally be scaled equally: for every doubling of model size the number of training tokens should also be doubled.[10]
Larger AI models trained on more data and using more computational resources, tend to perform better.[11][10] This happens even if the algorithms themselves remain unchanged.[10]
As early as 2018, OpenAI noted the exponential increase in compute to be have a key role in AI progress.[4] OpenAI considers three factors drive the advance of AI: algorithmic innovation, data, and the amount of compute available for training.[4] AI models with more compute not only improve in the tasks they were trained on but can develop emergent abilities.[12] Incremental improvements can lead to more abrupt leaps in capabilities.[11]
Increasing, promoting or constraining progress in artificial intelligence has often been done via controlling the amount of compute.[13] Policymarkers have enacted policies and provided support to make compute resources more accessible to domestic AI researchers.[13]
In a January 2022 report, the Center for Security and Emerging Technology (CSET) suggested to institutions that increasingly powerful and generalizable AI (AGI) will likely require other strategies than maximizing compute.[2][13] Some AI researchers are also concerned that government might exclusively focus on scaling compute instead of other strategies.[13]
The CSET has reported on the various bottlenecks which could explain why deep learning needs for compute have slow down:
For this goal, CSET advised policymakers to ensure that even researchers with smaller budgets could effectively contribute to AI research.[2] Other proposed strategies include using contemporary AI algorithms, managing modern AI infrastructure or focusing on interdisciplinary work between the AI field and other fields of computer science.[2]
A 2024 study on compute access found that academic-only AI research teams often have less compute intensive research topics, especially foundation models, compared to industry AI labs.[14] As a consequence, academia is likely to play a smaller role in advancing such techniques. The researchers suggest nationally-sponsored computing infrastructure as well as open science initiatives to boost academic compute access.[14]
Informasi ini disarikan dari Wikipedia dan disajikan kembali untuk tujuan edukasi. Konten tersedia di bawah lisensi CC BY-SA 3.0. Kami tidak bertanggung jawab atas ketidakakuratan data yang bersumber dari kontribusi publik tersebut.