Understanding and Fixing Bottlenecks in Optimization for Modern Machine Learning
Location: France
Source: EU Funding & Tenders Portal
Modern machine learning models have been successfully deployed across fields, from scientific studies to tech- nological developments in industry, but their development remains poorly understood. The training of a large language model such as GPT-3 is estimated to cost $4.6M, and public attempts to replicate the training process alone required teams of engineers to rotating on-call for months, mon
Project Information FAQ
Project Information
Want to explore the full details? View the full report
Participants
Sponsoring Agency | Obfuscated Data |
Company | Obfuscated Data |
Status
Original status | ongoing |
Taiyo status | Obfuscated Data |
Taiyo last update | 00-00-0000 |
Available timestamps | 00-00-0000 |
Available timestamp type | Obfuscated Data |
Contact
Contact name | Obfuscated Data |
Phone | 0000000000 |
ObfuscatedData@email.com | |
Address | Obfuscated Data, Obfuscated data, obfuscated data, Obfuscated data |
Description
Description | Modern machine learning models have been successfully deployed across fields, from scientific studies to tech- nological developments in industry, but their development remains poorly understood. The training of a large language model such as GPT-3 is estimated to cost $4.6M, and public attempts to replicate the training process alone required teams of engineers to rotating on-call for months, monitoring various statistics and constantly tweaking the training procedure when it broke. Existing theoretical frameworks offer limited insights into this process, as they do not capture the main difficulties that arise in practice when training neural networks, leaving practitioners to rely on error-prone heuristics and expensive trial-and-error. This leads not only to a large devel- opment cost dominated by wasted resources, but also limits the possible impacts of machine learning to areas considered profitable by industries that have the resources to carry this development. The objective of this project is to build a better understanding of how recently identified bottlenecks in neural network training slow down optimization and how to adress them. The specific aims are to: (a) Understand the impact of class imbalance on the dynamics of neural networks to identify where to allocate algorithmic resources. (b) Develop a theory to capture optimization difficulties early in training to guide the development of algorithms that improve performance during this crucial phase. (c) Identify new bottlenecks that arise from applications to new data types. The project combines experimental expertise of the postdoctoral and the theoretical expertise of the host insti- tution to identify and describe the real impact of data characteristics on neural network training. Understanding these bottlenecks will help develop more efficient and reliable algorithms and guidelines on best practices that depend on properties of the data. |
Original sub-sector | Obfuscated |
Original Currency | USD |
Original budget | 000000000000000 |
Procurement method | Obfuscated Data |
Budget | 000000000000000 |
Location
Region | Obfuscated |
Country | Obfuscated |
State | Obfuscated Data |
County | Obfuscated |
Location | Obfuscated Data, Obfuscated data, obfuscated data, Obfuscated data |
Source
Source reliability | High |
Data quality score | 100% |
Source | Obfuscated Data |
URL | obfuscated_data,obfuscateddata.com |
More Details
Project Type | Obfuscated Data |
Article Published Date | Obfuscated Data |
