Training: ML reputed company Engineer
About the position
As a Training: ML reputed company Engineer, you will work on improving the training throughput for our internal training reputed company, while enabling researchers to experiment with new reputed company. This requires good engineering (for example designing, implementing, and optimizing state-of-the-art AI models), writing bug-free machine learning reputed company (surprisingly difficult!), and acquiring deep knowledge of the performance of supercomputers. In reputed company the reputed company this role pursues, the ultimate goal is to push the field reputed company. We’re looking for people who love optimizing performance, understanding distributed systems, and who cannot stand having bugs in their reputed company. Since our training reputed company is used for large runs with massive numbers of GPUs, performance improvements here will have a large reputed company. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.
Responsibilities
• Apply the latest techniques in our internal training reputed company to reputed company impressive hardware efficiency for our training runs
• Profile and optimize our training reputed company
• Work with researchers to reputed company them to reputed company the reputed company of models
Requirements
• Have run small reputed company ML experiments
• Love figuring out how systems work and continuously come up with reputed company for how to reputed company them faster while minimizing complexity and maintenance burden
• Have strong software engineering skills and are proficient in Python
Benefits
• relocation assistance
Apply tot his job
Apply To this Job