DeepMind
Company
Software Engineer, Large Scale Pre-Training Performance
Job Description
This job posting has expired and no longer accepting applications.
We’re looking for a Software Engineer to re-define efficient training of frontier LLMs at massive scale. This role offers an opportunity to influence the design of frontier LLM models, and drive an effort to ensure efficient training and inference.
Key responsibilities:
-
Being responsible for Pre-Training efficiency and optimising the performance of the latest models on Google’s fleet of hardware accelerators - throughout the entire LLM research, training and deployment lifecycle.
-
Being responsible for guiding model design to ensure inference-efficiency.
-
Greatly improving the performance of LLM models on hardware accelerators by optimizing at all levels, including developing custom kernels when necessary.
-
Collaborating with the compiler, framework, and platform teams. And ensure efficient training at industry-largest scale.
-
Profile models to identify performance bottlenecks and opportunities for optimization.
-
Develop low-level custom kernels for maximum performance of the most critical operators.
-
Collaborating with research teams by enabling new critical operators in advance of their availability in frameworks and compilers.
You're an engineer looking to re-define efficient training of frontier LLMs at massive scale and have:
-
A proven track record of critical contributions to the distributed training of LLMs at 1e25 FLOPs scale on modern GPU/TPU clusters
-
Experience in programming hardware accelerators GPU/TPUs via ML frameworks (e.g. JAX, PyTorch) and low-level programming models (e.g. CUDA, OpenCL)
-
Experience in leveraging custom kernels and compiler infrastructure to improve performance on hardware
-
Experience with Python and neural network training (publications, open-source projects, relevant work experience, etc.)
The US base salary range for this full-time position is between $235,000 - $350,000 + bonus + equity + benefits. Your recruiter can share more about the specific salary range for your targeted location during the hiring process.
Application deadline: March 12, 2025
Note: In the event your application is successful and an offer of employment is made to you, any offer of employment will be conditional on the results of a background check, performed by a third party acting on our behalf. For more information on how we handle your data, please see our Applicant and Candidate Privacy Policyopen_in_new.
At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunity regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know.
Please mention that you found this job on MoAIJobs, this helps us grow. Thank you!
DeepMind
102 jobs posted
About the job
Similar Jobs
Discover more opportunities that match your interests
- 21 days ago
Research Engineer, Pre-training
Anthropic
RemoteView details - 21 days ago
Research Engineer, Pre-training
Anthropic
London, UKView details - 21 days ago
Research Engineer / Research Scientist, Pre-training
Anthropic
Zürich, CHView details - 2 days ago
Machine Learning Engineer - Pre-Training
Wayve
LondonView details - 11 days ago
Senior Software Engineer, ML Training Platform
DoorDash
San Francisco, CA; Sunnyvale, CA; Seattle, WAView details - 11 days ago
Senior/Staff Machine Learning Engineer, Training Runtime Performance
Nuro
Mountain View, California (HQ)View details - 29 days ago
Data Software Engineer - Data Scientist
Jump Trading
ShanghaiView details - 29 days ago
Software engineer, generative AI
Writer
San Francisco, CAView details - 22 days ago
Robot Network Software Engineer
Figure
San Jose, CAView details - 21 days ago
Software Engineer, AI Agents
Replicant
CanadaView details
Looking for something different?
Browse all AI jobs