Senior Software Engineer, TorchTPU
نظرة عامة على الوظيفة
-
تاريخ الإعلانمارس 26, 2026
-
الموقع
-
تاريخ إنتهاء الصلاحية--
المسمى الوظيفي
2026-03-17T14:59:08.343Z
106207909217477318
Minimum qualifications:
- Bachelor’s degree or equivalent practical experience.
- 5 years of experience with ML design and ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).
- 5 years of experience in software development.
- 5 years of experience testing, and launching software products, and 3 years of experience with software design and architecture.
Preferred qualifications:
- Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
- 8 years of experience with data structures and algorithms.
- 3 years of experience in a technical leadership role leading project teams and setting technical direction.
- 3 years of experience working in an organization involving cross-functional, or cross-business projects.
- Experience with compilers or ML frameworks.
About the job
The Core ML team contributes to frameworks and compilers that support the Google Cloud Platform (GCP) Cloud Tensor Processing Unit (TPU) service and related Machine Learning (ML) models and frameworks. The team provides ML infrastructure customers with large-scale, cloud-based access to Google’s first-party ML supercomputers to run training and inference workloads using PyTorch and JAX.
In this role, you will be responsible for the PyTorch ML framework, processes, ecosystem, and model performance, and engagements with customers who take advantage of Google’s TPUs to achieve scale and speed in their ML workloads.
We’re the driving force behind Google’s groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.
المسؤوليات
- Work on Artificial Intelligence (AI) framework development to enable PyTorch models to run on Google Cloud’s TPUs and GPUs and tune for performance.
- Provide support for ML frameworks and compilers on Cloud TPUs and Graphics Processing Units (GPUs), enabling the training and deployment of the most advanced machine learning models, managing innovation and breakthroughs.
- Enable PyTorch models for generative models, computer vision (e.g., image recognition, object detection, image generation), machine translation, language modeling, rankings and recommendations, speech recognition, etc.
- Collaborate with other Google teams and leading researchers across the industry to continuously bring ML capabilities to our PyTorch in Cloud offering.
- Design, develop, test, deploy, maintain, and improve software while contributing to open-source software development.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.