CareerScanCareerScan
JobsCompanies
BlogContact
For Employers
Sign InRegister Free
CareerScanCareerScan

India's verified job platform connecting candidates directly with employers. 100% free applications with instant ATS resume scoring.

Chennai, Bengaluru & Hyderabad
Jobs by location
Jobs in ChennaiJobs in BengaluruJobs in HyderabadJobs in PuneJobs in Mumbai
Popular roles
AR Caller JobsHealthcare Medical CodingReact / Full Stack DeveloperData & Power BI AnalystCustomer Support Executive
Top companies
TCS CareersCognizant JobsInfosys OpeningsApollo HospitalsOmega Healthcare
Career services
Free ATS Resume CheckerAI Resume Builder (Free)AI Job MatcherSalary Guide & BenchmarksJob Alerts on WhatsApp
© 2026 CareerScan India. All rights reserved.256-bit SSL encrypted & verified
Back to all jobs
  1. Home
  2. Jobs
  3. Principal ML Performance Engineer (GPU Optimization)
P
Proxima

Principal ML Performance Engineer (GPU Optimization)

New York
₹35/mo
6+ years exp
Full-time
Posted 3d ago
0 views
Actively Hiring Urgent Opening Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Principal ML Performance Engineer (GPU Optimization) role at Proxima.

Apply for this position

Apply on Company Website
Notice a broken link or wrong info?

Job Description

Principal ML Performance Engineer (GPU Optimization)

About Proxima

Proxima is a frontier AI and data generation company discovering the next generation of proximity therapeutics by making protein interactions programmable. Our platform brings together foundation-model machine learning, a scalable data generation engine, and a partnership track record exceeding $5B in collaborations across the world’s leading biopharma and tech organizations. We’ve recently closed an oversubscribed seed round with an elite group of VCs including DCVC, NVIDIA’s NVentures, AIX, Yosemite among others.

Neo-1 is our all-atom foundation model that combines state-of-the-art structure prediction and molecular generation in a single system. Neo-1 enables rapid exploration of chemical and structural space for high value, previously intractable targets, and in particular unlocks small molecule proximity therapeutics like molecular glues with AI for the first time.

In parallel, we are developing an advanced structural interactomics platform built on proprietary XLMS technology and a lab equipped with next-generation mass spectrometry instrumentation. This platform produces proteome-scale maps of protein interactions and helps identify small molecules that modulate proximity. Together with Neo-1, it creates an integrated system capable of co-folding protein complexes while generating candidate small molecules to influence those interactions.

Proximity-based therapeutics represent of the most promising frontiers in modern drug discovery with the potential to treat previously intractable diseases and target ‘undruggable’ proteins. We’re building the tech and the team to make that happen. Come join us!

What you'll do

  • Profile and optimize training and inference for structural and generative models, including transformers, diffusion, and geometric deep learning

  • Write and tune custom kernels (CUDA, Triton) and use compilers (torch.compile, TensorRT, XLA) when beneficial

  • Scale distributed training across 32-64 nodes, employing FSDP, DeepSpeed, tensor and pipeline parallelism, and mixed precision

  • Reduce inference cost by optimizing memory scaling for large complexes, improving diffusion sampling efficiency, batching ragged inputs, and maximizing throughput across up to 1000 GPUs

  • Manage GPU cluster efficiency on GCP, focusing on scheduling, utilization, spot strategy, and cost reporting

  • Develop benchmarks and profiling tools for the research team

What we need

  • Minimum of 6+ years experience in ML systems, HPC, or performance engineering, with a BS/MS/PhD in CS, EE, or related field

  • Demonstrated ability to set technical direction beyond coding: selecting infrastructure, influencing research teams, and mentoring engineers

  • Deep knowledge of PyTorch internals with hands-on experience profiling and fixing real bottlenecks

  • Experience with CUDA and Triton, skilled at reading Nsight output, and strong understanding of memory bandwidth and occupancy

  • Experience with distributed training at multi-node scale

  • Strong proficiency in Python and C++

  • Able to name a model they made materially faster and quantify the improvement

Nice to haves

  • Experience in geometric deep learning, equivariant networks, or protein structure models such as AlphaFold, ESM, or RFdiffusion

  • Experience writing kernels for structure-model primitives, including triangle attention, triangle multiplicative updates, cuEquivariance, or FlashAttention for pair bias

  • Experience orchestrating large batch inference and managing Kubernetes GPU scheduling

Frequently Asked Questions

How to apply for Principal ML Performance Engineer (GPU Optimization) at Proxima?

Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.

What is the salary for this role?

The salary for this role is $5 per annum.

What experience is required?

6+ years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

ApplicationActively Hiring
Apply on Company Website
Broken link or expired?
P

Proxima

Share this Opening

Job Alerts for management

Receive email alerts whenever new management roles in New York are posted.

Set Free Alert →

Similar Openings

Explore related active roles in management

View all
UrgentActively Hiring
Granicus
Principal Consultant - Communications & Experience Hybrid (Sacramento, CA based; client onsite)
Granicus Verified
5+ years
₹8.9L/mo
Remote, US
managementFull-timeRemote
Posted 1d ago
Apply Now
UrgentActively Hiring
Lts
Vice President - Financial Planning & Analysis
Lts Verified
12+ years
₹13.8L – ₹19L/mo
United States - Remote
managementFull-timeRemote
Posted 1d ago
Apply Now
UrgentActively Hiring
Jobber
Principal Engineer (AI, Office of the CEO)
Jobber Verified
0-2 Yrs
₹15.9L/mo
Boston
managementFull-time
Posted 1d ago
Apply Now

Principal ML Performance Engineer (GPU Optimization)

Proxima · New York

Apply on Company Website