NS Demo
Member of Technical Staff - ML Performance
Full Time · In Office · New York, New York (USA)
Posted Aug 5, 2026
Home Jobs Member of Technical Staff - ML Performance
M
Member of Technical Staff - ML Performance
Modal
On-site New York Full-time$200K – $350K
Offers Equity 30 days ago
Apply on Modal
About Us
AI needs a new infrastructure layer. We're building it at Modal.
Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now.
Our customers include category-defining companies like Lovable, Ramp, Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale.
We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September.
Our team includes creators of popular open-source projects (e.g.,Seaborn,Luigi), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience.
The Role
We are looking for strong engineers with experience in making ML systems performant at scale. If you are interested in contributing to open-source projects and Modal’s container runtime to push language and diffusion models towards higher throughput and lower latency, we’d love to hear from you!
Requirements
- 5+ years of experience writing high-quality, high-performance code.
- Experience working with torch, high-level ML frameworks, and inference engines (vLLM or TensorRT).
- Familiarity with Nvidia GPU architecture and CUDA.
- Experience with ML performance engineering (tell us a story about boosting GPU performance - debugging SM occupancy issues, rewriting an algorithm to be compute-bound, eliminating host overhead, etc).
- Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc).
Interested in this role?
Applications are handled on Modal's site.
Apply Now
Related Jobs
View all
AI Engineering
S
Senior Product Engineer, Growth & Lifecycle Infrastructure - Music & Audio
Stability AI
Remote
Full-time
T
AI Researcher, Core ML (Turbo)
Together AI
On-site
Full-time
C
Account Solution Architect
CoreWeave
On-site
Full-time
T
Backend Software Engineer - Data Platform & AI Data Products
Together AI
On-site
Full-time
S
AI Applications Ops Lead, GPS
Scale AI
On-site
Full-time
P
Member of Technical Staff (Machine Learning Engineer, Search)
Perplexity
Remote
Full-time
Back to all jobs
Related Jobs
View all
AI Engineering
S
Senior Product Engineer, Growth & Lifecycle Infrastructure - Music & Audio
Stability AI
Remote
Full-time
T
AI Researcher, Core ML (Turbo)
Together AI
On-site
Full-time
C
Account Solution Architect
CoreWeave
On-site
Full-time
T
Backend Software Engineer - Data Platform & AI Data Products
Together AI
On-site
Full-time
S
AI Applications Ops Lead, GPS
Scale AI
On-site
Full-time
P
Member of Technical Staff (Machine Learning Engineer, Search)
Perplexity
Remote
Full-time
Mention you found this on Data First Jobs — it helps us bring you more roles like this.
Member of Technical Staff - ML Performance
NS Demo
Similar Other Jobs
View all Other jobs→Everway
Director, Analytics & Decision Intelligence
LP Analyst
Private Equity Data Operations, Investor Services Lead
CACI International Inc
IT Data Specialist– eDiscovery
Harvey Nash
Head of Data
Trova
Business Intelligence Architect
BMO
Facility Operator Data Center
Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.
Free, no spam. Unsubscribe anytime.