Member of Technical Staff, ML Systems
Transparent Search Group · Menlo Park, California, United States ·
- Employment
- Full time
- Category
- Software engineering
- Experience
- 1+ years
- Visa sponsorship
- Yes
- Salary
- USD 180,000 – 230,000 / year
Transparent Search Group · Menlo Park, California, United States ·
LinkedIn · Mountain View, CA, us
General Motors · Sunnyvale, California, United States of America
preference-model · San Francisco
ML systems engineer at TensorScale who optimizes speed and efficiency across the training and inference stack for video, image, and world-model workloads. Day-to-day work involves writing CUDA/Triton GPU kernels, profiling with Nsight, and building distributed multi-node inference and training systems. Onsite in Menlo Park 5 days a week.
Company: TensorScale
Location: Menlo Park, CA (in office 5 days per week)
Compensation: $180,000 - $230,000 + competitive equity
Employment Type: Full-time
Visa Sponsorship: Visa transfers and new sponsorship (H-1B, TN)
TensorScale builds the training and inference stack for world models. Today's stack was built for language models; TensorScale is rebuilding it for video, image and world-model workloads by co-designing low-level GPU kernels, distributed systems and the models themselves. Its public benchmarks include MiniMax running roughly ten times faster at half the cost, 2K image generation in about four seconds for three cents, and LTX video running faster than real time.
TensorScale has five founders covering distributed systems, GPU kernel optimization, cloud infrastructure and research, with prior work at Fireworks AI, Meta, Google, Apple, Microsoft, Snowflake and Alibaba. It has raised a $10M seed.
TensorScale is hiring ML systems engineers to own speed and efficiency across its stack: low-level kernels, distributed inference engines and multi-node training and serving systems. You report directly to the cofounder and CEO. The work sits below the application layer (kernels, runtimes and distributed engines for video and world models), so this is not an agents or RAG role. You will feel at home if you would rather make a video model ten times faster than train one.
Hiring manager screen with the CEO (30 min), domain deep dive (60 min), system design (60 min), optional onsite.
CUDA, Triton, PyTorch, Nsight, NCCL, RDMA
Apple · New York City
Unconventional, Inc. · Mountain View, CA or Greater Seattle Area