System Software Engineer (gn) Cloud & Simulation
Wandelbots · Dresden ·
- Category
- Software engineering
- Company size
- 51-200
Wandelbots · Dresden ·
NavVis · Munich Hybrid (NavVis GmbH)
ncs3 · Singapore, sg
D3 Embedded · West Henrietta, NY
NVIDIA · US, CA, Santa Clara
Build and operate GPU-intensive infrastructure for robot simulation and 3D reconstruction. You will manage Kubernetes-based workloads, NVIDIA Omniverse/Isaac Sim platforms, and CI/CD pipelines using Python, IaC, and GitOps to productize simulation technologies for industrial robotics.
Build and operate the GPU infrastructure required by our simulation workloads, including NVIDIA GPU Operator, device plugins, driver lifecycle, and GPU sharing through MIG or time slicing
Work with the infrastructure team to integrate simulation workloads into our shared Kubernetes and cloud infrastructure
Make deployments declarative and reproducible using Infrastructure as Code and GitOps, and help consolidate the tooling used across today’s environments
Own CI/CD and the container and image lifecycle for CUDA- and NGC-based workloads
Design scheduling and autoscaling for two distinct workload profiles: latency-sensitive interactive streaming sessions and compute-intensive Gaussian Splatting jobs
Turn the Gaussian Splatting pipeline into a reproducible workflow, from data capture and training to OpenUSD assets
Operate and extend our Omniverse and Isaac Sim streaming platform, including Kit App Streaming, session lifecycle, and tenant isolation
Establish the platform as an internal standard through self-service workflows, golden paths, templates, onboarding, and documentation
Prepare the solution for operation by partners and customers through packaged deployments, versioned releases, upgrade paths, and actionable diagnostics
Support deployments across cloud, on-premises, and partner-managed environments
Build meaningful observability using metrics, logs, GPU telemetry, and actionable alerting
Define the simulation-specific security and access model, including ingress, TURN and STUN for WebRTC, RBAC, secrets management, SSO, and tenant boundaries
Work closely with our robotics, product, and platform teams to ensure that the solution supports real development and customer workflows
We are looking for an engineer who is interested in how complex systems are built, deployed, and operated, not only in delivering the next application feature.
You should bring:
Experience building or operating production systems in platform engineering, infrastructure, SRE, DevOps, or systems software
Practical Kubernetes experience, including resource management, workload scheduling, and debugging distributed systems under load
Experience with Infrastructure as Code, GitOps, and CI/CD, together with the ability to evaluate tools based on the problem rather than a fixed preference
Solid knowledge of Linux, containers, and networking
Good Python skills for automation, services, and data-processing workflows
Experience with GPU workloads, or a strong technical foundation and clear interest in working with drivers, CUDA, and GPU scheduling
The ability to turn working technology into a system that other teams can use independently
A strong sense of operational ownership, including documentation, diagnostics, maintainability, and reproducibility
Very good English; German is a plus
We do not expect candidates to have prior experience with every technology in this description. A strong systems foundation, the ability to learn unfamiliar components, and an interest in this problem space are more important than matching every item. The scope and responsibilities of the role will grow with your experience.
Experience turning an internally developed system into a product that can be operated in customer-controlled environments
Hands-on experience with NVIDIA Omniverse, Isaac Sim, Kit SDK, or OpenUSD
Experience with Gaussian Splatting, NeRF, photogrammetry, or comparable 3D reconstruction pipelines
Knowledge of WebRTC, pixel streaming, or other low-latency video technologies
Python, Rust, or Go skills sufficient to contribute to existing services and infrastructure tooling
Experience with robotics or simulation, such as ROS 2 or sim-to-real workflows
Experience with multi-tenant platforms, air-gapped deployments, or customer environments with specific security and compliance requirements
Bot Auto · Houston, TX or San Francisco Bay Area