Software Engineer Data Platform
SpaceXAI · Palo Alto, CA ·
- Category
- Data engineering
You will design, build, and operate distributed systems for large-scale data movement and compute. You will create high-throughput ingestion systems, scale Kafka infrastructure, tune processing engines, build data interfaces and pipelines, optimize reliability and performance, and collaborate on critical data workflows.
Responsibilities
- Design and implement high-throughput, low-latency data ingestion and transport systems
- Scale and optimize multi-tenant Kafka infrastructure
- Extend and tune Spark, Flink, and Trino for production pipelines
- Build interfaces, APIs, and pipelines for petabyte-scale data processing
- Debug and optimize distributed systems for reliability and performance
- Collaborate with ML, product, and infrastructure teams on data workflows
Requirements
- Expertise in distributed systems, stream processing, or large-scale data platforms
- Proficiency in Rust, Go, Scala, or similar systems languages
- Production experience with Kafka, Flink, Spark, Trino, or Hadoop
- Debugging, profiling, and performance optimization skills
- Track record of shipping and maintaining critical infrastructure
Benefits
- Equity
- Medical coverage
- Vision coverage
- Dental coverage
- 401(k) retirement plan
- Short-term disability insurance
- Long-term disability insurance
- Life insurance
- Discounts and perks