You will design, build, deploy, and monitor scalable safety infrastructure for content moderation, abuse detection, and agent guardrails. You will create APIs, data pipelines, and service architectures for real-time and batch workflows, establish observability and performance standards, and translate ML research into robust production systems.
Responsibilities
- Design and build scalable backend infrastructure for content moderation, abuse detection, and agent guardrails
- Deploy AI and ML models into production systems
- Architect APIs, data pipelines, and service architectures for real-time and batch moderation workflows
- Implement monitoring, alerting, and observability systems
- Establish SLIs, SLOs, and performance benchmarks
- Partner with ML engineers to translate research models into production-ready systems
- Drive technical decisions and contribute to the safety roadmap
Requirements
- 6+ years of backend software engineering experience building production systems at scale
- Experience with distributed systems, APIs, data pipelines, and Python
- Experience with asynchronous Python and backend frameworks
- Proficiency with cloud platforms such as AWS or GCP
- Proficiency with containerization tools such as Docker or Kubernetes
- Experience with CI/CD pipelines
- Experience with monitoring tools such as Prometheus and Grafana
- Experience deploying or working alongside ML or AI systems in production
Benefits
- Annual discretionary professional development stipend
- Annual discretionary social travel stipend
- Annual company offsite
- Monthly co-working stipend for employees not near a main hub