Senior Software Engineer, Core Infrastructure
Oracle · Nashville, TN, United States ·
- Seniority
- Senior
- Category
- Software engineering
- Experience
- 4+ years
Oracle · Nashville, TN, United States ·
Ll Oefentherapie · Dover, Canada
puck · Worthing, South Dakota, United States
OCLC · Minnesota
LinkedIn · Mountain View, CA, us
Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability. Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers’ implementations for scalability compliance. Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery‑oriented principles, and implements retries, circuit breakers, and timeouts. Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs. Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met.
As a Senior Core Infrastructure Engineer, you will build and operate the foundational distributed systems behind OCI. You will develop scalable, resilient components for high-volume data retrieval, storage, and processing, with a focus on correctness, security, observability, and dependable operation at global scale.
This position is office-based and requires onsite presence in Nashville, Tennessee. Relocation assistance may be available in accordance with Oracle's relocation policies.
What You'll Do
Design, implement, test, and optimize components of large-scale distributed systems
and data-plane services.
Build for horizontal and vertical scale, using distributed state, replication,
synchronization, and high-volume data-processing patterns.
Strengthen service reliability through redundancy, automatic failover, recovery-oriented
design, retries, circuit breakers, and timeouts.
Develop and execute performance, load, and fault-injection tests to validate scalability,
correctness, availability, and operational readiness.
Create dashboards, alarms, telemetry, runbooks, and automation that proactively detect
issues and accelerate recovery.
Diagnose production issues, participate in on-call rotations and incident response, and
contribute to root-cause analyses and durable corrective actions.
Develop infrastructure automation and Infrastructure as Code (IaC) for safe
maintenance, troubleshooting, patching, updates, and rollbacks.
Apply strong security controls—including encryption, access controls, and vulnerability
remediation—in multi-tenant environments.
What You'll Bring
Bachelor's or master's degree in Computer Science, Computer Engineering, or a related
field, or equivalent practical experience.
4+ years of professional software-development experience; advanced degree holders
may have fewer years of experience.
Proficiency in at least one object-oriented or systems programming language, such as
Java, C++, C#, or Go.
Experience building or operating scalable, highly available distributed systems or cloud
infrastructure.
Experience with system-level testing and automation, including performance, load,
reliability, or fault-injection testing.
Understanding of distributed-systems concepts, data structures, algorithms, operating systems, networking, and secure software-development practices.
Strong debugging, problem-solving, communication, and cross-functional collaboration skills.
Preferred Qualifications
Experience with cloud platforms, such as Oracle Cloud, AWS, Azure, or Google Cloud.
Experience with data-plane platforms, distributed storage, microservices, state
management, replication, or large-scale data processing.
Experience with observability, incident management, Infrastructure as Code, and
production service operations.
Familiarity with compliance requirements and security controls for cloud infrastructure.
Rocket Companies · Remote - Michigan