Cloud SRE & Automation Engineer - High Availability
Euronext · Bardi, Emilia-Romagna ·
- Work mode
- Onsite
- Category
- SRE
Euronext · Bardi, Emilia-Romagna ·
Nexus Recruitment Group · Pasig, Metro Manila, Philippines
instrumental-inc- · Palo Alto, USA
Trimble · UK - Newcastle
Devexperts · Porto, Porto District, pt
Euronext is seeking an experienced Site Reliability Engineer to ensure the reliability and performance of production systems while driving automation and operational excellence. You will collaborate with development, QA and infra teams to implement best practices, respond to incidents and improve CI/CD, configuration management and IaC in a cloud/containerized environment.
The role involves on-call duties, root cause analysis and ongoing improvements to monitoring, incident practices and
Provide day-to-day operational support for production environments to ensure high availability of critical services Develop, maintain and enhance automation scripts/tools using Bash, Python and Ansible to streamline tasks and incident response Monitor system performance, identify issues proactively and implement preventive solutions Collaborate with development, QA and infrastructure teams on deployment, monitoring and incident practices Participate in on-call rotation and perform root cause analysis to drive resolution Maintain and improve configuration management, CI/CD pipelines and infrastructure as code practices Document operational processes, troubleshooting steps and automation workflows Proven experience in a production support or SRE role in a high-availability environment. Strong automation skills with Bash, Python and Ansible. Experience with monitoring/alerting tools such as Prometheus, Grafana, Elastic Stack, Datadog. Solid Linux/Unix systems administration and troubleshooting. Familiarity with AWS and containerization (Docker, Kubernetes). Knowledge of IaC and networking, security and incident management.Devexperts · Tbilisi, Tbilisi, ge