Ir para o conteúdo

Informações gerais

Location
Istanbul, Turquia
Arranjo de trabalho
Tempo integral
Modelo de Trabalho
Híbrido
Assistência de realocação disponível
Não
Data de publicação
13-Jul-2026
ID da vaga
18491

Descrição e requisitos

Site Reliability Engineer (SRE) – Observability & Elastic

 

The Team You Will Join
You will be part of a high-performing engineering organization responsible for delivering resilient, secure, and observable platforms. Our team works at the intersection of software engineering, infrastructure, and security, ensuring that critical systems are highly available, well-monitored, and continuously optimized.
You will collaborate closely with product teams, security experts, and platform engineers to build a strong observability and reliability culture across the organization.

 

The Opportunity
• Work on enterprise-scale, mission-critical systems serving real business operations
• Build and enhance observability capabilities (logs, metrics, traces) to improve system reliability and transparency
• Utilize AI-powered analytics on observability data (logs, metrics, traces) to detect anomalies, accelerate root cause analysis, and improve operational intelligence
• Design and implement scalable and resilient platform solutions in cloud and hybrid environments
• Collaborate with cross-functional teams (development, infrastructure, security) to improve system reliability and performance
• Contribute to automation-first operations, reducing manual effort and increasing efficiency
• Gain hands-on experience with modern SRE practices including SLOs, incident management, and reliability engineering
• Participate in building a data-driven engineering culture using observability insights
• Be part of a global organization with modern engineering standards, tools, and practices

 

How You’ll Help Us Build a Confident Future (Key Responsibilities)

Observability & Monitoring
• Design, implement, and manage end-to-end observability solutions (metrics, logs, traces)
• Build and maintain Elastic Stack (ELK / OpenSearch) based logging and monitoring platforms
• Develop dashboards, alerts, and visualization layers for proactive issue detection
• Define and continuously improve SLIs, SLOs, and alerting strategies
• Enable log, metric, and trace correlation to improve troubleshooting efficiency

Reliability Engineering
• Ensure high availability, scalability, and performance of distributed systems
• Drive adoption of reliability practices such as incident retrospectives and proactive monitoring
• Participate in incident response, root cause analysis, and resilience improvement initiatives
• Implement automated remediation and self-healing mechanisms

Security & DevSecOps
• Integrate security monitoring and logging (SIEM-like use cases) into observability platforms
• Collaborate with security teams on threat detection, anomaly monitoring, and audit logging
• Contribute to DevSecOps practices, embedding security into CI/CD pipelines
• Support audit readiness and compliance reporting through structured logging and monitoring

Automation & Platform Engineering
• Automate operational workflows to reduce toil and increase efficiency
• Contribute to the improvement of CI/CD pipelines and release processes
• Support on-call operations and continuously improve alert quality and signal-to-noise ratio
• Develop Python scripts for synthetic monitoring and testing

 

What You Need to Succeed (Required Qualifications)
• Bachelor’s degree in Computer Science, Engineering, or related field
• 3+ years of experience in SRE, DevOps, or production engineering roles
• Good command of English
• Strong hands-on experience with Elastic Stack (Elasticsearch, Logstash, Kibana)
• Proficiency in other monitoring tools (e.g., Prometheus, Grafana, Azure Monitor, App Insights, Splunk)
• Experience with observability frameworks (metrics, distributed tracing, logging)
• Experience working with cloud platforms (Azure preferred)
• Strong scripting/programming skills (Python, Bash, etc.)
• Understanding of distributed systems and microservices architecture
• Solid understanding of security logging, audit trails, and system hardening

 

What Can Give You an Edge (Additional Skills)
• Experience in tools such as Visual Studio, Azure DevOps, GitHub Enterprise, GitLab, CI/CD
• Experience working in Financial Services / Insurance sector is an advantage
• Experience building advanced automation scripts or tooling is a plus
• Experience of working in an Agile environment and using Agile methodologies

 

What We Offer
• Opportunity to work on enterprise-scale, mission-critical systems
• Ownership of advanced observability and monitoring platforms
• A culture of engineering excellence, automation, and continuous improvement
• Collaboration with global teams and exposure to modern SRE practices
• Continuous learning and professional growth opportunities

Benefits We Offer

Our benefits are designed to care for your holistic well-being with programs for physical and mental health, financial wellness, and support for families. 

We offer private health insurance for you and your family, life insurance, employer pension plan, meal and transportation allowance, as well as a work from home allowance. We also provide a cultural Heritage Day off, and “back to school” and “school report day” leaves and much more!


About MetLife

Recognized on Fortune magazine's list of the "World's Most Admired Companies" and Fortune World’s 25 Best Workplaces™, MetLife, through its subsidiaries and affiliates, is one of the world’s leading financial services companies; providing insurance, annuities, employee benefits and asset management to individual and institutional customers. With operations in more than 40 markets, we hold leading positions in the United States, Latin America, Asia, Europe, and the Middle East.

Our purpose is simple - to help our colleagues, customers, communities, and the world at large create a more confident future. United by purpose and guided by our core values - Win Together, Do the Right Thing, Deliver Impact Over Activity, and Think Ahead - we’re inspired to transform the next century in financial services. At MetLife, it’s #AllTogetherPossible. Join us!