Site Reliability Engineer
Marmara Bölgesi, Turkey (Türkiye)
Listed on 2026-08-14
-
IT/Tech
SRE/Site Reliability
Site Reliability Engineer (SRE) – Observability & Elastic
Site Reliability Engineer (SRE) – Observability & Elastic
The Team You Will JoinYou will be part of a high-performing engineering organization responsible for delivering resilient, secure, and observable platforms. Our team works at the intersection of software engineering, infrastructure, and security, ensuring that critical systems are highly available, well‑monitored, and continuously optimized. You will collaborate closely with product teams, security experts, and platform engineers to build a strong observability and reliability culture across the organization.
TheOpportunity
- Work on enterprise‑scale, mission‑critical systems serving real business operations
- Build and enhance observability capabilities (logs, metrics, traces) to improve system reliability and transparency
- Utilize AI‑powered analytics on observability data (logs, metrics, traces) to detect anomalies, accelerate root cause analysis, and improve operational intelligence
- Design and implement scalable and resilient platform solutions in cloud and hybrid environments
- Collaborate with cross‑functional teams (development, infrastructure, security) to improve system reliability and performance
- Contribute to automation‑first operations, reducing manual effort and increasing efficiency
- Gain hands‑on experience with modern SRE practices including SLOs, incident management, and reliability engineering
- Participate in building a data‑driven engineering culture using observability insights
- Be part of a global organization with modern engineering standards, tools, and practices
- Design, implement, and manage end‑to‑end observability solutions (metrics, logs, traces)
- Build and maintain Elastic Stack (ELK / Open Search) based logging and monitoring platforms
- Develop dashboards, alerts, and visualization layers for proactive issue detection
- Define and continuously improve SLIs, SLOs, and alerting strategies
- Enable log, metric, and trace correlation to improve troubleshooting efficiency
- Ensure high availability, scalability, and performance of distributed systems
- Drive adoption of reliability practices such as incident retrospectives and proactive monitoring
- Participate in incident response, root cause analysis, and resilience improvement initiatives
- Implement automated remediation and self‑healing mechanisms
- Integrate security monitoring and logging (SIEM‑like use cases) into observability platforms
- Collaborate with security teams on threat detection, anomaly monitoring, and audit logging
- Contribute to Dev Sec Ops practices, embedding security into CI/CD pipelines
- Support audit readiness and compliance reporting through structured logging and monitoring
- Automate operational workflows to reduce toil and increase efficiency
- Contribute to the improvement of CI/CD pipelines and release processes
- Support on‑call operations and continuously improve alert quality and signal‑to‑noise ratio
- Develop Python scripts for synthetic monitoring and testing
- Bachelor’s degree in Computer Science, Engineering, or related field
- 3+ years of experience in SRE, Dev Ops, or production engineering roles
- Good command of English
- Strong hands‑on experience with Elastic Stack (Elasticsearch, Logstash, Kibana)
- Proficiency in other monitoring tools (e.g., Prometheus, Grafana, Azure Monitor, App Insights, Splunk)
- Experience with observability frameworks (metrics, distributed tracing, logging)
- Experience working with cloud platforms (Azure preferred)
- Strong scripting/programming skills (Python, Bash, etc.)
- Understanding of distributed systems and microservices architecture
- Solid understanding of security logging, audit trails, and system hardening
- Experience in tools such as Visual Studio, Azure Dev Ops, Git Hub Enterprise, Git Lab, CI/CD
- Experience working in Financial Services / Insurance sector is an advantage
- Experience building advanced automation scripts or tooling is a plus
- Experience of working in an Agile environment and using Agile methodologies
- Opportunity to work on enterprise‑scale, mission‑critical systems
- Ownership of advanced observability and monitoring platforms
- A culture of engineering excellence, automation, and continuous improvement
- Collaboration with global teams and exposure to modern SRE practices
- Continuous learning and professional growth opportunities
Our benefits are designed to care for your holistic well‑being with programs for physical and mental health, financial wellness, and support for families.
We offer private health insurance for you and your family, life insurance, employer pension plan, meal and transportation allowance, as well as a work from home allowance. We also provide a cultural Heritage Day off, and “back to school” and “school report day” leaves and much more!
#J-18808-LjbffrTo Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: