Vacancy Description
Responsibilities
- Design, implement, and continuously improve Site Reliability Engineering (SRE) practices to ensure highly available, scalable, and resilient production systems.
- Provide L2 production support by troubleshooting application, infrastructure, and platform issues while ensuring compliance with SLA and SLO commitments.
- Monitor, maintain, and optimize Microsoft Azure environments, including Azure Kubernetes Service (AKS), Azure Monitor, and Log Analytics.
- Support and troubleshoot Java and Spring Boot applications by analyzing logs, JVM performance, REST API issues, configuration problems, and application performance bottlenecks.
- Build, configure, and maintain observability solutions using AppDynamics, Dynatrace, Splunk, Grafana, and Prometheus to improve system visibility and proactive monitoring.
- Develop, enhance, and maintain CI/CD pipelines using Jenkins, Bitbucket, and Azure DevOps to support reliable and ...
Ready to Apply?
अभी आवेदन करें
Submit your application for Site Reliability Engineer (Dynatrace, GenAI, Prometheus, L2 Support, Splunk) at NEPTUNEZ SINGAPORE PTE. LTD.
Apply for this Position