
Want to know if this job is worth applying to?
Hyderabad, Telangana, India
Hybrid
• 2-4 years of experience designing, implementing, and maintaining CI/CD pipelines (e.g., Harness, GitHub Actions, ArgoCD or similar tools).
• Hands‑on experience with automation tools and Infrastructure as Code / Configuration as Code (CloudFormation, Terraform, Ansible).
• Strong understanding of Infrastructure as Code and Configuration as Code principles and patterns.
• Solid grasp of the software development lifecycle and modern SRE/DevOps practices.
• AWS administration and architecture experience, including networking, security, IAM, and core services.
• Experience operating Kubernetes clusters (EKS or other distributions) and containerized workloads.
• Deep experience with monitoring and observability tools such as OpenTelemetry, Groundcover, CloudWatch, Datadog, Prometheus, New Relic, or equivalent, including metrics, logs, and traces.
• Ability to define and track SLIs/SLOs and use them to guide reliability improvements.
• Proficiency in Linux administration, including system configuration, troubleshooting, and performance tuning.
• Programming/scripting skills in at least one language such as Python, Go, or Rust for automation, tooling, and observability integrations.
• Solid understanding of networking, load balancing, and performance tuning.
• Experience troubleshooting complex distributed systems, supporting incident response, and driving root cause analysis.
• Familiarity with risk mitigation, backup, and disaster recovery concepts.
Preferred
• Experience building unified observability platforms or standardized dashboards for multiple services/teams.
• Experience with GitOps workflows and tools for declarative infrastructure and application delivery.
• Background in incident command and post‑mortem frameworks.
• Experience integrating observability and reliability practices into microservices and/or serverless architectures.
• Experience integrating testing, security and compliance checks into CI/CD pipelines.