App LogoApp name
DP

Devarapu Praveen

Open to work2 years experience

Site Reliability Engineer

Bengaluru, IndiaTarget Roles: Backend Engineer • Full-Stack Developer • Frontend Specialist

Site Reliability Engineer | Kubernetes | OpenShift | DevOps | Linux | Observability

Standing Rank

Rank Not Available

Developer Badges

No badges earned yet

Skills & Technologies

21 skills
kubernetesopenshiftdockerhelmrhel linuxbashpythongitlab ci/cdgitterraformprometheusgrafanalokielk stackelastic apmawsazuremariadbsqlkafkaredis

GitHub Activity Evidence

GitHub Footprint Not Connected

Recruiters value verified code contributions. Linking a GitHub account displays live heatmaps, active contribution metrics, repository languages, and push activity tags.

Work Experience

Site Reliability Engineer

Olive Crypto Systems

Jan 2024 - Present Bengaluru, India
  • Operated production infrastructure for a bank’s CBDC platform across bare-metal and OpenShift environments, supporting reliable service delivery for a nationally scaled digital currency system handling 3L+ transactions/day.
  • Managed Kubernetes/OpenShift workloads including namespaces, deployments, StatefulSets, services, routes, ConfigMaps, Secrets, Helm releases, rolling updates, and failover readiness across 40+ services in UAT, Production, and DR.
  • Built and maintained observability using ELK Stack, Prometheus, and Grafana to centralize logs, metrics, dashboards, health checks, and alerts, reducing average incident detection time by 40%.
  • Led production release validation across DTSP and RTSP layers using log, metric, and service-health checks to verify deployments and support safe rollback when required.
  • Served as an SRE point of contact during critical incidents, coordinating with NPCI, bank infrastructure, network, security, and application teams to restore services within 30-minute SLA.
  • Performed root cause analysis for recurring production issues and translated findings into preventive fixes, monitoring improvements, and operational runbooks.
  • Automated certificate rotation checks, health checks, log collection, gateway validation, and routine support tasks using Bash, Python, and GitLab CI/CD pipelines, cutting manual effort by 30%.
  • Planned and executed backup validation, replication checks, failover testing, and disaster recovery drills across bare-metal, OpenShift, and blockchain nodes.
  • Supported server hardening, vulnerability remediation, VAPT closure, access reviews, and secure configuration of production systems within defined compliance timelines.
  • Deployed and supported Hyperledger Fabric peer nodes, ordering services, certificate authorities, certificates, and network services with node-level monitoring.
  • Participated in on-call rotations, triaged alerts, documented incident timelines, and contributed to blameless post-incident reviews and corrective actions.
  • Contributed to Infrastructure as Code and deployment standardization using Terraform, Helm, Git, and GitLab CI/CD.

Projects

AI-Driven Cloud-Native Observability Platform for Kubernetes

KubernetesOpenShiftDockerHelmJavaRedisKafkaPrometheusGrafanaLokiPromtailELK StackElastic APMAlertmanagerTerraformGitLab CI/CDBashPython
  • Built a production-inspired SRE platform around a containerized Java microservice with Redis caching, Kafka event streaming, and SQL-backed services deployed on Kubernetes/OpenShift.
  • Implemented end-to-end observability with Prometheus, Grafana, Loki, Promtail, ELK Stack, and Elastic APM for metrics, logs, dashboards, tracing, and application performance monitoring.
  • Configured Prometheus service discovery, metric scraping, alert rules, SLI/SLO monitoring, and Alertmanager notifications for Slack, Microsoft Teams, email, and PagerDuty.
  • Built Grafana dashboards for Kubernetes cluster health, JVM metrics, application latency, error rates, pod resource usage, and availability indicators.
  • Implemented Loki and Promtail for label-based Kubernetes log aggregation from application pods and cluster nodes, while using ELK for search, RCA, and operational analysis.
  • Automated deployments using Helm, Terraform, and GitLab CI/CD while managing Deployments, StatefulSets, Services, Ingress, ConfigMaps, Secrets, RBAC, Persistent Volumes, and HPA.
  • Designed AI-assisted workflows for anomaly detection, predictive alerting, root cause analysis, and automated remediation using observability data.
  • Validated resilience through liveness/readiness probes, rolling updates, horizontal scaling, backup checks, and disaster recovery testing.

Leaderboard Standings

Leaderboard Position Pending

Global test scores, peer standing percentiles, and algorithm leaderboard ranks are updated dynamically.

Assessment Highlights

Assessments Not Completed

Coding evaluations, system assessment results, and conceptual score badges will appear here after taking a test.

AI Collaboration Score

AI Collaboration Score Pending

Developer coding behavior, assistant cooperation, and AI pair-programming indicators are evaluated during live coding sessions.

Role Compatibility Profile

Role Compatibility Analysis Pending

Custom matching reports, candidate role compatibility percentiles, and core engineer strength profiles are processed once conceptual code screenings are complete.

Achievements

CBDC Platform Reliability Support

Supported production reliability operations for a bank’s CBDC platform, part of a national digital currency initiative showcased at India’s G20 presidency.

Automation Efficiency Gain

Reduced manual operational effort by 30% through Bash/Python automation of health checks, log collection, and validation.

Monitoring and Alerting Improvement

Strengthened monitoring and alerting coverage using Prometheus, Grafana, and ELK, improving production visibility and incident-response turnaround by 35%.

About Details

Professional Bio

Site Reliability Engineer with 2.6+ years supporting mission-critical banking and blockchain platforms across bare-metal, OpenShift, and hybrid-cloud environments. Expert in incident response, Kubernetes operations, observability, and automating manual tasks through Bash/Python.

B.Tech in Information Technology

GMR Institute of Technology, Rajam, India (2019 - 2023)

Languages: English, Hindi