App LogoApp name
OJ

Om Jaju

Open to work2 years experience

Site Reliability Engineer

Hyderabad, IndiaTarget Roles: Backend Engineer • Full-Stack Developer • Frontend Specialist

DevOps / Site Reliability Engineer

GitHub

Standing Rank

Rank Not Available

Developer Badges

No badges earned yet

Skills & Technologies

17 skills
awsazureterraformkubernetesdockerhelmjenkinsgithub actionsargocdprometheusgrafanasplunkpythonbashlinuxgitsql

Work Experience

Site Reliability Engineer I

Techdome

Jul 2026 - Present Hyderabad, India
  • Built 6 monitoring dashboards in Azure Monitor / Application Insights, uncovering 14+ critical bugs across production systems within the first weeks of onboarding.
  • Logged and triaged 14+ issues via Jira, assigning to dev teams for remediation; now leading a company-wide infrastructure audit to drive observability coverage toward 99.95%.

Site Reliability Engineer

Morgan Stanley (via mthree)

Apr 2025 - May 2026 Mumbai, India
  • Increased monitoring coverage from 30% → 99% by implementing Prometheus-based monitoring, alerting, and health checks across 7,560+ critical production systems.
  • Eliminated 120+ recurring production alerts through root cause analysis (RCA), remediation, and reliability improvements.
  • Reduced incident resolution time (TTR) by 60%+ by authoring 79+ SOPs and runbooks.
  • Automated directory provisioning and access management using Python, eliminating 16–20 manual operational requests/month.

Product Specialist (L2/L3 Support)

DarwinBox

May 2024 - Jan 2025 Hyderabad, India
  • Resolved 1,000+ production support tickets, performing triage, root cause analysis, and cross-team coordination in a distributed system environment.
  • Maintained 85–95% SLA adherence while resolving L2/L3 issues within a 3-hour SLA window for enterprise clients.

Projects

Full Stack DevOps Project

View Project
GitHub ActionsFlaskDockerKubernetesHelmArgoCD
  • Built an end-to-end CI/CD pipeline using GitHub Actions, containerized a Flask application with Docker, and deployed to Kubernetes using Helm and ArgoCD GitOps.

AWS Infrastructure Provisioning with Terraform

View Project
TerraformAWSS3DynamoDB
  • Built 4 reusable Terraform modules to provision highly available AWS infrastructure with multi-environment strategy; implemented remote state management using S3 and DynamoDB.

Leaderboard Standings

Leaderboard Position Pending

Global test scores, peer standing percentiles, and algorithm leaderboard ranks are updated dynamically.

Assessment Highlights

Assessments Not Completed

Coding evaluations, system assessment results, and conceptual score badges will appear here after taking a test.

AI Collaboration Score

AI Collaboration Score Pending

Developer coding behavior, assistant cooperation, and AI pair-programming indicators are evaluated during live coding sessions.

Role Compatibility Profile

Role Compatibility Analysis Pending

Custom matching reports, candidate role compatibility percentiles, and core engineer strength profiles are processed once conceptual code screenings are complete.

Achievements

Monitoring and Incident Resolution Improvements

Increased monitoring coverage from 30% to 99% and reduced incident resolution time (TTR) by 60%+.

About Details

Professional Bio

DevOps / Site Reliability Engineer with production experience at Morgan Stanley, focused on improving system reliability, observability, and operational automation. Experienced in Kubernetes, Terraform, CI/CD pipelines, and production reliability engineering.

B.Tech in Electronics & Computer Science Engineering (ECM)

Sreenidhi Institute of Science and Technology, Hyderabad (2020 - 2024)

Languages: English, Hindi
Om Jaju - Profile | Swiftcruit