Senior Site Reliability Engineer at StraitsX

Company: StraitsX

Location: Jakarta, Jakarta, Indonesia

Type: Full-time

Apply for this position

Frequently asked questions

Is this job still open?

GMI Jobs is currently showing this Senior Site Reliability Engineer role at StraitsX as an active listing from employer-controlled hiring sources. Always confirm final availability on the employer application page before applying.

How do I apply for this role?

Use the application link on this page to apply directly with StraitsX. GMI Jobs keeps the canonical job page, company context, and market links together so candidates can compare before leaving the site.

What should I compare before applying?

Compare the role location Jakarta, Jakarta, Indonesia, any listed pay signal, company profile, related crypto jobs, and salary context before applying. Listed pay signals are not guaranteed offers; confirm compensation with the employer.

Job Description

<p><strong>About The Role</strong></p> <p>The Site Reliability Engineering (SRE) team architects, builds, and maintains the rock-solid infrastructure that applications rely on. At the Senior Level, you own reliability, performance, and cost outcomes for the systems under your area end-to-end, not just executing well-defined tasks, but deciding between tradeoffs, scoping ambiguous problems, and driving process and system improvements that span teams. You'll work closely with development, security, and product teams, and mentor other engineers as a technical point of reference for the team.</p> <p><strong>What You Will Do</strong></p> <ul> <li>Own the availability, performance, scalability, and security of production systems end-to-end, across cloud (AWS/GCP) and on-premises environments.</li> <li>Design and evolve Kubernetes deployment strategy for production workloads.</li> <li>Own CI/CD and GitOps pipelines in production (ArgoCD or equivalent) and the Terraform that provisions the infrastructure behind them.</li> <li>Diagnose and resolve database performance issues.</li> <li>Build and maintain observability that surfaces problems before they become incidents.</li> <li>Seek out and implement process and system improvements affecting performance and security, coordinating with multiple stakeholders.</li> <li>Scope and lead medium-to-large infrastructure initiatives: gather requirements, prioritize by business impact, and communicate impact to stakeholders.</li> <li>Negotiate technical tradeoffs with stakeholders to meet business SLAs.</li> <li>Provide technical guidance and mentorship to peers and junior engineers; promote best practices and standards across the team.</li> <li>Maintain documentation and process discipline for the systems and incidents you own.</li> <li>Lead structured incident investigation, isolating server, database, and application layers with a metrics-first approach, including on-call during high-traffic events.</li> </ul> <p><strong>What Are We Looking For</strong></p> <ul> <li>At least 4 years of experience in either SRE, DevOps, MLOps, or platform engineering, including senior-level scope at a high-traffic company.</li> <li>Deep expertise in one major cloud provider (preferably AWS), with a proven ability to ramp up on the other quickly.Production experience & expertise with Kubernetes & Linux fundamentals</li> <li>CI/CD & GitOps (ArgoCD or other equivalent stacks)</li> <li>Database performance analysis & monitoring (MySQL, Postgres)</li> <li>Observability tooling & standards (Datadog, OpenTelemetry)</li> <li>Infrastructure as Code (Terraform)</li> <li>Str

More jobs at StraitsX

Browse More Jobs

Related searches for this role

Use these GMI Jobs search paths to compare similar openings before applying.

Popular job searches

Explore related crypto job pages with live listings and salary context.