Site Reliability Engineer

Acquire Intelligence

📍 Taguig Salary not disclosed Posted 2 Jul 2026 Active listing

About this role

Summary from our partner feed — apply on the employer site for the full posting.

A successful Site Reliability Engineer will have: Experience• Minimum 3+ years of hands-on experience running AWS production systems atscale• Proven expertise with AWS EKS (Elastic Kubernetes Service) or similar and MSK(Managed Streaming for Kafka) in production environments as well as databaseperformance diagnostics (MySQL, Postgres, MongoDB) in multi-TB scale databases• Strong background in Infrastructure as Code, preferably with Pulumi usingTypeScript or equivalent Terraform experience• Demonstrated experience participating in incident management (ideally as anincident commander with a track record of leading post-mortem processes)• Experience with high-volume data processing systems, ideally IoT telemetry orstreaming pipelines processing ≥50k messages per second• Background in implementing and maintaining observability solutions usingPrometheus, Grafana, PagerDuty, or similar tools Experience with CI/CD pipelinemanagement and deployment automation using GitLab, or similar platforms• Exposure to Hypervisors (VMWare, Hyper V), Microsoft Server stack, SAN/NAS, L2/3Networking Layers, Firewalls (Palo Alto), Switching (Aruba, Juniper) consideredadvantageous.

Technical Skills &

jsservices, AWS Lambda functions, and operational tooling• Deep understanding of AWS services ecosystem, with particular expertise incontainer orchestration, messaging systems, and content delivery• Strong networking fundamentals including TCP/IP, DNS, TLS, HTTP protocols, andcontainer networking (CNI)• Proficiency with monitoring and observability tools including Prometheus, Grafana,and incident management platforms• Experience with Infrastructure as Code tools, particularly Pulumi with TypeScript forcomprehensive AWS resource management• Understanding of security best practices including least-privilege access, IAM policymanagement, and compliance frameworks Behaviours• Systems thinking – Able to understand complex distributed systems and identifypotential failure points and optimization opportunities• Automation-first mindset – Consistently seeks to eliminate manual processes andbuild scalable, repeatable solutions• Incident leadership – Calm under pressure with strong communication skills duringhigh-stress situations and post-incident analysis• Collaborative approach – Works effectively with development teams to buildreliability into systems from the ground up• Continuous improvement focus – Proactively identifies opportunities for operationalenhancement and drives them to completion• Detail-oriented execution – Maintains high standards for documentation,monitoring, and operational procedures.

Motivation and interestsFounding influence – Join as one of the first SREs where your tooling choices andoperational processes become the organizational standard• Protected focus time – Error-budget policy guarantees at least 10% of timededicated to "make tomorrow better" reliability work• Sustainable on-call – True follow-the-sun rotation with no permanent night shifts,only regional time zone coverage• Professional growth – Opportunity to shape reliability engineering practices in ahigh-growth technology company• Global impact – Your work directly enables thousands of.

At a glance

Employer
Acquire Intelligence
Location
Taguig
Posted
2 Jul 2026
Source
en-ph.whatjobs.com

Apply on employer site ← All Private Sector Government jobs hub

Same employer, location, or role family from our partner feed.