职位详情
国籍要求:马来西亚
职位描述
The Role
Unlike traditional "Ops" roles, you will be a key member of a specific development squad, either Customer Experience or Back-Office Risk. You won't just manage servers; you will dive into the codebase, perform code reviews, and drive independent observability initiatives to ensure the platform remains scalable and resilient.
Key Responsibilities
Design-Phase Advisory: Partner with developers during the feature proposal stage to provide suggestions on monitoring, reliability, and system resilience.
Hands-on Implementation: Independently drive the integration of monitoring tools like Sentry directly into the codebase without waiting for feature developers.
Infrastructure as Code: Design and implement IaC using Terraform or CloudFormation to automate scalable, highly available environments on AWS.
Pipeline Optimization: Build and refine GitLab CI/CD pipelines to ensure reliable, zero-downtime software delivery.
Observability Mastery: Utilize our internal stack (Grafana, Prometheus, Loki, and Open Telemetry) to proactively detect and resolve performance bottlenecks.
Incident Leadership: Lead root cause analysis (RCA) and blameless post-mortems to drive a significant reduction in recurring incidents.
Common Requirements
Experience: 2+ years of proven experience in SRE, DevOps, or similar software engineering roles.
The "Bridge" Mindset: A strong rejection of the "us vs. them" mentality; you are eager to explore the application codebase and learn new languages (like C#) to unblock your team.
Automation: Proficiency in any CI/CD system (GitLab CI/CD is a major plus) and Infrastructure as Code (Terraform/CloudFormation).
Container Management: Proven experience in deploying and managing containerized applications and general orchestration.
Monitoring Stack: Familiarity with modern observability principles; experience with Prometheus, Grafana, Loki, or Sentry is a strong advantage.
Specialized Skills (Team Dependent)
We are looking for specialists to join 2 distinct squads:
Frontend/Customer-Facing Focus: * Proficiency in JavaScript and TypeScript (React/Vue experience preferred).
Plus: Experience with Serverless architectures, statically generated SPAs, and CDNs/web caching.
Backend/Back-Office Focus: * Proficiency in C# (or a strong backend background with a willingness to learn) to manage core risk dashboards and tracking systems.