Skip to content

Senior Staff Cloud Infrastructure Engineer

  • Hybrid
    • Bangalore, Karnātaka, India
    • Kuala Lumpur, Wilayah Persekutuan Kuala Lumpur, Malaysia
    +1 more
  • Infrastructure

Job description

At Toku, we create bespoke cloud communications and customer engagement solutions to reimagine customer experiences for enterprises. We provide an end-to-end approach to help businesses overcome the complexity of digital transformation and deliver mission-critical CX through cloud communication solutions. Toku combines local strategic consulting expertise, bespoke technology, regional in-country infrastructure, connectivity, and global reach to serve the diverse needs of enterprises operating at scale. Headquartered in Singapore, Toku supports customers across APAC and beyond, with a growing footprint across global markets.

 

We are looking for a highly experienced, hands-on Cloud Infrastructure Engineer to take on the most senior individual-contributor role within our Infrastructure team. Working alongside the VP of Infrastructure, you will set technical direction, architecture and engineering standards across our infrastructure platforms, while remaining deeply involved in solving our most complex technical challenges. You will lead the design and implementation of modern cloud infrastructure across AWS and Azure and drive greater automation, Infrastructure as Code and engineering maturity across the team. You will thrive in this role if you combine deep infrastructure expertise with the engineering mindset and hands-on capability to turn better ideas into better platforms.

Job requirements

What you will be doing

 

  • Cloud architecture & engineering: Design, build and evolve secure, scalable and highly available cloud infrastructure, translating business and engineering requirements into robust technical solutions across AWS and Azure.

  • Technical direction & architecture: Act as a senior technical authority for Infrastructure, owning key architectural decisions, establishing engineering standards and challenging existing approaches where better technical solutions can be introduced.

  • AWS platform architecture: Evolve our AWS estate across areas such as multi-account / AWS Organizations architecture, networking, IAM, security guardrails and core production services, ensuring the platform remains scalable, resilient and operationally effective.

  • Azure capability: Help establish and mature our Azure environment, including Azure Landing Zone architecture, governance, networking, security and reusable infrastructure patterns.

  • Infrastructure as Code: Own and improve Terraform-based infrastructure provisioning, developing reusable and tested modules, consistent multi-environment patterns and appropriate policy controls so infrastructure can be deployed reliably and repeatably.

  • Automation & tooling: Use Python, Go/Golang and scripting to automate infrastructure operations, eliminate repetitive manual work and build internal tooling that connects cloud services, deployment pipelines, observability platforms and operational workflows.

  • Platform engineering: Build reusable infrastructure capabilities that make it easier for engineering teams to provision, deploy, test, operate and recover applications without unnecessary manual dependency on the Infrastructure team.

  • CI/CD & GitOps: Design and improve infrastructure delivery pipelines, integrating Infrastructure as Code, version control, automated testing, policy enforcement and deployment automation into reliable engineering workflows.

  • Kubernetes: Architect, operate and improve production Kubernetes environments, particularly Amazon EKS, ensuring container platforms are scalable, resilient, secure and operationally robust.

  • Observability & operational automation: Build and configure effective monitoring and observability capabilities, including dashboards, monitors, alerting and integrations, while introducing opportunities for automated remediation and self-healing.

  • Reliability & resilience: Engineer infrastructure for high availability, disaster recovery, scalability and production reliability, contributing to capacity planning, incident resolution and root-cause analysis while continuously improving operational resilience.

  • Security by design: Apply strong cloud security practices across infrastructure architecture, IAM, networking, access controls, hardening and Infrastructure as Code, collaborating with security stakeholders to ensure appropriate controls and guardrails are embedded into the platform.

  • Engineering standards: Drive greater consistency in how infrastructure is designed, provisioned, changed and operated, using automation, policy and engineering practices to reduce configuration drift, manual intervention and operational risk.

  • Technology & cost optimisation: Evaluate cloud services, tooling and architectural approaches with consideration for performance, reliability, maintainability and cost, identifying opportunities to simplify the platform and improve infrastructure efficiency.

  • Team capability: Mentor infrastructure engineers through hands-on collaboration, technical and design reviews, and knowledge sharing, raising the team's cloud, automation and engineering capability while remaining an individual contributor.

  • Cross-functional engineering: Work closely with Software Engineering and other technical teams to understand application requirements, challenge assumptions constructively and design appropriate infrastructure solutions rather than simply implementing requested configurations.

 

We’d love to hear from you if you have

 

  • Cloud infrastructure: Extensive hands-on experience designing, building and operating production cloud infrastructure, with the technical depth expected of a Staff, Senior Staff, Principal or equivalent senior individual contributor.

  • AWS: Strong hands-on AWS architecture and production experience across services and capabilities such as EC2, ECS/EKS, VPC and networking, IAM, S3, RDS, CloudWatch, CloudFront and Route53, with experience of multi-account / AWS Organizations environments strongly valued.

  • Microsoft Azure: Strong hands-on Azure capability, ideally including Azure Landing Zone architecture, governance, networking and security, with the ability to help establish and mature cloud environments rather than only administer existing services.

  • Terraform / Infrastructure as Code: Advanced hands-on Terraform experience, including reusable modules, remote/state management, multi-environment infrastructure, testing and scalable IaC patterns and standards.

  • Programming & scripting: Strong hands-on capability with Python and/or Go (Golang) for infrastructure automation and tooling, alongside practical Bash/Shell or similar scripting experience. We are looking for genuine engineering capability beyond occasional or incidental scripting.

  • Automation: A strong track record of engineering repetitive operational activities out of infrastructure through automation, reusable tooling, API integrations, auto-remediation and self-service capabilities.

  • Kubernetes: Strong production Kubernetes experience, preferably including Amazon EKS, with the ability to architect, troubleshoot and improve containerised production environments.

  • CI/CD & GitOps: Hands-on experience designing and operating infrastructure CI/CD workflows using GitHub Actions, GitLab CI, Jenkins, Azure DevOps or similar. Experience with GitOps technologies such as ArgoCD or Flux, and Helm, would be an advantage.

  • Observability: Hands-on experience configuring enterprise monitoring and observability platforms such as Datadog or similar, including monitors, dashboards, alerting and third-party integrations rather than only consuming existing dashboards.

  • Cloud architecture: Strong architectural judgement across scalability, availability, resilience, disaster recovery, capacity and production reliability, combined with the hands-on depth to implement and validate your designs.

  • Networking: Solid infrastructure networking knowledge covering routing, DNS, load balancing, firewalls, cloud networking and connectivity between cloud and other environments.

  • Security: Strong knowledge of cloud security and IAM principles, including secure infrastructure design, access controls, hardening and security-conscious automation. Experience working with SIEM/SOAR platforms or closely with Security/SOC teams would be an advantage.

  • Engineering governance: Experience using automated testing, static analysis or policy-as-code to enforce infrastructure standards. OPA, Sentinel or similar technologies would be particularly relevant.

  • Broader platform tooling: Experience with technologies such as service meshes (Istio, Linkerd), alternative IaC approaches such as Pulumi, AWS CDK or CloudFormation, or comparable modern cloud-platform tooling would be an advantage.

  • Technical leadership: Experience influencing architecture and engineering standards, mentoring other engineers and raising technical capability while remaining deeply hands-on as an individual contributor.

  • Additional advantages: Experience with FinOps or cloud cost-management tooling, hybrid/on-premise infrastructure, and telecom, CPaaS or similarly regulated/data-residency-sensitive environments would be beneficial.

  • Location: This role can be based in Bangalore – India, or KL - Malaysia. It will operate on a mostly WFH basis for the time being, but in the future will require a hybrid WFH/WFO model.

Toku has been recognised as a LinkedIn Top Startup and by the Financial Times as one of APAC’s Top 500 High Growth Companies. If you’re looking to be part of a company on a strong growth trajectory while working on meaningful, real-world challenges, we’d love to hear from you.

Hybrid
  • Bangalore, Karnātaka, India
  • Kuala Lumpur, Wilayah Persekutuan Kuala Lumpur, Malaysia
+1 more
Infrastructure

or

Apply with Indeed unavailable