Back to Jobs
Development Just now

Platform Infrastructure Engineer

CanadaCanada
Full-time
112,000 CAD - 168,000 CAD
Mid-Level

Job Description

Key Skills Required

Master these to land this role

DevOpsBestseller ๐Ÿ”ฅ
Learn in 63 Hours
TerraformGoogle Cloud PlatformKubernetesAWS

Want to know if you're a match for this job?

Calculate My Match Score

Summary

Platform Infrastructure Engineering builds and operates Menlo Security's Infrastructure Platform, enabling our customers to connect to the Internet without compromise. As a Platform Infrastructure Engineer, you'll join a globally distributed team of experienced engineers building and managing the company's core infrastructure services on a cloud-native platform built on Google Kubernetes Engine and VMs spanning multiple regions and environments. The team manages infrastructure as code with Terraform and Spacelift, deploys with Helm, and emphasizes security-first design, comprehensive observability, and multi-region resilience. The team also uses AI-assisted development and code-review tools, including Gemini Code Assist, as part of the standard engineering workflow, and this role is expected to use LLM-based tooling to build and troubleshoot infrastructure code efficiently.

Outcomes & KPIs

Key Outcome(s) Owned:

  • Reliable, secure, and scalable infrastructure across GCP and AWS supporting Menlo's platform globally.
  • Reduced operational toil and incident recurrence through automation and Infrastructure as Code practices.
  • Comprehensive, end-to-end observability framework providing deep platform visibility, proactive health monitoring, and accelerated incident detection and resolution.

Success Metrics / KPIs:

  • Infrastructure uptime/availability across regions (e.g., 99.9%+)
  • Mean time to detect (MTTD) and mean time to resolve (MTTR) for incidents
  • Percentage of infrastructure changes deployed via IaC (Terraform) vs. manual changes
  • On-call incident volume and reduction in repeat/preventable incidents
  • Lead time for provisioning new infrastructure

What You'll Do

  • Implement, deploy, and maintain VM and Kubernetes infrastructure on GCP and AWS across dozens of clusters spanning development, staging, and production environments in multiple regions
  • Build and maintain Infrastructure as Code using Terraform modules and Spacelift (or equivalent TACOS), provisioning networking, compute, storage, and security components, and implementing multi-layer configuration management workflows
  • Implement and maintain observability solutions using Grafana Cloud, Prometheus/Mimir, and OTel collectors, designing dashboards and alerting rules across all platform components
  • Manage certificate lifecycle, DNS automation, ingress controllers, and service mesh networking with Cilium
  • Partner with peers and across Engineering, Product, Compliance, and Security teams to align on requirements and consult on capacity planning, disaster recovery, and architectural decisions
  • Identify and eliminate toil through automation โ€” writing scripts, building CI/CD pipelines, and using AI-assisted coding tools to move faster
  • Participate in a 24x7 on-call rotation as part of a globally distributed team, responding to incidents and driving post-incident reviews

Functional Competencies

Required:

  • Bachelor's degree in Computer Science, a related technical field, or equivalent practical experience
  • Proficiency in common programming and scripting languages, particularly Python, Bash, and Go
  • Understanding of network topologies, communication protocols (e.g., TCP/IP, HTTP/S, UDP, TLS), and enterprise-grade connectivity solutions
  • Kubernetes expertise, including cluster administration, RBAC, networking, workload management, and troubleshooting in production environments
  • Proven experience with Terraform for infrastructure provisioning and management
  • Knowledge of Google Cloud Platform services including GKE, VPC networking, Cloud DNS, Artifact Registry, Secret Manager, IAM, Gemini Code Assist, and Workload Identity
  • Clear understanding of how to use LLM-based code-assist tools to effectively build and troubleshoot software

Preferred / Nice to Have:

  • Experience with GitOps methodologies and tools

Why Menlo?

Our culture is collaborative, inclusive, and fun! We have five core values: Stay Aligned, Get It Done, Customer Empathy, Think Creatively and Help Each Other Out. We believe in open communication, supporting new ideas, and sharing a mutual mindset of what weโ€™re aiming to achieve together. There are tremendous opportunities to take initiative, implement new ideas, and have a hand in building a legacy.

How would you rate this job post?

See what other professionals think about this role.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More