Staff Engineer - Cloud Platform
AI Summary
Staff Engineer - Cloud Platform designs, architects, and operates highly scalable, secure, and resilient cloud platform infrastructure for broadband services, combining deep hands-on engineering with technical leadership across cloud, networking, and observability.
About this role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer - Cloud Platform based in Canada.
This is a senior individual contributor role focused on building and operating highly scalable, secure, and resilient cloud platforms for next-generation broadband services.
You will shape platform architecture, modernization, automation, and reliability strategies across complex cloud and telecom environments.
The role combines deep hands-on engineering with technical leadership across cloud infrastructure, networking, messaging, and observability.
You will work extensively with Google Cloud, Kubernetes, infrastructure as code, CI/CD, and distributed systems.
Your work will help engineering teams deliver products faster while maintaining carrier-grade reliability and operational excellence.
You will collaborate across Engineering, SRE, Network Operations, Product, and Customer Success teams in a highly distributed environment.
The position offers significant scope to influence technical direction, mentor engineers, and drive strategic platform initiatives.
Accountabilities
-
Design, architect, and operate highly available, scalable, secure, and resilient cloud platform infrastructure, while driving modernization through cloud-native technologies and automation.
-
Define and implement strategies for platform reliability, observability, disaster recovery, capacity management, and operational excellence.
-
Troubleshoot complex end-to-end service issues across cloud infrastructure, network services, broadband access environments, and customer-facing applications.
-
Design and operate NATS messaging infrastructure and HAProxy-based traffic management, including load balancing, routing, SSL termination, high availability, and performance optimization.
-
Build scalable service communication frameworks using event-driven architectures, service discovery, fault tolerance, and resiliency patterns.
-
Design and manage Google Cloud infrastructure, including GKE, Compute Engine, Cloud Storage, BigQuery, Pub/Sub, Cloud SQL, and Composer/Airflow.
-
Implement Infrastructure as Code using Terraform and Terragrunt, automating provisioning, deployment, compliance, and operational workflows.
-
Develop reusable platform services and self-service capabilities that improve engineering productivity and accelerate application delivery.
-
Design and maintain enterprise-scale CI/CD pipelines using tools such as Jenkins, GitLab CI, and GitHub Actions.
-
Establish and enhance monitoring, logging, tracing, and alerting using technologies such as Prometheus, Grafana, ELK/OpenSearch, VictoriaMetrics/VictoriaLogs, and cloud monitoring services.
-
Lead root cause analysis and incident response for critical platform and network issues, identifying opportunities for proactive prevention.
-
Partner with Product, Engineering, SRE, Network Operations, technical support, and Customer Success teams to deliver strategic platform initiatives.
-
Develop technical roadmaps, contribute to architecture governance and design reviews, establish engineering standards, and mentor engineers across cloud, networking, reliability, and telecom domains.
-
Bachelor’s degree in Computer Science, Engineering, Telecommunications, or a related discipline, or equivalent practical experience.
-
8+ years of experience designing and operating large-scale distributed systems in production environments.
-
5+ years of hands-on experience with public cloud platforms, preferably Google Cloud Platform.
-
Strong hands-on expertise with Kubernetes and containerized platforms.
-
Deep understanding of networking fundamentals, including TCP/IP, routing and switching, NAT, DNS, DHCP, firewalls, VPN technologies, and load balancing.
-
Practical experience configuring, tuning, troubleshooting, and operating HAProxy in highly available environments.
-
Hands-on experience with NATS and event-driven architectures.
-
Experience supporting telecommunications or broadband infrastructure, including OLTs, ONTs, routers, and access-network technologies.
-
Experience implementing Infrastructure as Code with Terraform.
-
Proficiency in Python, Go, Bash, or comparable scripting and programming languages.
-
Experience designing and operating CI/CD pipelines and DevOps platforms.
-
Strong knowledge of monitoring, telemetry, observability, incident management, and production troubleshooting.
-
Strong problem-solving skills and the ability to investigate complex technical issues from infrastructure through application layers.
-
Ability to work effectively across engineering disciplines and communicate technical concepts clearly with both technical and non-technical stakeholders.
-
Experience in broadband access networks, PON technologies, or service-provider environments is an advantage.
-
Experience with Kubernetes networking, service mesh technologies, Kafka, Redis, Elasticsearch/OpenSearch, or distributed databases is beneficial.
-
Knowledge of cloud and telecom security best practices and Google Cloud professional certifications are considered advantages.
-
Previous experience in SRE, Platform Engineering, or Telecom Cloud Operations environments is a plus.
-
Annual base salary range of CAD $141,000–$240,000, depending on geographic location and factors such as job-related knowledge, skills, and experience.
-
Potential eligibility for a bonus as part of the overall compensation package.
-
Remote work opportunity within Canada.
-
Opportunity to work on large-scale cloud, networking, and broadband infrastructure supporting next-generation digital services.
-
Exposure to modern technologies including Google Cloud, Kubernetes, Terraform, NATS, HAProxy, CI/CD, and advanced observability platforms.
-
Opportunities to influence technical strategy, architecture, engineering standards, and platform modernization.
-
Career development and mentoring opportunities within a multidisciplinary engineering environment.
-
Collaboration with distributed teams across cloud engineering, SRE, networking, operations, product, and customer-facing functions.
-
Comprehensive benefits package; specific benefits details are shared during the recruitment process.
Requirements
Benefits
Skills
Explore related jobs
More jobs at Jobgether
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.
