Canada•Brazil•Argentina•Las Vegas•Charlotte•Tampa•Colombia
Remote
Senior
Full Time
29 days ago
💰$ 90,000 - $ 125,000
remotekubernetessreplatform engineeringdevexci/cdautomationgolang
Requirements
- •6–9+ years in SRE / Platform / Infrastructure Engineering
- •Proven experience scaling Kubernetes in high-throughput production environments
- •Deep Kubernetes expertise including internals, scheduler behavior, custom resources, and cluster-scale failure diagnosis
- •Experience building platform infrastructure, control planes, or Kubernetes Operators
- •Strong distributed systems and production reliability experience
- •Terraform/GitOps ownership and automation design
- •Experience with GitOps workflows (Flux / ArgoCD)
- •Hands-on experience with CI/CD platforms at scale including GitHub Actions, GitHub Apps, and build systems
- •AWS/cloud infrastructure production experience
- •Proficiency in Go or another systems language
- •Track record of building infrastructure primitives with an automation-first mindset
What You'll Do
- •Design and build infrastructure primitives for CI/CD platform, build systems, and developer environments
- •Build and operate Kubernetes-based control plane for CI/CD platform
- •Develop GitHub Actions self-hosted runner infrastructure with autoscaling, isolation, and cost/performance tuning
- •Manage GitHub Apps and GitHub-as-code including permissions, webhooks, and org-wide automation
- •Ensure secure network access for CI/CD and remote dev environments using Tailscale
- •Implement GitOps-driven deployment of platform services using Flux
- •Create ephemeral/on-demand developer environments and build systems
- •Develop Kubernetes Operators and scaling automation for product teams
- •Build systems that define how engineering teams build, test, and deploy to improve reliability and scalability of developer experience
Nice to Have
- •Experience with Tailscale or similar zero-trust/overlay networking for CI/CD or remote dev environments
- •Experience designing SLO strategies and error-budget usage
- •Experience improving diagnosability and observability frameworks (OpenTelemetry or similar)
- •Experience building internal developer platforms (IDP) or ephemeral/on-demand dev environments
- •Experience building multi-tenant platform infrastructure
- •Experience contributing to Kubernetes ecosystem projects
- •Experience working in high-ambiguity environments
- •Experience with web3 concepts
