LATAM•Brazil•Uruguay•Peru•Colombia•Argentina•Paraguay•South Africa•Canada•Mexico
Remote
Mid Level
Full Time
3 days ago
KubernetesInfrastructure EngineerSite Reliability EngineerDevOpsCloudPythonBashGoDistributed Systems
Requirements
- •3+ years experience as Site Reliability Engineer, Infrastructure/Platform/DevOps Engineer, Software Engineer or similar roles
- •Strong passion for providing technical guidance to stakeholders with excellent communication skills
- •Strong understanding of distributed systems fundamentals
- •Strong Linux systems knowledge including shell, processes, networking basics, file systems, permissions
- •Networking fundamentals: TCP/IP, DNS, ports, IP addressing, basic routing, TLS, PKI concepts
- •Understanding of Kubernetes internals (control plane, scheduling, networking, storage) and concepts
- •Familiarity with orchestration or advanced Kubernetes networking (Cilium)
- •Ability to leverage AI tools and agents such as Claude and OpenAI
- •Scripting or programming experience (Python, Bash, Go, or similar)
What You'll Do
- •Gain hands-on experience across Kubernetes cluster operations, deployments, networking, and scalability
- •Help build, and operate Kubernetes clusters across multiple datacenters
- •Triage and organize infrastructure issues, learning systems and best practices
- •Work with senior engineers to solve complex problems, reducing operational burden
- •Build automation and tooling to improve team efficiency and reduce toil
- •Participate in on-call rotation to maintain platform reliability
Nice to Have
- •Working experience running Kubernetes clusters
- •Exposure to cloud platforms (AWS, GCP, Azure) or on-premises infrastructure
- •Knowledge of Terraform or other Infrastructure as Code tools
- •Experience with storage systems (e.g., Ceph/Rook)
