Senior DevOps Engineer
Ciklum
58 minutos atrás
•Nenhuma candidatura
Sobre
- Ciklum is looking for a Senior DevOps Engineer to join our team in Brazil.
- We are a custom product engineering company that supports both multinational
- organizations and scaling startups to solve their most complex business
- challenges. With a global team of over 4,000 highly skilled developers,
- consultants, analysts and product owners, we engineer technology that redefines
- industries and shapes the way people live.
About the role
- We are looking for a talented Senior DevOps Engineer to support our engineering
- work, with a strong focus on the underlying infrastructure that we need to get
- from source code to production, plus keep it happy once it's there. We need
- someone who understands software development and enjoys working with developers
- on all the things necessary to improving, deploying, monitoring, and operating
- production services.
Responsibilities
- * Build complex, highly available, and cost-optimized cloud infrastructure
- solutions on Azure and in Kubernetes (AKS)
- * Design, build, and maintain production-grade AKS clusters, including private
- cluster networking, node pool sizing and autoscaling, workload identity,
- ingress, certificate management, and cluster upgrade lifecycle
- * Own infrastructure as code end to end in Terraform: author and refactor
- reusable modules, manage per-environment workspaces and state, and drive
- changes through pull request, plan review, and apply
- * Design and implement advanced CI/CD pipelines ensuring quality, security, and
- efficiency, including container image build/publish and progressive promotion
- from lower environments to production
- * Establish comprehensive monitoring and observability strategies across
- cluster, application, and Azure platform services
- * Proactively identify and mitigate risks, leading troubleshooting efforts for
- complex issues across compute, networking, identity, and data layers
- * Implement high-level security practices, integrating vulnerability scanning
- and threat modeling, and manage secrets through a centralized secrets store
- with identity-based (rather than credential-based) access
- * Maintain and exercise multi-region high availability and disaster recovery,
- including failover and failback procedures for Kubernetes workloads and
- managed databases
- * Work with application engineering to cultivate operational standards
Requirements
- We know that sometimes, you can't tick every box. We would still love to hear
- from you if you think you're a good fit!
- * Bachelor's degree in Computer Science, Software Engineering, or a related
- field, or equivalent work experience
- * 5+ years of experience in a DevOps engineering role with significant Azure
- experience
- * Proven track record of designing and implementing large-scale Azure solutions
- * Hands-on experience running AKS in production, not just in development or
- proof-of-concept environments — including private clusters, cluster upgrades,
- autoscaling, and incident response on live workloads
- * Advanced, production-level Terraform experience: module design, state and
- workspace management, and multi-environment/multi-subscription deployments
- * Demonstrated ability to solve complex technical problems and architect robust
- solutions
- * Excellent communication and collaboration skills
- * Azure certifications (e.g., Azure Solutions Architect Expert, Azure DevOps
- Engineer Expert, Certified Kubernetes Administrator)
Desirable
- * Cloud: Deep knowledge of Azure services, architectures, and design patterns,
- including Entra ID, RBAC, managed identities and workload identity
- federation, Key Vault, and subscription/landing zone structure
- * Infrastructure as Code (IaC): Advanced skills in Terraform, or other IaC
- tools
- * CI/CD: Strong experience with pipeline-as-code (Azure Pipelines, GitHub
- Actions, or equivalent), service connections and federated pipeline
- authentication, and GitOps-style continuous delivery to Kubernetes (Argo CD,
- Flux, or similar)
- * Programming and Scripting: Strong scripting skills (PowerShell, Bash, Python)
- and proficiency with one or more programming languages (e.g., Go, C#/.NET,
- Node.js)
- * Containerization and Microservices: Extensive experience with Docker,
- Kubernetes (AKS), Helm, container registries, and microservices architectures
- * Complex Configuration Management: Expertise in configuration management tools
- (Ansible, Chef, Puppet, etc.)
- * Robust Networking: In-depth knowledge of networking concepts, Azure virtual
- network and subnet design, private endpoints and private DNS zones, hybrid
- connectivity (site-to-site VPN, network virtual appliances/firewalls), and
- network security
- * Data and Messaging Services: Familiarity with Azure managed data services
- such as SQL Managed Instance (including failover groups), Cosmos DB,
- Redis/managed cache, Storage, and Functions
- * Observability: Experience with Azure Monitor/Log Analytics and KQL, plus at
- least one third-party observability platform (Datadog, Grafana, or similar)
- * Resilience: Experience designing and testing multi-region redundancy,
- RTO/RPO-driven DR runbooks, and failover automation
- * Security Focus: Strong understanding of cloud security principles,
- penetration testing, and compliance standards (e.g., SOC 2, ISO 27001, PCI
- DSS)
- * Multi-cloud: ScriptCycle's footprint spans both Azure and Google Cloud, with
- an active workload migration track. Exposure to GCP and GKE (including
- Autopilot) and to cross-cloud networking and identity is a strong plus,
- though Azure is the primary focus of this role; experience supporting a GCP
- migration
- What’s in it for you?
- * Care: your mental and physical health is our priority. We ensure
- comprehensive company-paid medical insurance and mental health programs, 5
- undocumented sick-leave days per year
- * Tailored education path: boost your skills and knowledge with our regular
- internal events (meetups, conferences, workshops), Udemy license, language
- courses and company-paid certifications
- * Growth environment: share your experience and level up your expertise with a
- community of skilled professionals, locally and globally
- * Long-term employment with 20 working-days paid vacation and local bank
- holidays
- * Flexibility: 100% remote work mode
- * Opportunities: we value our specialists and always find the best options for
- them. Our Internal Mobility Program helps change a project if needed to help
- you grow, excel professionally and fulfill your potential
- * Global impact: work on large-scale projects that redefine industries with
- international and fast-growing clients
- * Welcoming environment: feel empowered with a friendly team, open-door policy,
- informal atmosphere within the company and regular team-building events
About us
- At Ciklum, we are always exploring innovations, empowering each other to achieve
- more, and engineering solutions that matter. With us, you’ll work with
- cutting-edge technologies, contribute to impactful projects, and be part of a
- One Team culture that values collaboration and progress.
- As we expand into Latin America, every Ciklumer is helping to shape our story.
- Collaborate with seasoned experts and make a global impact backed by two decades
- of industry leadership.
- Explore, empower, engineer with Ciklum!
- Interested already? We would love to get to know you! Submit your application.
- We can’t wait to see you at Ciklum.
- #LI-IK1



