Senior DevOps Engineer

Senior DevOps Engineer

Senior DevOps Engineer

Ciklum

58 minutos atrás

Nenhuma candidatura

Sobre

  • Ciklum is looking for a Senior DevOps Engineer to join our team in Brazil.
  • We are a custom product engineering company that supports both multinational
  • organizations and scaling startups to solve their most complex business
  • challenges. With a global team of over 4,000 highly skilled developers,
  • consultants, analysts and product owners, we engineer technology that redefines
  • industries and shapes the way people live.

About the role

  • We are looking for a talented Senior DevOps Engineer to support our engineering
  • work, with a strong focus on the underlying infrastructure that we need to get
  • from source code to production, plus keep it happy once it's there. We need
  • someone who understands software development and enjoys working with developers
  • on all the things necessary to improving, deploying, monitoring, and operating
  • production services.

Responsibilities

  • * Build complex, highly available, and cost-optimized cloud infrastructure
  • solutions on Azure and in Kubernetes (AKS)
  • * Design, build, and maintain production-grade AKS clusters, including private
  • cluster networking, node pool sizing and autoscaling, workload identity,
  • ingress, certificate management, and cluster upgrade lifecycle
  • * Own infrastructure as code end to end in Terraform: author and refactor
  • reusable modules, manage per-environment workspaces and state, and drive
  • changes through pull request, plan review, and apply
  • * Design and implement advanced CI/CD pipelines ensuring quality, security, and
  • efficiency, including container image build/publish and progressive promotion
  • from lower environments to production
  • * Establish comprehensive monitoring and observability strategies across
  • cluster, application, and Azure platform services
  • * Proactively identify and mitigate risks, leading troubleshooting efforts for
  • complex issues across compute, networking, identity, and data layers
  • * Implement high-level security practices, integrating vulnerability scanning
  • and threat modeling, and manage secrets through a centralized secrets store
  • with identity-based (rather than credential-based) access
  • * Maintain and exercise multi-region high availability and disaster recovery,
  • including failover and failback procedures for Kubernetes workloads and
  • managed databases
  • * Work with application engineering to cultivate operational standards

Requirements

  • We know that sometimes, you can't tick every box. We would still love to hear
  • from you if you think you're a good fit!
  • * Bachelor's degree in Computer Science, Software Engineering, or a related
  • field, or equivalent work experience
  • * 5+ years of experience in a DevOps engineering role with significant Azure
  • experience
  • * Proven track record of designing and implementing large-scale Azure solutions
  • * Hands-on experience running AKS in production, not just in development or
  • proof-of-concept environments — including private clusters, cluster upgrades,
  • autoscaling, and incident response on live workloads
  • * Advanced, production-level Terraform experience: module design, state and
  • workspace management, and multi-environment/multi-subscription deployments
  • * Demonstrated ability to solve complex technical problems and architect robust
  • solutions
  • * Excellent communication and collaboration skills
  • * Azure certifications (e.g., Azure Solutions Architect Expert, Azure DevOps
  • Engineer Expert, Certified Kubernetes Administrator)

Desirable

  • * Cloud: Deep knowledge of Azure services, architectures, and design patterns,
  • including Entra ID, RBAC, managed identities and workload identity
  • federation, Key Vault, and subscription/landing zone structure
  • * Infrastructure as Code (IaC): Advanced skills in Terraform, or other IaC
  • tools
  • * CI/CD: Strong experience with pipeline-as-code (Azure Pipelines, GitHub
  • Actions, or equivalent), service connections and federated pipeline
  • authentication, and GitOps-style continuous delivery to Kubernetes (Argo CD,
  • Flux, or similar)
  • * Programming and Scripting: Strong scripting skills (PowerShell, Bash, Python)
  • and proficiency with one or more programming languages (e.g., Go, C#/.NET,
  • Node.js)
  • * Containerization and Microservices: Extensive experience with Docker,
  • Kubernetes (AKS), Helm, container registries, and microservices architectures
  • * Complex Configuration Management: Expertise in configuration management tools
  • (Ansible, Chef, Puppet, etc.)
  • * Robust Networking: In-depth knowledge of networking concepts, Azure virtual
  • network and subnet design, private endpoints and private DNS zones, hybrid
  • connectivity (site-to-site VPN, network virtual appliances/firewalls), and
  • network security
  • * Data and Messaging Services: Familiarity with Azure managed data services
  • such as SQL Managed Instance (including failover groups), Cosmos DB,
  • Redis/managed cache, Storage, and Functions
  • * Observability: Experience with Azure Monitor/Log Analytics and KQL, plus at
  • least one third-party observability platform (Datadog, Grafana, or similar)
  • * Resilience: Experience designing and testing multi-region redundancy,
  • RTO/RPO-driven DR runbooks, and failover automation
  • * Security Focus: Strong understanding of cloud security principles,
  • penetration testing, and compliance standards (e.g., SOC 2, ISO 27001, PCI
  • DSS)
  • * Multi-cloud: ScriptCycle's footprint spans both Azure and Google Cloud, with
  • an active workload migration track. Exposure to GCP and GKE (including
  • Autopilot) and to cross-cloud networking and identity is a strong plus,
  • though Azure is the primary focus of this role; experience supporting a GCP
  • migration
  • What’s in it for you?
  • * Care: your mental and physical health is our priority. We ensure
  • comprehensive company-paid medical insurance and mental health programs, 5
  • undocumented sick-leave days per year
  • * Tailored education path: boost your skills and knowledge with our regular
  • internal events (meetups, conferences, workshops), Udemy license, language
  • courses and company-paid certifications
  • * Growth environment: share your experience and level up your expertise with a
  • community of skilled professionals, locally and globally
  • * Long-term employment with 20 working-days paid vacation and local bank
  • holidays
  • * Flexibility: 100% remote work mode
  • * Opportunities: we value our specialists and always find the best options for
  • them. Our Internal Mobility Program helps change a project if needed to help
  • you grow, excel professionally and fulfill your potential
  • * Global impact: work on large-scale projects that redefine industries with
  • international and fast-growing clients
  • * Welcoming environment: feel empowered with a friendly team, open-door policy,
  • informal atmosphere within the company and regular team-building events

About us

  • At Ciklum, we are always exploring innovations, empowering each other to achieve
  • more, and engineering solutions that matter. With us, you’ll work with
  • cutting-edge technologies, contribute to impactful projects, and be part of a
  • One Team culture that values collaboration and progress.
  • As we expand into Latin America, every Ciklumer is helping to shape our story.
  • Collaborate with seasoned experts and make a global impact backed by two decades
  • of industry leadership.
  • Explore, empower, engineer with Ciklum!
  • Interested already? We would love to get to know you! Submit your application.
  • We can’t wait to see you at Ciklum.
  • #LI-IK1