Source description
About the role
Position Overview: The Senior DevOps Engineer is responsible for leading the design, implementation, and maintenance of scalable, secure, and reliable infrastructure and deployment pipelines. This role ensures that development and operations teams work together effectively by automating processes, improving system reliability, and enabling continuous integration and continuous delivery (CI/CD). The ideal candidate will have a strong technical background in cloud platforms, infrastructure-as-code, automation, and a passion for system performance and scalability. Key Responsibilities:
Design, build, and maintain scalable and secure infrastructure across cloud and on-premises environments (AWS, Azure, GCP, etc.).
Implement and manage CI/CD pipelines to automate testing, integration, and deployment of applications.
Champion Infrastructure-as-Code (IaC) practices using tools like Terraform, CloudFormation, or Ansible.
Monitor system performance, availability, and security, implementing improvements where necessary.
Develop and maintain configuration management solutions.
Collaborate closely with software engineering teams to optimize application performance and delivery.
Identify and resolve bottlenecks in development and deployment workflows.
Manage container orchestration environments (Kubernetes, Docker Swarm, ECS, etc.).
Implement and manage monitoring, logging, and ing systems (Prometheus, Grafana, ELK Stack, Datadog, etc.).
Advocate and implement security best practices in all DevOps processes (IAM, encryption, vulnerability scanning).
Participate in incident response, troubleshooting, and root cause analysis.
Mentor junior DevOps engineers and promote best practices across the organization.
Drive continuous improvement in infrastructure, deployment processes, and tooling.
Document system architecture, standards, and procedures for knowledge sharing and auditing.
Key Skills and Qualifications:
Bachelor’s Degree in Computer Science, Engineering, Information Systems, or a related field (or equivalent work experience).
5+ years of hands-on experience in DevOps, Site Reliability Engineering, or related roles.
Strong expertise in cloud infrastructure (AWS, Azure, or GCP certification is a plus).
Proficiency with IaC tools like Terraform, CloudFormation, or Pulumi.
Deep experience with containerization and orchestration tools (Docker, Kubernetes).
Strong scripting skills (Bash, Python, Go, or similar).
Hands-on experience with version control (Git) and DevOps CI/CD tools (Jenkins, GitLab CI, CircleCI, ArgoCD).
Familiarity with monitoring and logging solutions.
Understanding of networking concepts (DNS, TCP/IP, VPN, Load Balancing, etc.).
Knowledge of security standards and best practices in DevOps environments.
Strong problem-solving, collaboration, and communication skills.
More at rhsandbox