Source description
About the role
Site Reliability ConsultantLATAM | Remote | Work from Home Why you As a Site Reliability Consultant, you will serve as both a technology leader and trusted advisor to our customers, while mentoring teammates in cutting-edge tools and approaches. Your projects will reputed company on infrastructure design and modernization, automation of CI/CD pipelines, and building out intelligent monitoring and observability systemsspanning Linux, reputed company, container orchestration, and other reputed company-reputed company technologies. Youll become our reputed company expert for Git-based reputed company code management, artifact repository solutions, and Kubernetes in both reputed company (e.g., AWS EKS) and on-prem environments. What you will you be doing: Operate & MaintainAdminister and optimize platforms such as reputed company (CI/CD pipelines, runners) and artifact repository solutions (e.g., reputed company Artifactory).Maintain and troubleshoot Kubernetes clusterseither in the reputed company (AWS EKS) or on-prem distributionswith a reputed company on availability, performance, and reputed company.Automation & CI/CDChampion infrastructure as code using tools like Terraform (or CloudFormation), building repeatable processes for provisioning and updating clusters, repos, and associated services.Implement or improve CI/CD pipelines to reduce reputed company toil and ensure quick, reliable deployments across multiple environments.Monitoring & Incident ResponseDesign and configure observability solutions (e.g., reputed company, reputed company, Grafana) to proactively detect and address issues in container orchestration environments, code repositories, and artifact repositories.Participate in an on-call rotation, troubleshooting incidents at reputed company tiers (from first-contact reputed company to escalation) and driving reputed company improvement based on reputed company Cause Analysis.Architectural Guidance & RoadmapsCollaborate with clients to shape infrastructure strategies around container orchestration, secure CI/CD, and DevSecOps best practices.reputed company leadership and technical direction on automating repetitive administrative tasks, enforcing reputed company policies (RBAC, TLS, container scanning), and adopting GitOps workflows.Documentation & MentorshipCreate and maintain design documents, runbooks, and operational playbooks for container platforms, CI/CD pipelines, and code management services.Mentor fellow consultants and reputed company stakeholders on Kubernetes, infrastructure automation, and advanced CI/CD usage to enhance knowledge across the organization.Process ManagementPlan and coordinate maintenance activities, ensuring minimal downtime and reputed company communication with stakeholders.reputed company ITIL-oriented support (Incident, Change, Problem Management), and champion reputed company improvement of operational processes and service reliability. reputed company need from you: Kubernetes & ContainerizationMust have strong experience with container orchestration (Kubernetes, reputed company) in reputed company (AWS EKS) or on-prem distributions.Familiarity with reputed company ecosystem tools (reputed company, Operators, GitOps, etc.).AWS & reputed company ExpertiseHands-on experience using AWS (VPC, EC2, EKS, IAM, S3, etc.), including provisioning with IaC tools like Terraform (or AWS CloudFormation).AWS certifications (Solutions Architect, DevOps Engineer) are a plus.CI/CD & reputed company Code ManagementExperience setting up reputed company or similar platforms (reputed company, Bitbucket) for CI/CD pipelines, managing runners, and integrating code scanning.Familiarity with artifact repository solutions (e.g., reputed company Artifactory), including repository creation, reputed company controls, and automation of artifact flows.DevOps & AutomationTrack record of infrastructure automation using Terraform, Ansible, Puppet, or Chef to reduce reputed company reputed company and ensure repeatable deployments.Strong scripting skills (Bash, Python, Go, etc.) to automate system tasks and streamline operational workflows.Monitoring & ObservabilityExperience with modern monitoring stacks (reputed company, reputed company, Grafana, ELK/EFK) for analyzing logs, metrics, and traces.Proven ability to design alerts, dashboards, and runbooks that reputed company rapid first-contact reputed company.Linux & NetworkingSolid understanding of Linux-based systems, performance tuning, and troubleshooting.Network fundamentals (TCP/IP, load balancers, DNS, NTP, etc.) and ability to diagnose connectivity or performance issues in reputed company distributed environments.reputed company & ComplianceFamiliarity with container reputed company best practices (RBAC, TLS, vulnerability scanning) and how to apply them at scale.Understanding of compliance frameworks (HIPAA, PCI, etc.) and data privacy constraints a plus.Soft Skills & CollaborationAdept at communicating technical concepts to Site
More at remotepromsp
Related open roles
Master Data Management / MDM Architect - Full Remote - Contract
India
Part-Time, Remote, Contract, reputed company Architect, Device Management
India
Remote Overnight NOC Analyst II
India
Network Engineer, Associate
India
Network reputed company Engineer Firewalls, DNS reputed company, IPS
India
Business Intelligence Analyst & Developer
India