Source description
About the role
Design and implement scalable, fault-tolerant infrastructure for our web applications and blockchain systems Build and maintain CI/CD pipelines, deployment automation, and infrastructure as code Develop comprehensive monitoring, alerting, and observability systems to ensure system reliabilityLead incident response efforts and conduct post-incident reviews to improve system resilience Implement security best practices across all infrastructure and deployment processes Collaborate with engineering teams to optimize application performance and reliability Establish SLIs, SLOs, and error budgets to measure and improve service reliability Automate operational tasks and reduce manual toil through tooling and process improvements Manage capacity planning and cost optimization for cloud resources
More at Superstate
