Source description
About the role
Job description The core premise for the our Platform Engineer lies in designing building and maintaining internal platforms to provide developers with selfservice tools and automation for deploying and running software efficiently and reliably. We code our way out of problems where operations are concerned addressing availability scalability latency and efficiency challenges within the vast infrastructure here. You will impact millions of people all over the globe with your creative solutions You work in one of the biggest ecommerce companies in the world You will solve exciting problems at scale by writing and deploying code across tens of thousands of servers Ensuring an everything as code mindset for yourself and your team You will have the opportunity to collaborate with many of the worlds leading SREs SWEs and Platform Engineers You will be free to launch your own ideas and solutions within our sophisticated production environment Here are some of the tools and technologies we use to achieve this Python Go Puppet Kubernetes Elasticsearch Prometheus HAProxy Cassandra Kafka etc What youll be doing Design develop and implement highly distributed large scale platforms that have a major impact on developers Deliver a fully integrated endtoend developer focused engineering experience to make common things simple and complex things possible Embed operational resilience and service reliability directly into the development lifecycle Build effective monitoring to supervise the health of your system and jump in to handle outages Build and run capacity tests to manage the growth of your systems Embed Business Continuity and IT Disaster Recovery DR planning into the full engineering lifecycle Embed security and privacy directly into service bootstrapping and delivery pipelines by automating controls and reducing manual decision points Be an advocate of engineering standard processes Share the oncall rotation and be an escalation contact for incidents Contribute to oraganization growth through interviewing onboarding or other recruitment efforts What youll bring 5-8 years of hands-on experience in software platform or site reliability engineering within the technology sector coupled with expertise in building operating and maintaining sophisticated and scalable systems Solid experience in at least one programming language we use Go Python Java Ruby Perl Experience with Infrastructure as Code technologies Knowledge of cloud computing fundamentals Solid foundation in Linux administration and troubleshooting Understanding of Service level agreements and objectives Additional experience in AWS solutions Kubernetes Networking Security or Storage is desirable Supervising observability technologies like Prometheus Graphite Grafana Kibana Elasticsearch are a plus Disclaimer : This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
More at Rarr Technologies