Source description
About the role
Specific Responsibilities include: Gaining an understanding of our Current State and then defining the Target State Architecture and Strategy. Work as the Technical Product Owner in partnership with our Business Product Owners to define the technical strategy for our data driven products and applications. Drive towards adoption of the target state architecture by executing on the strategy. Must possess the leadership qualities necessary to drive change and adoption of the strategy. Develop Architectures for highly scalable and fault-tolerant applications using Cloudera/Hadoop and Relational database technologies. Provide technical and architectural oversight for systems and projects that are required to be reliable, massively scalable, highly available (99.999% uptime), and maintainable. Introduce best practices and principles to enable consistent delivery and enable alignment with long term direction. Lead and mentor other team members. Foster development best practices within the team. Identify and drive process improvements. Facilitate communication across groups. Work with our product organization to develop business requirements into architecture and integrate into our long-term platform strategy. Define solution level architecture for engineering teams including guidance on development tools, target platforms, operations, and security. Provide leadership to the development teams for successful project implementation on the selected data platform. Provide expertise to team engineers as needed. Stay up to date on new tools techniques in the data space. Conduct proof of concept activities with key business users in support of advanced use cases. Qualifications BS or MS in Computer Science or related degree from an accredited university. 10 + years of experience architecting, designing and developing large scale data solutions utilizing a mixture of Big Data and Relational database platforms. Direct experience with Cloudera s Hadoop based technologies. Proven experience leading teams resulting in the successful deployment of applications built on data platforms. Advanced Relational Database Experience (RDBMS) in one or more of the following: Microsoft SQL Server. Oracle. MySQL. Experience as the technical lead, organizing and mentoring junior and intermediate level developers/DBAs. Experience developing software with Java. Proven Linux experience including; Basic Administration. Files and Permissions. Directory Navigation. Job Scheduling. Shell Scripts. Hadoop Distributed Data Files System experience including: Use and set up of Blocks, Name nodes, Data nodes. File Systems Interfaces, parallel copies, cluster balance and archiving. Scaling Out including data flow, combiner functions, running distributed jobs. Data "Ingestion" techniques such as streaming, ETL, etc. Hadoop Pipes. Sorting, joins and side data distributions Experience with the wider ecosystem of Hadoop based Technologies, such as: Hive/Impala and other Map Reduce techniques and approaches. Spark. Kafka. Sqoop and Flume. Oozie. Presenting Results to Business users: Selecting appropriate visualization paradigms and tools. Building Search UI with HUE. Powering custom Web applications with Impala and Search. Traditional BI tools for Batch and Hive presentation.
More at Koo