Source description
About the role
Title: Data/Database Engineer Location: Remote – EST Duration:12 Months
Must Haves: Senior Database Engineer, who will support administration teams on both SQL Server (SQL, Stored Procedures) and Big Data Engineering platforms (Python/Pyspark/Datababricks). can lean more SQL Server, SSIS-ETL, Stored Procedures, Performance Tuning with some newer experience in data engineering using Python and/or Databricks currently supporting SQL Server on-premise but they are moving to cloud and big data (Python/Pyspark/Databricks) SQL Data Engineer with some recent big data experience
Day to Day Responsibilities Support the administration team supporting SQL Server Databases, Databricks Development and or administration, and be tasked with creating necessary automation processes via scripts/procedures to support the business need (usually around system compliance policies). Demonstrates a knowledge of the general development life cycle with ability to lead design review sessions, testing/validation feedback sessions, and own the entirety of a project from requirement gathering to internal and end user facing documentation. Assemble large, complex datasets that meet functional / non-functional business requirements. Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability, etc. Build the infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources including but not limited to SQL, Databricks, and Pyspark.
Qualifications
Bachelor’s Degree and minimum of 9-12+ years of experience of Data background including both database development and data engineering. 8+ years of hands-on experience with a wide variety of database development with Microsoft SQL Server including SQL database design. 5+ years’ experience with writing SQL queries and Stored Procedures. 3+ years’ experience with SSIS for Extract, Transfer, Load (ETL) solutions. 2+ years of experience with object-oriented Python programming or 1-2 years’ experience with Databricks. 1+ years’ experience with Spark (Pyspark). Experience with PowerShell scripting for automation. Exposure to data models, data mining, and segmentation techniques. General understanding of data pipelines; should be able to work with REST, SOAP, FTP, HTTP, and ODBC. Capable of supporting and working with cross-functional teams in a dynamic environment Familiarity with Cloud platforms such as AWS, Azure or GCP desired.
More at Thoughtwave Software and Solutions