Source description
About the role
Senior Data Analyst (Machine Learning) role.
Please ensure candidates are thoroughly screened before submission, with a strong emphasis on hands-on SQL, PySpark, Pandas, Python, and Machine Learning experience. Please do not submit candidates who cannot demonstrate strong hands-on SQL and PySpark coding skills. These areas will be heavily evaluated throughout the interview process.
Interview Process: Quick turnaround interviews
ANY VISA IS 0K
TOP REQUIREMENT – SQL WRITING QUERIES (Non-Negotiable)
This is the Hiring Manager's #1 requirement.
Candidates must be able to independently write, execute, and troubleshoot SQL queries. They will be technically screened by our team prior to submission and will also be tested onsite by the client.
Required skills include:
Data extraction
Joins, aggregations, and filtering
Basic to intermediate data transformations
Strong hands-on coding ability
Interview Format
In-person interviews only (NO video interviews)
Candidates must be local and available to attend onsite interviews
Live SQL and PySpark coding exercises will be conducted during the interview process
Many candidates struggle with hands-on coding, so please validate technical skills thoroughly before submission
Certifications
Databricks Data Engineer Certification is highly preferred and considered the most important certification
Certificatio1n can be completed post-onboarding if not currently held
Google Cloud (GCP) Certification is also acceptable and may be completed after hire
PySpark (Highest Priority Technical Skill)
Candidates must have strong hands-on experience with:
Building and maintaining data pipelines
Large-scale data processing
Data transformation and optimization
Working within distributed data environments
Python / Pandas
Required experience includes:
Data cleansing and preparation
Data transformation
Exploratory data analysis
Data manipulation using Pandas
Machine Learning (Approximately 20% of the Role)
Candidates must already possess hands-on Machine Learning experience. No ML training will be provided.
Required experience includes:
Classification models
Regression models
Clustering techniques
Model evaluation and performance measurement
Data Analytics Focus
The majority of the role will involve:
SQL-based data extraction and analysis
PySpark data processing
Pandas-driven analysis
Generating business insights from large datasets
End-to-End Analytics Experience
Candidates should understand complete workflows that integrate:
SQL
PySpark
Python/Pandas
Machine Learning models and analytics
Team Structure
10 Data Engineers
6 Data Scientists
More at 3B Staffing