Source description
About the role
Job description Job Title: Solution Architect (Python, Spark, AWS) Company: RSquareSoft Technologies Location: Pune Employment Type: Full-Time Experience Level: 6-10 years Job Overview: We are looking for a Solution Architect with a strong Data Engineering background to drive the architecture, design, and technical direction of complex, cloud-based data platforms. This role requires deep hands-on expertise in big data technologies along with the ability to design scalable, secure, and high-performance data solutions aligned with business needs. Key Responsibilities: · Own end-to-end solution architecture for cloud-native data platforms and big data ecosystems. · Design scalable, resilient, and secure architectures using AWS and distributed data processing frameworks. · Lead architecture and design discussions, including high-level and low-level design reviews. · Provide hands-on technical guidance in Python, PySpark, Spark, and Hadoop ecosystem. · Design and optimize data pipelines, ETL/ELT workflows, and large-scale data processing systems. · Define and enforce architectural standards, best practices, and design principles for data engineering. · Collaborate with engineering, data, product, and business stakeholders to translate requirements into data-driven solutions. · Review critical technical designs and implementations to ensure alignment with target data architecture. · Identify and mitigate architectural risks related to performance, scalability, data quality, and cost optimization. · Support teams in resolving complex data engineering and system design challenges. Required Skills and Qualifications: · Strong experience in Cloud Architecture (AWS) including data services (S3, EMR, Glue, Redshift, etc.). · Solid background in Big Data technologies and distributed systems. · Strong programming experience in Python with expertise in PySpark/Spark. · Expertise in building and optimizing large-scale data pipelines and batch/stream processing systems. · Strong SQL skills for data modeling, transformation, and performance optimization. · Proven expertise in system design and data architecture for scalable and distributed systems. · Experience creating and reviewing HLDs and LLDs.