Source description
About the role
Role Overview: As a Python Automation Engineer specializing in web scraping and AI, you will be responsible for developing and maintaining scalable web scraping and browser automation systems for structured and unstructured web data extraction. Your role will involve building advanced scraping workflows, handling dynamic JavaScript-based websites, and creating AI-powered extraction and automation workflows. Additionally, you will work on optimizing crawler performance, collaborating with cross-functional teams, and monitoring scraping jobs to ensure automation efficiency. Key Responsibilities: - Develop and maintain scalable web scraping and browser automation systems for structured and unstructured web data extraction. - Build advanced scraping workflows using Playwright, Selenium, Scrapy, APIs, and Python automation frameworks. - Handle dynamic JavaScript-based websites, browser rendering challenges, anti-bot protections, proxies, and session management. - Build AI-powered extraction and automation workflows using LLMs, AI agents, and modern AI automation frameworks. - Work on unstructured content extraction, intelligent parsing, classification, summarization, and data enrichment workflows using AI models. - Optimize crawler performance, improve scraping reliability, and maintain high-quality data pipelines. - Collaborate with AI, engineering, and product teams for scalable data acquisition and automation projects. - Monitor scraping jobs, debug failures, and improve automation efficiency across large-scale scraping systems. Qualifications: - Bachelor's degree in Computer Science, Engineering, IT, or related field. - Strong proficiency in Python programming and browser automation frameworks. - Good understanding of JavaScript-rendered websites, APIs, automation workflows, and scalable scraping systems. - Experience using LLM APIs and AI automation tools for unstructured data workflows is highly preferred. - Familiarity with Docker, Linux, Git, AWS/GCP, or cloud deployment environments is a plus. - Experience using Playwright, Selenium, Scrapy, APIs, and Python automation frameworks. Role Overview: As a Python Automation Engineer specializing in web scraping and AI, you will be responsible for developing and maintaining scalable web scraping and browser automation systems for structured and unstructured web data extraction. Your role will involve building advanced scraping workflows, handling dynamic JavaScript-based websites, and creating AI-powered extraction and automation workflows. Additionally, you will work on optimizing crawler performance, collaborating with cross-functional teams, and monitoring scraping jobs to ensure automation efficiency. Key Responsibilities: - Develop and maintain scalable web scraping and browser automation systems for structured and unstructured web data extraction. - Build advanced scraping workflows using Playwright, Selenium, Scrapy, APIs, and Python automation frameworks. - Handle dynamic JavaScript-based websites, browser rendering challenges, anti-bot protections, proxies, and session management. - Build AI-powered extraction and automation workflows using LLMs, AI agents, and modern AI automation frameworks. - Work on unstructured content extraction, intelligent parsing, classification, summarization, and data enrichment workflows using AI models. - Optimize crawler performance, improve scraping reliability, and maintain high-quality data pipelines. - Collaborate with AI, engineering, and product teams for scalable data acquisition and automation projects. - Monitor scraping jobs, debug failures, and improve automation efficiency across large-scale scraping systems. Qualifications: - Bachelor's degree in Computer Science, Engineering, IT, or related field. - Strong proficiency in Python programming and browser automation frameworks. - Good understanding of JavaScript-rendered websites, APIs, automation workflows, and scalable scraping systems. - Experience using LLM APIs and AI automation tools for unstructured data workflows is highly preferred. - Familiarity with Docker, Linux, Git, AWS/GCP, or cloud deployment environments is a plus. - Experience using Playwright, Selenium, Scrapy, APIs, and Python automation frameworks.
More at AIMLEAP