Source description
About the role
Full Description About the Role We are seeking an experienced Principal Software Engineer to join our platform team focusing on AI/ML Platform (AMP). As a key member of this team, you will be responsible for designing, implementing and managing software systems for the AI/ML Platform and orchestrating the full ML development lifecycle for partner teams. Key Responsibilities: System Design: You will design high-scale deployment architectures and observability for AI/ML models. Mentoring: Share knowledge, best practices, and participate in design reviews to step up expertise at the team level. Multi-cloud Architecture: Define components that leverage strengths from multiple cloud platforms (e.g., AWS, Azure) to optimize performance, cost, and scalability. AI/ML Observability: Build systems for monitoring performance of AI/ML models and finding insights on underlying data such as drift detection, data fairness/bias, and anomalies. ML Solution Deployment: Develop tools for building and deploying ML artefacts in production environments and facilitating a smooth transition from development to deployment. Big Data Management: Automate and orchestrate tasks related to managing big data transformation and processing and build large-scale data stores for ML artifacts. Scalable Services: Design and implement low-latency, scalable prediction, and inference services to support diverse needs of users. Cross-Functional Collaboration: Collaborate across diverse teams, including machine learning researchers, developers, product managers, software architects, and operations, fostering a collaborative and cohesive work environment. End-to-end Ownership: Take end-to-end ownership of components and work with other engineers in the team, including design, architecture, implementation, rollout, and onboarding support to partner teams, production on-call support, testing/verification, investigations, etc. ## Full Description About the Role We are seeking an experienced Principal Software Engineer to join our platform team focusing on AI/ML Platform (AMP). As a key member of this team, you will be responsible for designing, implementing and managing software systems for the AI/ML Platform and orchestrating the full ML development lifecycle for partner teams. Key Responsibilities: System Design: You will design high-scale deployment architectures and observability for AI/ML models. Mentoring: Share knowledge, best practices, and participate in design reviews to step up expertise at the team level. Multi-cloud Architecture: Define components that leverage strengths from multiple cloud platforms (e.g., AWS, Azure) to optimize performance, cost, and scalability. AI/ML Observability: Build systems for monitoring performance of AI/ML models and finding insights on underlying data such as drift detection, data fairness/bias, and anomalies. ML Solution Deployment: Develop tools for building and deploying ML artefacts in production environments and facilitating a smooth transition from development to deployment. Big Data Management: Automate and orchestrate tasks related to managing big data transformation and processing and build large-scale data stores for ML artifacts. Scalable Services: Design and implement low-latency, scalable prediction, and inference services to support diverse needs of users. Cross-Functional Collaboration: Collaborate across diverse teams, including machine learning researchers, developers, product managers, software architects, and operations, fostering a collaborative and cohesive work environment. End-to-end Ownership: Take end-to-end ownership of components and work with other engineers in the team, including design, architecture, implementation, rollout, and onboarding support to partner teams, production on-call support, testing/verification, investigations, etc.
More at Autodesk