苹果AIML- Infrastructure Systems Engineer (Machine Learning), Machine Learning Platform Technologies
任职要求
Minimum Qualifications • Master or PhD degree in Computer Science, Electrical Engineering or equivalent • 5+ years of Systems or AIML production-service experience, commensurate with running cutting-edge hybrid cloud services in China and the rest of the world • Solid understanding of system architecture and large-scale service or computational platform operations • Demonstrated understanding of system management, covering aspects of configuration and usage accounting • Proficiency in coding with scripting and programming languages, including Bash, Python, Golang and Java - while having the ability to select the proper language as a tool to solve a certain problem • Experience in large-scale service and job deployment, using an orchestration framework (Kubernetes) and cloud services for large-scale projects • Experience in observability of system behaviors (e.g. Prometheus, Grafana) Preferred Qualifications • Self-motivated and proactive, with demonstrated creative and critical thinking capabilities • Ability to identify problems in depth, distinguishing purposes vs. measures without confusion • Strong sense of thoroughness, driving details, delive…
工作职责
The Infrastructure Systems Engineer will do the following tasks, through collaboration with team members in China and around the world. - Analyze the requirements, demands, constraints and challenges of machine learning in local or global environments, design or re-design platform architecture to improve its scalability and agility, and to enable new, high-impact use cases - Develop and implement the above design, bringing it to an internal product, with observability to support efficient system management - Design and/or enhance automation of operations for infrastructure and platforms, including tools and processes of monitoring, logging and alerting, to improve scalability in both system construction and runtime operations - Support Dev and Eng efforts through provisioning operational solutions, co-design ML application architecture and drive the coordination among local and global, internal and cross-functional groups to achieve the result of success - Create performance profile for platforms and services, defining service level objectives (SLO) and driving the measurement, monitoring and evaluation over these objectives - Lead constant evaluation on system performance and reliability, discover potential faults, drive RCA and fixes
• The Infrastructure Systems Engineer Intern will do the following tasks, through collaboration with team members in China and around the world. • - Analyze the requirements, demands, constraints and challenges of machine learning platform in local or global environments. Design or re-design platform architecture to improve its scalability and agility, and to enable new, high-impact use cases • - Investigate new technologies to enhance system performance, reliability and redundancy. Create performance profile for platforms and services, defining service level objectives (SLO) and driving the measurement, monitoring and evaluation over these objectives • - Improve automation of operations for infrastructure and platforms, including tools and processes of monitoring, logging and alerting, to improve scalability in both system construction and runtime operations • - Develop and implement the above design, bringing it to an internal product, with observability to support efficient systems management
As an AIML Data Operations Team Lead, you’ll be responsible for providing daily leadership and promoting the development of 25-30 Annotation Analyst team members. You are self-motivated, friendly and have a passion to support your team’s development through critical and creative thinking. You’ll identify, promote, and implement innovative ideas to meet and exceed performance and quality goals set by leadership. You will support employees’ success through establishing relationships and seeking to understand what motivates individuals. You will oversee performance management, ensuring daily, monthly, quarterly operational metrics are met. You’ll effectively execute on management and administrative tasks such as leading staff meetings, conducting regular one-on-one’s, hiring, training, and development of employee performance. You will drive operational improvements, team collaboration, and recommend innovative solutions to business challenges. You’ll establish clear and effective working relationships and communication channels with internal partners. This position comes with competitive pay, great benefits, eligibility to participate in our company stock plan, time off, an employee discount, and dedicated resources to support your ongoing growth and career development.
As part of the AIML Data Operations Team, you play a central role in enhancing the user experience and team collaboration. As an SPR supporting the AIML Data Operations organization you will engage with Account and Crowd Managers to define and align on project deliverables. Manage resource plan for analyst headcount, partnering with Account, Crowd Managers, Team Leads and Capacity Planning to align targets and goals across portfolio. Draft and implement policies related to performance and project management of all projects delivered from your site. Effectively communicate project status/metrics and insights to management. Partner with Account and Crowd Managers to ensure data accuracy and timeliness of data. Manage support for project approvals, prioritization and flag outliers for leadership. Coordinate site project rollout and execution (platform changes, HC, training, metrics). Bring in consistency in project delivery via standardized templates and project management best practices. Provide Tech support to analysts as needed. Work with tool owners to update/fix tools as projects come in. Successful candidates are highly organized, highly analytical, independent, able to make decisions based on data driven insights, and ready to dig into the details and get the job done despite practical or technical hurdles.
This role will collaborate closely with experts or engineers at Information Intelligence. You will receive hands-on mentorship and guidance throughout the project exploration, as well as gaining significant experience in software engineering and natural language processing. We anticipate you could use your deep knowledge to tackle meaningful technical problems, transfer your ideas into solutions in the next generation of Siri product.