Senior AI Platform Engineer Opportunity Remote Considered
Quick Overview
Job Description
Senior AI Platform Engineer
Full Time-Office Located in Mountain View, CA (100% remote considered)
Opportunity:
Are you ready to be at the forefront of integrating machine learning with healthcare technology? We are seeking a dynamic and innovative ML Ops Engineer. The ideal candidate is self-driven, versatile in handling multiple projects, and a collaborative team player. You will be instrumental in developing our cutting-edge machine learning platform and enhancing our existing healthcare solutions. We value individuals who are adept at working with complex systems and possess exceptional communication and leadership skills.
Responsibilities:
- Architect, design, and build robust and efficient AI/ML systems for a production environment, focusing on backend distributed systems, microservices, and ensuring system accuracy.
- Lead cross-functional initiatives end-to-end (scoping, timelines, dependencies), driving alignment across Data, Infra and ML Engineering.
- Collaborate closely with AI/ML engineers to optimize workflows for model training, real-time inference, monitoring, and troubleshooting
- Be a subject matter expert on ML infrastructure, providing guidance to both internal teams and external stakeholders.
- Ensuring operational excellence and reliability of ML systems. Define and enforce SLAs around system performance, including latency, throughput, and resource utilization.
- Develop tools for effective model management, continuous monitoring, and enhancing the efficiency and effectiveness of the entire ML lifecycle.
- Actively engage in mentorship and knowledge sharing to promote a culture of continuous learning and improvement within the team.
Qualifications:
- Bachelor's or Master's degree in Computer Science, Engineering, or a related field with 5+ years of industry experience
- Familiarity with architectural frameworks of large, distributed, and high-scale ML applications. Experience in the implementation of applications using LLM s and GenAI is a huge plus.
- Solid understanding of MLOps, data structures, and software design principles.
- Proficiency in programming with Python and experience in other languages like Java or Go.
- Strong knowledge in deploying scalable machine learning models, including experience with Docker, Kubernetes, and microservices architecture.
- Experience with database technologies (e.g., SQL, NoSQL) and big data processing frameworks (e.g., Spark) is a plus
Skills
Similar jobs
Site Reliability Engineer (SRE)- Kubernetes
Spiceorb · United States
23 minutes agoDEVOPS ENGINEER
AaraTechnologies Inc · Virginia Beach, United States
25 minutes agoRelease Engineer (CI/CD, Infrastructure & Automation)
Openmind Technologies · United States
27 minutes agoNetwork Engineer
TEKsystems c/o Allegis Group · Indian Head, United States
5 hours ago$45 - $50/hrSenior DevOps Architect
Select Solutions Group, LLC · United States
5 hours agoSenior Devops Engineer/Lead
Infinite Computer Solutions (ICS) · Dallas, United States
5 hours ago$80k/yr