ServiceTitan is hiring a Senior AI Engineer to build scalable, low-latency machine learning systems and production ML services using Python and Azure.
Responsibilities
- Design and implement scalable, low-latency ML systems and services, covering model training, inference, and deployment infrastructure
- Partner with product teams to integrate, operationalize, and deploy machine learning models into production, including foundational models (such as LLMs), traditional ML, and specialized models for speech, NLP, and forecasting
- Improve performance and real-time responsiveness of ML pipelines to support strong customer experience with minimal latency
- Communicate effectively with engineers, product managers, customers, and partners
Requirements
- 5+ years of experience writing production-level code in Python
- 2+ years of experience deploying machine learning models (NLP, speech, forecasting, recommendation systems) to production at scale
- Experience designing, building, and deploying scalable ML inference services for real-time or high-throughput applications
- Advanced knowledge of machine learning methods and algorithms, including traditional ML and deep learning theory and techniques
- Experience with reinforcement learning (RL) techniques and implementing continuous learning or online optimization systems for production
- Strong understanding of the machine learning lifecycle from experimentation to production deployment
- Experience with data platforms such as SQL Server , PostgreSQL , Redshift , and Snowflake
- Experience with public cloud environments such as Azure or AWS
- Experience with microservices and asynchronous messaging for high-volume, real-time systems, including Kafka and Azure Service Bus (critical)
- Experience with serverless architecture such as Azure Functions and AWS Lambda
- Experience with Azure Cognitive Services or similar cloud AI offerings is a plus
- Experience with developer tools and practices including Git , unit testing, debugging, profiling, and JIRA
- Comfort working in a fast-paced environment with a globally distributed team
Technologies
- Python, Azure Cloud, AIOps, MLOps
- LLMs, NLP, speech, forecasting
- SQL Server, PostgreSQL, Redshift, Snowflake
- Azure, AWS
- Kafka, Azure Service Bus
- Azure Functions, AWS Lambda
- Azure Cognitive Services
- Git, JIRA, microservices, serverless architecture
Benefits
- Flextime, recognition, and support for autonomous work
- Flexible time off with ample learning and development opportunities
- Comprehensive onboarding program
- Leadership training for Titans at all levels
- Bonusly, peer-nominated awards, and more
- Company-paid medical, dental, and vision (with 100% employer paid options and 90% coverage for dependents)
- FSA and HSA
- 401k match
- Telehealth options including memberships to One Medical
- Parental leave and support
- Up to $20k in fertility services (IUI and IVF)
- Surrogacy and adoption reimbursement
- On demand maternity support through Maven Maternity
- Free breast milk shipping through Maven Milk
- Pet insurance
- Legal advisory services
- Financial planning tools
Compensation and Location
- Location: California (onsite)
- Salary: USD 168,200 - 269,900 per year
- Minimum experience: 5 years
Recruitment AI Use
- ServiceTitan uses automated and AI-assisted tools to support certain aspects of its recruitment process
- AI tools are not used to make hiring decisions; hiring decisions are made by the hiring teams
Senior AI Engineer in san francisco at Unknown Company
This position is listed as full time and onsite.