Amazon.com Services LLC is seeking a Senior Inference Engineer to own the inference stack for real-time multimodal conversational AI. You will shape model architectures for servable deployment, build real-time runtimes within tight latency budgets, and develop offline systems for training and RL integration.
You will collaborate with scientists and hardware teams to ensure fast, cost-effective inference at scale, across training, evaluation, and deployment pathways.
#J-18808-LjbffrSenior Real-Time Multimodal Inference Engineer in boston at Unknown Company
This position is listed as full time and onsite.