Tether's AI model team seeks an engineer to drive model serving and inference architectures for advanced AI systems. You will optimize deployment across resource-constrained devices and edge platforms to deliver high throughput and low latency in real-world scenarios.
Collaborating with cross-functional teams, you will design robust inference pipelines, establish performance metrics, and push innovations in diffusion models, vision transformers, pruning, and quantization.
#J-18808-Ljbffr