Unknown Company

Remote AI Model Compression & Quantization Engineer

Remote • Posted 3 days ago
Onsite Full Time Engineering

Tether's AI model team seeks an engineer to drive model serving and inference architectures for advanced AI systems. You will optimize deployment across resource-constrained devices and edge platforms to deliver high throughput and low latency in real-world scenarios.

Collaborating with cross-functional teams, you will design robust inference pipelines, establish performance metrics, and push innovations in diffusion models, vision transformers, pruning, and quantization.

#J-18808-Ljbffr
Back to Job Search