Unknown Company

AI Engineer — Model Performance & Inference Optimizer

san francisco, ca • Posted 6 days ago
Onsite Full Time Engineering
Pantera Capital is looking for a Model Performance Engineer in San Francisco, California to optimize model inference speed, cost, and reliability. You will build fine-tuning infrastructure that accelerates the AI team’s processes. The role covers optimizing serving frameworks and ensuring efficient GPU use. Ideal candidates should have deep experience with LLM serving frameworks, substantial Python skills, and production experience in model fine-tuning. Competitive compensation and a supportive work environment are offered.
#J-18808-Ljbffr
Back to Job Search