Centific AI Research seeks a PhD Research Intern to design and evaluate reinforcement learning systems for agentic AI workflows. You will build RL environments, reward models, and post-training pipelines for LLM-based agents, translating research into practical enterprise solutions.
Location options include Palo Alto or Remote work, duration 3–6 months, with a competitive stipend and mentorship from researchers and engineers.
#J-18808-Ljbffr