Unknown Company

Remote Research Intern - Agentic RL for Enterprise AI

Remote • Posted Yesterday
Remote Full Time Other

Centific AI Research seeks a PhD Research Intern to design and evaluate reinforcement learning systems for agentic AI workflows. You will build RL environments, reward models, and post-training pipelines for LLM-based agents, translating research into practical enterprise solutions.

Location options include Palo Alto or Remote work, duration 3–6 months, with a competitive stipend and mentorship from researchers and engineers.

#J-18808-Ljbffr
Back to Job Search