返回全部职位

    RL AI Research Scientist

    ResearchRemote (US/Singapore Preferred)Full-time

    Design, implement, and scale novel reinforcement learning algorithms that form the core of Pokee's AI agent platform.

    关于该职位

    As an RL Research Scientist, you will design, implement, and scale novel reinforcement learning algorithms that form the core of Pokee’s AI agent platform. You’ll work at the frontier of RL applied to real-world enterprise tasks—developing methods for context selection, long-horizon planning, and reward shaping that enable agents to operate reliably at scale.

    你将做什么

    • Design and implement novel RL algorithms for training AI agents on complex, multi-step enterprise workflows
    • Develop and refine reward modeling, context selection, and policy optimization techniques that improve agent accuracy over extended task horizons
    • Run large-scale experiments, analyze results rigorously, and translate research findings into production-ready components
    • Collaborate closely with infrastructure engineers to ensure research prototypes scale efficiently on both cloud and on-device hardware
    • Contribute to the company’s intellectual property through publications, patents, and open-source contributions
    • Stay current with the latest advances in RL, LLM fine-tuning, and AI agent architectures, and propose new research directions

    我们在寻找

    必填

    • PhD (or equivalent research experience) in Reinforcement Learning, Machine Learning, or a closely related field
    • Strong publication record at top venues (NeurIPS, ICML, ICLR, AAAI, or equivalent)
    • Deep expertise in RL fundamentals: policy gradient methods, value-based methods, model-based RL, multi-agent RL, or RLHF/RLAIF
    • Proficiency in Python and at least one deep learning framework (PyTorch strongly preferred)
    • Experience training and fine-tuning large language models is a significant plus
    • Demonstrated ability to take research from prototype to production

    加分项

    • Experience with on-device or edge inference optimization (quantization, distillation, MoE architectures)
    • Familiarity with enterprise software deployment, compliance, or regulated industries
    • Track record of open-source contributions in RL or LLM ecosystems
    • Experience with distributed training at scale (FSDP, DeepSpeed, Megatron)

    你是谁

    You want to join a small, elite team solving one of the hardest problems in AI—building agents that actually work in the real world. You’ll have direct impact on the product, access to cutting-edge research, and the opportunity to shape the future of enterprise AI from the ground up.

    申请 RL AI Research Scientist

    准备好加入了吗?填写下方表单即可申请。

    你从哪里得知这个机会?(可多选)
    在 LinkedIn 上关注
    Pokee Logo

    Pokee AI

    企业级 AI 智能体,部署在你自己的基础设施中。

    解决方案

    金融医疗健康电商法务教育制造业科技

    公司

    招聘安全联系我们

    关注我们

    TwitterLinkedInRedditDiscord

    资源

    API 文档

    法律

    服务条款隐私政策无障碍系统状态

    © 2026 Pokee AI。保留所有权利。

    条款与条件隐私政策