CareerPlanSign in

智能体-强化学习算法研究员-CodeBuddy/WorkBuddy

Shenzhen, China💼 Full-time🗓 2026-09-28

Core

Researching and designing Agentic Workflow and Agentic Memory solutions to solve problems in the code domain, with a focus on Reinforcement Learning (RL) that outperforms Supervised Fine-Tuning (SFT).

Role type

Senior Research Scientist (Agentic RL & Engineering)

Builds

Scalable Agentic Workflow solutions for code assistance

Domain

Artificial Intelligence / Reinforcement Learning / Software Engineering

Deliverable

production ML models

Required skills

Reinforcement Learning, Agentic Workflow design, Agentic Memory architecture, Deep Learning optimization, LLM inference optimization, Python, C/C++, Golang, Java, JavaScript, TypeScript

Preferred skills

Publications in top-tier conferences (ACL, EMNLP, NeurIPS, ICML, ICLR)

Responsibilities

Designing memory access logic for agents; Adjusting training corpora and objectives based on user data and agent implementation details; Designing frontend and backend architectures to optimize LLM inference costs; Collaborating with agent developers to align model training with agent design.

Seniority

Senior, hands-on IC

Sourced via tencent · Listed on CareerPlan, which tracks 844,000+ jobs from 20+ sources.