招聘中 社招

大模型算法工程师(行程规划方向)(MJ031345)

内容团队

  • 上海
  • AI & BI

收录时间

岗位职责

  1. 参与构建旅游领域AI Agent系统,主导旅游垂类模型的训练与调优,解决多城市、多约束条件下的动态路线优化问题。
  2. 设计强化学习(RL)框架,结合PPO、GRPO等算法实现,在行程规划领域达到优于顶尖LLM或Agent能力。
  3. 构建端到端模型评测体系:设计多维度评估指标,研发大模型与传统运筹算法的融合架构。
  4. 开发个性化推荐能力,整合文本、POI、用户行为数据,推动生成式AI在旅游场景的落地应用。 【

职位要求

】 基础要求:

  1. 计算机科学、应用数学或运筹学硕士及以上学历,1年以上大模型全流程实战项目经验(数据构建→训练→评测→部署)。
  2. 深度掌握模型训练与评测技术:
  3. 精通大模型微调技术及分布式训练框架(DeepSpeed、Megatron)
  4. 具备强化学习实战经验,熟悉PPO、GRPO等算法在决策优化场景的应用
  5. 掌握模型评测方法论(自动指标+人工评估)及A/B测试设计 优先条件:
  6. RL项目经验:有基于强化学习的决策优化系统(如路径规划)落地经验者优先
  7. 大模型训练经验:主导过模型的训练/微调/评测全流程者优先 Job Responsibilities Participate in building the AI Agent system in the tourism field, lead the training and tuning of tourism vertical models, and solve dynamic route optimization problems under multi-city and multi-constraint conditions. Design reinforcement learning (RL) frameworks, implement them with algorithms such as PPO and GRPO, and achieve performance superior to top-tier LLMs or Agents in the itinerary planning field. Construct an end-to-end model evaluation system: design multi-dimensional evaluation metrics and develop a fusion architecture of large models and traditional operations research algorithms. Develop personalized recommendation capabilities, integrate text, POI, and user behavior data, and promote the practical application of generative AI in tourism scenarios. Qualifications Basic Requirements: Master's degree or above in Computer Science, Applied Mathematics, Operations Research or related majors, with more than 1 year of practical experience in the full process of large model projects (data construction → training → evaluation → deployment). Have a deep grasp of model training and evaluation technologies. Proficient in large model fine-tuning technologies and distributed training frameworks (DeepSpeed, Megatron). Possess practical experience in reinforcement learning, and be familiar with the application of algorithms such as PPO and GRPO in decision optimization scenarios. Master model evaluation methodologies (automatic metrics + manual evaluation) and A/B test design. Preferred Qualifications: RL project experience: Priority is given to candidates with practical experience in deploying decision optimization systems (such as path planning) based on reinforcement learning. Large model training experience: Priority is given to candidates who have led the full process of model training/fine-tuning/evaluation.

相似岗位

有些机会,只在官网短暂出现

真正值得关注的岗位,常常只在企业官网短暂开放,可能两三天后就下线,也未必会同步到综合招聘平台。没有持续关注,你甚至不会知道它曾经出现。职先机持续聚合并核验官网岗位,帮你抓住职场先机,快人一步。

微信小程序 / 当前岗位

在微信里继续看这个岗位

内容团队大模型算法工程师(行程规划方向)(MJ031345)

扫码直达当前岗位收藏、浏览记录和下线提醒留在微信里
电脑端可直接微信扫码;手机端可保存小程序码后在微信中识别。
微信扫码在职先机小程序查看大模型算法工程师(行程规划方向)(MJ031345)正在准备岗位码
职先机微信小程序一岗一码 · 正式版直达保存小程序码

岗位提醒 · 01

收藏当前岗位,下线及时通知

当前岗位大模型算法工程师(行程规划方向)(MJ031345)内容团队 · 上海

01

扫码进入当前岗位无需重新搜索,直接打开当前岗位详情。

02

收藏并开启提醒在小程序内收藏后,岗位下线或变化时及时通知。

03

继续查看相似机会原岗位变化时,可继续浏览相关在招岗位。

职先机会持续核验公开岗位;岗位状态与最终招聘结果仍以企业招聘官网为准。

微信小程序扫码打开当前岗位

扫码打开当前岗位,完成收藏并开启下线提醒。

微信扫码在职先机小程序收藏大模型算法工程师(行程规划方向)(MJ031345)正在准备岗位码

微信小程序 · 当前岗位

请前往微信小程序提交内推申请

内容团队大模型算法工程师(行程规划方向)(MJ031345)社招 · 上海

01进入当前岗位扫码或打开后直达本岗位内推申请页。

02提交 PDF 简历简历替换、处理进度和消息提醒在小程序内同步。

手机端将尝试直接打开微信;电脑端请使用微信扫码。

微信扫码进入大模型算法工程师(行程规划方向)(MJ031345)内推申请页正在准备岗位码
当前岗位专属码扫码后无需重新搜索岗位保存小程序码

内推提示

暂不支持该岗位内推

如需投递,建议前往官网进行投递。