大模型算法工程师-后训练
北京、上海
社招
全职
研发 - 算法
大模型算法
职位描述
我们致力于提升大模型的通用能力,让模型更好地理解用户意图、处理复杂信息,并通过推理、工具使用与多轮协作完成真实任务。你将参与通用模型的后训练研发,探索训练方法、能力协同优化与评测反馈机制,持续提升模型的整体表现和实际任务完成质量。根据个人经验与兴趣,你将重点负责以下一个或多个方向: 1. 研发监督微调、强化学习与偏好优化等后训练方法,提升模型的指令遵循、推理、工具使用和复杂任务执行能力。 2. 探索高质量数据构建、奖励设计与反馈机制,推动数据、训练策略和模型能力的持续迭代。 3. 研究多能力协同优化,优化数据配比与训练流程,缓解不同能力之间的冲突、遗忘与回退,提升模型整体表现。
职位要求
投递

MiniMax is a leading global technology company and one of the pioneers of large language models (LLMs) in Asia. Our mission is to build a world where intelligence thrives with everyone.
MiniMax develops proprietary LLMs across various modalities, including a trillion-parameter MoE model, a speech model with low latency and native support for major Asian languages, and a state-of-the-art text-to-speech and text-to-video models. Experience it now at https://hailuoai.com/
Leveraging these multi-modality general-purpose models, the MiniMax API Platform offers enterprises and developers secure, flexible, and reliable API services, enabling the rapid deployment of AI applications.