ShengShu

Agent算法实习生

ShengShu  •  Onsite  •  17 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

Agent算法实习生
北京
实习
研发 - 算法
职位描述
Agentic Video Generation 架构设计与研发:
1. 负责多模态视频创作 Agent 系统的核心架构设计,包括 Task Planning(任务拆解与规划)、Memory(长短期记忆)、Reflection(自我反思与纠错)以及 Tool-Calling(工具链调度)
2. 建立 Agent 生成视频的多维度评价指标(画质、连贯性、故事表达、指令遵循度),并通过 Human-in-the-loop(人机协同)数据反馈持续迭代 Agent 决策模型。
职位要求
学历背景:计算机、人工智能、自动化、软件工程或数学相关专业,硕士及以上学历(博士优先)。
1. 熟悉大模型 Agent 架构
2. 多模态/视频领域:深入理解 Multimodal LLM(VLM)及 Video Diffusion/DiT 原理,对视频生成或图像生成技术有扎实的理论基础和实战经验。
3. 精通 Python,熟练掌握 PyTorch 开发框架,具备优秀的工程抽象与系统设计能力。
4. 对 AIGC 领域(尤其是 AI 视频生成)有极高的技术热情,具备优秀的技术洞察力、复杂问题拆解能力与自驱力。
投递
ShengShu

About ShengShu

We are the first General World Model ("GWM") company globally that built a GWM unifying the digital world and the physical world. We are dedicated to building a unified intelligence framework capable of modeling, reasoning, predicting and acting upon the underlying rules that govern both digital and physical worlds. Guided by first principles thinking, we use visual and auditory information, which naturally encodes the physical world, to train our foundation world model and replicate the human process of perceiving, simulating and interacting with the world, and ultimately, to enable AGI that connects the digital world with the physical world.

Industry
Unknown
Company Size
11-50 employees
Headquarters
Beijing, CN
Year Founded
2023
Social Media