ShengShu

视频生成算法实习生 (蒸馏)

ShengShu  •  Onsite  •  18 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

视频生成算法实习生 (蒸馏)
北京
实习
研发 - 算法
职位描述
负责视频生成模型的压缩、蒸馏,重点提升生成速度、成本效率和产品可用性:
1. 负责视频 diffusion / flow 模型的步数蒸馏。
2. 设计不同训练和推理策略,对比速度、质量、稳定性和可控性。
3. 与评测团队协作,评估蒸馏模型在动态、主体一致性、文字、人物、复杂 prompt 等维度的退化和收益。
4. 支持模型交付,沉淀版本、配置、速度、成本和效果结论。
职位要求
1. 熟悉 diffusion sampling、ODE/SDE solver,以及consistency model、MeanFlow、DMD、rCM等经典扩散蒸馏算法。
2. 有大规模视频模型蒸馏经验者优先。
3. 对画质、动态、稳定性和速度的权衡敏感,在调参过程中建立对算法和数据的直觉。
4. 能把实验结论转化为可交付推理配置。
投递
ShengShu

About ShengShu

We are the first General World Model ("GWM") company globally that built a GWM unifying the digital world and the physical world. We are dedicated to building a unified intelligence framework capable of modeling, reasoning, predicting and acting upon the underlying rules that govern both digital and physical worlds. Guided by first principles thinking, we use visual and auditory information, which naturally encodes the physical world, to train our foundation world model and replicate the human process of perceiving, simulating and interacting with the world, and ultimately, to enable AGI that connects the digital world with the physical world.

Industry
Unknown
Company Size
11-50 employees
Headquarters
Beijing, CN
Year Founded
2023
Social Media