关于本岗位:
我们正在寻找多模态算法实习生,参与多模态大模型在内容理解方向的前沿探索与应用落地。
岗位职责:
· 前沿技术探索:跟踪多模态大模型方向的前沿技术,参与多模态内容理解任务的技术应用,持续提升模型效果;
· 模型开发与应用:参与多模态大模型的结构设计、训练与微调,支持下游功能与应用开发。
任职要求:
· 计算机科学与技术、人工智能等相关专业硕士或博士在读;
· 有多模态、强化学习、NLP或CV相关研究或项目经验;
· 对大模型及新兴技术领域有浓厚兴趣,具备较强的学习能力和研究潜力;
· 具备良好的数学与编程基础,熟悉PyTorch深度学习框架;
· 具备良好的沟通能力和团队协作精神。
加分项:
· 在NLP、CV、ML相关核心会议或期刊发表过论文。
About This Role:
We're looking for a Multimodal Algorithm Intern to join Meitu's AI team. You'll work hands-on with multimodal large models — tracking frontier techniques and applying them to real multimodal content understanding across our imaging products. What you build will shape how we improve our AI-powered tools, not sit on a shelf.
What You'll Own:
· Track and evaluate the latest advances in multimodal large models, and translate promising techniques into real content understanding tasks;
· Contribute to multimodal large model design, training, and fine-tuning — from architecture decisions through downstream feature and application development;
· Iterate systematically to improve model performance, with clear metrics and reproducible results that strengthen the team's technical foundation.
What You Bring:
· Currently pursuing a Master's or PhD in Computer Science, AI, or a related field;
· Hands-on research or project experience in at least one of: multimodal learning, reinforcement learning, NLP, or computer vision;
· Strong foundation in mathematics and programming, with working proficiency in PyTorch;
· Genuine curiosity about large models and emerging AI techniques — you can quickly absorb new methods and turn them into working experiments;
· Comfortable sharing research findings and collaborating with teammates to move work forward.
Nice to Have:
· Publications at top NLP, CV, or ML conferences or journals.