jobs in WECHAT INTERNATIONAL PTE. LTD.

全职 Multimodal Reinforcement Learning Algorithm Researcher 工作, 薪水 up to SGD 8,000, WECHAT INTERNATIONAL PTE. LTD. Central Region (Singapore) 公司招聘中 - Ricebowl

Multimodal Reinforcement Learning Algorithm Researcher

WECHAT INTERNATIONAL PTE. LTD.

SGD8,000 - SGD8,000 每月

Central Region (Singapore)

分享
保存

工作地点

  • 10 ANSON ROAD Central Region (Singapore) Singapore

职位描述

岗位职责

Roles &Responsibilities

1.Conduct research on reinforcement learning algorithms formultimodal models, including diffusion models for image and video generation,autoregressive models for multimodal understanding, and cutting-edge unifiedmultimodal frameworks.

2.Design and develop reinforcement learning trainingframeworks and reward modeling strategies to enable efficient large-scaletraining, improve training stability, and address issues such as rewardhacking.

3.Explore next-generation reinforcement learning paradigmsthat enable more direct and efficient learning from environmental feedback.

Skills Required

1.Bachelor’s degree or above in Computer Science or relatedfields.

2.Excellent research capabilities with publications in topconferences including ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, SIGGRAPH, etc.

3.Strong engineering and programming skills, withexperience in deep learning system implementation, model training and inferenceoptimization, CPU/GPU acceleration, and distributed training and inference.

4.Preference given to candidates with experience indiffusion models, autoregressive models, text-to-image / text-to-videogeneration.

5.Preference given to candidates withparticipation experience in ACM/NOIP (Informatics Olympiad).

重要安全守则

申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。

了解更多