アラートの作成
似たような求人をメールで送る

Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型

部署紹介/Business Unit

LIGHTSPEED STUDIOS は、物語性の高いストーリー、没入感のあるゲームプレイ、最先端のテクノロジーを駆使し、ゲーム開発をアートとサイエンスの融合として追求し続けています。世界中のプレイヤー、そしてあらゆるデバイスに次世代のゲームエクスペリエンスを提供することを使命としています。
LIGHTSPEED STUDIOS is made up of passionate players who advance the art & science of game development through great stories, great gameplay, and advanced technology. We are focused on bringing next generation experiences to gamers who want to enjoy them anywhere, anytime, across multiple genres and devices.

チーム紹介/About the Hiring Team

Lightspeed Tech Center は、PUBG モバイルを始めとする様々な高品質ゲームを手掛けた Lightspeed Studios の研究開発部門です。Tech Center は技術研究とイノベーションを牽引し、ゲーム作品にそのライフサイクルを通して、エンジン、音響、QA、AI、次世代技術、技術提携などのテクニカルサポートを提供しています。
Lightspeed Tech Center is a R&D department under Lightspeed Studios which develop PUBG Mobile and other high-quality games. Our Tech Center leads the research, exploration, and discovery of innovative technologies and provides technical services for all games during all phases of life cycle, including engine, audio, QA, AI, next generation game, technical cooperation, etc.

仕事内容/What the Role Entails

Job Responsibilities

●Research and develop advanced speech synthesis and generation algorithms (e.g., TTS, voice conversion, sound/music generation) based on LLMs, multimodal/omnimodal LLMs.

●Develop and optimize speech and audio synthesis systems for online applications, improving effectiveness, efficiency, and scalability.

●Explore and advance full-duplex/streaming multimodal LLM capabilities in speech understanding, generation, and real-time spoken interaction.

●Collaborate cross-functionally with research and engineering teams from prototyping to production.

岗位职责

● 基于大语言模型(LLM)及多模态/全模态大模型,研究和开发前沿语音合成与生成算法(如 TTS、语音转换、声音/音乐生成等)。

● 开发和优化面向线上应用的语音及音频合成系统,提升效果、效率和可扩展性。

● 探索并推进全双工/流式多模态大模型在语音理解、语音生成及实时语音交互方面的能力。

● 与研究和工程团队跨职能协作,推动项目从原型验证到生产落地。

応募資格/Who We Look For

Requirements

●Ph.D. in Computer Science, Electrical Engineering, Signal Processing, or a closely related field.

●Strong foundation in speech/audio processing and modern generative models (e.g., diffusion, flow matching, autoregressive, codec-based approaches).

●Hands-on experience extending LLMs to speech/audio modalities (e.g., speech tokenizers, multimodal adapters, speech-text joint training). Experience with full-duplex or streaming spoken dialogue systems and real-time interaction modeling is a plus.

●Proficient in Python and deep learning frameworks (e.g., PyTorch); experience with distributed training is a plus.

●Track record of publications at top-tier venues (e.g., ICML, NeurIPS, ICLR, ACL, ICASSP, Interspeech).

●Mandarin Chinese / Business-level Japanese or English

任职要求

● 计算机科学、电子工程、信号处理或相关领域博士学位。

● 具备扎实的语音/音频处理基础,熟悉现代生成模型(如扩散模型、流匹配、自回归模型、基于编解码器的方法等)。

● 具有将大语言模型扩展至语音/音频模态的实际经验(如语音分词器、多模态适配器、语音-文本联合训练等);有全双工或流式语音对话系统及实时交互建模经验者优先。

● 熟练使用 Python 及深度学习框架(如 PyTorch);有分布式训练经验者优先。

● 在顶级学术会议发表过论文(如 ICML、NeurIPS、ICLR、ACL、ICASSP、Interspeech 等)。

●中文/商务日语或英文

一緒に働く魅力/Why Join Us?

賃金
年収額:応相談(昇給/賞与あり)
※経験・スキル等を考慮し、当社規定により優遇します。
※一般社員の場合、固定残業代月35時間分が含まれます。
管理職の場合、固定残業代10時間分が含まれます。
労働時間・休日
就業時間 09:30 ~ 18:30 休憩時間 60分
フレックスタイムあり
完全週休2日制(土日・祝日)
有給休暇:入社初年度は、入社月により期間按分付与。在籍年数踏まえて最大22日
年末年始休暇(12/30~1/3)
病気休暇(有給):12日/年(有給休暇とは別途付与)
福利厚生
<各種保険>
健康保険 、雇用保険、 労災保険、 厚生年金保険、 総合福祉団体定期保険(GTL)
業務災害安心総合保険(GPA)、団体長期障害所得補償保険(GLTD)
海外旅行保険
<各種手当>
通勤手当(上限5万円/月) 、通信手当(上限1万円/月)、 自己啓発手当(上限2万円/月)、 平日に出社した場合食事補助1800円/日、 定期健診手当(上限7万5千円/年1回)
<支援制度・その他待遇>
フリードリンクコーナー 、ギフトカード(お盆、正月の2回支給、2万円/1回)
従業員支援プログラム(EAP) 、慶弔見舞金 、永年勤続賞/休暇
この求人広告に記載されているポリシー(給与、福利厚生、労働条件などを含むがこれに限らない)は、会社のリアルタイムポリシーに応じて調整される可能性があります。

平等な雇用機会を重視するテンセント/Equal Employment Opportunity at Tencent

テンセントは、雇用機会の均等を最重要視しています。私たちは、多様性がイノベーションを促進し、ユーザーやコミュニティに対してより良いサービスを提供する原動力になると確信しています。
社員一人ひとりが十分なサポートを受け、自身の目標と会社の目標を共に達成できるよう、モチベーションを高められる環境づくりを大切にしています。

As an equal opportunity employer, we firmly believe that diverse voices fuel our innovation and allow us to better serve our users and the community. We foster an environment where every employee of Tencent feels supported and inspired to achieve individual and common goals.


Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型

企業サイトでの申請
Back to search page