招聘中 社招

Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型

IEG

  • 东京
  • 技术

收录时间

岗位职责

部署紹介/Business Unit LIGHTSPEED STUDIOS は、物語性の高いストーリー、没入感のあるゲームプレイ、最先端のテクノロジーを駆使し、ゲーム開発をアートとサイエンスの融合として追求し続けています。世界中のプレイヤー、そしてあらゆるデバイスに次世代のゲームエクスペリエンスを提供することを使命としています。 LIGHTSPEED STUDIOS is made up of passionate players who advance the art & science of game development through great stories, great gameplay, and advanced technology. We are focused on bringing next generation experiences to gamers who want to enjoy them anywhere, anytime, across multiple genres and devices. チーム紹介/About the Hiring Team Lightspeed Tech Center は、PUBG モバイルを始めとする様々な高品質ゲームを手掛けた Lightspeed Studios の研究開発部門です。Tech Center は技術研究とイノベーションを牽引し、ゲーム作品にそのライフサイクルを通して、エンジン、音響、QA、AI、次世代技術、技術提携などのテクニカルサポートを提供しています。 Lightspeed Tech Center is a R&D department under Lightspeed Studios which develop PUBG Mobile and other high-quality games. Our Tech Center leads the research, exploration, and discovery of innovative technologies and provides technical services for all games during all phases of life cycle, including engine, audio, QA, AI, next generation game, technical cooperation, etc. 仕事内容/What the Role Entails Job Responsibilities●Research and develop advanced speech synthesis and generation algorithms (e.g., TTS, voice conversion, sound/music generation) based on LLMs, multimodal/omnimodal LLMs. ●Develop and optimize speech and audio synthesis systems for online applications, improving effectiveness, efficiency, and scalability. ●Explore and advance full-duplex/streaming multimodal LLM capabilities in speech understanding, generation, and real-time spoken interaction. ●Collaborate cross-functionally with research and engineering teams from prototyping to production. 岗位职责 ● 基于大语言模型(LLM)及多模态/全模态大模型,研究和开发前沿语音合成与生成算法(如 TTS、语音转换、声音/音乐生成等)。 ● 开发和优化面向线上应用的语音及音频合成系统,提升效果、效率和可扩展性。 ● 探索并推进全双工/流式多模态大模型在语音理解、语音生成及实时语音交互方面的能力。 ● 与研究和工程团队跨职能协作,推动项目从原型验证到生产落地。 応募資格/Who We Look For Requirements ●Ph.D. in Computer Science, Electrical Engineering, Signal Processing, or a closely related field. ●Strong foundation in speech/audio processing and modern generative models (e.g., diffusion, flow matching, autoregressive, codec-based approaches). ●Hands-on experience extending LLMs to speech/audio modalities (e.g., speech tokenizers, multimodal adapters, speech-text joint training). Experience with full-duplex or streaming spoken dialogue systems and real-time interaction modeling is a plus. ●Proficient in Python and deep learning frameworks (e.g., PyTorch); experience with distributed training is a plus. ●Track record of publications at top-tier venues (e.g., ICML, NeurIPS, ICLR, ACL, ICASSP, Interspeech). ●Mandarin Chinese / Business-level Japanese or English

职位要求

详情请见上方岗位职责内说明

有些机会,只在官网短暂出现

真正值得关注的岗位,常常只在企业官网短暂开放,可能两三天后就下线,也未必会同步到综合招聘平台。没有持续关注,你甚至不会知道它曾经出现。职先机持续聚合并核验官网岗位,帮你抓住职场先机,快人一步。

微信小程序 / 当前岗位

在微信里继续看这个岗位

IEGSenior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型

扫码直达当前岗位收藏、浏览记录和下线提醒留在微信里
电脑端可直接微信扫码;手机端可保存小程序码后在微信中识别。
微信扫码在职先机小程序查看Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型正在准备岗位码
职先机微信小程序一岗一码 · 正式版直达保存小程序码

岗位提醒 · 01

收藏当前岗位,下线及时通知

当前岗位Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型IEG · 东京

01

扫码进入当前岗位无需重新搜索,直接打开当前岗位详情。

02

收藏并开启提醒在小程序内收藏后,岗位下线或变化时及时通知。

03

继续查看相似机会原岗位变化时,可继续浏览相关在招岗位。

职先机会持续核验公开岗位;岗位状态与最终招聘结果仍以企业招聘官网为准。

微信小程序扫码打开当前岗位

扫码打开当前岗位,完成收藏并开启下线提醒。

微信扫码在职先机小程序收藏Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型正在准备岗位码

微信小程序 · 当前岗位

请前往微信小程序提交内推申请

IEGSenior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型社招 · 东京

01进入当前岗位扫码或打开后直达本岗位内推申请页。

02提交 PDF 简历简历替换、处理进度和消息提醒在小程序内同步。

手机端将尝试直接打开微信;电脑端请使用微信扫码。

微信扫码进入Senior Researcher, Speech Synthesis and Multimodal LLM|高级研究员 - 语音合成与多模态大模型内推申请页正在准备岗位码
当前岗位专属码扫码后无需重新搜索岗位保存小程序码

内推提示

暂不支持该岗位内推

如需投递,建议前往官网进行投递。