岗位职责
部署紹介/Business Unit LIGHTSPEED STUDIOS は、物語性の高いストーリー、没入感のあるゲームプレイ、最先端のテクノロジーを駆使し、ゲーム開発をアートとサイエンスの融合として追求し続けています。世界中のプレイヤー、そしてあらゆるデバイスに次世代のゲームエクスペリエンスを提供することを使命としています。 LIGHTSPEED STUDIOS is made up of passionate players who advance the art & science of game development through great stories, great gameplay, and advanced technology. We are focused on bringing next generation experiences to gamers who want to enjoy them anywhere, anytime, across multiple genres and devices. チーム紹介/About the Hiring Team Lightspeed Tech Center は、PUBG モバイルを始めとする様々な高品質ゲームを手掛けた Lightspeed Studios の研究開発部門です。Tech Center は技術研究とイノベーションを牽引し、ゲーム作品にそのライフサイクルを通して、エンジン、音響、QA、AI、次世代技術、技術提携などのテクニカルサポートを提供しています。 Lightspeed Tech Center is a R&D department under Lightspeed Studios which develop PUBG Mobile and other high-quality games. Our Tech Center leads the research, exploration, and discovery of innovative technologies and provides technical services for all games during all phases of life cycle, including engine, audio, QA, AI, next generation game, technical cooperation, etc. 仕事内容/What the Role Entails Job Responsibilities●Research and develop advanced speech synthesis and generation algorithms (e.g., TTS, voice conversion, sound/music generation) based on LLMs, multimodal/omnimodal LLMs. ●Develop and optimize speech and audio synthesis systems for online applications, improving effectiveness, efficiency, and scalability. ●Explore and advance full-duplex/streaming multimodal LLM capabilities in speech understanding, generation, and real-time spoken interaction. ●Collaborate cross-functionally with research and engineering teams from prototyping to production. 岗位职责 ● 基于大语言模型(LLM)及多模态/全模态大模型,研究和开发前沿语音合成与生成算法(如 TTS、语音转换、声音/音乐生成等)。 ● 开发和优化面向线上应用的语音及音频合成系统,提升效果、效率和可扩展性。 ● 探索并推进全双工/流式多模态大模型在语音理解、语音生成及实时语音交互方面的能力。 ● 与研究和工程团队跨职能协作,推动项目从原型验证到生产落地。 応募資格/Who We Look For Requirements ●Ph.D. in Computer Science, Electrical Engineering, Signal Processing, or a closely related field. ●Strong foundation in speech/audio processing and modern generative models (e.g., diffusion, flow matching, autoregressive, codec-based approaches). ●Hands-on experience extending LLMs to speech/audio modalities (e.g., speech tokenizers, multimodal adapters, speech-text joint training). Experience with full-duplex or streaming spoken dialogue systems and real-time interaction modeling is a plus. ●Proficient in Python and deep learning frameworks (e.g., PyTorch); experience with distributed training is a plus. ●Track record of publications at top-tier venues (e.g., ICML, NeurIPS, ICLR, ACL, ICASSP, Interspeech). ●Mandarin Chinese / Business-level Japanese or English
职位要求
详情请见上方岗位职责内说明