← 发现更多职位
字节跳动
正式

Research Scientist Graduate (Foundation Model-Speech-Interaction & Learning) - 2026 Start (PhD)

面议
美国 · San Jose(圣何塞) · 经验要求见详情
研发正式A47929美国Seed Foundation Model Campus Recruitment - Graduates国际招聘

关于这个机会

About the Team Established in 2023, the ByteDance Seed team is dedicated to pioneering new paths toward artificial general intelligence. We aspire to advance the frontier of intelligence to drive progress for both technology and society. With a long-term vision for the AI sector, the Seed team's research spans MLLM, GenMedia, AI for Science, and Robotics. We maintain a global presence with laboratories and career opportunities across China, Singapore, and the United States. To date, we have launched industry-leading general foundation models and cutting-edge multimodal capabilities. Our technology powers over 50 application scenarios — including Doubao, Jimeng, TRAE, Dola and Dreamnia — and serves enterprise customers through Volcano Engine and BytePlus. Third-party data shows that the Doubao App ranks first in user volume in the Chinese market, while Doubao foundation models lead the industry in average daily token consumption. The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning. We are looking for talented individuals to join our team. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Successful candidates must be able to commit to an onboarding date by the end of the year. Please state your availability and graduation date clearly in your resume. Responsibilities - Contribute cutting-edge research to ByteDance product evolution (e.g., Douyin, Capcut, and more) to impact billions of users worldwide. - Work on advanced science and technology in audio processing and generation (e.g., Dialogue Systems, Audio-Video Models, Speech Synthesis, Voice Conversion, Audio Codec Learning, Audio Language Modeling, etc.) - Research, model, design, develop and evaluate novel machine learning models and algorithms. - Collaborate with globally based researchers and engineering teams in developing machine learning models and algorithms.

任职要求

Minimum Qualifications - Individuals who are completing or have recently completed a PhD degree in Computer Science, Electrical Engineering, Electrical and Computer Engineering, Physics, Mathematics, or a related discipline. - Good knowledge of theoretical and empirical research in addressing research problems - Solid knowledge and experience with at least one popular deep learning framework (e.g., PyTorch, TensorFlow) and familiarity with deep neural network architectures - Experience in both neural and non-neural, classical machine learning models and algorithms Preferred Qualifications - Research experience in one or more of the following fields: speech synthesis, audio generation, large language model, computer vision, generative models - Strong first-author publications at accredited conferences such as(e.g., NeurIPS, ICML, ICLR, ACL, EMNLP, NAACL etc.) - Proficient in C / C + +, Python, and shell programming languages, and have a deep understanding of data structure and algorithm design As a condition of employment, all successful candidates must be able to establish authorization to work in the United States. For this position, the Company does not provide sponsorship or any immigration-related benefits.

官方来源与核验

字节跳动官方招聘 · 职位编号 7671032227920431413

最近核验:2026-09-16T14:05:07.210038+00:00

查看官方职位详情 ↗

内推申请说明

本站为独立内推协助平台。申请会交由管理员核实岗位与内推渠道,不等于已在公司官网投递;薪资、岗位状态和实际招聘流程以官方信息为准。