← 发现更多职位
字节跳动
正式

Software Development Engineer-AI/LLM Network-Global Frontier Tech Research Program-2027 Start

面议
美国 · San Jose(圣何塞) · 经验要求见详情
研发正式A74270美国Global Frontier Tech Recruitment Program - 2027 Grad国际招聘

关于这个机会

We are looking for talented individuals to join our team in 2027. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Launch your career where inspiration is infinite at our Company. Successful candidates must be able to commit to an onboarding date by end of year 2027. Please state your availability and graduation date clearly in your resume. Team Introduction: ByteDance Networking brings together innovative ideas and technologies from network architecture, software defined networking (SDN), network virtualization, switch software and hardware co-design, and high-speed networking, to create hyperscale data-center networking solutions that power several of the most popular apps of the world such as Douyin and TikTok which serve hundreds of millions of users around the globe. ByteDance Networking is responsible for designing, building, and operating the global, intelligent network infrastructure to meet the requirements of high availability, scalability, and high-performance. By joining this team, you will gain marketable software development and/or network operation experiences in data center networking at massive scale. Topic Content: With the large-scale adoption of LLMs and AI agents, traditional cloud-native infrastructure can no longer meet the ultra-high performance and elasticity requirements of AI workloads. Network and Observability: Research intelligent fault localization and root cause analysis for large-scale AI clusters, combined with intelligent tuning of time-series databases to improve cluster stability. This topic aims to build a next-generation AI-native infrastructure to support the deployment of LLMs and AI agents, improve resource utilization, reduce costs, support elastic scaling, and drive the technological evolution of AI infrastructure. Responsibilities: - Design, implementation and deployment of high-speed network technologies to support AI/LLM applications. - Design and development of platforms/systems for monitoring, analysis and diagnosis of large scale AI/LLM network. - Research and development of high-performance AI communication framework, network protocol stacks, and codesign optimization of host-network-application to improve the scalability, reliability and performance of AI/LLM network. - Building next generation AI network infrastructure supporting large scale heterogeneous network hardware with innovative and deployable solutions.

任职要求

Minimum Qualifications: - Individuals who are completing or recently completed a PhD in Software Development, Computer Science, Computer Engineering, or a related technical discipline. Preferred Qualifications: - Proficiency in computer network and network programming. - Proficiency in one or several mainstream programming languages, including C/C++, Python, Go and so on. - Be familiar with the latest advances in the area of high-speed network systems, including RDMA, congestion control, AI network optimization and so on. - Experience in developing high performance communication frameworks(including NCCL, MPI and RPC libraries) is a plus. - Experience in developing software systems for AI network diagnosis and performance optimization is a plus.

官方来源与核验

字节跳动官方招聘 · 职位编号 7629215845797218613

最近核验:2026-09-16T14:05:07.210038+00:00

查看官方职位详情 ↗

内推申请说明

本站为独立内推协助平台。申请会交由管理员核实岗位与内推渠道,不等于已在公司官网投递;薪资、岗位状态和实际招聘流程以官方信息为准。