← 发现更多职位
字节跳动
正式

Machine Learning Engineer - AI Compiler Optimization

面议
美国 · San Jose(圣何塞) · 经验要求见详情
研发正式A86940美国国际招聘

关于这个机会

The mission of our AML team is to push the next-generation AI infrastructure and recommendation platform for the ads ranking, search ranking, live & ecom ranking in our company. We also drive substantial impact on core businesses of the company. Currently, we are looking for Machine Learning Engineer in AI Compiler Optimization to join our team to support and advance that mission. Responsibilities: - Responsible for building and implementing the compilation optimization system for the recommendation machine learning engine. Design and implement full-stack optimization solutions at the graph, operator, and memory levels specifically for recommendation model scenarios, including but not limited to graph-operator fusion and automatic operator generation, to maximize hardware computing limits. - Collaborate closely with hardware and algorithm teams to carry out hardware-software co-design. Optimize compilation strategies based on hardware characteristics to improve the efficiency of hardware-software synergy. - Responsible for the compilation adaptation of recommendation models from the PyTorch framework to the engine. Optimize the entire process of model import, conversion, and code generation to simplify the model deployment process and enhance development efficiency.

任职要求

Minimum Qualifications: - Proficient in one of the mainstream AI compiler frameworks (e.g., Triton, MLIR, TVM), with practical project experience in customized compilation optimization and Pass development based on the framework. - Experience in GPU/NPU compilation optimization, mastering core techniques such as loop optimization, memory optimization, and operator optimization, with the ability to independently perform performance bottleneck analysis and technical optimization. - Familiar with common model structures and compilation adaptation logic of deep learning frameworks such as PyTorch and TensorFlow, capable of designing targeted optimization solutions. Preferred Qualifications: - Familiar with the architecture design of recommendation machine learning engines. Solid implementation experience in compilation optimization and low-latency inference optimization for large-scale recommendation systems, with the ability to handle compilation optimization needs in high-concurrency scenarios. - Experience contributing to open-source AI compiler projects (e.g., TVM, MLIR), or possess technical expertise in the compilation adaptation of large models for recommendation scenarios and the automatic generation of sparse operators.

官方来源与核验

字节跳动官方招聘 · 职位编号 7642171416270260485

最近核验:2026-09-16T14:05:07.210038+00:00

查看官方职位详情 ↗

内推申请说明

本站为独立内推协助平台。申请会交由管理员核实岗位与内推渠道,不等于已在公司官网投递;薪资、岗位状态和实际招聘流程以官方信息为准。