← 发现更多职位
字节跳动
Regular

Backend Software Engineer, TikTok Data Ecosystem (Data Lake)

面议
新加坡 · Singapore · 经验要求见详情
R&DRegularA199110新加坡国际招聘TikTok

关于这个机会

About The Team The TikTok Data Ecosystem Team has the vital role of crafting and implementing a storage solution for offline data in TikTok's recommendation system, which caters to more than a billion users. Their primary objectives are to guarantee system reliability, uninterrupted service, and seamless performance. They aim to create a storage and computing infrastructure that can adapt to various data sources within the recommendation system, accommodating diverse storage needs. Their ultimate goal is to deliver efficient, affordable data storage with easy-to-use data management tools for the recommendation, search, and advertising functions. What you will be doing: 1. Design and implement an offline/real-time data architecture for large-scale recommendation systems. 2. Design and implement a flexible, scalable, stable, and high-performance storage system and computation model. 3. Troubleshoot production systems, and design and implement necessary mechanisms and tools to ensure the overall stability of production systems. 4. Build industry-leading distributed systems such as offline and online storage, batch, and stream processing frameworks, providing reliable infrastructure for massive data and large-scale business systems.

任职要求

Minimum Qualifications: - Bachelor's Degree or above, majoring in Computer Science, or related fields, with 1+ years of experience building scalable systems; - Proficiency in common big data processing systems like Spark/Flink at the source code level is required, with a preference for experience in customizing or extending these systems; - A deep understanding of the source code of at least one data lake technology, such as Hudi, Iceberg, or DeltaLake, is highly valuable and should be prominently showcased in your resume, especially if you have practical implementation or customisation experience; - Knowledge of HDFS principles is expected, and familiarity with columnar storage formats like Parquet/ORC is an additional advantage; - Prior experience in data warehousing modeling; - Proficiency in programming languages such as Java, C++, and Scala is essential, along with strong coding skills and the ability to troubleshoot effectively; Preferred Qualifications: - Experience with other big data systems/frameworks like Hive, HBase, or Kudu is a plus; - A willingness to tackle challenging problems without clear solutions, a strong enthusiasm for learning new technologies, and prior experience in managing large-scale data (in the petabyte range) are all advantageous qualities.

官方来源与核验

字节跳动官方招聘 · 职位编号 7274869648948185400

最近核验:2026-09-16T12:31:41.068845+00:00

查看官方职位详情 ↗

内推申请说明

本站为独立内推协助平台。申请会交由管理员核实岗位与内推渠道,不等于已在公司官网投递;薪资、岗位状态和实际招聘流程以官方信息为准。