Big Data Engineer, Data Lake / Feature Store

5 Months ago • 2 Years + • Monetization

Job Summary

Job Description

The batch processing team at ByteDance is responsible for offline data processing and distributed training. You will be developing and optimizing the in-house Feature Store functionality based on Iceberg, participating in optimizing the integration of Iceberg with various upper-level computing engines, and being involved in platform-related infrastructure development.
Must have:
  • Bachelor's Degree or above in Computer Science or related fields
  • 2+ years of relevant development experience
  • Strong programming ability in Java, Python, C++
  • Experience with large-scale distributed systems
  • In-depth knowledge of data lake formats like Delta, Hudi, or Iceberg
Good to have:
  • In-depth research or practical experience in Hadoop, Spark, Flink, Presto
  • Experience with open-source big data computing frameworks

Job Details

Responsibilities
About ByteDance Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content. Why Join Us Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible. Together, we inspire creativity and enrich life - a mission we aim towards achieving every day. To us, every challenge, no matter how ambiguous, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always. At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve. Join us. About The Team The batch processing team is responsible for the company's offline data processing and distributed training, supporting various business scenarios such as offline ETL and machine learning within the company. The components involved include the offline computing engine Spark, the in-house distributed training framework Primus, feature storage solutions like Iceberg and Hudi, as well as Ray, a next-generation distributed application framework. Faced with massive-scale scenarios, extensive functional and performance optimizations have been carried out in Spark, Primus, Feature Store, and support for the adoption of the new-generation distributed application framework Ray in relevant company scenarios. What you will be doing: - Responsible for the development and performance optimisation of the in-house Feature Store functionality based on Iceberg; - Participant in optimisation of the integration of Iceberg with various upper-level computing engines; - Involve in platform-related infrastructure development.
Qualifications
Minimum Qualifications - Bachelor's Degree or above, majoring in Computer Science, or related fields, with 2+ years of relevant development experience in the field with a strong programming ability, and proficiency in Java, Python, C++, with the ability to develop and optimize large-scale distributed systems. - In-depth research and relevant experience in one or more data lake formats such as Delta, Hudi, or Iceberg. Preferred Qualifications - In-depth research or practical experience in open-source big data computing frameworks and scenarios like Hadoop, Spark, Flink, Presto, and more. ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too. #LI-CT

Similar Jobs

ByteDance - Network Reliability Engineer - Physical Network Infrastructure

ByteDance

Singapore (On-Site)
3 Months ago
The Walt Disney Company - Senior Software Engineer, Data Reliability

The Walt Disney Company

Santa Monica, California, United States (On-Site)
1 Week ago
Canva - Security Engineer Internship 2025/26 - ANZ

Canva

Auckland, Auckland, New Zealand (Remote)
3 Weeks ago
Google - Software Engineer III, Android, Google Store

Google

Bengaluru, Karnataka, India (On-Site)
1 Week ago
Google - Student Researcher, PhD, Winter/Summer 2025

Google

Mountain View, California, United States (On-Site)
5 Months ago
ByteDance - Music Partnerships Manager - SoundOn

ByteDance

Jakarta, Jakarta, Indonesia (On-Site)
5 Months ago
ByteDance - Machine Learning Scientist Graduate, Scaling AI for Biology (AML - AI-for-Science) - 2025 Start (PhD)

ByteDance

Seattle, Washington, United States (On-Site)
6 Months ago
ByteDance - Software Engineer, Architecture and Infrastructure

ByteDance

Seattle, Washington, United States (On-Site)
5 Months ago
ByteDance - Corporate Real Estate Strategy Specialist

ByteDance

Dubai, Dubai, United Arab Emirates (On-Site)
2 Months ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

PwC - Senior Associate_Azure Data Engineer_Data & Analytics_Advisory_PAN  India

PwC

Kolkata, West Bengal, India (On-Site)
6 Months ago
Rush Street Interactive - Senior Server Engineer

Rush Street Interactive

Estonia (Hybrid)
2 Months ago
Google - Senior Solutions Acceleration Architect, Application

Google

Singapore (On-Site)
1 Week ago
Google - Engineering Manager

Google

Bengaluru, Karnataka, India (On-Site)
1 Week ago
Next Level Business Services - Java Developer

Next Level Business Services

Dallas, Texas, United States (On-Site)
6 Months ago
Digital Extremes - AI Programmer

Digital Extremes

London, Ontario, Canada (Hybrid)
1 Month ago
Interactive Brokers - Senior Software Engineer

Interactive Brokers

Greenwich, Connecticut, United States (On-Site)
6 Months ago
Luxoft - Senior Java Developer

Luxoft

Pune, Maharashtra, India (On-Site)
5 Months ago
Google - Strategic Cloud Engineer, Application Modernization, Technical Delivery

Google

Washington, District Of Columbia, United States (On-Site)
1 Week ago
Next Level Business Services - BigData Architect

Next Level Business Services

Bentonville, Arkansas, United States (On-Site)
6 Months ago

Get notifed when new similar jobs are uploaded

Jobs in Singapore

ByteDance - Senior Software Engineer - Stability Platform

ByteDance

Singapore (On-Site)
5 Months ago
ByteDance - Site Reliability Engineer, SealSuite

ByteDance

Singapore (On-Site)
2 Weeks ago
Tencent - Senior Regional Manager of WeChat Overseas Payments

Tencent

Singapore (On-Site)
7 Months ago
ByteDance - Payroll Analyst - HR Operations - Singapore

ByteDance

Singapore (On-Site)
5 Months ago
ByteDance - Backend Engineer Intern, Video-On-Demand

ByteDance

Singapore (On-Site)
1 Month ago
ByteDance - Production System Engineer, Infrastructure Engineering

ByteDance

Singapore (On-Site)
5 Months ago
ByteDance - Payment Strategy Intern (LATAM)

ByteDance

Singapore (On-Site)
2 Weeks ago
ByteDance - Strategy Manager - BytePlus

ByteDance

Singapore (On-Site)
5 Months ago
Tencent - Senior Legal Counsel, Privacy & Data Protection

Tencent

Singapore (On-Site)
6 Months ago

Get notifed when new similar jobs are uploaded

Monetization Jobs

ByteDance - Strategy Leader-Information System

ByteDance

Singapore (On-Site)
3 Months ago
ByteDance - Backend Software Engineer Graduate (Global E-commerce-US) - 2025 Start (BS/MS)

ByteDance

Seattle, Washington, United States (On-Site)
5 Months ago
ByteDance - Senior Software Engineer, Global Payment Risk & Compliance

ByteDance

San Jose, California, United States (On-Site)
5 Months ago
ByteDance - Product Operations, Search Ads AI Data Service - Trust & Safety

ByteDance

Pasig, Metro Manila, Philippines (On-Site)
2 Months ago
ByteDance - Experienced Enterprise Internal Control Partner

ByteDance

Singapore (On-Site)
5 Months ago
Anzuio - Sales Account Manager

Anzuio

England, United Kingdom (Hybrid)
1 Month ago
Playtika - Technical Operation Specialist - Temporary

Playtika

Israel (On-Site)
5 Months ago
Warner Bros Games - Senior Manager, Ad Monetization

Warner Bros Games

San Francisco, California, United States (Hybrid)
3 Weeks ago
ByteDance - ByteDance Back-end Engineer Graduate Program (Dubai 2025)

ByteDance

Dubai, Dubai, United Arab Emirates (On-Site)
2 Weeks ago
ByteDance - Software Engineer Intern (On-Device AI - Intelligent Creation-AI Platform)

ByteDance

San Jose, California, United States (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

About The Company

Where imagination meets innovation, delivering limitless gaming experiences.

San Diego, California, United States (On-Site)

San Jose, California, United States (On-Site)

Dubai, Dubai, United Arab Emirates (On-Site)

New York, New York, United States (On-Site)

San Jose, California, United States (On-Site)

San Jose, California, United States (On-Site)

Seattle, Washington, United States (On-Site)

View All Jobs

Get notified when new jobs are added by ByteDance

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug