Big Data Engineer, Data Lake / Feature Store

3 Months ago • 2 Years + • Monetization

Job Summary

Job Description

The batch processing team at ByteDance is responsible for offline data processing and distributed training. You will be developing and optimizing the in-house Feature Store functionality based on Iceberg, participating in optimizing the integration of Iceberg with various upper-level computing engines, and being involved in platform-related infrastructure development.
Must have:
  • Bachelor's Degree or above in Computer Science or related fields
  • 2+ years of relevant development experience
  • Strong programming ability in Java, Python, C++
  • Experience with large-scale distributed systems
  • In-depth knowledge of data lake formats like Delta, Hudi, or Iceberg
Good to have:
  • In-depth research or practical experience in Hadoop, Spark, Flink, Presto
  • Experience with open-source big data computing frameworks

Job Details

Responsibilities
About ByteDance Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content. Why Join Us Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible. Together, we inspire creativity and enrich life - a mission we aim towards achieving every day. To us, every challenge, no matter how ambiguous, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always. At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve. Join us. About The Team The batch processing team is responsible for the company's offline data processing and distributed training, supporting various business scenarios such as offline ETL and machine learning within the company. The components involved include the offline computing engine Spark, the in-house distributed training framework Primus, feature storage solutions like Iceberg and Hudi, as well as Ray, a next-generation distributed application framework. Faced with massive-scale scenarios, extensive functional and performance optimizations have been carried out in Spark, Primus, Feature Store, and support for the adoption of the new-generation distributed application framework Ray in relevant company scenarios. What you will be doing: - Responsible for the development and performance optimisation of the in-house Feature Store functionality based on Iceberg; - Participant in optimisation of the integration of Iceberg with various upper-level computing engines; - Involve in platform-related infrastructure development.
Qualifications
Minimum Qualifications - Bachelor's Degree or above, majoring in Computer Science, or related fields, with 2+ years of relevant development experience in the field with a strong programming ability, and proficiency in Java, Python, C++, with the ability to develop and optimize large-scale distributed systems. - In-depth research and relevant experience in one or more data lake formats such as Delta, Hudi, or Iceberg. Preferred Qualifications - In-depth research or practical experience in open-source big data computing frameworks and scenarios like Hadoop, Spark, Flink, Presto, and more. ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too. #LI-CT

Similar Jobs

Microsoft - Software Engineering IC2

Microsoft

Prague, Prague, Czechia (On-Site)
• 1 Month ago
Nielsen Holdings - Senior Software Engineer (Java/Scala, Spark, Kubernetes, AWS)

Nielsen Holdings

Bengaluru, Karnataka, India (Hybrid)
• 4 Months ago
Next Level Business Services - CQ5 Developer/Architect

Next Level Business Services

Sunnyvale, California, United States (On-Site)
• 3 Months ago
Microsoft - Member of Technical Staff, AI - Reinforcement Learning Systems

Microsoft

Mountain View, California, United States (Hybrid)
• 1 Day ago
Luxoft - Senior Mobile QA Automation

Luxoft

Pune, Maharashtra, India (On-Site)
• 3 Months ago
Red Games Co - Senior Economy Designer/Monetization Designer

Red Games Co

(Remote)
• 3 Weeks ago
ByteDance - Video Codec Algorithm Engineer - Multimedia Lab

ByteDance

Seattle, Washington, United States (On-Site)
• 3 Months ago
InMobiInMobi - Senior Product Manager - InMobi DSP

InMobiInMobi

Bengaluru, Karnataka, India (On-Site)
• 5 Days ago
Plarium - Monetization Product Manager

Plarium

Helsinki, Uusimaa, Finland (Hybrid)
• 22 Hours ago
Xsolla - Director of User Acquisition – Game Developers (Offerwall)

Xsolla

Los Angeles, California, United States (Remote)
• 4 Days ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

ByteDance - Research Engineer in Large Model System

ByteDance

San Jose, California, United States (On-Site)
• 3 Months ago
Nagarro - Associate Principal Engineer - Salesforce Solutions Architect

Nagarro

Colombia (Remote)
• 2 Months ago
TOPPAN Edge  Inc  - Senior QA Engineer I

TOPPAN Edge Inc

Bengaluru, Karnataka, India (On-Site)
• 3 Months ago
Google - Fullstack Software Engineer

Google

(On-Site)
• 2 Months ago
PwC - Backend Developer/Consultant (freelance)

PwC

Warsaw, Masovian Voivodeship, Poland (Hybrid)
• 4 Months ago
Microsoft - Platform Engineering Manager

Microsoft

Redmond, Washington, United States (Hybrid)
• 2 Days ago
VGW - Senior Engineer

VGW

Perth, Western Australia, Australia (On-Site)
• 6 Days ago
Paypal - Senior AI Machine Learning Engineer

Paypal

San Jose, California, United States (On-Site)
• 4 Months ago
Google - Software Engineering Manager II, Infrastructure, Google Cloud

Google

Durham, North Carolina, United States (On-Site)
• 1 Month ago
Nagarro - Senior Staff Engineer (Scrum Master)

Nagarro

Johannesburg, Gauteng, South Africa (On-Site)
• 4 Months ago

Get notifed when new similar jobs are uploaded

Jobs in Singapore

Saviynt - Director, Partner Enablement

Saviynt

Singapore (Hybrid)
• 4 Months ago
ByteDance - Merchant Financing Product Manager - Global Payment

ByteDance

Singapore (On-Site)
• 3 Weeks ago
PwC - Digital Tax - MA/SM

PwC

Singapore (On-Site)
• 4 Months ago
The Walt Disney Company - Marketing Ops Consultant - Contract

The Walt Disney Company

Singapore, Singapore (On-Site)
• 3 Months ago
ByteDance - Global SRE Lead, Security Engineering

ByteDance

Singapore (On-Site)
• 3 Months ago
ByteDance - Security Software Engineer

ByteDance

Singapore (On-Site)
• 3 Months ago
Keywords Studios (Player Support) - Chinese to English Game translators (Freelance/remote job)

Keywords Studios (Player Support)

Singapore (Remote)
• 3 Months ago
ByteDance - Benefits Operations Specialist, APAC - Singapore

ByteDance

Singapore (On-Site)
• 2 Months ago
ByteDance - Incident Response Manager - Infrastructure Engineering

ByteDance

Singapore (On-Site)
• 3 Months ago

Get notifed when new similar jobs are uploaded

Monetization Jobs

ByteDance - Software Development Engineer - Machine Learning System

ByteDance

San Jose, California, United States (On-Site)
• 3 Months ago
ByteDance - Data center Technical Project Manager, Data Center Development

ByteDance

Singapore (On-Site)
• 3 Months ago
ByteDance - Research Scientist in Foundation Model (Music) - 2025 Start (PhD)

ByteDance

San Jose, California, United States (On-Site)
• 3 Months ago
ByteDance - Business Operations Manager, ISV Commerce Partnerships – Partner Product

ByteDance

Seattle, Washington, United States (On-Site)
• 1 Month ago
Voodoo - Lead Creative Manager - Gaming

Voodoo

İstanbul, Türkiye (On-Site)
• 5 Months ago
ByteDance - Experienced Enterprise Internal Control Partner - E-commerce - Singapore

ByteDance

Singapore (On-Site)
• 3 Months ago
InMobiInMobi - Account Manager - Ad Monetization

InMobiInMobi

Bengaluru, Karnataka, India (On-Site)
• 1 Month ago
ByteDance - Fraud Strategy Expert - Global Payment - Singapore

ByteDance

Singapore (On-Site)
• 3 Months ago
ByteDance - E-commerce - Partner Operations Program Manager(Service)

ByteDance

Taguig, Metro Manila, Philippines (On-Site)
• 3 Months ago
ByteDance - KOL Business Development Manager - DCar (Third-party Contractor)

ByteDance

Los Angeles, California, United States (On-Site)
• 3 Months ago

Get notifed when new similar jobs are uploaded

About The Company

Where imagination meets innovation, delivering limitless gaming experiences.

Taguig, Metro Manila, Philippines (On-Site)

Singapore (On-Site)

Dubai, Dubai, United Arab Emirates (On-Site)

State Of São Paulo, Brazil (On-Site)

Seattle, Washington, United States (On-Site)

San Jose, California, United States (On-Site)

San Jose, California, United States (On-Site)

View All Jobs

Get notified when new jobs are added by ByteDance

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug