Data Engineer Lead

3 Months ago • 5-8 Years • Data Analyst

Job Summary

Job Description

Glean seeks a Data Engineer Lead to build and lead its first data engineering group in Bangalore. This hands-on role initially involves contributing to urgent business needs, improving data availability, partnering with cross-functional teams (Product Engineering, Go-to-Market, Finance), architecting data models, ensuring data quality and availability, and improving ETL tooling (dbt, BigQuery, Metabase). The successful candidate will lead the development of best practices, collaborate across time zones, and scale data infrastructure. Long-term, this role will evolve into managing and growing the data engineering team.
Must have:
  • 8+ years data engineering/software engineering experience
  • 1+ year tech lead experience
  • Full-cycle data warehousing project experience
  • Database design, architecture, and cost-efficient scaling
  • Proficiency in SQL, Python/Java/Golang
  • Experience with BigQuery, dbt
Good to have:
  • Salesforce, Marketo, Google Analytics experience
  • Distributed data processing & storage (HDFS)
  • Data privacy experience
  • Experience with Beam, Spark, Apache Kafka, Stitch, Hevo Data, Fivetran, GCP/AWS

Job Details

About Glean

We’re on a mission to make knowledge work faster and more humane. We believe that AI will fundamentally transform how people work. In the future, everyone will work in tandem with expert AI assistants who find knowledge, create and synthesize information, and execute work. These assistants will free people up to focus on the higher-level, creative aspects of their work.

We’re building a system of intelligence for every company in the world. On the surface, you can think of it as Google + ChatGPT for the enterprise. Under the hood, our platform is the connective tissue between AI and knowledge. It brings all of a company’s knowledge together, understands it at a deep level, provides industry-leading search relevance over it, and connects it to generative AI agents and applications.

Glean was founded by a seasoned team of former Google search and Facebook engineers who saw a need in the enterprise space for their technical depth and passion for AI. We’re a diverse team of curious and creative people who want to help each other get big things done—so we can help other teams do the same. 

We're backed by some of the Valley's leading venture capitalists—including Sequoia, Kleiner Perkins, Lightspeed, and General Catalyst—and have assembled a world-class team with senior leadership experience at Google, Slack, Facebook, Dropbox, Rubrik, Uber, Intercom, Pinterest, Palantir, and others.

Data Engineering Role:

Glean is building a world-class Data Organization composed of data science, applied science, data engineering and business intelligence groups. Our data engineering group will be based in our Bangalore, India office. We are hiring our first data engineer. In this role, you will:

  • Start as a fully hands-on individual contributor. If you deliver on the most urgent business needs with high quality hands-on execution, and showcase your leadership skills as an IC by effective collaboration with your XFNs as well as your manager and the rest of the company’s leadership, this role would evolve into you forming Glean’s first data engineering group within the Data org. 
  • Help improve the availability of high-value upstream raw data by 
    • channeling inputs from data science and business intelligence to identify biggest gaps in data foundations
    • partnering with Product Engineering teams as they craft product logging initiatives & processes
    • partnering with Go-to-Market & Finance operations groups to create streamlined data management processes in enterprise apps like Salesforce, Marketo and various accounting software
  • Architect and implement key tables that transform structured and unstructured data into usable models by the data, operations, and engineering orgs.
  • Ensure and maintain the quality and availability of Glean’s data within reasonable SLAs
  • Own and improve the reliability, efficiency and scalability of ETL tooling, including but not limited to dbt, BigQuery, Metabase. 
  • Partner with Business Intelligence to improve the reliability, scalability and usability of our business intelligence & visualization tools like Metabase for Data, product, engineering and operations teams.
  • Implement and disseminate developer-friendly best practices for our data stack to ensure that data, operations, and engineering can efficiently write source-controlled and adhoc SQL code and other ETL jobs.

You will thrive at this role if:

  • You have 8+ yrs of work experience in data engineering /software engineering as a bachelor degree holder. This requirement is 7+ for masters degree holders and 5+ for PhD Degree holders.
  • You have 1+ year of tech lead management experience and have mentored several data engineers before.
  • You have experience in full cycle data warehousing projects inclusive of requirements analysis, proof-of-concepts, design, development, testing and implementation
  • You have experience in database designing, architecting and cost efficient scaling
  • You have experience in architecting end to end cloud solutions for internal and third party data products
  • You have a high degree of proficiency with SQL and are able to set best practices and up-level our growing SQL user base within the organization
  • You are proficient in at least one of Python, Java and Golang
  • You have experience with cloud based data tools like BigQuery and dbt
  • You have experience with large scale data processing tools like Beam and Spark.
  • You have experience with data pipelining tools like Apache, Stitch, Hevo Data and Fivetran
  • You are familiar with cloud computing services like GCP and/or AWS.
  • You are concise and precise in written and verbal communication. Technical documentation is your strong suit. 
  • You have experience working with a large array of cross-functional partners ranging from product and engineering/research to go-to-market and finance
  • You have experience working with stakeholders and peers in different time zones 

You are a particularly good fit if:

  • You have experience with Salesforce, Marketo, and Google Analytics.
  • You have experience in distributed data processing & storage, e.g. HDFS
  • You have experience in data privacy, e.g. data access governance.
  • You have experience forming the data engineering charter in a startup

Similar Jobs

Meta - Software Engineer (University Grad)

Meta

San Francisco, California, United States (On-Site)
3 Months ago
Trellix - Sr Software Development Engineer ,Data Protection

Trellix

Bengaluru, Karnataka, India (On-Site)
3 Months ago
PwC - IN-Manager _Technical Delivery Manager_ Emerging Technologies_ Advisory_ Bengaluru

PwC

Bengaluru, Karnataka, India (On-Site)
4 Months ago
Google - Senior Quantitative UX Researcher, Search AI Overviews

Google

New York, New York, United States (On-Site)
3 Months ago
Zeta - Manager - Software Development

Zeta

Hyderabad, Telangana, India (On-Site)
4 Months ago
PwC - Corp Managed Svcs BOS RCMS- Senior Associate-Business Analyst - Operate

PwC

Bengaluru, Karnataka, India (On-Site)
4 Months ago
G5 Games - BI Analyst - BI Engineer

G5 Games

Astana, Astana, Kazakhstan (Remote)
3 Months ago
Dream11 - SDE 3 - ML & Data Platform

Dream11

Mumbai, Maharashtra, India (On-Site)
4 Months ago
Optum - Data Scientist

Optum

Noida, Uttar Pradesh, India (On-Site)
4 Months ago
Google - Staff Data Scientist, Research

Google

Mountain View, California, United States (On-Site)
3 Months ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Google - Research Scientist, Gemini

Google

New York, New York, United States (On-Site)
3 Months ago
Nasdaq - Senior Java Engineer

Nasdaq

Lisbon, Lisbon, Portugal (Hybrid)
4 Months ago
Info Stretch - Java Developer

Info Stretch

Lansing, Michigan, United States (On-Site)
3 Months ago
City State Entertainment - Senior Server Engineer (Remote)

City State Entertainment

Bothell, Washington, United States (Remote)
7 Months ago
Saviynt - Engineer, CloudOps

Saviynt

Atlanta, Georgia, United States (On-Site)
4 Months ago
ION - Principal Technical Consultant - Endur

ION

Berlin, Berlin, Germany (On-Site)
4 Months ago
Next Level Business Services - Java Full Stack Developer

Next Level Business Services

Tulsa, Oklahoma, United States (On-Site)
3 Months ago
ness - Sr Technical Application Architect

ness

Bengaluru, Karnataka, India (On-Site)
3 Months ago
PublicisGroupe - Senior Associate Data Engineering L2 DE - Big Data AWS

PublicisGroupe

Gurugram, Haryana, India (On-Site)
4 Months ago

Get notifed when new similar jobs are uploaded

Jobs in Bengaluru, Karnataka, India

PwC - Senior Associate-D365 Technical

PwC

Mumbai, Maharashtra, India (On-Site)
4 Months ago
Brillio - Engineering CoE Leader, Cloud Transformation - R01540491

Brillio

Bengaluru, Karnataka, India (Hybrid)
4 Months ago
STAGE - Marketing Manager - Haryana

STAGE

Noida, Uttar Pradesh, India (On-Site)
5 Months ago
Logitech - Salesforce CPQ Developer

Logitech

Chennai, Tamil Nadu, India (On-Site)
3 Months ago
Nasdaq - Oracle - Database Administrator Specialist

Nasdaq

Mumbai, Maharashtra, India (On-Site)
4 Months ago
gigamon - Sr. Cost Accounting Manager

gigamon

Chennai, Tamil Nadu, India (On-Site)
3 Months ago
MIPS - Embedded Software Engineer – RTOS – CPU/Platform Software Team

MIPS

Pune, Maharashtra, India (On-Site)
4 Months ago
Principal Global Services - Architect - Engineering

Principal Global Services

Hyderabad, Telangana, India (On-Site)
4 Months ago
Google - Business Intelligence Analyst

Google

Hyderabad, Telangana, India (On-Site)
3 Months ago

Get notifed when new similar jobs are uploaded

Data Analyst Jobs

PlayStation Global - Senior Portfolio Analyst

PlayStation Global

Aliso Viejo, California, United States (Hybrid)
3 Months ago
Peak - Data Scientist (New Grad)

Peak

(On-Site)
4 Months ago
DAZN - Streaming Data Analyst

DAZN

Hyderabad, Telangana, India (On-Site)
4 Months ago
Xsolla - Researcher/Analyst

Xsolla

Baku, Azerbaijan (Remote)
3 Months ago
Lulalend - Head of Credit Data Science

Lulalend

Cape Town, Western Cape, South Africa (On-Site)
4 Months ago
Meta - GRC Analyst

Meta

Menlo Park, California, United States (On-Site)
3 Months ago
Playrix - Senior Big Data Engineer

Playrix

Montenegro (Remote)
3 Months ago
Dream Game Studios - Senior ML Scientist

Dream Game Studios

Mumbai, Maharashtra, India (On-Site)
3 Months ago
Easy Brain - Middle/Senior Data Analyst

Easy Brain

Limassol, Limassol, Cyprus (Hybrid)
4 Months ago
Meta - Data Engineer Intern

Meta

Menlo Park, California, United States (On-Site)
3 Months ago

Get notifed when new similar jobs are uploaded

About The Company

Singapore (Remote)

Melbourne, Victoria, Australia (Remote)

Houston, Texas, United States (Remote)

Dallas, Texas, United States (Remote)

Nashville, Tennessee, United States (On-Site)

London, England, United Kingdom (Remote)

Bengaluru, Karnataka, India (On-Site)

Bengaluru, Karnataka, India (On-Site)

View All Jobs

Get notified when new jobs are added by Glean

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug