Data Engineer Lead

1 Month ago • 5-8 Years • Data Analyst

About the job

Job Description

Glean seeks a Data Engineer Lead to build and lead its first data engineering group in Bangalore. This hands-on role initially involves contributing to urgent business needs, improving data availability, partnering with cross-functional teams (Product Engineering, Go-to-Market, Finance), architecting data models, ensuring data quality and availability, and improving ETL tooling (dbt, BigQuery, Metabase). The successful candidate will lead the development of best practices, collaborate across time zones, and scale data infrastructure. Long-term, this role will evolve into managing and growing the data engineering team.
Must have:
  • 8+ years data engineering/software engineering experience
  • 1+ year tech lead experience
  • Full-cycle data warehousing project experience
  • Database design, architecture, and cost-efficient scaling
  • Proficiency in SQL, Python/Java/Golang
  • Experience with BigQuery, dbt
Good to have:
  • Salesforce, Marketo, Google Analytics experience
  • Distributed data processing & storage (HDFS)
  • Data privacy experience
  • Experience with Beam, Spark, Apache Kafka, Stitch, Hevo Data, Fivetran, GCP/AWS
Not hearing back from companies?
Unlock the secrets to a successful job application and accelerate your journey to your next opportunity.

About Glean

We’re on a mission to make knowledge work faster and more humane. We believe that AI will fundamentally transform how people work. In the future, everyone will work in tandem with expert AI assistants who find knowledge, create and synthesize information, and execute work. These assistants will free people up to focus on the higher-level, creative aspects of their work.

We’re building a system of intelligence for every company in the world. On the surface, you can think of it as Google + ChatGPT for the enterprise. Under the hood, our platform is the connective tissue between AI and knowledge. It brings all of a company’s knowledge together, understands it at a deep level, provides industry-leading search relevance over it, and connects it to generative AI agents and applications.

Glean was founded by a seasoned team of former Google search and Facebook engineers who saw a need in the enterprise space for their technical depth and passion for AI. We’re a diverse team of curious and creative people who want to help each other get big things done—so we can help other teams do the same. 

We're backed by some of the Valley's leading venture capitalists—including Sequoia, Kleiner Perkins, Lightspeed, and General Catalyst—and have assembled a world-class team with senior leadership experience at Google, Slack, Facebook, Dropbox, Rubrik, Uber, Intercom, Pinterest, Palantir, and others.

Data Engineering Role:

Glean is building a world-class Data Organization composed of data science, applied science, data engineering and business intelligence groups. Our data engineering group will be based in our Bangalore, India office. We are hiring our first data engineer. In this role, you will:

  • Start as a fully hands-on individual contributor. If you deliver on the most urgent business needs with high quality hands-on execution, and showcase your leadership skills as an IC by effective collaboration with your XFNs as well as your manager and the rest of the company’s leadership, this role would evolve into you forming Glean’s first data engineering group within the Data org. 
  • Help improve the availability of high-value upstream raw data by 
    • channeling inputs from data science and business intelligence to identify biggest gaps in data foundations
    • partnering with Product Engineering teams as they craft product logging initiatives & processes
    • partnering with Go-to-Market & Finance operations groups to create streamlined data management processes in enterprise apps like Salesforce, Marketo and various accounting software
  • Architect and implement key tables that transform structured and unstructured data into usable models by the data, operations, and engineering orgs.
  • Ensure and maintain the quality and availability of Glean’s data within reasonable SLAs
  • Own and improve the reliability, efficiency and scalability of ETL tooling, including but not limited to dbt, BigQuery, Metabase. 
  • Partner with Business Intelligence to improve the reliability, scalability and usability of our business intelligence & visualization tools like Metabase for Data, product, engineering and operations teams.
  • Implement and disseminate developer-friendly best practices for our data stack to ensure that data, operations, and engineering can efficiently write source-controlled and adhoc SQL code and other ETL jobs.

You will thrive at this role if:

  • You have 8+ yrs of work experience in data engineering /software engineering as a bachelor degree holder. This requirement is 7+ for masters degree holders and 5+ for PhD Degree holders.
  • You have 1+ year of tech lead management experience and have mentored several data engineers before.
  • You have experience in full cycle data warehousing projects inclusive of requirements analysis, proof-of-concepts, design, development, testing and implementation
  • You have experience in database designing, architecting and cost efficient scaling
  • You have experience in architecting end to end cloud solutions for internal and third party data products
  • You have a high degree of proficiency with SQL and are able to set best practices and up-level our growing SQL user base within the organization
  • You are proficient in at least one of Python, Java and Golang
  • You have experience with cloud based data tools like BigQuery and dbt
  • You have experience with large scale data processing tools like Beam and Spark.
  • You have experience with data pipelining tools like Apache, Stitch, Hevo Data and Fivetran
  • You are familiar with cloud computing services like GCP and/or AWS.
  • You are concise and precise in written and verbal communication. Technical documentation is your strong suit. 
  • You have experience working with a large array of cross-functional partners ranging from product and engineering/research to go-to-market and finance
  • You have experience working with stakeholders and peers in different time zones 

You are a particularly good fit if:

  • You have experience with Salesforce, Marketo, and Google Analytics.
  • You have experience in distributed data processing & storage, e.g. HDFS
  • You have experience in data privacy, e.g. data access governance.
  • You have experience forming the data engineering charter in a startup
View Full Job Description

Add your resume

80%

Upload your resume, increase your shortlisting chances by 80%

About The Company

Bengaluru, Karnataka, India (On-Site)

Bengaluru, Karnataka, India (On-Site)

United States (Remote)

Palo Alto, California, United States (Hybrid)

Bengaluru, Karnataka, India (On-Site)

Bengaluru, Karnataka, India (On-Site)

Melbourne, Victoria, Australia (On-Site)

Palo Alto, California, United States (On-Site)

Bengaluru, Karnataka, India (On-Site)

View All Jobs

Get notified when new jobs are added by Glean

Similar Jobs

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Google - Research Scientist, Gemini

Google, United States (On-Site)

Nasdaq - Senior Java Engineer

Nasdaq, Portugal (Hybrid)

Info Stretch - Java Developer

Info Stretch, United States (On-Site)

City State Entertainment - Senior Server Engineer (Remote)

City State Entertainment, United States (Remote)

Saviynt - Engineer, CloudOps

Saviynt, United States (On-Site)

Next Level Business Services - Java Full Stack Developer

Next Level Business Services, United States (On-Site)

Get notifed when new similar jobs are uploaded

Jobs in Bengaluru, Karnataka, India

Get notifed when new similar jobs are uploaded

Data Analyst Jobs

PlayStation Global - Senior Portfolio Analyst

PlayStation Global, United States (Hybrid)

DAZN - Streaming Data Analyst

DAZN, India (On-Site)

Xsolla - Researcher/Analyst

Xsolla, Azerbaijan (Remote)

Lulalend - Head of Credit Data Science

Lulalend, South Africa (On-Site)

Meta - GRC Analyst

Meta, United States (On-Site)

Playrix - Senior Big Data Engineer

Playrix, Montenegro (Remote)

Dream Game Studios - Senior ML Scientist

Dream Game Studios, India (On-Site)

Easy Brain - Middle/Senior Data Analyst

Easy Brain, Cyprus (Hybrid)

Meta - Data Engineer Intern

Meta, United States (On-Site)

Get notifed when new similar jobs are uploaded