Logo for Rockset

Data Engineer, People Innovation Labs

Role overview

Qualifications

  • 3+ years of experience as a data engineer and 8+ years of any software engineering experience (including data engineering)
  • Proficiency in Python, Scala, or Java
  • Experience with data warehousing technologies such as Databricks and Snowflake, and ETL schedulers such as Fivetran, Airflow, Dagster, Prefect, or similar
  • Experience with distributed processing technologies and frameworks such as Spark, Hadoop, Flink and distributed storage systems (e.g., HDFS, S3)

Responsibilities

  • Design, build and manage people data pipelines, ensuring all data is seamlessly integrated into our Databricks warehouse
  • Develop canonical datasets to track key people metrics and People Innovation Labs product metrics
  • Implement robust and fault-tolerant systems for data ingestion and processing
  • Participate in data architecture and engineering decisions, bringing your strong experience to bear as the primary data engineering expert on the team

About the company

Rockset logo

Rockset

Data Analytics & Business Intelligence

Our Vision We believe that a data-driven world has the potential to make life better for everyone. Enterprises are still struggling to use complex data primarily because real-world data is messy and cannot be put to use easily. We are bridging the gap by changing the way data is stored, processed and accessed for making better, faster data-driven decisions and data powered apps. Empowering enterprises to unleash all their data is a difficult challenge that inspires us every day. Our Team Rockset's team has deep expertise in storage, data management and distributed systems. Members of our team started the Hadoop File System project back in 2006 that helped ignite the big data movement. We previously founded and led the creation of Facebook's online social graph serving engine and graph search projects - TAO, and Unicorn - that power all of Facebook's user facing and search products. Our team also helped build the original backend for Gmail at Google. On the enterprise side, members of our team have experience launching VMware's vSAN and building the industry's first nested virtualization in the cloud at Ravello. We intimately understand data, cloud and scale as well as the challenges and opportunities it creates for enterprises.

Company details

Company typeScaleup
IndustryData Analytics & Business Intelligence
Company size51 - 200

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

About the Team

At OpenAI, we’re building the connective tissue between our mission and our people. People Innovation Labs is a fast-moving engineering team embedded in the People organization, focused on rethinking how we find and retain the best talent and empower everyone to do their best work. From recruiting to culture, we’re designing systems that give our People Team a significant edge by infusing OpenAI’s models and first-principles thinking into every aspect of our work. Our projects range from greenfield 0-1 products like OpenHouse (our internal knowledge hub) to AI-powered automations and scalable recruiting tools. We’re defining the future of work at OpenAI, creating a blueprint for how AI can supercharge productivity, culture, and innovation.

About the Role

We are looking for a hands-on engineering manager to lead technical and product strategy and execution for People Innovation Labs’ OpenHouse pod. OpenHouse is our flagship employee-facing product, serving as a culture and communication hub and an organization-wide front door into all other aspects of People Innovation Labs’ work. The OpenHouse pod is composed of full stack product engineers who are deeply curious about culture, recruiting and people development, and want to know everything from the business strategy and metrics down through the code that gets us there. In this role, you will work with People Innovation Labs leadership to build and grow the team and OpenHouse product, innovating on how we apply LLMs along the way.

We’re seeking a Data Engineer to build data-intensive systems that will power People Innovation Labs’ internal products and enable the People Analytics function to do their best work. These data pipelines are crucial for our build-out of people products backed by business systems of record and for ongoing people data analytics.

One example of an employee-facing product you’ll help us build is OpenHouse, which serves as a culture and communication hub and an organization-wide front door into all other aspects of People Innovation Labs’ work. OpenHouse and other products in our portfolio are built by full stack product engineers who are deeply curious about culture, recruiting and people development, and want to know everything from the business strategy and metrics down through the code that gets us there. In this role, you will work with People Innovation Labs leadership and software engineers and the People Analytics team to build the data systems that enable this work.

In this role, you will:

  • Design, build and manage people data pipelines, ensuring all data is seamlessly integrated into our Databricks warehouse.

  • Develop canonical datasets to track key people metrics and People Innovation Labs product metrics.

  • Work collaboratively with various teams, including, Data Platform, Data Science, People Analytics, and Compensation and Equity to understand their data needs and provide solutions.

  • Implement robust and fault-tolerant systems for data ingestion and processing.

  • Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear as the primary data engineering expert on the team.

  • Ensure the security, integrity, and compliance of data according to industry and company standards.

Your background might look something like:

  • Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience (including data engineering).

  • Proficiency in at least one programming language commonly used within Data Engineering, such as Python, Scala, or Java.

  • Experience with data warehousing technologies such as Databricks and Snowflake, and expertise with ETL schedulers such as Fivetran, Airflow, Dagster, Prefect, or similar.

  • Experience with distributed processing technologies and frameworks, such as Spark, Hadoop, Flink and distributed storage systems (e.g., HDFS, S3).

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Data Engineer Related jobs

Other jobs at Rockset

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.