Skip to content
Skip to content
AI Engineer Jobs
rPotential

Sr. AI Product Data Engineer

rPotential

Location
Hybrid (San Francisco, California)
Compensation
$190k - $245k/yr
Employment
Full-time
Level
Senior Level
Posted 3 days ago

About the Role

rPotential is a small, fast-moving team building an enterprise intelligence platform that helps Fortune 500 companies understand workforce changes and AI integration. This role involves taking ownership of production data pipelines to power customer-facing products.

Skills

Data Engineering Databricks Unity Catalog Python SQL Postgres Pipeline Development Data Quality Monitoring Azure Terraform Entity Resolution Identity Matching AI LLM

Perks

  • Hybrid Work

Full job details

Sr. AI Product Data Engineer


rPotential | SF Bay Area | 3 days/week in person


About us

rPotential is building a platform that helps large companies understand where to apply AI, how work is changing, and where people can be redeployed to higher-value work. We work directly with Fortune 500 companies and leading AI companies.

We're a small team, move quickly, and expect everyone to be hands-on.


About the role

Our product depends on data. We combine proprietary, third-party, and AI-generated datasets in Databricks and Postgres to power customer-facing products.

We have production pipelines today, but they were built quickly and vary in how they are structured, tested, monitored, and operated. We're looking for a senior data engineer to take ownership of this layer, improve what exists, and establish a consistent way to build and run data pipelines going forward.

This role also requires understanding the business context behind the data. You'll need to understand what the data represents, where it came from, where it can be misleading, and how it is used in the product.


What you'll do

  • Own and improve our production data pipelines.
  • Standardize pipeline structure, scheduling, retries, testing, monitoring, and backfills.
  • Build monitoring around freshness, coverage, failures, and data quality.
  • Own Databricks across jobs, Unity Catalog, permissions, environments, and cost.
  • Work in our product monorepo alongside the engineering team.
  • Work with product and business teams to understand new datasets and resolve data issues.
  • Make it easier to take new data sources from prototype to production.


What we're looking for

  • 5+ years of data engineering experience with production pipeline ownership.
  • Strong DataBricks and Unity Catalog experience, including workspace administration.
  • Strong Python and SQL.
  • Experience improving an existing data environment while it remained live.
  • Strong business judgment around data and an interest in understanding what the data actually means.
  • Comfortable working through ambiguous or messy data with product and business teams rather than waiting for a finished spec.
  • Comfortable with AI coding tools
  • Bay Area based or commutable and comfortable working in person three days per week.

Nice to ave

  • Azure, terraform, Postgres
  • Entity resolution or identity matching.
  • Labor-market, company, people, or other large third-party datasets.
  • Data systems supporting AI or LLM products.


What we offer

High ownership, direct work with the founder and engineering team, and the opportunity to define how data engineering is done at rPotential as the team grows.