Plaid

Senior Data Engineer - Data Engineering

Plaid$130K — $155K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years of experience in data engineering, focusing on large-scale data pipelines.
  • Experience building data models and pipelines for datasets ranging from 500TB to petabytes.
  • Proficient with SQL and modern orchestration tools like DBT, Mode, and Airflow.
  • Familiarity with data warehouses and lakes such as Redshift, Snowflake, and Databricks.
  • Skilled in building batch and real-time data pipelines using technologies like Spark and Kafka.
  • Expertise in schema design for unstructured data and evolving analytics schemas.
  • Strong ability to collaborate with stakeholders to align data solutions with business needs.

Responsibilities

  • Understand Plaid’s product and strategy to inform data choices and usage principles.
  • Design data solutions with a focus on quality and performance.
  • Lead data engineering projects that foster cross-company collaboration.
  • Own and manage SQL and Python data pipelines for data lakes and warehouses.
  • Ensure documentation on dataset quality and uptime is comprehensive.

Benefits

  • Comprehensive benefit plan including medical, dental, vision, and 401(k).
  • Support for diversity and inclusion within the workplace.
  • Encouragement for candidates with diverse experiences to apply.
Full Job Description
Making data-driven decisions is key to Plaid's culture. To support that, we need to scale our data systems while maintaining correct and complete data. We provide golden datasets and tooling to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively. In addition, Plaid will not be successful if we can't move quickly. We build the data systems and tools that enable everyone at Plaid to be data-driven, making analytics easy, obvious, and proactive across the company.

Data Engineers heavily leverage SQL and Python to build data workflows that integrate with our Golang applications. We use tools like DBT, Airflow, Redshift, Atlan, and Retool to orchestrate data pipelines and define workflows. We work with engineers, product managers, business intelligence, data analysts, and many other teams to build Plaid's data strategy and a data-first mindset.

You will be in a high impact role that will directly enable business leaders to make faster and more informed business judgements based on the datasets you build. You will have the opportunity to carve out the ownership and scope of internal datasets and visualizations across Plaid which is a currently unowned area that we intend to take over and build SLAs on. You will have the opportunity to learn best practices and up-level your technical skills from our strong DE team and from the broader Data Platform team. You will collaborate with and have strong and cross functional partnerships with literally all teams at Plaid from Engineering to Product to Marketing/Finance etc.

Responsibilities
  • Understanding different aspects of the Plaid product and strategy to inform golden dataset choices, design and data usage principles.
  • Have data quality and performance top of mind while designing datasetsLeading key data engineering projects that drive collaboration across the company.
  • Advocating for adopting industry tools and practices at the right time.
  • Owning core SQL and python data pipelines that power our data lake and data warehouse.
  • Well-documented data with defined dataset quality, uptime, and usefulness.


Qualifications
  • 4+ years of dedicated data engineering experience, solving complex data pipelines issues at scale.
  • You've have experience building data models and data pipelines on top of large datasets (in the order of 500TB to petabytes)
  • You value SQL as a flexible and extensible tool, and are comfortable with modern SQL data orchestration tools like DBT, Mode, and Airflow.
  • You have experience working with different performant warehouses and data lakes; Redshift, Snowflake, Databricks.
  • You have experience building and maintaining batch and realtime pipelines using technologies like Spark, Kafka.
  • You appreciate the importance of schema design, and can evolve an analytics schema on top of unstructured data.
  • You are excited to try out new technologies. You like to produce proof-of-concepts that balance technical advancement and user experience and adoption.
  • You like to get deep in the weeds to manage, deploy, and improve low level data infrastructure.
  • You are empathetic working with stakeholders. You listen to them, ask the right questions, and collaboratively come up with the best solutions for their needs while balancing infra and business needs.
  • You are a champion for data privacy and integrity, and always act in the best interest of consumers.

Our mission at Plaid is to unlock financial freedom for everyone. To support that mission, we seek to build a diverse team of driven individuals who care deeply about making the financial ecosystem more equitable. We recognize that strong qualifications can come from both prior work experiences and lived experiences. We encourage you to apply to a role even if your experience doesn't fully match the job description. We are always looking for team members that will bring something unique to Plaid!

Additional compensation in the form(s) of equity and/or commission are dependent on the position offered. Plaid provides a comprehensive benefit plan, including medical, dental, vision, and 401(k). Pay is based on factors such as (but not limited to) scope and responsibilities of the position, candidate's work experience and skillset, and location. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation or benefit plans.

About Plaid

Plaid is a financial services company based in New York City. The company builds a technology platform, which enables applications to connect with users' bank accounts. Plaid focuses on enabling consumers and businesses to interact with their bank accounts, check balances, and make payments through financial technology applications. The company was founded in 2013 by Zach Perret and William Hockey. In January 2020, Visa announced that it would acquire Plaid for $5.3 billion. The acquisition was completed in January 2021.
Learn more about Plaid
Size
600 employees
Industry
Founded
2011

Similar Jobs

More Jobs at Plaid

More Information Technology Jobs

Find similar Senior Data Engineer - Data Engineering jobs: