Senior Software Engineer - Data Infrastructure
AI summary of the role
Senior Software Engineer on Plaid's Data Infrastructure team, responsible for scaling data systems (warehouse, lakehouse, Spark, streaming) and building abstractions that enable other engineers to safely and quickly derive insights from consumer-permissioned financial data.
What you’ll do
- Scale data systems while maintaining correct and complete data
- Provide tooling and guidance to teams across engineering, product, and business to explore data quickly and safely
- Build data and machine learning infrastructure for prototyping and iterating on products built on consumer-permissioned financial data
- Scale existing data pipelines in a performant and cost efficient way
What you’ll bring
- Domain expertise in Data Warehouse, Data Lakehouse, Spark, Workflow Orchestration, and Streaming technologies
- Experience scaling data pipelines with a focus on performance and cost efficiency
- Ability to build abstractions and platforms for other engineers
Technologies
Data Warehouse · Data Lakehouse · Spark · Workflow Orchestration · Streaming · Machine Learning Infrastructure
About Plaid
API platform connecting consumer bank accounts to fintech apps; powers account linking, payments, identity, and fraud detection for thousands of fintechs and banks.
Private Late
Source and classification
Internal deployment & tooling · Evidence for this classification:
maintaining correct and complete data. We provide tooling and guidance to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively. We build the data and machine learning infrastructure to enable Plaid engineers to prototype and iterate on products and features built on top of consumer-permissioned financial data. Engineers on Data Infrastructure are domain experts in Data Warehouse, Data Lakehouse, Spark, Workflow Orchestration, and Streaming technologies. We scale our existing data pipelines in a performant and cost efficient way while creating the necessary abstractions to make developing on top of this platform extremely simple for other engineers at Plaid.
More from the job description
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Making data driven decisions is key to Plaid's culture. To support that, we need to scale our data systems while maintaining correct and complete data. We provide tooling and guidance to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively. We build the data and machine learning infrastructure to enable Plaid engineers to prototype and iterate on products and features built on top of consumer-permissioned financial data. Engineers on Data Infrastructure are domain experts in Data Warehouse, Data Lakehouse, Spark, Workflow Orchestration, and Streaming technologies. We scale our existing data pipelines in a performant and cost efficient way while creating the necessary abstractions to make developing on top of this platform extremely simple for other engineers at Plaid.
Employer postings · Data from · Sources