Tech Stack
Job Description, Responsibilities & Requirements
About the Position
We are seeking a Senior Data Engineer to design and develop the data management layer for our platform at Alpaca, a US-headquartered self-clearing broker-dealer. Our team is dedicated to opening financial services to everyone on the planet, and we're backed by top-tier global investors. We're a dynamic team of 380+ globally distributed members who thrive working from our favorite places around the world.
Responsibilities
- Design and oversee key forward- and reverse-ETL patterns to deliver data to relevant stakeholders.
- Develop scalable patterns in the transformation layer to ensure repeatable integrations with BI tools across various business verticals.
- Expand and maintain the Alpaca Data Lakehouse architecture's constantly evolving elements.
- Collaborate closely with sales, marketing, product, and operations teams to address key data flow needs.
- Operate the system and manage production issues in a timely manner.
Requirements
- 7+ years of experience in data engineering, including 2+ years of building scalable, low-latency data platforms capable of handling >100M events/day.
- Proficiency in at least one programming language, with strong working knowledge of Python and SQL.
- Experience with cloud-native technologies like Docker, Kubernetes, and Helm.
- Strong hands-on experience with relational database systems and object storage implementations like Apache Iceberg.
- Strong hands-on experience with Google Cloud Platform and its various data-related services (Composer, Dataproc, Datastream, etc.).
- Experience in building scalable transformation layers, preferably through formalized SQL models (e.g., dbt).
- Ability to work in a fast-paced environment and adapt solutions to changing business needs.
- Experience with ETL orchestrators / frameworks like Apache Airflow and Airbyte.
- Production experience with streaming systems like Apache Kafka.
- Exposure to infrastructure, DevOps, and Infrastructure as Code (IaaC), like Terraform.
- Deep knowledge of distributed systems, storage, transactions, and query processing utilizing open-source distributed query engines like Trino (formerly PrestoSQL).
Nice to Have
- Experience with dbt.
- Experience with Apache Iceberg.
- Experience with Trino.
- Experience with Terraform.
We Offer
- Competitive Salary & Stock Options
- Health Benefits
- New Hire Home-Office Setup: One-time USD $500
- Monthly Stipend: USD $150 per month via a Brex Card
About the Company
Alpaca is a US-headquartered self-clearing broker-dealer and brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more. Our recent Series D funding round brought our total investment to over $320 million, fueling our ambitious vision. Amongst our subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial institutions across 40 countries with our institutional-grade APIs. This includes broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges, totaling over 9 million brokerage accounts.
Alpaca is proudly backed by top-tier global investors, including Portage Ventures, Spark Capital, Tribe Capital, Social Leverage, Horizons Ventures, Unbound, SBI Group, Derayah Financial, Elefund, and Y Combinator.
Alpaca is proud to be an equal opportunity workplace dedicated to pursuing and hiring a diverse workforce.