Staff Data Engineer
Courtyard Inc. · United States
Technology, Information and Internet · 51-200 employees
About the role
The Staff Data Engineer will own the end-to-end data foundation, including external market data ingestion, entity resolution, and warehouse modeling. They will also set engineering standards and provide reliable data infrastructure for pricing and machine learning models.
What they look for
Requirements
Candidates must have 5+ years of experience in data engineering with a focus on production pipelines and handling third-party data at scale. Proficiency in SQL, Python, dbt, and cloud warehouse technologies like BigQuery is required.
Benefits
Full description
About Courtyard
Courtyard.io is the fastest-growing collectibles startup, ever. We are revolutionizing the world of collectibles trading by enabling instant liquidity and delivering a high-velocity, immersive experience. From trading cards to comics, we're redefining how people discover, collect, and unlock value.
Courtyard is not just another marketplace. All assets are securely vaulted and fully insured, giving collectors peace of mind alongside unmatched speed and simplicity. Whether you're investing, discovering, or curating your dream collection, we've built a platform that's trusted, simple, and built for speed.
And we're just getting started. We're a remote-first company hiring across all functions to push the boundaries of what's possible in collectibles and digital ownership.
Job Summary
Courtyard is hiring a Staff Data Engineer to own the data foundation our pricing runs on. We price and buy back assets in real time off a fragmented external market, and turning that into something a model can safely consume is the data problem here. You'll be the senior-most data engineering voice in a lean data org — owning architecture and standards, not working a ticket queue.
What You'll Do
- Own external market data ingestion end to end — acquisition, normalization, and delivery into the warehouse
- Own the canonical item model and the entity resolution that maps messy inbound records onto it
- Build quality and monitoring that catches gaps and degradation before a stakeholder notices
- Own the warehouse layer: dbt modeling, tests, and documentation across marketplace, vault, payments, and event data
- Set schema contracts with product engineering so upstream changes break the build, not the metrics
- Serve pricing and ML with reliable training and feature tables
- Set the engineering bar for the data org and level up DS on production-path work
What We're Looking For
- 5+ years in data engineering with real production pipeline ownership, at senior or staff scope
- Experience taming third-party data at volume — unstable sources, schema drift, no advance notice
- Entity resolution in practice, including the judgment to know when a confident match is wrong
- Deep SQL and Python; production dbt with tests and incremental models you actually maintained
- Strong dimensional modeling — SCD Type 2, incremental loads, backfills, late-arriving data
- Cloud warehouse depth. We run BigQuery and GCP; Redshift, and dbt
- Marketplace or e-commerce data models: orders, payments, refunds, promos, attribution
- Broad enough to navigate backend systems and pair with AI on the hard parts
- Bonus: pricing or valuation systems, LLM-assisted extraction, a real interest in collectibles
What You'll Get In Return
- A dynamic and engaging environment focused on fostering real growth and innovation
- Opportunities to create amazing products that our customers truly love and value
- Comprehensive health insurance packages with dependent coverage
- Competitive salary with ample opportunities for career advancement and development
- Enjoy the flexibility of a fully remote work environment
- 401(k) plan with a 4% employer match to help you plan for the future
- $100 monthly dogfooding stipend to support trying out our products firsthand
- Access to employee wellness programs designed to support your overall well-being