Design, build, and operate scalable data transformation pipelines and dimensional models using dbt, SQL, and AWS. Implement infrastructure-as-code (AWS CDK), CI/CD, and query performance optimizations while ensuring data quality, governance, and collaboration with cross-functional and offshore teams.
Overview /Objective:
We are seeking a seasoned Data Engineer to join our Sports Analytics & Engineering Practice. This role is pivotal in shaping and implementing our client’s vision for a cutting-edge, cloud-native data ecosystem. You will architect and build scalable data infrastructure that transforms raw data into high-value assets, powering analytics across digital products, fan engagement, and marketing domains. Your work will directly contribute to the development of a world-class customer data platform.
ResponsibilitiesResponsibilities:
- Design and build robust, scalable data transformation pipelines using SQL, DBT, and Jinja templating
- Develop and maintain data architecture and standards for Data Integration and Data Warehousing projects using DBT and Amazon Redshift
- Collaborate with cross-functional teams to gather requirements and deliver dimensional data models that serve as a single source of truth
- Own the full stack of data modeling in DBT to empower analysts, data scientists, and BI engineers
- Enhance and maintain the analytics codebase, including DBT models, SQL scripts, and ERD documentation
- Ensure data quality, governance alignment, and operational readiness of data pipelines
- Apply software engineering best practices such as version control, CI/CD, and code reviews
- Optimize SQL queries for performance, scalability, and maintainability across large datasets
- Implement best practices for SQL performance tuning, including partitioning, clustering, and materialized views
- Build and manage infrastructure as code using AWS CDK for scalable and repeatable deployments. Integrate and automate deployment workflows using AWS CodeCommit, CodePipeline, and related DevOps tools
- Support Agile development processes and collaborate with offshore teams
Required Qualifications:
- Bachelor’s or Master’s (preferred) degree in a quantitative or technical field such as Statistics, Mathematics, Computer Science, Information Technology, Computer Engineering or equivalent
- 5+ years of experience in data engineering and analytics on modern data platforms
- 3+ years’ extensive experience with DBT or similar data transformation tools, including building complex & maintainable DBT models and developing DBT packages/macros
- Deep familiarity with dimensional modeling/data warehousing concepts and expertise in designing, implementing, operating, and extending enterprise dimensional models
- Understand change data capture concepts
- Experience working with AWS Services (Lambda, Step Functions, MWAA, Glue, Redshift)
- Hands-on experience with AWS CDK, CodeCommit, and CodePipeline for infrastructure automation and CI/CD
- Python proficiency or general knowledge of Jinja templating in Python and/or PySpark
- Agile experience and willingness to work with extended offshore teams and assist with design and code reviews with customer
- A great teammate and self-starter, strong detail orientation is critical in this role.
Similar Jobs
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Senior Data Engineer responsible for Epic and EHR integrations supporting value-based care, risk and quality workflows, member attribution, roster management, and care-gap tools. The role designs and executes EHR development tasks, coordinates with clinical, data, operations, and development teams, documents business and data flows, evaluates AI and automation opportunities, and develops reporting to identify data-quality issues and prevent outages.
Top Skills:
Ai ToolsCaboodleClarityEhr IntegrationsEpicHealthy PlanetSQL
Big Data • Cloud • Productivity • Software • Database • Analytics • Automation
Build and maintain Databricks-based data platforms, including ingestion, transformation, storage, governance, data modeling, and serving pipelines. Establish medallion architecture standards, canonical data models, quality controls, lineage, schema evolution, and reliable batch or incremental processing. Improve pipeline observability, scalability, idempotency, and recoverability while moving curated data to systems such as ClickHouse. Collaborate across application and analytics teams to create durable, governed production datasets.
Top Skills:
Amazon AuroraAmazon RdsApache AirflowSparkBigQueryCdcClickhouseCloud Object StorageDatabricksDelta LakeIamOpenmetadataPostgresSnowflakeUnity Catalog
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Design and scale lakehouse architecture, streaming and batch pipelines, and reliable data platforms using Kafka, Spark, Airflow, Iceberg, Databricks, and related technologies. Build Medallion-layer data systems, manage open table formats, optimize distributed queries, monitor platform reliability, resolve data quality issues, and collaborate with analysts, data scientists, and product teams.
Top Skills:
Apache AirflowApache HudiApache IcebergApache KafkaSparkDatabricksDelta LakePythonSQLStarburstTrino
What you need to know about the Boston Tech Scene
Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.
Key Facts About Boston Tech
- Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
- Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
- Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
- Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories



