GoGuardian Logo

GoGuardian

Data Engineer II

Posted Yesterday
Remote
Hiring Remotely in United States
130K-150K Annually
Mid level
Remote
Hiring Remotely in United States
130K-150K Annually
Mid level
Designs and optimizes ETL pipelines supporting analytics, data science, and machine learning. Builds labeling, retraining, MLOps, model monitoring, and production inference workflows. Contributes to lakehouse architecture, data modeling, governance, lineage, and quality. Uses infrastructure as code and containerization to create reproducible environments, collaborates with data scientists and BI teams, participates in code reviews, and improves engineering practices.
The summary above was generated by AI
What We Do
 
At GoGuardian, we’re helping build a future where all learners are ready and inspired to solve the world’s greatest challenges. Our award-winning system of learning solutions is purpose-built for K-12 and trusted by school leaders to promote effective teaching and equitable engagement while helping empower educators to keep students safe. 
 
What It’s Like to Work at GoGuardian

We are an outcomes-focused learning company with a steadfast focus on improving learning environments, one classroom at a time. Working with us means joining a remote team of diverse, committed, mission-driven employees who are inspired by our vision, dedicated to our customers, and ready to roll up their sleeves. Guardians put their heads together to solve problems, learn together from experiments that fail, and stand together by their work with full accountability. We balance our diligence with an inclusive culture that invites everyone to bring their whole self to work. Join us and learn why “I love the people here” is one of the most frequent comments we hear from Guardians.

The Role

We’re looking for a Data Engineer II to help design, build, and continuously improve the GoGuardian Analytics and AI/ML ecosystem. This position sits on the Data Engineering team, a group responsible for building and maintaining the core data platform that powers analytics, product insights, and machine learning across the company. You’ll collaborate closely with Data Science, Business Intelligence, and other teams to enable the next generation of data-driven products and AI capabilities.

The ideal candidate combines strong software engineering and data architecture skills with curiosity about machine learning systems and a drive to automate, optimize, and scale data workflows.

What You'll Do

  • Design, build, and optimize ETL pipelines that power analytics, data science, and ML workflows using tools such as Databricks, PySpark, and Airflow.
  • Develop and maintain labeling and retraining pipelines for machine learning models, ensuring quality, reproducibility, and observability.
  • Implement and support MLOps practices, including model versioning, CI/CD for ML, and model monitoring in production environments.
  • Collaborate with data scientists to productionize and scale model training, inference, and evaluation pipelines.
  • Contribute to the design and evolution of the data lakehouse, including schema design, partitioning strategies, and performance optimization.
  • Document and communicate data architecture, lineage, and dependencies to ensure transparency and maintainability across teams.
  • Champion data quality and governance, ensuring that datasets are accurate, well-structured, and compliant with organizational standards.
  • Leverage infrastructure-as-code and containerization to build reproducible, maintainable environments.
  • Participate in code reviews and continuous improvement of engineering best practices within the team.

Who You Are

  • Bachelor’s degree in Computer Science, Engineering, or related field.
  • 2–4 years of experience building and operating large-scale data systems, ideally supporting analytics and ML workloads.
  • Proficiency in Python and SQL, with experience in PySpark, pandas, or similar data processing frameworks.
  • Experience with DBT
  • Experience with modern data warehousing and lakehouse platforms, preferably Databricks.
  • Hands-on experience with workflow orchestration tools such as Airflow, Dagster, or Prefect.
  • Strong understanding of data modeling, ETL design, and distributed data systems.
  • Experience with AWS data and compute services (S3, Lambda, ECS, CloudWatch, etc.) or equivalent cloud platforms.
  • Familiarity with MLOps concepts (e.g., feature stores, model registries, CI/CD for ML).
  • Experience using Infrastructure as Code, preferably Terraform.
  • Excellent problem-solving, collaboration, and communication skills; comfortable working in a dynamic, fast-paced environment.

 What We Offer 

  • Competitive pay, complete health insurance, 401(k) matching, and an employee equity plan.
  • Flexible time off, paid holidays, paid parental leave, and a paid year-end holiday break.
  • A robust catalog of benefits that support your professional growth and personal wellbeing, including work from home funds, fertility & adoption reimbursement, and more…

Plus the intangible:

  • A varied and challenging role in an innovative, global company.
  • Supportive, driven colleagues who have your back and share your passion.

The typical base salary range for this position is $130,000 - $150,000 per year. The range displayed on this job posting reflects the minimum and maximum target for new hire base pay for this position and your pay will be determined by a variety of factors, including your primary work location, skills, qualifications and experience. Additional benefits information is listed on our careers page.


Please share this with your friends or co-workers who may be interested in working at GoGuardian! We have multiple openings and are always looking for talented people. 

 
GoGuardian is an equal opportunity employer and makes employment decisions on the basis of merit and business needs. GoGuardian does not discriminate against employees, applicants, interns or volunteers on the basis of race, religion, color, national origin, ancestry, physical disability, mental disability, medical condition, pregnancy, marital status, sex, age, sexual orientation, military and veteran status, registered domestic partner status, genetic information, gender, gender identity, gender expression, or any other characteristic protected by applicable law.
 
GoGuardian's Job Applicant Privacy Policy is located here
 
#BI-Remote

Similar Jobs

2 Days Ago
Remote
USA
176K-207K Annually
Senior level
176K-207K Annually
Senior level
Big Data • Food • Mobile • Payments
Senior Software Engineer joining Fetch’s Data Platform team to architect and evolve scalable data pipelines, governed data access, delivery infrastructure, and partner integrations. The role owns ambiguous technical problems, leads multi-system initiatives and migrations, improves reliability and engineering efficiency, shapes technical direction, collaborates across teams, and mentors engineers. Candidates need strong backend or platform engineering experience, data infrastructure knowledge, cross-team leadership, operational judgment, and excellent communication skills.
Top Skills: Ai-Assisted Development WorkflowsCi/CdEvent StreamingKubernetesObservability Tooling
6 Days Ago
Remote
U.S.
95K-119K Annually
Mid level
95K-119K Annually
Mid level
Consumer Web • Digital Media • Information Technology • Music • Software
Build and maintain production data pipelines and models using SQLMesh, BigQuery, and Python. Improve observability, reliability, and performance; support finance and revenue pipelines; participate in on-call rotation and own features end-to-end.
Top Skills: AirflowBigQueryDagsterDbtPrefectPythonSQLSqlmesh
10 Days Ago
Remote
United States
109K-125K Annually
Mid level
109K-125K Annually
Mid level
Cloud • Healthtech • Information Technology • Software • Consulting
Design, implement, and maintain data interactions and backend systems using Python and Apache Spark. Build APIs, optimize ETL and distributed data processes, contribute to UI and automated testing, and support CI/CD-driven deployments within an agile team focused on healthcare quality reporting.
Top Skills: SparkAPIsCi/Cd PipelinesData ModelingData WarehousesDatabasesDistributed ComputingETLOrchestrationPython

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account