Maximum of 25 job preferences reached.
Top Remote Senior Machine Learning Engineer Jobs in Boston, MA
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Design, build, and support high-throughput, low-latency LLM inference and data pipelines; collaborate with data scientists to post-train and evaluate models; implement scalable, production-ready systems with strong testing, monitoring, and performance focus; mentor engineers and iterate on architecture, reliability, and user-facing AI applications.
Top Skills:
AnsibleAWSCassandraChefDockerElasticsearchGCPGpuGpu-ClustersJvmKafkaKubernetesMaasPythonSparkTerraform
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Build and operate production machine learning systems for Safety AI, including low-latency APIs, data pipelines, model serving, evaluation, monitoring, and rollout infrastructure. Process large-scale camera and telematics data, optimize cloud and edge-to-cloud execution, track model performance and drift, and partner with applied scientists, firmware engineers, platform teams, and product managers to deliver reliable safety features.
Top Skills:
SparkAWSAzureC++Ci/CdComputer VisionDockerGCPGoGpusGrafanaInfrastructure As CodeJavaKubernetesMachine LearningMlflowPythonPyTorchRayRay ServeScala
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Lead end-to-end LLM inference performance across serving systems, models, GPU hardware, and workloads. Define rigorous benchmarks, diagnose bottlenecks from scheduling through kernels and interconnects, optimize single- and multi-node deployments, and deliver production-ready runtimes and configurations. Collaborate with product and infrastructure teams, evaluate inference technologies, and implement runtime fixes. The role requires strong Python engineering, production experience with serving engines such as vLLM or SGLang, GPU profiling, and knowledge of modern inference optimization techniques.
Top Skills:
CudaDistributed ServingGpu ProfilingGpu SystemsHigh-Speed NetworkingPythonQuantizationSglangSpeculative DecodingTritonVllm
Artificial Intelligence
Design, train, and deploy scalable ML models (generative, classification, regression). Define feature roadmaps, collaborate with engineering to implement code and APIs, adapt ML methods for distributed/GPU environments, and create evaluation/annotation programs for fine-tuning and performance measurement.
Top Skills:
Apache BeamSparkApi DevelopmentDistributed SystemsGpuPythonPyTorchTensorFlow
Artificial Intelligence • Information Technology • Software • Cybersecurity
Own the full lifecycle of production classification and detection models, including transformer fine-tuning, entity detection, behavioral risk scoring, training, serving, versioning, evaluation, drift detection, and retraining. Build SageMaker-based ML systems and define output contracts supporting audit and policy enforcement. Lead the evaluation harness and mentor a co-op while operating as the company’s sole ML engineer.
Top Skills:
Amazon EksAmazon SagemakerAWSGrafanaHelmHugging Face TransformersKafkaKinesisPrometheusPythonPyTorchSQLTerraform
Real Estate • Travel • PropTech
Lead development of cutting-edge LLM and Agentic AI systems for Airbnb customer support, from research and prototyping to production deployment. Collaborate with product, design, and engineering to build scalable ML infrastructure, improve model performance, and implement evaluation, fine-tuning, and guardrails for chat and voice assistants.
Top Skills:
Agentic AiAutogenDistillationGrpoLanggraphLlmLlm Evaluation FrameworksMl InfrastructureMlopsModel ServingMulti-Agent OrchestrationPretrainingPrompt EngineeringRagReactRlhfSft
Artificial Intelligence • Professional Services • Consulting • Automation
Design and deploy production recommendation and retrieval engines, build data orchestration pipelines for multiple modalities, implement agentic workflow frameworks, and prototype, optimize, and scale models with monitoring, governance, and low-latency SLAs.
Top Skills:
AirflowSparkEmbeddingsKafkaKubernetesLangchainPythonPyTorchRagRetrieval LibrariesServerlessTensorFlowTransformer-Based Models
Artificial Intelligence • Cybersecurity
Own the production lifecycle for machine learning and LLM-backed systems, including training and post-training pipelines, inference serving, model releases, monitoring, retrieval infrastructure, and cost and latency optimization. Partner with AI researchers and engineering teams to move prototypes into reliable production capabilities across customer tenants. Develop ETL and GraphQL features, enforce tenant isolation, and support deployments in regulated or customer-controlled environments.
Top Skills:
Amazon BedrockAWSAzureDatadogDockerETLGCPGrafanaGraphQLKubernetesNeo4JPostgresPrometheusPythonSQLTgiVllm
Automotive • Big Data • Insurance • Software • Transportation
Design, develop, and deploy machine learning models and pipelines to optimize operations. Lead full ML project lifecycle, ensure model evaluation/monitoring, collaborate cross-functionally, mentor junior engineers, and drive continuous improvement in ML applications and processes.
Top Skills:
AirflowAws EcrAws S3Aws SagemakerCi/CdDvcNumpyPandasPythonRestful ApisScikit-LearnSQL
Artificial Intelligence • Fintech • Information Technology • Machine Learning • Financial Services
Design and deploy machine learning models for financial document processing, optimize large language models, collaborate with cross-functional teams, and guide junior peers in a fast-paced environment.
Top Skills:
DockerKubernetesPythonPyTorchTensorFlow
Consumer Web • Digital Media • Marketing Tech • Analytics
Designs, evaluates, and improves ranking and recommendation systems through experimentation, feature engineering, and production machine learning. Builds reliable data and feature pipelines, incorporates multimodal signals, manages uncertainty, calibration, and model drift, and monitors system performance. Partners cross-functionally with Engineering, Product, Business, and Analytics to deliver scalable, latency-sensitive solutions with measurable business impact.
Top Skills:
Ci/CdDistributed SystemsFeature StoresPythonSQL
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Software • Generative AI
The Sr. Machine Learning Engineer will develop and deploy ML solutions for healthcare, manage data pipelines, and work with large datasets to enhance healthcare delivery.
Top Skills:
AWSC++KubernetesPythonPyTorchScikit-LearnSparkTensorFlow
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Artificial Intelligence • Consumer Web • Digital Media • Machine Learning • Software
Develop and optimize large-scale advertising ranking models for CTR and CVR prediction, calibration, user representations, and sequence modeling. Own machine learning systems end to end, including data pipelines, feature engineering, training, evaluation, deployment, and maintenance. Collaborate with platform, product, data science, and engineering teams to run experiments and improve advertiser outcomes, revenue, and user relevance. Apply deep learning and recommendation techniques within production latency, reliability, and cost constraints.
Top Skills:
PythonPyTorchTensorFlow
AdTech • Fintech • Marketing Tech
Design, build, deploy, and operate production machine learning systems for personalization, recommendations, optimization, decisioning, experimentation, and measurement. Build reusable ML platforms, pipelines, and tooling; manage model monitoring, quality, drift detection, retraining, governance, and reliability. Partner with data scientists, engineers, and product teams to productionize models, make architectural decisions, and provide technical leadership through mentorship, reviews, and engineering standards.
Top Skills:
Amazon EksSparkAWSDatabricksFeature StoresKubernetesMlflowModel-Serving PlatformsPythonTerraform
Artificial Intelligence • Gaming • Software • Generative AI
Senior AI/ML Engineer responsible for translating federal supply chain and logistics problems into analytical solutions, developing forecasting, classification, and NLP models, and deploying them into production. The role owns model development, data pipelines, monitoring, retraining, drift detection, validation, documentation, and stakeholder communication. Work includes operating within government-controlled environments and meeting federal bias evaluation and documentation requirements.
Top Skills:
Amazon SagemakerAws GovcloudAzure GovernmentGpu-Accelerated TrainingKerasKubeflowMlflowPyTorchTensorFlow
Cloud • Software
Design, optimize, and deploy deep learning models for customer behavior prediction. Perform feature engineering, implement research-informed architectures, scale training pipelines, collaborate with client teams, document findings, and uphold engineering best practices.
Top Skills:
Distributed TrainingGoogle Cloud PlatformMlopsPythonPyTorchTensorFlow
Artificial Intelligence • Industrial
Build and deploy production machine-learning systems within client environments, covering data discovery, modeling, validation, deployment, scheduling, monitoring, and retraining. Develop forecasting and detection models for imperfect time-series data, establish honest evaluation and KPI frameworks, and ship usable APIs or interfaces. Work directly with client IT and data teams through discovery, demos, and feedback cycles. The role also involves data engineering, software development, basic cloud and DevOps work, and advising clients when ML is not appropriate.
Top Skills:
AWSAzureDockerDuckdbFastapiGradient BoostingLlmsPandasPolarsPostgisPostgresPythonReactScikit-LearnSQLStatsmodelsTimescaledbTypescript
Artificial Intelligence • Software
Build and operate scalable ML infrastructure and platform capabilities across cloud and on-premises environments. Develop Python-based developer tooling, services, and production systems supporting experimentation, distributed training, model deployment, and operations. Lead complex initiatives from architecture through rollout, improve reliability and developer productivity, and partner with ML and infrastructure teams. The role emphasizes Kubernetes, AWS, distributed compute, GPU-intensive workloads, observability, and production excellence.
Top Skills:
AirflowAWSCloud ComputingComputer VisionDistributed SystemsKubeflowKubernetesMachine Learning InfrastructurePythonRayRoboticsSpark
Biotech • Agriculture
Lead the design, development, deployment, and monitoring of production AI/ML applications for dairy-farm intelligence. Responsibilities include data wrangling, feature engineering, model selection, experimentation, MLOps, agent evaluation, observability, fine-tuning, AI governance, and integration with cloud infrastructure. The role also drives AI adoption, establishes engineering practices, collaborates cross-functionally, and may build React and TypeScript user experiences.
Top Skills:
AWSDatabricksMcpReactTypescript
Healthtech • Software
Designs, builds, and operates scalable machine learning platforms, pipelines, and infrastructure for model training, deployment, serving, and monitoring. Develops MLOps tooling, CI/CD workflows, model registries, feature stores, and experiment tracking. Ensures production reliability, observability, security, compliance, and cost efficiency across ML systems. Collaborates with product and engineering teams, integrates platforms with existing systems and data sources, and mentors engineers on reusable ML platform practices.
Top Skills:
AWSAzureAzure Machine LearningCi/CdDatabricksDockerExperiment TrackingFeature StoresGCPJavaKubeflowKubernetesMlflowModel RegistriesPythonRay
Agriculture
Lead the design, development, deployment, and monitoring of production AI/ML applications for dairy farm intelligence. Responsibilities include data wrangling, feature engineering, model selection, experimentation, MLOps, agent evaluation, observability, fine-tuning, AI governance, and adoption enablement. The role also involves collaborating on React and TypeScript user experiences, translating ambiguous requirements into scalable systems, and communicating technical tradeoffs across product, engineering, and dairy science teams.
Top Skills:
Agent Orchestration FrameworksAWSDatabricksDeep Learning FrameworksMcpMlopsReactTypescript
Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Generative AI
Develop and maintain open-source Python frameworks and ML developer tools, architect scalable libraries, grow early-stage projects, build real-time streaming capabilities, integrate modern frontend technologies, and collaborate with open-source contributors and the broader developer community.
Top Skills:
JavaScriptPythonReactSvelteTypescriptWebrtcWebsockets
Biotech
Build and scale machine learning infrastructure, training systems, inference workflows, and agentic AI tools for protein design research. Improve distributed training performance, GPU utilization, reliability, and scalability while partnering with AI scientists and protein engineers. Translate research prototypes into reusable production systems, contribute to technical direction, evaluate emerging ML tools, and support scientific communication.
Top Skills:
Ci/CdCudaCustom KernelsDistributed TrainingDockerGpu ProfilingKubernetesMlopsModel MonitoringModel RegistriesModel VersioningRayTriton
AdTech • Big Data • Marketing Tech • Mobile • Grocery & Supermarkets • Consumer Packaged Goods (CPG)
Build and improve machine learning systems for ad ranking, relevance, and optimization. Responsibilities include developing active learning and LLM-assisted labeling workflows, owning experimentation and evaluation, and deploying low-latency inference services. The role partners with product, data, and platform teams to improve advertiser performance and user engagement while ensuring production reliability, scalability, latency, and data quality. Candidates will use Python, Go, AWS, distributed systems, and AI-assisted development tools.
Top Skills:
AWSAws BedrockChatgptClaudeGithub CopilotGoLangchainLlmsPythonVector Databases
Software
Own the full lifecycle of clinical text extraction models, from problem framing and annotation strategy through production deployment, evaluation, monitoring, and retraining. Build clinical NLP pipelines, establish defensible quality and error-analysis systems, and apply LLMs using rigorous evaluation, observability, and cost controls. Partner with clinical data and backend engineering teams while addressing PHI protection, de-identification, auditability, and access controls. As the first ML hire, set technical direction and potentially build and lead the future ML team.
Top Skills:
Clinical NlpCptFhirFine-TuningIcd-10Information ExtractionLarge Language ModelsLoincModel EvaluationModel MonitoringNamed Entity RecognitionOcrPrompt EngineeringPythonRetrievalRxnormSequence LabelingSnomed CtStructured Output
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Boston, MA Companies Hiring Remote Senior Machine Learning Engineers
See AllPopular Job Searches
All Filters
Total selected ()
No Results
No Results





_1.png)



.png)

.png)














_1.png)







