Top AI & Machine Learning Jobs in Boston, MA

5 Days AgoSaved
In-Office
Boston, MA, USA
191K-239K Annually
Expert/Leader
191K-239K Annually
Expert/Leader
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Own DigitalOcean’s enterprise AI-native architecture, including agent orchestration, shared AI platform services, tool and capability layers, enterprise-system integrations, identity, authorization, governance, observability, evaluation, and cost controls. Build production prototypes and reference implementations, write code, establish technical standards, mentor distributed engineers, and re-engineer business workflows with functional teams. Partner across platform engineering, security, identity, data, and business technology to deliver governed, scalable agentic systems.
Top Skills: APIsArizeAutogenBoomiBraintrustCrewaiEu Ai ActEvent-Driven ArchitectureGoGraphragGreenhouseIso/Iec 42001JavaKubernetesLangfuseLanggraphLangsmithModel Context Protocol (Mcp)MulesoftN8NNetSuiteNist Ai RmfOpenai Agents SdkOpentelemetryOwasp Agentic AiPythonSalesforceSemantic KernelServerlessServicenowSpiffe/SpireTypescriptVector StoresWorkatoWorkday
Reposted 7 Days AgoSaved
In-Office
Boston, MA, USA
191K-239K Annually
Senior level
191K-239K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Lead architecture and low-level optimizations for LLM inference to maximize throughput and minimize latency. Profile and optimize GPU kernels, implement quantization and parallelization strategies (FP8/INT8/FP4), tune libraries (AITER, Triton, CUDA, ROCm), and mentor engineers while translating hardware limits into shippable platform features.
Top Skills: Ai/Ml InferenceAmd AiterAmd Mi355XBf16CudaDeepseekFlashattentionFp4Fp8Glm-5Gpu Kernel DevelopmentInt8MoeMulti-Node Gpu ClusteringOpenai TritonQwen3-235BRmsnormRocmTensorrtTflopsTriton Compiler
Reposted 8 Days AgoSaved
In-Office
Boston, MA, USA
139K-174K Annually
Senior level
139K-174K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Lead design and implementation of high-scale, multi-tenant AI inference data plane services. Optimize distributed LLM serving (prefill/decode disaggregation, KV-cache routing, autoscaling, flow control), contribute upstream to open-source inference projects, mentor engineers, and operate resilient, observable production services running on Kubernetes and specialized AI hardware.
Top Skills: GoGrpcKserveKubernetesLlm-DModular MaxNixlNvidia DynamoPythonRay ServeSglangTensorrtTensorrt-LlmTgiVllm
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account