Design, optimize, and deploy ML models for resource-constrained edge devices. Apply compression, quantization, and hardware-aware tuning; build cross-platform inference runtimes; enable on-device updates, telemetry, and secure execution; collaborate with hardware and product teams and maintain documentation and benchmarks for production edge AI.
Edge AI Engineer – Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: Edge AI Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$155,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary:
We are looking for an Edge AI Engineer to design, optimize, and deploy machine learning models that run efficiently on resource-constrained edge devices, including mobile platforms, embedded systems, and specialized accelerators. The role requires deep expertise in model compression, quantization, and hardware-aware optimization, along with strong systems engineering skills to ship reliable AI capabilities outside the data center. The ideal candidate has shipped edge AI in production environments where compute, memory, energy, and connectivity constraints fundamentally shape the engineering trade-offs.
Key Responsibilities
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3899. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: Edge AI Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$155,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary:
We are looking for an Edge AI Engineer to design, optimize, and deploy machine learning models that run efficiently on resource-constrained edge devices, including mobile platforms, embedded systems, and specialized accelerators. The role requires deep expertise in model compression, quantization, and hardware-aware optimization, along with strong systems engineering skills to ship reliable AI capabilities outside the data center. The ideal candidate has shipped edge AI in production environments where compute, memory, energy, and connectivity constraints fundamentally shape the engineering trade-offs.
Key Responsibilities
- Design and implement edge AI solutions optimized for diverse hardware including mobile SoCs, NPUs, and embedded accelerators.
- Apply quantization, pruning, distillation, and architectural optimization to fit models within edge constraints.
- Tune model performance for latency, energy efficiency, and memory footprint on target hardware.
- Build cross-platform inference runtimes leveraging frameworks such as TensorFlow Lite, ONNX Runtime, and Core ML.
- Optimize models for specific accelerator backends including DSPs, NPUs, and mobile GPUs.
- Implement on-device model update, versioning, and rollback workflows that allow safe staged rollouts to large device populations and rapid recovery if a model release behaves unexpectedly in the field.
- Design hybrid edge-cloud architectures that gracefully degrade based on connectivity and device capability.
- Build telemetry pipelines that respect privacy while enabling continuous improvement.
- Collaborate with hardware, firmware, and product teams to align AI capabilities with device constraints.
- Implement secure execution paths, model protection, and integrity verification on edge devices.
- Develop benchmarking suites that characterize accuracy, latency, and energy trade-offs across devices.
- Drive responsible AI considerations including on-device privacy and bias evaluation.
- Maintain comprehensive, current technical documentation — including architecture diagrams, design decisions, configuration references, runbooks, and operational procedures — so that the system remains supportable, auditable, and easy to onboard new engineers onto over time.
- Stay current with edge AI hardware and software developments, regularly review release notes and community discussions, and translate noteworthy advances into concrete recommendations and adoption proposals for the team.
- Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
- Six or more years of experience in ML engineering, with significant work on edge or mobile AI.
- Strong proficiency in Python and C++.
- Hands-on experience with model compression, quantization, and pruning techniques.
- Experience with at least one major edge inference framework.
- Solid understanding of mobile and embedded hardware architectures.
- Experience deploying ML models to production on mobile or embedded platforms.
- Strong performance engineering and profiling skills.
- Familiarity with on-device privacy and security considerations.
- Strong communication and cross-functional collaboration skills.
- Experience with custom NPU or DSP toolchains.
- Familiarity with federated learning or on-device personalization.
- Exposure to safety-critical or industrial edge deployments.
- Open-source contributions to edge AI frameworks.
- Experience optimizing LLMs for on-device inference.
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3899. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Similar Jobs
Aerospace • Artificial Intelligence • Cloud • Software
Build and scale high-performance platform software for AI-driven radar systems across edge and cloud. Implement C++, Go, and Python components, distributed services, APIs, and deployment tooling; improve performance, reliability, and observability; collaborate with AI and embedded teams and mentor junior engineers.
Top Skills:
Ai/Ml PlatformsAPIsC++Ci/CdDistributed SystemsDockerEdge ComputingEmbedded SystemsGoInfrastructure-As-CodeKubernetesLinuxNetworkingPython
Consumer Web • Internet of Things • Security
Lead on-device ML inference deployment and performance optimization for outdoor camera/doorbell workloads. Optimize latency, throughput, memory, power and stability across CPU/DSP/NPU/GPU. Perform kernel/operator-level tuning, runtime integration in C/C++, quantization validation, and build profiling/benchmarking tooling. Drive cross-functional work with ML, firmware, and hardware teams and provide staff-level technical leadership and mentoring.
Top Skills:
CC++Cmsis-NnDeimDfineDmaDspEmbedded LinuxFlame GraphsFp16GpuHardware CountersInt8IspNeonNpuOnednnOnnx RuntimePerfQnnpackRt-DetrRtosSimdSveTensorrtTfliteTracingVendor RuntimesVideo DecodeVideo EncodeXnnpackYoloZero-Copy Buffers
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
As a Staff Machine Learning Engineer, you will design AI products on Edge devices, optimize ML model performance, and work with large-scale datasets, collaborating across teams to deliver impactful solutions.
Top Skills:
C++PythonRayRustSpark
What you need to know about the Boston Tech Scene
Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.
Key Facts About Boston Tech
- Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
- Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
- Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
- Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories



