Design, prototype, and productionize applied ML solutions across NLP, vision, recommendation, and structured data. Lead experimentation, distributed/mixed-precision training, model optimization, observability, data tooling, and responsible AI evaluations. Collaborate with platform and product teams, mentor engineers, and influence the AI roadmap.
AI Research Engineer – Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: AI Research Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$150,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary:
We are seeking an AI Research Engineer to bridge cutting-edge applied research and production engineering, designing and shipping advanced machine learning systems that solve high-impact business problems. The role blends scientific rigor with practical software engineering, requiring deep understanding of modern ML and deep learning techniques alongside the ability to build robust, scalable, and well-instrumented production pipelines. The ideal candidate stays current with the rapidly evolving AI research landscape, can critically evaluate new techniques for real-world applicability, and is comfortable operating across the full lifecycle from problem framing and experimentation to deployment and continuous improvement.
Key Responsibilities
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3899. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: AI Research Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$150,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary:
We are seeking an AI Research Engineer to bridge cutting-edge applied research and production engineering, designing and shipping advanced machine learning systems that solve high-impact business problems. The role blends scientific rigor with practical software engineering, requiring deep understanding of modern ML and deep learning techniques alongside the ability to build robust, scalable, and well-instrumented production pipelines. The ideal candidate stays current with the rapidly evolving AI research landscape, can critically evaluate new techniques for real-world applicability, and is comfortable operating across the full lifecycle from problem framing and experimentation to deployment and continuous improvement.
Key Responsibilities
- Design, prototype, and evaluate applied AI solutions across natural language, vision, recommendation, and structured data domains.
- Translate ambiguous business problems into well-scoped ML formulations with clear success metrics and evaluation strategies.
- Stay current with the latest research in deep learning, large language models, and adjacent areas, and assess applicability to internal use cases.
- Implement rigorous experimentation workflows including baselines, ablations, and statistically sound evaluation methodology.
- Build production-quality training and inference pipelines using modern ML frameworks and orchestration tools.
- Collaborate with ML platform engineers to ensure efficient use of compute, storage, and accelerator resources.
- Optimize models for accuracy, latency, throughput, and cost based on production requirements.
- Develop tooling for dataset construction, labeling, validation, and ongoing monitoring of data quality.
- Partner with product, design, and domain experts to ensure model behavior aligns with user needs and policy requirements.
- Implement safety, fairness, and reliability evaluations and incorporate findings into model selection decisions.
- Document research findings, design decisions, and operational characteristics clearly for both technical and non-technical audiences.
- Mentor engineers on applied ML methodology, evaluation rigor, and responsible deployment.
- Contribute to internal knowledge sharing, reading groups, and prototype-to-production playbooks.
- Influence the broader AI roadmap based on research insight, capability gaps, and emerging opportunities.
- Master’s or PhD in Computer Science, Machine Learning, Statistics, or a closely related field; or equivalent applied experience.
- Six or more years of combined research and applied ML engineering experience.
- Strong proficiency in Python and modern ML frameworks such as PyTorch or JAX.
- Hands-on experience training, fine-tuning, and evaluating deep learning models at non-trivial scale.
- Solid grounding in mathematics, statistics, and the theoretical foundations of modern ML.
- Experience taking ML models from research prototype to production with appropriate observability and safeguards.
- Familiarity with distributed training, mixed-precision training, and accelerator hardware.
- Strong written and verbal communication skills, including ability to explain complex methods clearly.
- Demonstrated ability to read, evaluate, and adapt techniques from current research literature.
- Track record of shipping impactful applied AI projects.
- Published research at top-tier AI/ML venues.
- Experience with large language model training, fine-tuning, or evaluation.
- Familiarity with retrieval-augmented generation, agentic systems, or multimodal architectures.
- Exposure to responsible AI, model evaluation, and alignment practices.
- Experience contributing to open-source ML projects.
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3899. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Similar Jobs
Artificial Intelligence • Information Technology • Software
12-16 week internship implementing and optimizing GPU/accelerator kernels and custom operators, profiling and improving LLM inference and runtime performance, building benchmarks, fixing regressions, and shipping code to production or open-source AI infra projects.
Top Skills:
Amd GpusAttentionAws TrainiumC++CudaDynamoFlashattentionGemmHipLmcacheMoeNeuron ProfilerNeuron SdkNsightNvidia GpusPagedattentionPythonPyTorchPytorch ProfilerPytorch/XlaQuantizationRocmRocm ProfilerSglangTritonVllm
Hardware • Manufacturing
Lead R&D on LLM training and inference optimization, train and evaluate large-scale models on Tenstorrent hardware, improve performance via techniques like speculative decoding, quantization, kernel fusion, flash attention and distributed training, diagnose system bottlenecks, and translate research into scalable production solutions through cross-functional collaboration.
Top Skills:
Ai AcceleratorsCompilersDistributed TrainingFlashattentionKernel FusionLlmsPythonPyTorchQuantizationRisc-VSpeculative Decoding
Artificial Intelligence • Information Technology • Software
Design and optimize high-performance kernels and custom operators for attention, MoE, GEMM, quantization and collective communication across NVIDIA, AMD and AWS Trainium. Improve LLM inference runtimes, develop distributed training/inference solutions at scale, use compilers and SDKs (Neuron, Torch Dynamo, PyTorch/XLA), contribute to open-source, and publish technical findings.
Top Skills:
Amd GpusAws NeuronAws TrainiumC++CudaFlashattentionLmcacheMoeNeuron CompilerNeuron ProfilerNeuron SdkNsightNvidia GpusPagedattentionPythonPytorch/XlaRlhfRocm ProfilerRocm/HipSglangTensorrt-LlmTorch DynamoTritonVllm
What you need to know about the Boston Tech Scene
Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.
Key Facts About Boston Tech
- Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
- Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
- Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
- Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

