d-Matrix Logo

d-Matrix

Principal Architect, Performance Analysis and Modeling

Reposted One Month Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in Santa Clara, CA
195K-285K Annually
Expert/Leader
Remote or Hybrid
Hiring Remotely in Santa Clara, CA
195K-285K Annually
Expert/Leader
The Principal Architect will analyze ML workloads, contribute to hardware/software features for inference accelerators, and develop analytical performance models.
The summary above was generated by AI

At d-Matrix, we are focused on unleashing the potential of generative AI to power the transformation of technology. We are at the forefront of software and hardware innovation, pushing the boundaries of what is possible. Our culture is one of respect and collaboration.

We value humility and believe in direct communication. Our team is inclusive, and our differing perspectives allow for better solutions. We are seeking individuals passionate about tackling challenges and are driven by execution.  Ready to come find your playground? Together, we can help shape the endless possibilities of AI. 

Working onsite at our Santa Clara, CA headquarters 3 days per week hybrid. Will consider remote in the United States.

The role: Principal Architect, Performance Analysis and Modeling

d-Matrix is seeking outstanding computer architects to help accelerate AI application performance at the intersection of both hardware and software, with particular focus on emerging hardware technologies (such as DIMC, D2D, 3D-DRAM, etc.) and emerging workloads (such as generative inference, etc.). Our acceleration philosophy cuts through the system, ranging from efficient tensor cores, storage, and data movements along with co-design of dataflow and collective communication techniques.

What you will do:

  • As a member of the architecture team, you will analyze the latest ML workloads (multi-modal LLMs, CoT reasoning models, video/audio generation)

  • You will contribute hardware and software features that power the next generation of inference accelerators in datacenters

  • This role requires keeping up with the latest research in ML architecture and algorithms and collaborating with different partner teams, including Product, Hardware Design, Compiler, Inference Server, and Kernels

  • Your day-to-day work will include (1) analyzing the properties of emerging machine learning algorithms and workloads and identifying functional, performance implications (2) Creating analytical models to project performance on current and future generations of d-matrix hardware (3) proposing new HW/SW features to enable or accelerate these algorithms

What you will bring:

Minimum:

  • BSEE with 10+ years of industry experience or MSEE preferred with 8+ years of industry experience

  • Solid grasp through academic or industry experience in multiple of the relevant areas, computer architecture, hardware-software codesign, performance modeling, ML fundamentals (particularly DNNs)

  • Programming fluency in C/C++ or Python

  • Experience with developing analytical performance models, architecture simulators for performance analysis

  • Research background with publication record in top-tier architecture, or machine learning venues is a huge plus (such as ISCA, MICRO, ASPLOS, HPCA, DAC, MLSys etc.)

  • Self-motivated team player with strong sense of collaboration and initiative

Equal Opportunity Employment Policy

d-Matrix is proud to be an equal opportunity workplace and affirmative action employer. We’re committed to fostering an inclusive environment where everyone feels welcomed and empowered to do their best work. We hire the best talent for our teams, regardless of race, religion, color, age, disability, sex, gender identity, sexual orientation, ancestry, genetic information, marital status, national origin, political affiliation, or veteran status. Our focus is on hiring teammates with humble expertise, kindness, dedication and a willingness to embrace challenges and learn together every day.

d-Matrix does not accept resumes or candidate submissions from external agencies. We appreciate the interest and effort of recruitment firms, but we kindly request that individual interested in opportunities with d-Matrix apply directly through our official channels. This approach allows us to streamline our hiring processes and maintain a consistent and fair evaluation of al applicants. Thank you for your understanding and cooperation.

Similar Jobs

29 Days Ago
Remote or Hybrid
175K-285K Annually
Senior level
175K-285K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Analyze emerging AI workloads and build analytical performance models and architecture simulators for current and future AI inference accelerator hardware. Collaborate with hardware, compiler, inference server, kernel, and product teams to validate assumptions, identify HW/SW optimization opportunities, incorporate architecture research, and document methodologies. The role is a hands-on technical contributor focused on performance analysis across the hardware/software boundary.
Top Skills: 3D-DramAi Inference AcceleratorsAnalytical Performance ModelingArchitecture SimulatorsC/C++Computer ArchitectureD2DDeep Neural NetworksDimcHw/Sw Co-DesignMachine LearningMultimodal LlmsPython
12 Minutes Ago
In-Office or Remote
California, USA
126K-253K Annually
Mid level
126K-253K Annually
Mid level
Artificial Intelligence • Cloud • Information Technology • Consulting
Manages HPE channel partners to drive revenue, profitability, pipeline growth, and partner loyalty. Develops joint business plans, communicates HPE technology and sales strategies, coordinates marketing and sales activities, forecasts performance, supports partner compliance, and tailors solutions to customer needs. The role builds relationships with VARs, distributors, service providers, and other partners while achieving assigned sales quotas in the SMB segment.
12 Minutes Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
126K-248K Annually
Senior level
126K-248K Annually
Senior level
Big Data • Cloud • Software • Database
Build and maintain MongoDB’s distributed database infrastructure for cluster scalability. Design data partitioning, data movement, dynamic scaling, workload execution, and resilient distributed protocols. Solve complex performance and low-latency systems problems, contribute production-quality C++ code, participate in design and code reviews, lead new feature development, and mentor engineers.
Top Skills: C++Data PartitioningDistributed SystemsLow-Latency NetworkingMongoDBSystem Observability

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account