Performance Engineer, Hardware

External job listingat River AI

River AI seeks hardware performance engineers to model custom silicon, develop AI hardware simulators, optimize kernels, and analyze performance.

External source - not verified3 weeks agoOpen until: Jan 5, 2027

Salary

Not provided

Location

Palo Alto, United States

Employment type

Full time

Workplace

Not provided

Performance Engineer, Hardware

Palo Alto, United States

Job description

About the role

River AI is creating personal AI owned and shaped by each individual by rewriting the entire stack, including personal hardware for local inference, custom training infrastructure, next-generation UIs, and frontier deep learning research.

River AI seeks exceptional hardware performance engineers to architect, model, and correlate high-performance custom silicon. The role develops high-fidelity simulators to predict how an AI accelerator architecture and SoC system handle real-world AI models. It includes ownership of performance models, ISA and kernel optimization, and FPGA emulators for software development, in collaboration with compiler, IR, software, and RTL design teams.

Responsibilities

  • Design and implement high-performance and functional models of complex hardware using C++ and/or SystemC.
  • Conduct what-if studies of architectural changes, including cache sizes, branch predictors, pipeline depths, scatter/gather, and matmul shaping, and evaluate their impact on IPC, MFU, TTFT, and total execution time.
  • Analyze and profile AI kernels and software stacks to generate representative traces that stress-test hardware models.
  • Collaborate with compiler and kernel teams to optimize software mapping to hardware and ensure the architecture supports emerging algorithmic breakthroughs efficiently.
  • Validate performance models against RTL and pre-silicon emulators, keeping model accuracy within strict tolerance levels.
  • Identify performance bottlenecks.

Requirements

The source listing does not specify requirements.

Key facts

  • The role involves architecting, modeling, and correlating high-performance custom silicon.
  • The work includes simulator development, architectural studies, workload characterization, and performance correlation.
  • The role involves collaboration with compiler, IR, software, kernel, and RTL design teams.

Frequently asked questions

  • What does the role involve?

    It involves modeling high-performance custom silicon and developing simulators to predict how AI hardware handles real-world AI models.

  • What kinds of work will I do?

    Responsibilities include architectural studies, workload characterization, hardware/software co-design, and performance correlation.

  • Which teams will I collaborate with?

    The role collaborates with compiler, IR, software, kernel, and RTL design teams.

Similar jobs