inhousefyi
← Back to listings

Member of Technical Staff, Hardware, Compiler Engineer

River AI Inc.Austin, Texas, United States · Posted 3 months ago
Full-timeEst. 310,000 USD
Apply now

Description

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructure, next-generation UIs, and frontier deep learning research.

Who we are

We are scientists, engineers, and builders from the industry's top tech companies and AI labs. We bring a proven track record of scaling consumer systems for hundreds of millions of users and architecting the pre-training infrastructure behind today's frontier models.

About the Role

We are looking for exceptional AI compiler engineers to build the software bridge between the newest AI models and our high-performance custom silicon. You will create and build the compiler stack from PyTorch graphs all the way to optimized custom ISA assembly code. You will take ownership of kernel algorithms, intermediate representations, and even modify the ISA as necessary to achieve a flexible and high performance compiler stack. You will be collaborating both up and down the stack with AI researchers and modelers, as well as with performance engineers and silicon architects.

What You’ll Do

  • Graph Lowering & Optimization: Design and implement compiler passes to lower PyTorch models into custom hardware, leveraging MLIR dialects and LLVM frameworks.
  • Custom Backend Development: Develop and maintain the backend toolchain for our custom silicon, including instruction scheduling, register allocation, and hardware-specific code generation.
  • Memory & Loop Transformations: Design sophisticated tiling and fusion strategies to maximize bandwidth utilization and minimize on-chip memory movement.
  • Kernel Integration: Collaborate with software and hardware teams to integrate high-performance kernels (Triton/CUDA-like) into the automated compiler flow.
  • Performance Profiling: Identify "compilation gaps" where the compiler fails to achieve peak hardware performance, and collaborate with the performance team for targeted optimizations to close those gaps.
  • HW/SW Co-Design: Partner with the RTL and Architecture teams to change the custom ISA definitions.

Skills and Qualifications

Minimum Qualifications:

  • Bachelor’s degree in Electrical Engineering or Computer Engineering, and 5+ years practical industry experience working with advanced process nodes (7nm or below).
  • Deep hands-on experience with MLIR or XLA for deep learning workloads.
  • Expert-level understanding of PyTorch internals and how they interface with external backends.
  • Proficiency in modern C/C++ for building robust, scalable, and high-performance compiler infrastructure.
  • Advanced knowledge in Computer Architecture, especially the Programming Model, of at least one style of chip, including SoCs, CPUs, GPUs, or AI accelerators
  • A highly collaborative mindset to push boundaries and co-design effectively with other engineers.

Preferred Qualifications: (We encourage you to apply even if you don't meet all of these)

  • Hands-on experience in post-Silicon firmware and model update patches
  • Experience defining and implementing custom dialects, lowering passes, and graph rewrites in an LLVM-based ecosystem.
  • Knowledge of the tradeoffs between static and runtime environments, including JITs and ABIs

Logistics & Benefits

  • Location: This role is based in Austin, Texas or Palo Alto, California.
  • Compensation: Depending on background, skills, and experience, the expected annual salary range for this position is $200,000 - $420,000 USD.
  • Visa Sponsorship: We sponsor visas. We can't guarantee success for every candidate or role, but if you're the right fit, we're committed to working through the visa process.
  • Benefits: River AI offers generous health, dental, and vision benefits, unlimited PTO, and relocation support as needed.

Similar jobs

River AI Inc.Austin, Texas, United States

Est. 310,000 USD

At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructure,…

Full-time
River AI Inc.Austin, Texas, United States

Est. 310,000 USD

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructur…

Full-time
River AI Inc.Palo Alto, California, United States

Est. 310,000 USD

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructur…

Full-time
River AI Inc.Palo Alto, California, United States

Est. 310,000 USD

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu…

Full-time
River AI Inc.Austin, Texas, United States

Est. 310,000 USD

At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure,…

Full-time
River AI Inc.Austin, Texas, United States

Est. 310,000 USD

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu…

Full-time
River AI Inc.Palo Alto, California, United States

Est. 310,000 USD

At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure,…

Full-time
River AI Inc.Palo Alto, California, United States

Est. 310,000 USD

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu…

Full-time
EnCharge AIRemote

Est. 222,500 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
DensityAIMountain View, California, United States

Est. 257,500 USD

About the role Own MLIR dialect design and lowering passes for our AI accelerator — defining the high-level tensor IR, async / streaming semantics, and sharded-tensor types that bridge ML frameworks to silicon. Work with…

Full-time

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
DensityAIMountain View, California, United States

Est. 267,500 USD

About the role Own LLVM backend development for our custom accelerator ISA — instruction selection, register allocation, scheduling, and linker support across scalar, vector, and floating-point pipelines. Work with chip-…

Full-time
Tenstorrent University JobsAustin, Texas, United States

Est. 100,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentAustin, Texas, United States

Est. 250,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentToronto, Ontario, Canada

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
River AI Inc.Palo Alto, California, United States

Est. 250,000 USD

At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure,…

Full-time
Tenstorrent University JobsSanta Clara, California, United States

Est. 104,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
GraphcoreBristol, England, United Kingdom

Est. 80,000 GBP

About the job Build the framework support that helps AI developers unlock Graphcore hardware. You will help Graphcore accelerators work seamlessly with state-of-the-art ML frameworks, including Triton and PyTorch. Report…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentRemote

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-timeRemote
TenstorrentAustin, Texas, United States

Est. 250,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
DensityAIMountain View, California, United States

Est. 285,000 USD

About the role You will write, evaluate, and profile specialized compute kernels that run on a custom AI accelerator. This is the critical interface between high-level ML workloads and silicon — your code directly determ…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
Samsung SemiconductorSan Jose, California, United States

Est. 208,000 USD

Please Note: To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period. Advancing the World’s Technology Toget…

Full-time
TenstorrentRemote

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-timeRemote
PyTorch Engineer4 months ago
GraphcoreBristol, United Kingdom

Est. 80,000 GBP

About the job Help make Graphcore hardware feel native inside the ML frameworks engineers use every day. As a Software Engineer in our PyTorch team, you will help build the software that connects Graphcore accelerators w…

Full-time
TenstorrentRemote

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-timeRemote