inhousefyi
← Back to listings

AI Compiler Engineer

EnCharge AIRemote · Posted 2 months ago
Full-timeRemoteEst. 222,500 USD
Apply now

Description

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute efficiency and density compared to today’s best-in-class solutions. The high-performance architecture is coupled with seamless software integration and will enable the immense potential of AI to be accessible in power, energy, and space constrained applications. EnCharge AI launched in 2022 and is led by veteran technologists with backgrounds in semiconductor design and AI systems.

About the Role

EnCharge AI is seeking a highly skilled and experienced AI Compiler Engineer to spearhead the efforts in developing and optimizing graph compilers tailored to cutting-edge AI and ML workloads. You will collaborate with hardware architects, and AI researchers to enhance performance, optimize computation graphs, and enable efficient model deployment on EnCharge’s Inference Accelerators.

Responsibilities

Architect, design, and implement optimizations for AI model execution on graph compilers to improve performance, reduce latency, and maximize hardware utilization.

  • Work closely with ML researchers, hardware engineers, and software developers to design and deploy AI models, understanding and addressing hardware-specific challenges.
  • Work on performance optimizations for neural network models, such as layer fusion, operator fusion, and graph-level transformations.
  • Develop compiler optimizations and passes that convert high-level AI models (e.g., from TensorFlow, PyTorch) into intermediate representations (IR).
  • Implement parsing, semantic analysis, and IR generation for deep learning frameworks.
  • Research and integrate the latest advancements in compiler design, ML model optimizations, and hardware acceleration into graph compilers.
  • Provide leadership, mentorship, and technical guidance to a team of engineers focused on graph compiler optimizations.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or related field (Ph.D. preferred).
  • 3+ years in compiler development, with a strong focus on AI or ML graph compilers.
  • Proficiency in AI graph compiler frameworks (e.g., MLIR, Torch-FX)
  • Solid background in hardware architectures (e.g., GPUs, TPUs, ASICs) and optimization techniques such as fusion, quantization, and tiling.
  • Familiarity with neural networks operators and code generation.
  • Strong understanding of intermediate representations, code parsing, and semantic analysis in compiler design.
  • Proficiency in C++, Python, or other programming languages commonly used in compiler development.
  • Open-source contributions to AI software frameworks and libraries is a plus
  • Demonstrated experience leading and mentoring engineering teams with successful project delivery.

EnCharge AI is an equal employment opportunity employer in the United States.

The salary range for this position is $190,000 to $255,000 USD per year. (Per Year: $195,000 to $265,000 CAD | €110,000 to €160,000 EUR | 1,206,458 to 1,754,848 NOK)
Actual compensation offered will be determined based on job-related knowledge, skills, and experience.

Similar jobs

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-time
EnCharge AIRemote

Est. 210,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 140,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 225,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 210,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 124,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 225,000 USD

EnCharge AI is building a new generation of AI compute systems designed to deliver dramatically higher energy efficiency for AI inference. Our technology combines innovative compute architectures with a full-stack softwa…

Full-timeRemote
EnCharge AIRemote

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIRemote

Est. 225,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
DensityAIMountain View, California, United States

Est. 257,500 USD

About the role Own MLIR dialect design and lowering passes for our AI accelerator — defining the high-level tensor IR, async / streaming semantics, and sharded-tensor types that bridge ML frameworks to silicon. Work with…

Full-time
Tenstorrent University JobsToronto, Ontario, Canada

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentToronto, Ontario, Canada

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
DensityAIMountain View, California, United States

Est. 267,500 USD

About the role Own LLVM backend development for our custom accelerator ISA — instruction selection, register allocation, scheduling, and linker support across scalar, vector, and floating-point pipelines. Work with chip-…

Full-time
EnCharge AIBangalore, India

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-time
GraphcoreGdańsk, Pomeranian Voivodeship, Poland

Est. 306,300 PLN

Senior: PLN 260,400 - 352,200Staff: PLN 350,700 - 474,400Subject to alignment to the responsibilities and duties of the role - we currently have multiple positions available at both Senior and Staff levelAbout Graphcore…

Full-time
EnCharge AIRemote

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
EnCharge AIUnited States

Est. 175,000 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
TenstorrentAustin, Texas, United States

Est. 300,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
EnCharge AIGermany

Est. 80,000 EUR

Research Engineer, Applied AI Location: Germany About EnCharge AI: EnCharge AI is building the next generation AI platform. Our novel in-memory-computing architecture delivers a 10x step-function improvement in compute e…

Full-time
EnCharge AIRemote

Est. 137,500 USD

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute…

Full-timeRemote
GraphcoreLondon, United Kingdom

Est. 80,000 GBP

About the job Validate the ML stack that turns accelerator hardware into trusted AI performance. This role sits where modern ML models meet Graphcore’s software and hardware stack. You will test, benchmark and validate c…

Full-time
GraphcoreCambridge, United Kingdom

Est. 80,000 GBP

About the job Validate the ML stack that turns accelerator hardware into trusted AI performance. This role sits where modern ML models meet Graphcore’s software and hardware stack. You will test, benchmark and validate c…

Full-time
GraphcoreLondon, United Kingdom

Est. 80,000 GBP

About the job Explore how a next-generation AI architecture behaves from rack scale to data centre scale. As a Senior Systems Engineer, you will investigate a novel AI computing architecture under demanding real-world wo…

Full-time
FigureSan Jose, California, United States

Est. 227,500 USD

Figure is an AI robotics company developing autonomous general-purpose humanoid robots. The goal of the company is to ship humanoid robots with human level intelligence. Its robots are engineered to perform a variety of…

Full-time
GraphcoreLondon, United Kingdom

Est. 80,000 GBP

About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to…

Full-time
GraphcoreBristol, United Kingdom

Est. 80,000 GBP

About the job Validate the ML stack that turns accelerator hardware into trusted AI performance. This role sits where modern ML models meet Graphcore’s software and hardware stack. You will test, benchmark and validate c…

Full-time
Tenstorrent University JobsSanta Clara, California, United States

Est. 104,000 USD

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify inn…

Full-time
GraphcoreCambridge, England, United Kingdom

Est. 80,000 GBP

About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to…

Full-time