Compiler Runtime Engineer
Oho Group · San Francisco Bay Area
Apply & track with Apply EdgeCompiler Runtime EngineerI’m working with an ambitious AI infrastructure company developing a compiler and runtime platform that enables machine-learning workloads to execute efficiently across different hardware architectures.They are looking for a Runtime Engineer to build the systems layer responsible for turning optimized compiler output into reliable, high-performance execution. You will work closely with compiler and hardware engineers across runtime architecture, parallel execution, scheduling and performance analysis.You will:Design a portable runtime for AI workloadsBuild workload partitioning, parallel execution and kernel-scheduling capabilitiesPrototype new runtime techniques and evaluate them against real hardwareBenchmark compiled workloads and diagnose performance bottlenecksDevelop profiling tools that inform compiler and runtime improvementsHelp evolve the platform around real ML workloads and deployment requirementsThey are looking for:4+ years in compiler, runtime or low-level systems engineeringStrong modern C++Deep knowledge of concurrency and asynchronous executionUnderstanding of memory hierarchies and vector/scalar computeExperience with OS kernels, hypervisors or similarly low-level systemsExperience with CUDA, ROCm, GPU programming, HPC, large compute clusters, PyTorch, JAX or Triton would be particularly valuable.This is an opportunity to own a foundational part of a new AI compute platform, working directly between compiler output, hardware execution and distributed systems.