
Founding ML Engineer
Base Compute
Founding ML Engineer
Base Compute is an AI inference lab focused on bringing AGI on-device. We are looking for a Founding ML Engineer to work at the frontier of on-device AI, building and scaling our custom inference engine and optimizing for diverse silicon platforms. The role requires 3+ years of experience in ML engineering or systems programming, expertise in GPU programming, and a strong sense of ownership.
Founding ML Engineer
Base Compute is an AI inference lab focused on bringing AGI on-device. We are looking for a Founding ML Engineer to work at the frontier of on-device AI, building and scaling our custom inference engine and optimizing for diverse silicon platforms. The role requires 3+ years of experience in ML engineering or systems programming, expertise in GPU programming, and a strong sense of ownership.
Salary
Core Qualifications
Technical (Must-have)
Soft Skills
Preferred Qualifications
Technical (Nice-to-have)
Key Responsibilities
- Building and scaling our custom inference engine, handling weight loading, KV-cache management, and request scheduling
- Writing and tuning custom kernels and leveraging hardware-specific instructions for diverse architectures (Apple Silicon, NVIDIA, AMD, Snapdragon, edge platforms)
- Developing robust, low-latency serving runtimes in C++ for model routing, continuous batching, and novel decoding strategies
- Identifying and eliminating bottlenecks across the stack, from memory bandwidth to kernel interleaving