/#jobs
#Platform#About Us#Employers#AI Jobs
Get Started
Melbourne
3 Aug 2026
Founding ML Engineer logo

Founding ML Engineer

Base Compute

Melbourne
3 Aug 2026
Apply

Founding ML Engineer

Base Compute is an AI inference lab focused on bringing AGI on-device. We are looking for a Founding ML Engineer to work at the frontier of on-device AI, building and scaling our custom inference engine and optimizing for diverse silicon platforms. The role requires 3+ years of experience in ML engineering or systems programming, expertise in GPU programming, and a strong sense of ownership.

Core AIHybridFull-timeMid LevelRustC++

Founding ML Engineer

Base Compute is an AI inference lab focused on bringing AGI on-device. We are looking for a Founding ML Engineer to work at the frontier of on-device AI, building and scaling our custom inference engine and optimizing for diverse silicon platforms. The role requires 3+ years of experience in ML engineering or systems programming, expertise in GPU programming, and a strong sense of ownership.

Apply
Core AIHybridFull-timeMid LevelRust

Salary

Not specified

Work Location

Melbourne, Victoria, Australia, AU

Work Model

Hybrid: in-person from the office most days

Experience Required

3 years

Employment Type

Full-time

Experience Level

3+ years of experience in ML engineering or systems programming

Core Qualifications

Technical (Must-have)
RustC++CUDAROCmMetalTritonGPU programmingLLM architecturesquantizationspeculative decodingKV-cache managementcontinuous batchingperformance profiling
Soft Skills
ownershipautonomycommunicationhonest feedbackdocumentation

Preferred Qualifications

Technical (Nice-to-have)
ML compilerstorch.compilecustom operatorslow-precision inferenceINT8FP8FP4Edge LLMOps

Key Responsibilities

  • •Building and scaling our custom inference engine, handling weight loading, KV-cache management, and request scheduling
  • •Writing and tuning custom kernels and leveraging hardware-specific instructions for diverse architectures (Apple Silicon, NVIDIA, AMD, Snapdragon, edge platforms)
  • •Developing robust, low-latency serving runtimes in C++ for model routing, continuous batching, and novel decoding strategies
  • •Identifying and eliminating bottlenecks across the stack, from memory bandwidth to kernel interleaving
AIMachine LearningInferenceOn-device AISystems EngineeringGPU ProgrammingC++RustFounding EngineerHybrid
/#jobs

Your gateway to a successful career. Show your growth. Be ready for your next step. Capture and seize the best opportunities.

  • Data
  • FAQ
  • Articles
  • AI Jobs
  • Platform
  • Employers
  • About Us
  • Legal
© 2026/#jobsAll rights reserved.

For queries/support, email jobs.support@slashhash.ai