Clera home
·Dashboard

Jobs at Base Compute (Now Hiring) — 2 open

Base Compute logoBase Compute

Founding ML Engineer

Melbourne, Victoria, Australia · On-site

Mid level

About Us Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device. We’re building the …

Skills: ML Engineering, Systems Programming, Rust, C++, GPU Programming

Base Compute logoBase Compute

Founding ML Researcher

Melbourne, Victoria, Australia · On-site

Senior

About Us Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device. We’re building the …

Skills: ML Research, LLM Architectures, AI Inference, Speculative Decoding, Quantization Theory

Base Compute logo

Founding ML Engineer

Base Compute

Melbourne, Victoria, Australia • On-site

Apply
Mid level

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

  • Full-time
  • Founding Team Equity, Strong Base Salary
  • Posted 23d ago
  • ~40 hrs/week

Responsibilities

Develop and scale a custom inference engine and low-latency serving runtimes in C++. Optimize kernels across diverse hardware architectures to maximize on-device AI performance.

Requirements

Requires 3+ years of experience in ML engineering or systems programming with expertise in GPU programming and hardware optimization. Must have a strong understanding of modern LLM architectures and the ability to drive research into production infrastructure.

Full job description

About Us

Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.

We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.

We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.

The Role

We’re looking for a Founding ML Engineer to work at the frontier of on-device AI. This role is for someone who lives at the intersection of systems engineering and machine learning, turning state-of-the-art research into hyper-optimized, production-ready infrastructure.

You’ll have significant ownership over our entire inference stack and direct influence on the technical bets the company makes.

What You’ll Work On

  • Inference engine development: Building and scaling our custom inference engine, handling everything from weight loading and KV-cache management to efficient request scheduling

  • Cross-platform silicon optimization: Writing and tuning custom kernels and leveraging hardware-specific instructions to squeeze maximum performance out of diverse architectures, including Apple Silicon, NVIDIA, AMD, Snapdragon, and other edge platforms

  • Systems architecture: Developing robust, low-latency serving runtimes in C++ to manage model routing, continuous batching, and novel decoding strategies under strict thermal and memory constraints

  • Performance profiling: Identifying and eliminating bottlenecks across the entire stack, from memory bandwidth ceilings to kernel interleaving

What We’re Looking For

  • 3+ years of experience in ML engineering or systems programming (Rust, C/C++), with a strong track record of building performance-critical software

  • Expertise in GPU programming and hardware optimization across various platforms (CUDA, ROCm, Metal, Triton, or similar)

  • Solid understanding of modern LLM architectures, including parsing formats and implementing optimization techniques (quantization, speculative decoding, etc.)

  • A strong sense of ownership and autonomy: the ability to take ambiguous architectural challenges and drive them from research translation directly into production-ready infrastructure

  • Good communication: the ability to explain complex architectural decisions simply, give honest feedback and document systems cleanly

  • Nice-to-haves:

    • Familiarity with ML compilers (torch.compile, custom operators)

    • Experience with low-precision inference (INT8/FP8/FP4)

    • Knowledge of Edge LLMOps

What We Offer

  • Founding team equity and strong base salary

  • Direct influence on technical direction: your ideas will shape the roadmap

  • Work on genuinely hard problems that haven't been solved yet

  • Small team, fast iteration, low bureaucracy

Location

The team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.

Related keywords

AI InferenceAGIOn-device AISilicon OptimizationDistributed InferenceKV-cacheRequest SchedulingApple SiliconNVIDIAAMDSnapdragonContinuous BatchingMemory BandwidthKernel InterleavingTorch.compileINT8

About Base Compute

LinkedInVisit site

Run AGI on-device at enterprise scale.

Industry
Artificial Intelligence
Company size
2-10 employees
Founded
2026
Headquarters
Melbourne, Victoria
LinkedIn followers
186

Base Compute is an AI inference lab headquartered in Melbourne, Australia. Our mission is to run AGI on-device. We build the runtimes and infrastructure that make powerful AI run on device. Fast, private, and at near-zero marginal cost. We’re building a small, senior team of researchers and engineers focused on deep systems work across inference, optimisation, and AI infrastructure. We operate close to the metal and from first principles. We invest heavily in research, with a focus on automated AI research as a key driver of future breakthroughs in how intelligent systems are built and improved. We’re looking for exceptional researchers and engineers who care about performance, rigor, and building foundational systems that ship.

Offices: 287-293 Collins St, Melbourne, Victoria 3000, AU

AI inferenceAI ResearchLocal AIGPU KernelsEdge AIOn-device AISovereign AILLMand Private AI
View all jobs at Base Compute

About Base Compute

LinkedInVisit site

Run AGI on-device at enterprise scale.

Industry
Artificial Intelligence
Company size
2-10 employees
Founded
2026
Headquarters
Melbourne, Victoria
LinkedIn followers
186

Base Compute is an AI inference lab headquartered in Melbourne, Australia. Our mission is to run AGI on-device. We build the runtimes and infrastructure that make powerful AI run on device. Fast, private, and at near-zero marginal cost. We’re building a small, senior team of researchers and engineers focused on deep systems work across inference, optimisation, and AI infrastructure. We operate close to the metal and from first principles. We invest heavily in research, with a focus on automated AI research as a key driver of future breakthroughs in how intelligent systems are built and improved. We’re looking for exceptional researchers and engineers who care about performance, rigor, and building foundational systems that ship.

Offices: 287-293 Collins St, Melbourne, Victoria 3000, AU

AI inferenceAI ResearchLocal AIGPU KernelsEdge AIOn-device AISovereign AILLMand Private AI
View all jobs at Base Compute

Similar companies hiring

Navan (87)Litera (35)Socure (26)Edison Scientific (20)Abridge (15)Ataraxis AI (12)Fractile (6)Praxent (2)Pipl (2)Sage Care (2)Physical Superintelligence (1)NexTexAI (1)
Clera home

Your AI-talent agent. Connecting talents with dream jobs.

Earn $5,000

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

© 2026 Clera Labs, Inc.

PrivacyTermsBug Bounty