🇩🇪 Berlin, Germany · 8h ago
Founding ML Researcher
Base Compute
LinkedInleadEnglish-friendly
About UsBase Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.The RoleWe’re looking for a Founding ML Researcher to work at the frontier of on-device AI. This role is for someone who identifies problems and potentials, designs and executes experiments and derives insights that translate into real-world performance.You’ll have significant ownership over our research agenda and direct influence on the technical bets the company makes.What You’ll Work OnInference research: Identifying and validating new approaches to on-device efficiency, including speculative decoding variants, novel quantization schemes and entirely new techniques yet to be discoveredModel routing research: Building the intelligence that decides how requests are served between on-device vs. frontier API modelsAutoresearch pipelines: Designing systems that can autonomously explore, hypothesize and evaluate research ideas that accelerate our R&D loopEvaluations and benchmarks: Developing rigorous evals that measure performance in the real world, outside of clean academic settingsWhat We’re Looking ForPhD in ML or equivalent industry research experienceDeep understanding of LLM architectures and the principles of AI inferenceExpertise in a relevant topic, such as speculative decoding, quantization theory, model distillation, reinforcement learningA track record of producing results that people build on: research papers, open-source projects or blog posts that prove out novel ideasGood communication: the ability to explain complex ideas simply, give honest feedback and document findings in a reproducible wayNice-to-haves:Familiarity with GPU and accelerator architectures and kernel optimization (CUDA, ROCm, Metal, Triton, etc.)Experience deploying models under on-device constraints (memory bandwidth, latency budgets, and thermal and power ceilings)What We OfferFounding team equity and strong base salaryDirect influence on technical direction: your ideas will shape the roadmapWork on genuinely hard problems that haven't been solved yetSmall team, fast iteration, low bureaucracyLocationThe team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.Sourced from LinkedIn. Relocantly aggregates public job postings; apply on the original site.