Base Career helps you apply smarter for this job.
Key skills for this role
We run a multi-tenant LLM inference platform: customers send requests to an OpenAI-compatible gateway, and we handle routing, tenancy, access control, usage metering and billing on top of our own GPU fleet.
You would own backend services and the Kubernetes platform they run on. This is not a role where infrastructure is someone else’s problem, you write the Go service, the Helm chart, the network policy and the runbook, and you are the one who verifies it in production.
We are moving deliberately toward a Kubernetes-native architecture : less imperative tooling and hand-run scripts, more declarative APIs, custom resources and controllers that reconcile state. If you have wanted to build operators and control planes rather than consume them, that is the direction of this role.
We run a multi-tenant LLM inference platform: customers send requests to an OpenAI-compatible gateway, and we handle routing, tenancy, access control, usage metering and billing on top of our own GPU fleet.
You would own backend services and the Kubernetes platform they run on. This is not a role where infrastructure is someone else’s problem, you write the Go service, the Helm chart, the network policy and the runbook, and you are the one who verifies it in production.
We are moving deliberately toward a Kubernetes-native architecture : less imperative tooling and hand-run scripts, more declarative APIs, custom resources and controllers that reconcile state. If you have wanted to build operators and control planes rather than consume them, that is the direction of this role.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
London, GBR
London, GBR
London, GBR
London, GBR
London, GBR
London, GBR
London, GBR
London, GBR
4+ years of experience
Strong Go. You have shipped and maintained production Go services, and you are comfortable with concurrency, context propagation and error handling that fails loudly instead of silently.
Real Kubernetes depth. Not just kubectl apply. You understand the control loop, know why a pod is not ready without guessing, and have written Helm charts, network policies and RBAC that you then had to debug.
SQL and relational data modelling. PostgreSQL specifically. You can reason about transactions, indexes and migration safety.
You verify your work. You do not report something as working because it deployed and the health check is green. You go and prove it, and you say plainly what you did not test.
Clear written English. Design notes, runbooks, incident write-ups. Much of our engineering context lives in writing.
Building Kubernetes operators / controllers (controller-runtime, CRDs, kubebuilder, Operator SDK)
GitOps at scale - Argo CD or Flux, ApplicationSets, multi-cluster
LLM serving internals - vLLM, SGLang, TensorRT-LLM, KServe, or similar; GPU scheduling, batching, KV cache behaviour
Distributed messaging ( NATS , Kafka) and event-driven pipelines
Traefik or Envoy/Istio at the ingress layer
Time-series and log stores (VictoriaMetrics/VictoriaLogs, Prometheus, ClickHouse)
Frontend competence ( React + TypeScript ) - our operator console is ours to maintain, and being able to fix it end to end is valuable
Billing, metering or payments systems, or anything else where being wrong is expensive
Python, for the gateway extension layer
Cash and equity compensation along with various fringe benefits (healthcare, lunch, wellbeing, and more).
Profitable operations with rapid, sustained growth.
40+ nationalities, with 6 different ones on the management team.
A real chance to make an impact and work alongside world class engineers, researchers, and partners across the global AI ecosystem.
Work mode: Based in Helsinki / London or remote in Europe
Level: Senior
Employment type: Full time and permanent
We're building fast and this role needs the right person behind it. There's no artificial deadline, but when we find who we're looking for, we move. If this sounds like your next move, apply now.
Please submit your application through our Careers page. We don't accept applications sent by email.
European AI cloud provider delivering on-demand GPU infrastructure to developers and organizations from renewable-powered Finnish data centers.
Visit company websiteJobs and hiring trendsFull-time
Mid · 4+ years experience
Hybrid
Apply faster on company sites with our extension.