Senior DevOps / Platform Engineer, AI Infrastructure (m/f/x)

🏢 MAIA · company page
📍 Germany
💰 EUR 75,000 - 85,000 / annual
📅 Posted Sep 20, 2026 · via Himalayas
🏷 Platform Engineering, DevOps Engineer, AI Infrastructure, Site Reliability Engineering, Senior Platform Engineering, Senior AI Platform Engineer +7 more
Apply on original site ↗

MAIA is the AI platform built for companies where generic AI tools break down because the data is complex, the stakes are high, and precision matters.

We work on the difficult parts of enterprise AI: understanding complex documents, integrating AI into real workflows, making organisational knowledge accessible, and operating the underlying systems reliably and securely.

As our customer base, product capabilities, and engineering team grow, we are investing further in the platform beneath MAIA . We are looking for our first dedicated Senior DevOps / Platform Engineer to take technical ownership of our infrastructure and production operations.

You will develop an existing platform and shape its next stage across architecture, automation, observability, security, and AI infrastructure. Our environment combines self-administered Linux servers, European infrastructure providers, dedicated hardware, and services from AWS, Azure, and GCP. Over time, we will also expand our European infrastructure and inference options as one part of our platform strategy.

This is a hands-on senior individual contributor role without people management. You will initially be the only dedicated Platform Engineer, working closely with our software engineers and company leadership.

We are opening this role to remote candidates across Germany. For this setup to work, you need a strong track record of independently operating production systems, communicating proactively, and moving complex infrastructure work forward without constant coordination.
Tasks

-
You take technical ownership of MAIA ’s production infrastructure. You operate and evolve self-administered Linux systems across virtual machines and dedicated servers, including containers, networks, reverse proxies, and API gateways.

-
You make deployments safe, repeatable, and easy for engineers to use. You improve our Infrastructure as Code, GitHub Actions workflows, deployment processes, automated checks, versioning, and rollback capabilities.

-
You operate and improve PostgreSQL in production. This includes performance analysis, connection pooling, capacity planning, backups, and regularly tested restore procedures.

-
You develop our self-hosted observability stack. You connect metrics, logs, traces, and actionable alerts so that we understand system behaviour and identify problems before they affect more customers.

-
You strengthen security across our infrastructure. You implement and maintain IAM, least privilege, secrets management, TLS, vulnerability scanning, and patch management.

-
You implement technical controls for ISO 27001. You design them so that evidence is generated continuously and remains understandable and auditable.

-
You help shape our AI infrastructure. You integrate model and inference providers and evaluate them based on reliability, latency, throughput, cost, and operational effort.

-
You explore self-hosted LLM inference where it creates real value. Over time, this may include GPU infrastructure and serving technologies such as vLLM. Previous production experience in this area is helpful but not required.

-
You improve incident response and operational resilience. You investigate root causes, establish useful runbooks, document operational knowledge, and turn incidents into lasting improvements.

-
You improve the internal developer experience. You reduce manual work, create clear interfaces and workflows, and help engineers ship changes with confidence.

-
You make infrastructure and inference costs transparent. You identify relevant cost drivers and help us make informed build, buy, and hosting decisions.

You will help determine the initial priorities after assessing the existing platform. We expect you to identify the most relevant risks, explain the available options, and take improvements through to reliable production operation.
Requirements
Your technical experience

- You have several years of experience operating production SaaS

Flights + hotels

This role requires you to be in Germany. If that means relocating or flying in, it is worth checking fares before you commit to a start date.

Compare flights and hotels →

← All remote jobs

Comparing DevOps Engineer pay and openings — the live median is $119k?All remote DevOps Engineer jobs →DevOps Engineer salary data →
Want more like this? Browse every live remote developer role.All remote developer jobs →
Get new developer jobs by email
Daily email, only when there's something new. One click to stop.

Get remote developer jobs like this by email

10 hand-picked jobs, one email a day. No spam, unsubscribe anytime.

Similar for you