01
Kubernetes and developer platforms
When your clusters have grown over years, a platform needs rebuilding, or a migration has stalled.
The work. I design the target architecture, build or migrate clusters and set up GitOps, delivery, namespaces, access rights, autoscaling and observability.
Aim. Teams deploy on their own without every release becoming an ops ticket.
02
Cloud cost
When the cloud bill grows faster than usage and it is unclear which workloads are driving the cost.
The work. I break the cost down by workload, re-size resources, remove what is unused and review commitments, discounts, storage and traffic. Then I implement the changes.
Aim. Running cost falls at the same or better performance. For one previous client that came to around €10,000 per month.
03
Data platforms and databases
When queries are slowing down, a migration is due, or a database becomes unstable under load.
The work. I work on PostgreSQL, MongoDB, Elasticsearch and ClickHouse: sharding, replication, performance analysis, backup and restore, and migrations with a controlled cutover.
Aim. The data layer stays stable and predictable under real load.
04
Private AI infrastructure
When inference needs to move onto your own infrastructure because data cannot leave the company, or API costs are rising too steeply.
The work. I build the infrastructure for it: GPU nodes, scheduling, model serving, scaling and operational procedures for language, speech, vision, embeddings and LLM inference.
Aim. The models run in your environment. Data paths and running costs stay transparent. Running it yourself is not free, but it is controllable.
05
Troubleshooting and infrastructure security
When the same fault keeps coming back, or security hardening stays stuck in the backlog.
The work. I find the cause and implement the corrections. I harden clusters, secure network paths, clean up access rights and secrets, and create dependable routes for patches and updates.
Aim. The problem is fixed at the cause, and the protection fits your environment. At Rackspace the diagnosis and replacement plan were ready in two weeks. Implementation followed during the two-month engagement.
06
Ongoing operations and development
When the platform needs to be operated and developed further after the project.
The work. We work with clear responsibilities, agreed availability and a model that fits your team.
Aim. The platform stays maintained and keeps developing, even when internal capacity is short.