page
About
RetakeData is an independent infrastructure engineering practice led by a team with 10+ years across sysadmin, SRE, and platform engineering. Private cloud, observability, automation, and on-prem AI.
Independent infrastructure engineering practice
RetakeData is an independent infrastructure engineering practice. We help teams build and operate systems they can keep under their own control: bare metal, private cloud, hybrid platforms, observability, automation, and private internal tools.
The work is hands-on. We design, implement, test failure modes, document what changed, and stay involved where useful: maintenance, observability, optimization, and training.
Approach
Infrastructure problems rarely stay in one box. A migration can become a storage problem. A storage problem can become an observability problem. An observability problem can expose missing automation, weak runbooks, or a team that was never trained on the system.
That is where we fit best: crossing the layers without losing the details.
What We Work On
Bare metal, networking, storage, virtualization, Proxmox/Ceph, vSphere/OpenStack migrations, HA design.
Automation, deployment workflows, service discovery, observability, alerting, RUM, SLOs, and monitoring as code.
Internal apps, SRE tools, private AI, RAG, RBAC-aware assistants, and operational tooling that stays inside your infrastructure.
Practical tools around real operations, from SSH fleet access to Terraform providers and audit UIs.
How We Work
Understand first
We start by understanding the current system, not by replacing it. What runs where, who operates it, what breaks, what nobody wants to touch, and what needs to be measurable.
Implement with the team
Then we implement with the team: small enough to stay understandable, solid enough to survive production, documented and monitored enough to maintain.
Scale when needed
For larger missions, vetted engineers from our network can join the delivery. One point of contact, same operating style.
Background
retakedata is led by Sabri Mjahed, an infrastructure engineer with 10+ years across sysadmin, SRE, and platform engineering.
From racks and virtualization to observability, automation, private AI, and internal tools, the focus stays the same: build systems that are reliable in production and clear for your team to own.
Proof at Scale
11 TB/day across Loki, Elasticsearch, and Thanos. Grafana deployments serving 2000 users. Ingestion tuning, recording rules, SLO dashboarding, query optimization.
3000+ VMs delivered across 4 providers. 6-phase pipeline: Git PR, Terraform, Consul/NetBox, Ansible, HAProxy, Centreon. Team-operated, not solo-built.
100+ Proxmox nodes deployed with PXE automation. Production experience on ZFS, NFS, SAN, NVMe-oF, and Ceph. vSphere migrations, HA cluster design.
DDoS mitigation on F5, distributed firewalling across 1000+ VMs via Ansible. Cisco/Juniper networking, ExtraHop NDR, vulnerability management.
GPU servers running vLLM with private RAG over 1000+ documents. RBAC-aware SRE assistants operating entirely inside client infrastructure.
PostgreSQL HA (Patroni/etcd), large-scale Couchbase, Kafka pipelines, Elasticsearch clusters. Operated at production scale with proper failover and recovery.
Want to scope something?
Bring a platform, migration, observability, or private systems problem. We will help turn it into a concrete plan.