Mouhamadou Lamine Gueye, senior DevOps / Platform / SRE engineer. I deploy, secure and operate cloud & on-prem infrastructure in high availability: Kubernetes, IaC, CI/CD, observability. 5+ years, multi-cloud, banking-grade.
A national compliance-certificate platform buckling under event traffic. I re-architected it for high availability and taught it to scale before the storm, not after.
Stood up and now operate the bank's on-prem Data as a Service platform in HA, provisioned entirely as code, secured end to end.
Forensic investigation and recovery of a breached Kubernetes environment for an identity / eKYC platform, then hardening so it wouldn't happen twice.
I co-founded Telrai ↗, a teleradiology platform, and I run all of its infrastructure and operations.
Migrated a live production database from MySQL to MongoDB with no service interruption, I built the whole procedure end to end, from schema reshaping to a verified cutover.
Reverse-chronological, like any good log.
Deploy, secure and operate the bank's on-prem data platform in high availability. Kubernetes as code, full data stack run & reliability, DevSecOps and end-to-end observability.
Docker containerization and GitLab CI/CD; deploy & operate banking APIs (Spring Boot, FastAPI). PostgreSQL/Redis admin; interbank payment integrations. Seconded to CBAO from Dec 2025 to Sept 2026.
Re-architected an ECS Fargate + ALB platform for HA and event-scale: predictive + reactive autoscaling, p95 13s→~1s, major cost reduction, ALB/CloudWatch observability. Also ran a zero-downtime MySQL→MongoDB migration.
Built complete cloud infra (ECS, EC2, RDS, ALB, Route 53, ECR) for internal SaaS. ECS Fargate + Docker, GitLab CI/CD, NiFi clusters, Prometheus/Grafana; DigitalOcean & Vultr deployments.
Cloud infrastructure consulting for startups & SMEs across AWS, GCP, DigitalOcean and Vultr. End-to-end delivery: architecture, automation (Docker, Kubernetes, Terraform, CI/CD), security, monitoring and cost.
Co-founded Telrai, a teleradiology platform for remote sharing and interpretation of medical imaging (DICOM, OHIF viewer, AWS HealthImaging). As CTO I own the whole stack: Spring Boot backend, React front-ends, AWS infrastructure, CI/CD, Cloudflare edge and security.
Taught web development part-time, mentored students through hands-on projects, code reviews and live debugging, from fundamentals to shipping working apps. Explaining systems simply is half of good engineering.
Member of Senegal's national men's chess team (2021 & 2023), representing the country in Malawi and Budapest. The same discipline I bring to reasoning about failure domains and blast radius.
🎓 Engineering degree · École Supérieure Polytechnique (ESP), Dakar.
A 289-line production Dockerfile for a Node.js API with Chrome and native modules, read line by line: what multi-stage builds really buy, why the runtime is Debian slim and not Alpine, 84 build arguments that bake secrets into the image, a chown that doubles the size, and the rewrite.
A platform kept dropping requests after the scaling work was done. Load balancer logs showed a synchronous OCR job, not the infrastructure. Why metrics belong inside the deliverable, why the perimeter must be written down, and the three levels at which a system gets optimised.
A search cluster died and five production services went with it, none of which needed search to serve a page. A post-mortem, minute by minute: a non-vital dependency inside a liveness probe, and a watchdog I had shipped two days earlier with no blast radius.
More notes on the way: Kubernetes on bare metal, NiFi in production, zero-downtime migrations.
Alongside permanent roles I take one-off missions. Remote, in French or English, on an existing platform — usually one someone else built.
Someone left, documentation is thin, and nobody is sure what runs where. I inventory it, get access back under control, write down what I find, and fix what is dangerous first.
Done on a sabotaged Kubernetes estate: access recovered, data recovered, cluster back in production.Account to account, on-prem to cloud, one database engine to another. The plan states what breaks, what rolls back, and how you know it worked.
74 services rebuilt in a new AWS account in one night; a MySQL to MongoDB switch with zero downtime.Not a spreadsheet of theoretical savings: a measured baseline, the levers ranked by what they actually return, and the ones I apply.
One estate taken down 39 %. Found 170 USD/month of extended support on a 13 USD/month database.Alerting that fires on the right thing, dashboards someone reads, and enough captured history to reconstruct an incident to the minute rather than guess at it.
Read the post-mortem in the writing section — that timeline exists because the platform was instrumented for it.Typically a few days to a few weeks. Tell me the problem, not the job title — I will say if I am the wrong person for it.
Open to Senior DevOps / Platform / SRE roles and freelance missions, remote or relocation. If your platform needs to stay up under pressure, I'm your engineer.
What do you need help with? (pick one or more)
Tell me about you & the project
Thanks, your request is on its way. I read every message personally and I'll get back to you shortly.