SRE and data platform engineer working remotely from Tunis with a Paris-based team. I keep data pipelines fast, observable and cheap to run, on GCP, HashiCorp Nomad and ClickHouse, mostly in Go and Python.
Most of my work is for a company and private, so the details live on my portfolio: portfolio.khalilrezgui0.workers.dev
- Reliability: on-call, incident response, runbooks, and observability with Prometheus, Grafana, Graylog and Fluentd.
- Data infrastructure: MySQL to ClickHouse migrations, ingestion architecture and benchmarks, data-quality monitoring, and cost per event.
- Cloud: production on GCP (Pub/Sub, Cloud Run, BigQuery) and cost optimization.
- Go: event-processing services and tooling.
| Repo | What it shows |
|---|---|
| pulse-sre | Durable append-only event queue in Go: fsync tradeoffs, concurrency, race-tested, with a written bug log |
| debug-practice-go | Alerting utilities in Go (dedup, severity, retries, log parsing) with tests and a debug report |
| etl_devops | ETL into a Flask and Plotly dashboard with container monitoring, one command to run |
| port_sniffer | Multithreaded port scanner in Rust |
Open to remote roles with European teams (Italian citizen, no sponsorship needed). khalilrezgui0@gmail.com · LinkedIn

