Infrastructure Reliability & Security Engineer · 8 years
I build tools that detect, diagnose, and safely fix production failures across Kubernetes clusters and Linux fleets.
Compute Central · KubeRescue · Linux Vitals · IPMG
| Project | Problem it solves | Links |
|---|---|---|
| KubeRescue (Go) | Restarting crash-looping pods by hand hides the cause. KubeRescue records the evidence (exit code, restart count, owner) before acting, and every remediation is bounded and dry-run first. Pre-1.0. | Live · Repo |
| Linux Vitals (Ansible) | Health-checking a mixed RHEL, Ubuntu, and SUSE fleet without installing agents. Compares baseline to post-change state, writes one HTML report, and only fixes things when you opt in. | Live · Repo · Galaxy |
| IPMG (Python) | Finding which hosts went down since the last scan. Parallel ping sweeps, reverse DNS, and scan-to-scan diffs. brew install sameeralam3127/tap/ipmg |
Live · Repo |
| k8s-kubeadm-lab (Shell) | Practising etcd recovery, upgrades, and RBAC somewhere it's safe to break: a reproducible multi-node kubeadm cluster across macOS and Windows. | Repo |
| llm-dev-kit (TypeScript) | Chat and RAG over private PDFs without data leaving the machine: local-first on Ollama, with optional cloud model routing. | Repo |
Checked every 6 hours by a GitHub Actions workflow.
| Tool | Status | Response ms | Last checked (UTC) |
|---|---|---|---|
| IPMG | 🟢 Up | 167 | 2026-10-06 00:02 |
| KubeRescue | 🟢 Up | 158 | 2026-10-06 00:02 |
| Linux Vitals | 🟢 Up | 174 | 2026-10-06 00:02 |
- IBM/docling-pipelines — refactor(ollama): hoist repeated imports out of OllamaClient hot-path methods (#133)
Happy to talk about reliability, Kubernetes, Linux automation, and infrastructure security.



