DevOps & Site Reliability Engineer | AWS Certified Solutions Architect | Kubernetes & Cloud Infrastructure
DevOps and Site Reliability Engineer with 3+ years of experience supporting Linux platforms, Kubernetes/RKE2 environments, and distributed services across banking and public-health systems. I work on reliable delivery, incident troubleshooting, observability, infrastructure automation, networking, identity, data platforms, and clear operational documentation.
My strongest hands-on experience is in on-premises and cloud-native infrastructure. My AWS knowledge is certification-backed and reinforced through labs and architecture exercises; I do not present it as production AWS operating experience.
Based in Addis Ababa, Ethiopia, and open to international DevOps, Cloud, Platform Engineering, and Site Reliability Engineering opportunities.
- Platform operations: Linux, Kubernetes, RKE2, Rancher, Docker, Helm, Longhorn, networking, storage, upgrades, rollouts, and troubleshooting
- Reliability engineering: incident response, health checks, observability, capacity awareness, failure drills, runbooks, and evidence-based production readiness
- Delivery automation: GitHub Actions, GitOps, Argo CD, container pipelines, Kustomize, Ansible, and repeatable validation gates
- Cloud foundation: AWS Solutions Architect certification, VPC/IAM/EC2/S3/RDS/Route 53/CloudWatch knowledge, plus structured cloud and Terraform labs
| Project | What it demonstrates | Automated evidence |
|---|---|---|
| RKE2 High-Availability Platform Reference | Sanitized three-server etcd quorum, dedicated workers, Longhorn, security controls, Helm examples, and an operations runbook | Configuration and manifest validation in GitHub Actions |
| OpenSearch Observability Platform | OpenTelemetry-based logs, metrics, traces, service maps, SLO calculations, dashboards, and production gates for RKE2 | Functional lab validation and documented operating checks |
| Spring Boot CI/CD and GitOps Lab | Maven, Docker, Kubernetes/Kustomize, Argo CD, GitHub Actions, GHCR, probes, resource controls, and container hardening | Tests, image build, health smoke test, linting, and schema validation |
| Kafka and PostgreSQL Resilience Lab | Three-node Kafka KRaft quorum, PostgreSQL physical streaming replication, health checks, and controlled recovery boundaries | Cross-broker messaging, standby replication, and one-broker failure drill |
| Ansible Infrastructure Automation Lab | Reusable roles for Linux baselines, Docker, hardening, and RKE2 node preparation | YAML lint, Ansible lint, syntax, and idempotency checks |
Public repositories use sanitized or disposable lab configurations. They demonstrate engineering decisions and repeatable evidence without exposing employer, customer, network, credential, or production data.
| Area | Technologies |
|---|---|
| Platforms | Linux, Kubernetes, RKE2, Rancher, Docker, Helm, Longhorn, Harbor, Argo CD |
| Automation and delivery | Ansible, Terraform labs, Git, GitHub Actions, CI/CD, GitOps, Bash |
| Reliability and observability | OpenTelemetry, Prometheus, Grafana, OpenSearch, SigNoz, alerting, SLOs, incident response |
| Networking, identity, and edge | TCP/IP, DNS, TLS, Nginx, load balancing, Kubernetes RBAC, Keycloak, AWS IAM |
| Data, messaging, and storage | PostgreSQL, MySQL, YugabyteDB, Redis, Kafka, ActiveMQ, SeaweedFS, MinIO |
| AWS knowledge and labs | VPC, EC2, IAM, S3, RDS, Route 53, CloudWatch, architecture and Terraform exercises |
- AWS Certified Solutions Architect – Associate
- AWS Certified Cloud Practitioner