DevOps & Cloud Infrastructure Engineer
Job Summary Birrama Digital Commerce PLC is seeking a hands-on, highly skilled DevOps & Cloud Infrastructure Engineer to manage, secure, automate, and continuously scale the company's technology infrastructure. Working across Linux operating systems, local/private and public clou
Birrama Digital Commerce PLC is seeking a hands-on, highly skilled DevOps & Cloud Infrastructure Engineer to manage, secure, automate, and continuously scale the company's technology infrastructure. Working across Linux operating systems, local/private and public cloud platforms, networking, containerized workloads, Kubernetes, CI/CD pipelines, and observability stacks, this role ensures production environments remain secure, highly available, and reliable. Strong practical expertise in on-premises/private cloud infrastructure, production Linux administration, and container orchestration is core to this position.
Key Responsibilities1. Operating Systems, Systems Administration & Networking
Administer, patch, and maintain production Linux distributions (Ubuntu, RHEL, Alpine Linux).
Manage system users, access controls, permissions, package dependencies, and server performance optimization.
Configure and troubleshoot networking protocols, including DNS, TCP/IP, HTTP/HTTPS, SSH, and SSL/TLS certificate management.
Maintain firewall rules, network access controls, routing, and secure remote management access.
2. Cloud Infrastructure, Storage & Virtualization
Deploy, maintain, and scale virtual machines and compute resources across local/private cloud and public cloud platforms.
Configure and administer object storage platforms (S3-compatible, Google Cloud Storage, Azure Blob Storage).
Design and manage Virtual Private Clouds (VPCs), subnets, NAT gateways, load balancers, and IAM access policies based on the principle of least privilege.
Support backup, disaster recovery (DR), business continuity, and maintain strict isolation between dev, staging, and production environments.
3. Containerization, Kubernetes & Orchestration
Build, optimize, and secure Docker images using multi-stage builds and container vulnerability scanning.
Deploy and manage Kubernetes (K8s) workloads (Pods, Deployments, StatefulSets, Ingress Controllers, ConfigMaps, Secrets, Persistent Volumes).
Manage application deployments and release configurations using Helm charts.
Troubleshoot deployment failures, pod errors, resource limits, and network bottlenecks within containerized clusters.
4. CI/CD, Automation & Infrastructure Management
Maintain source code repositories and version control workflows in Git (GitHub, GitLab, Bitbucket).
Design, automate, and optimize CI/CD pipelines using GitHub Actions, GitLab CI/CD, Jenkins, and GitOps tools (ArgoCD).
Manage deployment secrets, environment variables, and rollback procedures safely across environments.
Automate repetitive infrastructure and operational tasks.
5. Monitoring, Observability & Incident Response
Implement continuous monitoring, alerting, and operational dashboards using Prometheus, Grafana, Datadog, or New Relic.
Establish centralized log aggregation and tracing stacks (ELK Stack, Fluentd/Fluent Bit, CloudWatch, GCP Cloud Logging).
Investigate production incidents using metrics, logs, and traces; conduct root-cause analysis (RCA) and implement preventive measures.
6. Infrastructure Security, Reliability & Documentation
Apply security best practices across Linux OS, networks, cloud resources, containers, and Kubernetes environments.
Maintain complete documentation for infrastructure architecture, runbooks, operational guides, and recovery procedures.
Collaborate closely with software development teams to enhance deployment pipelines, environment consistency, and system performance.
Bachelor's Degree in Computer Science, Software Engineering, Information Technology, Computer Engineering, Information Systems, or a closely related technical field.
Demonstrated hands-on experience in DevOps, Cloud Engineering, Systems Administration, Site Reliability Engineering (SRE), or Infrastructure Engineering.
Proven experience administering production Linux environments and networking infrastructure.
Practical experience deploying containerized applications, managing Kubernetes clusters, and building automated CI/CD pipelines.
Direct experience managing private/local cloud or virtualized infrastructure is highly preferred.
Operating Systems & Scripting: Strong command of Linux administration (Ubuntu, RHEL, Alpine), bash scripting, package management, and system troubleshooting.
Networking & Security: Practical knowledge of TCP/IP, DNS, HTTP/HTTPS, SSH, SSL/TLS, firewalls, VPCs, subnets, routing, and IAM principles.
Containers & Kubernetes: Hands-on experience with Docker, multi-stage builds, Kubernetes workloads, Ingress, and Helm charts.
CI/CD & Automation: Proficiency with Git, GitHub Actions, GitLab CI/CD, Jenkins, or GitOps (ArgoCD).
Monitoring & Observability: Experience setting up Prometheus, Grafana, ELK Stack, or Fluentd for centralized logging and metrics.
Troubleshooting: Analytical problem-solving skills for debugging production incidents, performance bottlenecks, and infrastructure faults.
Certifications: CKA/CKAD (Kubernetes), RHCSA/RHCE (Linux), AWS, Azure, GCP, or networking/security certifications.
Infrastructure as Code (IaC): Hands-on experience with Terraform or OpenTofu.
Configuration Management: Experience with Ansible, Puppet, or Chef.
Database Infrastructure: Experience with database cluster deployment, backup, and restore procedures.
ምንጭ · Source: GeezJobs · ማመልከትዎ በፊት ከቀጣሪው ያረጋግጡ።