DevOps Engineer (Cloud - Native | on - Premise)
We are seeking a DevOps Engineer with hands-on experience in on-premises Kubernetes and cloud-native infrastructure. The role focuses on building and managing fully self-hosted, open-source environments on bare metal or OpenStack, implementing automation, GitOps, monitoring, and secure production operations.
The core responsibilities for the job include the following:
Kubernetes and Infrastructure:
- Deploy and manage Kubernetes clusters on bare metal / OpenStack.
- Create and maintain Helm charts.
- Configure network policies and CNI plugins.
- Manage multi-cluster environments and upgrades.
CI/CD and GitOps:
- Build and maintain CI/CD pipelines using Jenkins and GitOps tools.
- Manage self-hosted runners and artifact repositories (Nexus, Artifactory).
- Integrate automated testing, security scanning (Trivy), and code quality checks (SonarQube).
- Implement progressive deployment strategies (Canary, Blue-Green).
Monitoring and Observability:
- Deploy and manage Prometheus and Grafana dashboards.
- Implement centralized logging (EFK stack or Loki).
- Configure distributed tracing using Jaeger.
- Set up AlertManager for alerting and incident response.
Networking and Platform Services:
- Configure Ingress controllers (NGINX, Traefik) and load balancers (MetalLB).
- Configure DNS and service discovery.
- Manage TLS/SSL certificates using cert-manager.
Security and Compliance:
- Implement RBAC and Kubernetes Pod Security Standards.
- Perform security audits (Kube Bench, Kube Hunter).
- Manage secrets using Vault or Sealed Secrets.
Requirements:
- Bachelor's degree in computer science, IT, or a related field (or equivalent experience).
- Strong hands-on experience with Kubernetes (on-premises).
- OpenStack infrastructure experience.
- Experience in Jenkins CI/CD pipeline implementation, Helm chart development, Linux administration and scripting, and infrastructure automation practices.
- Monitoring, logging, and observability tools.
- Kubernetes networking and storage fundamentals.
Nice to Have:
- Multi-cluster management tools.
- Backup and disaster recovery automation.
- Policy enforcement frameworks.
- Kubernetes certifications.
Soft Skills:
- Strong troubleshooting and problem-solving.
- Clear communication and documentation.
- Self-driven and collaborative mindset.
- Production support readiness.