Job Title: Devops Engineer
Location: Mumbai
Experience:
Employment Type: Full-Time
About Oneture:
Oneture is building intelligent, cloud-native solutions that drive real business impact. We are looking for a passionate Senior Data Scientist to join our growing team and lead predictive modeling initiatives.
About the Role:
Oneture Technologies is looking for a hands-on DevSecOps / Platform Engineer to join an on-site, air-gapped on-premise
Kubernetes engagement. The platform underpins a mission-critical, security-hardened system and runs entirely on-premise across bare-metal infrastructure — this is not a cloud-first role. The engineer will own day-to-day cluster operations, CI/CD tooling, and observability across Production, DR, and UAT environments.
Key Responsibilities
- • Own and operate the on-premise Kubernetes clusters across PR, DR, and UAT — upgrades, patching, capacity, and incident response.
- • Build and maintain on-prem CI/CD pipelines and secure air-gapped deployment workflows.
- • Implement and maintain the Prometheus/Grafana monitoring stack and an on-prem logging solution, with alerting for critical workloads.
- • Work closely with application and data teams (streaming, database) to keep the platform secure, available, and performant.
- • Document infrastructure, runbooks, and troubleshooting procedures for a security-hardened, air-gapped setup.
Primary Skills: Must Have:
1.On-Premise Kubernetes (End-to-End):
• Proven hands-on experience installing, configuring, and operating a Kubernetes cluster entirely on-premise / bare-metal (e.g. RKE2, kubeadm, k3s) — not managed/cloud K8s (EKS/GKE/AKS).
• Strong grasp of on-prem-specific components: CNI (e.g. Calico), MetalLB or equivalent for bare-metal load balancing, ingress controllers, and persistent storage (e.g. Longhorn, local-path-provisioner).
• Experience managing multi-node clusters across Production, DR, and UAT with high-availability and DR failover considerations.
• Comfortable troubleshooting at the cluster, node, and container-runtime level (e.g. containerd) in a fully air-gapped environment
2. On-Premise DevSecOps Tooling [MANDATORY] :
• Experience standing up and administering CI/CD tooling on-premise — Jenkins or equivalent (GitLab CI, Tekton, etc.), fully self-hosted with no external/SaaS dependency.
• Ability to design secure, air-gapped CI/CD and artifact/image workflows (e.g. offline image transfer and import pipelines) where the environment has no direct internet access.
• Working knowledge of security hardening for on-prem environments — access controls, sudo/user policy restrictions, firewall and network segmentation across subnets.
3. On-Premise Monitoring & Logging [MANDATORY]
• Hands-on experience deploying and managing a self-hosted monitoring and alerting stack — Prometheus and Grafana at minimum.
• Experience with an on-prem logging/log-aggregation stack (e.g. Loki, ELK/EFK, or equivalent) for cluster and application-level observability.
• Ability to build actionable dashboards and alerts for cluster health, node resource usage, and workload-level metrics without relying on cloud-native/SaaS monitoring tools.
4. Linux Fundamentals [MANDATORY]
• Strong Linux administration skills (RHEL/CentOS preferred) — Kubernetes runs on Linux, and day-to-day troubleshooting happens at the OS level.
• Comfortable with systemd, networking (iptables/nftables, routing), storage/disk management, and package management in an offline/air-gapped context.
• Solid shell scripting ability for automation and troubleshooting.