Rubiscape deploys to five distinct
topologies — SaaS, BYOC, On-Premises, Hybrid, and Air-Gap — and the Platform
Infrastructure Engineer is the engineer who makes all five reliable,
observable, and fast. You will own the Kubernetes platform, CI/CD pipelines
(ArgoCD + GitHub Actions), and the cloud infrastructure (AWS/Azure/GCP) that
runs Rubiscape’s six-studio platform for Fortune 500 enterprise customers. This
is an infrastructure role with a strong software engineering dimension: you
will write Terraform modules, Helm charts, and Go or Python automation tooling
— not just click through cloud consoles. If you want your infrastructure work
to directly translate into platform uptime, faster developer velocity, and the
ability to ship new enterprise customers to production in under 90 days, this
role is built for you.
· Design,
provision, and maintain multi-cloud Kubernetes clusters (EKS, AKS, GKE) using
Terraform and Helm, supporting SaaS, BYOC, and on-prem customer topologies from
a single control-plane model.
· Own the
GitOps delivery pipeline: manage ArgoCD application sets, define promotion
workflows across dev/staging/production environments, and enforce deployment
policies via OPA/Gatekeeper.
· Build and
maintain GitHub Actions CI pipelines — including build caching, container image
signing (Cosign), vulnerability scanning (Trivy), and automated integration
test orchestration.
· Define and
enforce platform security posture: Kubernetes RBAC, network policies, Pod
Security Standards, secret rotation via Vault, and audit logging pipelines to
SIEM.
· Implement
platform observability: Prometheus + Grafana dashboards, AlertManager routing,
distributed tracing with Jaeger/Tempo, and SLO-based alerting for all
production services.
· Manage the
data platform infrastructure layer: Kafka cluster operations (MSK/Confluent),
PostgreSQL HA (Patroni or RDS Multi-AZ), Redis cluster configuration, and Spark
on Kubernetes job scheduling.
· Act as the
infrastructure partner for product engineers — reviewing Kubernetes manifests,
advising on resource sizing, and building self-service tooling that reduces
infrastructure toil for the wider engineering team.
· Experience
delivering air-gap or on-premises enterprise deployments with offline container
registries, local Helm chart repositories, and network-restricted environments.
· Background
with service mesh technologies (Istio, Linkerd) for mTLS, traffic management,
and fine-grained observability between Rubiscape’s microservices.
· Familiarity
with FinOps practices — cloud cost allocation, Spot/Preemptible instance
strategies, and right-sizing recommendations for Kubernetes workloads.
· Kubernetes
operator development experience (Kubebuilder, Operator SDK) for automating
complex stateful application lifecycle management.
Rubiscape is India’s leading Decision
Intelligence Platform, unifying data engineering, BI, machine learning, and
agentic AI in a single governed platform. Built in Pune and trusted by Fortune
500 enterprises across BFSI, manufacturing, healthcare, and government. 8
international innovation patents. 10 Industry-Academia Labs & COEs. From BI
to AI — One Platform. Every Decision.
RequirementsRequirements
· 4+ years
of infrastructure or platform engineering experience, with at least 2 years
operating production Kubernetes clusters at enterprise scale.
· Deep
Terraform expertise: module authoring, remote state management, workspace
strategies for multi-environment/multi-cloud provisioning.
· Hands-on
ArgoCD or Flux experience for GitOps-based application delivery, including
multi-cluster federation and rollback automation.
· Strong
security engineering background: cloud IAM, Kubernetes RBAC, network policy
design, zero-trust networking principles, and compliance frameworks (SOC 2, ISO
27001).
· Scripting
and automation proficiency in Python or Go for building internal platform
tooling, operator controllers, or custom Kubernetes admission webhooks.
· Experience
managing stateful workloads on Kubernetes: PostgreSQL, Redis, Kafka — including
backup/restore strategies, failover testing, and capacity planning.



