DevOps / Site Reliability Engineer
Posted on 28 Jul 2026
DevOps Engineer Intern
Company: Valura.Ai Team: Platform & Infrastructure Location: Bangalore / GIFT City / Dubai / Remote Experience: 3 to 7 years Reports to: [Head of Engineering / CTO] Employment type: Full-time
About Valura
Valura.Ai is a regulated global-investing platform. We let investors in India (IFSCA / GIFT City, LRS route) and the UAE buy US equities, bonds, funds and gold through a single AI-native product. Behind the app sits a real financial system: a double-entry ledger, multiple broker and custodian integrations, KYC/AML pipelines, an AI multi-agent research service, and a partner API used by distributors.
We run two regulated books in two jurisdictions with different rules, different brokers and different reporting obligations. Infrastructure is not a support function here; it is the thing that keeps client money reconciled and auditors satisfied.
The role
You will own how our code gets from a pull request to production, and how we know production is healthy. That means a hybrid estate: managed AWS container workloads for production, a self-hosted platform for lower environments, and a mix of Node/TypeScript and Python services. Your job is to make that estate boring, reproducible and observable.
This is a hands-on role with real ownership. You will not be filing tickets for someone else to apply.
What you will do
CI/CD and release engineering
- Own GitHub Actions pipelines across a Turborepo / pnpm monorepo of around twenty applications plus Python services; keep build times and cache hit rates sane.
- Maintain and harden the branch-to-environment topology across development, staging and production, with the right gates, approvals and rollbacks.
- Reduce release risk: automated migration checks, blue/green or rolling deploys, one-command rollback.
Cloud and infrastructure as code
- Manage AWS: container orchestration, load balancing, managed Postgres, container registry, object storage, DNS, IAM and secrets management, across more than one region.
- Write and review Terraform; bring remaining manually provisioned resources into code with reviewable plans.
- Operate our self-hosted environments end to end, including reverse-proxy routing, TLS and wildcard certificates, and container lifecycle.
Databases and stateful services
- Operate Postgres, Redis and MongoDB: backups, tested restores, point-in-time recovery, connection pooling, upgrades.
- Guard the migration path. Schema migrations that run as part of a deploy must never be able to take an environment down.
Observability and on-call
- Own metrics, logs, traces and alerting. We have working uptime monitoring and chat-based alerting today; you will decide what the mature version looks like and build it.
- Define SLOs for the services that matter (order placement, deposits, ledger sync, reconciliation) and alert on symptoms, not noise.
- Build the incident process: paging, runbooks, blameless postmortems.
Security and compliance
- Own secrets management end to end: a single managed store, no secrets in source or CI, automated rotation, clear ownership.
- Enforce least-privilege IAM and network segmentation, and keep the public attack surface as small as the product allows.
- Support audit and regulatory requirements in both jurisdictions: access logs, change history, evidence of controls, data residency.
Developer experience
- Make local development fast and safe, with strong guardrails between developer environments and production data.
- Provide ephemeral or per-branch environments where they pay for themselves.
- Document the estate so onboarding takes days, not weeks.
What we are looking for
Must have
- 3+ years running production infrastructure for a real product, not just lab work.
- Strong Linux, networking and Docker fundamentals. You can debug a container that will not start, a TLS chain that will not validate, and a connection pool that is exhausted.
- Solid AWS experience, especially container workloads (ECS or EKS), load balancing, IAM and managed Postgres.
- Terraform in anger: modules, state management, drift, imports.
- CI/CD ownership on GitHub Actions (or equivalent) for a multi-service repo.
- Comfortable with Postgres operations: backups, restores, replication, migration safety.
- Scripting in Python and/or Bash; able to read Node/TypeScript and Python application code well enough to diagnose a problem and open the fix yourself.
- Clear written communication. Runbooks and postmortems are part of the job.
Nice to have
- Fintech, brokerage, payments or another regulated environment; familiarity with audit and change-control expectations.
- Self-hosted PaaS experience (Coolify, Dokku, CapRover) or bare-metal Docker fleets.
- Kubernetes, if we decide we need it.
- Observability stacks: Prometheus/Grafana, OpenTelemetry, Loki, Sentry, Datadog.
- An OIDC provider such as Keycloak, plus VPN and zero-trust access tooling.
- Cost engineering. Knowing what a workload should cost and why it does not.
- Experience supporting Python ML/LLM services in production.
How we will evaluate you
- Intro call (30 min) with engineering.
- Technical conversation (60 min): walk through an incident you owned end to end, and a system you built or rebuilt.
- Practical exercise (take-home, timeboxed to ~3 hours, or a live pairing session): containerize and deploy a small service with Terraform and a CI pipeline, then tell us how you would monitor it.
- Systems and judgement round (60 min): architecture, migration safety, secrets management, and regulated data across regions.
- Founder conversation.
Why this role is interesting
- Real ownership of a production estate from day one, with the mandate to redesign the parts that need it.
- Genuinely hard problems: two regulated jurisdictions, a double-entry ledger that must reconcile daily against external brokers, and a rapidly growing AI service tier.
- Small team, short decision path. You will ship in your first week.
- [Compensation band, ESOP, benefits to be inserted.]
To apply
Send your CV plus a short note on one piece of infrastructure you are proud of and one you would rebuild, to [careers@valura.ai].
Valura.Ai
Bangalore / GIFT City / Dubai
1.0 Exp.
Hybrid
Thank you, we have received your application
Our team will evaluate your application and get back to you