OPEN TO DEVOPS · SRE · PLATFORM ROLES

I keep production
fast, reliable &
observable.

$ whoami →

DevOps · SRE · Observability engineer with 9+ years running production at scale — from mission-critical trading platforms to multi-cloud Kubernetes fleets on AWS & GCP. I build platforms that ship faster, fail less, and tell you why before your customers do.

📍 Bengaluru, IN 9+ yrs production AWS · GCP certified
Kubernetes AWS · GCP SLO 99.9% GitOps · ArgoCD 📈Observability
SCROLL
0
Years in production
0
Availability targets owned
0
Cloud & network certs
0
Domains: fintech · gov · trading
01 · About

Reliability is a feature.
I engineer it.

Nine years, three domains, one obsession: production that stays up.

I started in 2016 wearing every hat at once — infrastructure, networks, firewalls, releases — the kind of environment where if something broke, you fixed it, because there was nobody else. That built an instinct I still rely on: understand the whole system, not just your layer.

From there I went deep into cloud-native: designing and running Kubernetes platforms on AWS EKS and GCP GKE, building GitOps pipelines with ArgoCD, Helm, Jenkins and GitHub Actions, wiring service meshes, and hardening supply chains for PCI, HIPAA and SOC compliance — while mentoring engineers along the way.

Today I lead observability for live Mutual Fund & Equity trading platforms — where seconds of latency are measured in money. I define the SLOs and SLIs, build the Splunk and Dynatrace dashboards leadership actually uses, and run the RCAs when things get interesting.

Platform builder

Multi-cloud Kubernetes, GitOps delivery, service mesh — designed and run in production, not in a lab.

Observability-first SRE

SLOs, error budgets, synthetic monitoring and predictive dashboards that catch issues before users do.

🛡

Security & compliance aware

Image scanning, hardened pipelines, PCI / HIPAA / SOC support — shipped fast without cutting corners.

Calm under fire

Incident response and RCA in high-pressure financial environments. Trading hours don't wait.

engineer-profile.yaml
apiVersion: career/v9 kind: Engineer metadata: name: rajnikanth-s location: bengaluru-in role: module-lead / sre spec: experience: 9y+ clouds: [aws, gcp] runtime: [k8s, eks, gke] delivery: [argocd, helm, jenkins] observability: - splunk - dynatrace - prometheus/grafana onCall: true # and calm about it status: phase: Running restarts: 0 conditions: - type: ProductionReady status: "True"
02 · Arsenal

Tech stack, battle-tested.

Every tool below has survived real production incidents with me.

Cloud Platforms

certified architect
AWS · EKSGCP · GKEIAM & VPCCost OptimizationRight-sizing

Containers & Orchestration

daily driver
KubernetesDockerHelmIstioNginx Ingress

CI/CD & GitOps

ship it safely
JenkinsGitHub ActionsGitLab CIArgoCDMaven / Gradle

Observability & SRE

core specialty
SplunkDynatracePrometheusGrafanaSLO / SLISynthetic MonitoringRCA

Automation & IaC

toil elimination
PythonBashAnsibleGitJira
🗄

Logging, Data & Network

full-stack depth
ELK / EFKGraylogPostgreSQLMongoDBCassandraBIG-IPpfSenseOpenVPN
03 · Track record

Nine years of keeping the lights on.

From bare-metal networks to multi-cloud Kubernetes to real-time trading observability.

Module Lead — SRE & Observability
Apr 2025 — Present
TEKsystems · Trading / Capital Markets
▲ Proactive alerting → faster response▼ Downtime on trading services◆ SLOs defined org-wide
  • Own reliability for live Mutual Fund & Equity trading platforms in a high-pressure financial environment.
  • Designed and implemented SLOs / SLIs for latency, availability and service health across critical trading services.
  • Built advanced Splunk dashboards with predictive insights — real-time monitoring for engineering and business visibility.
  • Created Dynatrace health/performance dashboards and proactive alerting, cutting response times and downtime.
  • Built synthetic monitors automating user-journey validation; drive RCA and continuous observability improvements.
SplunkDynatraceSLO/SLISynthetic MonitoringIncident Mgmt
DevOps & SRE Engineer
Oct 2022 — Apr 2025
Opt It Technologies · Multi-client production platforms
▲ Deploy frequency via GitOps▼ Cloud spend (right-sizing)◆ PCI · HIPAA · SOC supported
  • Designed DevOps solutions for multiple production-grade environments; managed Kubernetes on AWS EKS & GCP GKE.
  • Built CI/CD with Jenkins & GitHub Actions; automated deployments via ArgoCD + Helm (GitOps).
  • Implemented Prometheus/Grafana monitoring and centralized logging with ELK / EFK / Graylog.
  • Ran Istio and Nginx Ingress traffic management; image vulnerability scanning with Clair & Anchore.
  • Led cloud cost optimization and mentored engineers on DevOps practices and architecture.
EKS / GKEArgoCDHelmIstioPrometheusDevSecOps
IT Admin & DevOps Engineer
Aug 2016 — Sep 2022
Ensomerge Services · Infrastructure & releases
▲ HA network w/ dual-WAN failover◆ Minimal-downtime operations
  • Built and deployed Java applications with Maven / Gradle on Jenkins; managed Tomcat and production releases.
  • Designed pfSense firewall with dual-WAN failover and OpenVPN for secure remote access.
  • Managed DNS and load balancing with F5 BIG-IP; network monitoring via ntopng.
  • Owned patching, upgrades and security hardening across the estate with minimal downtime.
JenkinsTomcatpfSenseBIG-IPLinux
04 · Case studies

Problems solved, at production scale.

Selected work — architecture, problem, impact.

Trading Svcs Splunk Dynatrace Synthetics SLO Board Alerting

Trading Observability Command Center

Problem

Mutual Fund & Equity trading services had fragmented visibility — incidents were detected reactively, often by users, in an environment where latency = money.

Solution

Defined SLOs/SLIs for latency, availability and health. Built predictive Splunk dashboards, Dynatrace performance boards, proactive alerting and synthetic user-journey monitors.

Impact

▲ Issues caught before users notice · faster incident response · single pane of glass for engineering + business.

SplunkDynatraceSLO/SLISynthetics
Git Repo Helm ArgoCD EKS GKE

Multi-Cloud Kubernetes GitOps Platform

Problem

Multiple production environments across AWS and GCP with manual, drift-prone deployments and inconsistent monitoring.

Solution

Standardized on EKS + GKE with ArgoCD + Helm GitOps delivery, Jenkins/GitHub Actions CI, Istio traffic management, Prometheus/Grafana monitoring and centralized ELK/EFK logging.

Impact

▲ Deployment frequency · zero config drift (Git as source of truth) · consistent operations across clouds.

EKSGKEArgoCDIstioGrafana
Commit CI Build Scan 🛡 Prod Clair · Anchore PCI · HIPAA · SOC

Secure CI/CD Supply Chain (DevSecOps)

Problem

Container images shipped to regulated environments (PCI, HIPAA, SOC) without automated vulnerability gates — audit risk and manual review bottlenecks.

Solution

Embedded Clair and Anchore image scanning into Jenkins/GitHub Actions pipelines, enforced policy gates pre-deploy, and hardened the release path to support compliance initiatives.

Impact

▲ Vulnerabilities blocked before production · audit-ready pipelines · compliance shipped without slowing delivery.

ClairAnchoreJenkinsGH Actions
WAN 1 WAN 2 pfSense HA BIG-IP LB OpenVPN

High-Availability Network & Release Infrastructure

Problem

Single points of failure across internet uplinks, remote access and load balancing threatened business continuity.

Solution

Designed pfSense firewall with dual-WAN failover, OpenVPN remote access, F5 BIG-IP load balancing and DNS, ntopng network monitoring, plus hardened patching and release processes.

Impact

▲ Survived uplink failures with zero business interruption · secure remote workforce · minimal-downtime releases.

pfSenseBIG-IPOpenVPNLinux
05 · Credentials

Certified across cloud & network.

Validated by Google, Amazon and Cisco.

Professional Cloud Architect

GOOGLE CLOUD
2026

Solutions Architect — Associate

AWS
2026

CCNA

CISCO
Certified

AWS Associate Training

UDEMY
Completed
06 · Impact

Numbers that move the needle.

What nine years of reliability engineering looks like.

0
Years running production systems
0
Dashboards, monitors & alerts built
0
Deployments automated via CI/CD & GitOps
0
Engineers mentored on DevOps practices
07 · Interactive

Talk to my terminal.

Recruiter-friendly CLI. Try help, skills or sudo hire-me.

rajni@production: ~/portfolio — zsh
rajni@production:~$ ./welcome.sh
Hi, I'm Rajnikanth S — DevOps · SRE · Observability. Type help to see available commands.
rajni@production:~$
08 · Signals

What teams say about working with me.

From engineering peers, leads and stakeholders.

When trading hours are live and something looks off, Rajni's dashboards are the first place everyone looks — and usually the reason we caught it early.

PM

Platform Manager

Trading / Capital Markets

He didn't just set up our Kubernetes clusters — he made deployments boring. GitOps, monitoring, alerts: everything just works.

EL

Engineering Lead

Cloud Platform Team

Calm in incidents, thorough in RCAs, generous as a mentor. The engineers he coached still follow his runbooks.

DM

Delivery Manager

Fintech Programs

09 · Contact

Let's build something reliable.

Open to DevOps, SRE, Platform and Cloud engineering roles. Response SLO: < 24 hours.