🚀 Download Initiated Successfully.

Sadiq Karur

//
Get in Touch Talk to AI Blog
OPEN_TO_COLLAB :: 00:00:00 IST

7+

Years Exp

50+

Prod Nodes

99.95%

SLO Attained

I build calm systems for chaotic problems. Agents that explain themselves, pipelines that heal themselves, and infrastructure that lets people sleep.

01

Automate it twice

If a task happens twice, it becomes a script. If the script runs twice, it becomes a pipeline.

02

Trace every decision

An agent nobody can audit is a liability. Every tool call and model decision leaves a span in the trace.

03

Ship the prototype

A working demo in the customer's environment beats a hundred slides. Prototype first, polish under load.

Career Path

Professional Log

Forward Deployed Engineer, Agentic AI CoE

Tata Consultancy Services (TCS)
May 2026 to Present
  • Embedded with customers as an FDE in the Agentic AI Center of Excellence, turning ambiguous business problems into working agentic prototypes, fast.
  • Building ServiceNow POCs: agentic workflows for ticket triage, incident enrichment, and Now Assist style automation across ITSM processes.
  • Building Amazon Bedrock PoCs, including an autonomous observability agent using Strands Agents on Bedrock AgentCore Runtime, with OTEL/ADOT traces flowing into CloudWatch for full tool-call and model-decision visibility.
  • Applying advanced AWS agentic patterns: AgentCore Memory & Gateway, multi agent orchestration, RAG over enterprise knowledge, and human in the loop guardrails.
  • Designing and orchestrating multi agent systems with LangChain, CrewAI, and AutoGen for enterprise automation workflows across AWS/Azure cloud infra.

DevOps Cloud Engineer

Tata Consultancy Services (TCS)
Mar 2023 to Apr 2026
  • Engineered multi account AWS landing zones with Terraform and Control Tower, standardizing provisioning across client environments.
  • Built CI/CD pipelines with Jenkins and GitHub Actions deploying to EKS via Helm, including blue/green and canary release strategies.
  • Migrated on prem client workloads to AWS and Azure, combining lift and shift with selective refactoring and cutting infra costs by 30%.
  • Ran the observability stack: Prometheus, Grafana, and CloudWatch dashboards with SLO driven alerting that reduced MTTR by 45%.
  • Automated patching and configuration at fleet scale using Ansible and AWS Systems Manager, with tested DR runbooks per client.

Systems Engineer

CGI
Aug 2019 to Feb 2023
  • Administered Nginx/Apache servers across 50+ production nodes, maintaining SOC2 compliance.
  • Optimized Jenkins pipelines for Java/.NET apps, improving deployment speed by 40%.
  • Containerized legacy workflows using Docker, reducing environment inconsistencies by 60%.
  • Automated Prometheus/Node Exporter deployments via Ansible for real-time monitoring.
  • Automated filesystem management and repo setup using Shell scripting, reducing manual toil by 70%.

0+

Years in Production

0+

Nodes Managed

0%

MTTR Reduced

0

Incidents Documented

Showcase

Key Deployments

EKS AutoScaler

Architected high availability EKS clusters with Karpenter for rapid node provisioning. Reduced cloud costs by 25% via Spot instances.

Kubernetes AWS Go

Autonomous Observability Agent

Bedrock PoC: Strands Agents on AgentCore Runtime that investigate AWS infra anomalies autonomously. Every tool call and model decision traced via OTEL/ADOT into CloudWatch.

Bedrock Strands AgentCore OTEL

ServiceNow Agentic POC

Agentic ITSM workflows in ServiceNow: AI driven ticket triage, incident enrichment from telemetry, and automated resolution suggestions with human in the loop approval.

ServiceNow GenAI ITSM

Zero Drift Infra

Modular Terraform library for VPC, RDS, and ECS provisioning. Enforced policy-as-code using OPA/Checkov in CI pipelines.

Terraform HCL CI/CD

RAG Knowledge Mesh

Enterprise retrieval pipeline on Bedrock Knowledge Bases with OpenSearch vectors. Chunking strategies tuned per document type, citations returned with every answer.

Bedrock RAG OpenSearch

GitOps Delivery Platform

Multi cluster EKS delivery with ArgoCD and Kustomize. Every environment reconciled from git, drift detected and reverted automatically within minutes.

ArgoCD Kubernetes GitOps

Incident Copilot

On call agent living in Slack. Pulls CloudWatch metrics and Grafana panels on demand, drafts incident timelines, and suggests runbook steps for responders.

GenAI SRE Slack
Dispatches

Latest Incident Logs

View All Logs
Expertise

Technical Arsenal

Kubernetes

Kubernetes

Terraform

Terraform

Docker

Docker

Agentic AI

Agentic AI

ServiceNow

AWS Bedrock

Strands Agents

AgentCore

OpenTelemetry

Python

Python

Ansible

Ansible

Jenkins

Jenkins

Grafana

Grafana

Linux

Linux

SonarQube

SonarQube

Prometheus

Prometheus

Proof of Work

Credentials

CERTIFIED

AWS Certified Solutions Architect, Associate

Amazon Web Services

CERTIFIED

Azure DevOps Engineer

Microsoft

CERTIFIED

Claude Certified Architect (Foundations)

Anthropic

CERTIFIED

Agentic AI Engineer Professional

Professional Certification

How I Work

Engagement Protocol

01

Discover

Embed with your team. Map the workflow, the data, and the pain before writing a line of code.

02

Prototype

A working agent in your environment within days. Real data, real integrations, real feedback.

03

Harden

Guardrails, evals, and observability. Every decision traced, every failure mode rehearsed.

04

Handover

Docs, training, and a clean path to production. Your team owns it when I leave.

Questions

Frequently Queried

What does a Forward Deployed Engineer actually do?

I sit with the customer, not behind a ticket queue. FDEs take an ambiguous business problem, embed with the team that owns it, and build a working solution in their real environment. Less handoff, more shipping.

Can you build an agentic POC for my team?

That is the day job. Recent examples include ServiceNow ticket triage agents and an autonomous observability agent on Amazon Bedrock. Reach out with the problem, not the solution, and we will scope it together.

What is your default stack?

AWS first: Bedrock, AgentCore, Strands Agents, EKS, Terraform. Observability through OpenTelemetry into CloudWatch and Grafana. For ITSM work, ServiceNow. Pragmatic about everything else.

Where can I read your incident writeups?

The blog hosts 70+ deep dives on outages and breaches, from CrowdStrike to Salt Typhoon. Each covers the timeline, the technical root cause, and the lessons. Visit bl0g.surge.sh.

Neural Interface

AI Terminal

user@sk:~/neural-link
CONNECTED :: GROQ_API
> SYSTEM_INIT... OK
> LOADING_PROFILE_DATA... OK
> ESTABLISHING_NEURAL_LINK... SUCCESS
> Hello! I am a digital assistant for Sadiq Karur. How can I help you today? Would you like to know more about Sadiq's work or portfolio?
>