Site Reliability Engineer (SRE) Operations
teciem · TCMi - Bengaluru · India · On-site
Posted Aug 30, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
The Work We Do
Teciem designs, builds, and delivers treasury and capital markets software solutions for financial institutions worldwide. We serve banks of every size and geography, offering the right setup for the right need.
Our solutions are designed to replace multiple disconnected systems with one complete, front-to-back platform, helping customers to capture trading and business opportunities quickly, clearly and with control. We cover the entire trading lifecycle, ensuring that everything - from execution to position keeping, to risk management – runs smoothly.
With decades of experience and one of the largest, most diverse client bases in the industry, we turn deep industry knowledge into software that covers most asset classes, meets complex real-world treasury and capital market's needs, and adapts as markets evolve.

 About Kondor UP
Kondor UP is the SaaS edition of Kondor , Teciem's flagship treasury management platform. Operated by TCM (TeCIEM) on AWS , Kondor UP delivers treasury capabilities as a fully managed, always-on service - processing real financial value on behalf of tier-1 banks under strict regulatory requirements.
The platform runs on Amazon EKS (Kubernetes), backed by Amazon RDS for SQL Server (Multi-AZ), Amazon MSK (Kafka), Amazon MQ , and a comprehensive observability stack (OpenTelemetry, Prometheus, Grafana, Loki, Jaeger). Changes are delivered through GitOps (ArgoCD) and Terraform , with blue/green deployments and feature flags ensuring zero-downtime operations.
Segregation of duties, full audit trails, and always-on availability are not edge cases in this domain - they are the baseline.
Role Summary
As a Senior Site Reliability Engineer (SRE) - SRE Operations , you will be responsible for the reliability, availability, and performance of the Kondor UP production platform. This is an operations-focused engineering role : you own the production system, define and defend SLOs, lead on-call and incident response, and relentlessly drive down toil through automation and engineering.
You apply software engineering discipline to operational problems - turning production failures into systemic improvements, building the observability that gives teams real-time insight, and shifting reliability practices left into the delivery lifecycle.
The SRE Operations engineer works closely with the InfraOps , AppOps , NetOps , and SecOps teams, and is a senior contributor within the SRE workstream .
Key Responsibilities
1. Service Reliability & SLO Management
Define, instrument, and own Service Level Indicators (SLIs) , Service Level Objectives (SLOs) , and error budgets across critical Kondor UP services (deal capture, risk engine, market data feeds, API gateway).
Maintain the customer-facing SLA commitments; monitor error budget burn rates (2h / 24h windows) and trigger reliability work when burn thresholds are breached.
Produce monthly SLA packs and tenant-scoped reliability reports for customer…