Senior SRE Engineer
Leantechio
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
Description
Company Overview:
Global Technology Services is a rapidly expanding organization situated in Medellín, Colombia. We pride ourselves on possessing one of the most influential networks within software development and IT services for the entertainment, financial, and logistics sectors. Our corporate projections offer a multitude of opportunities for professionals to elevate their careers and experience substantial growth. Joining our team means engaging with expansive engineering teams across Latin America, Philippines and the United States, contributing to cutting-edge developments in multiple industries.
Position Title: Senior SRE Engineer
Location: LATAM (Remoto)
What you will be doing:
We are seeking a Senior Oracle Transportation Management (OTM) Site Reliability Engineer with extensive experience supporting mission-critical enterprise applications and transportation platforms. This role combines deep Oracle OTM expertise with strong SRE, cloud, automation, observability, and production operations capabilities.
The ideal candidate has at least 10 years of overall IT/SRE experience, including 5+ years of hands-on Oracle Transportation Management experience across both on-premises and SaaS/cloud environments. This person will be responsible for maintaining platform reliability, troubleshooting complex production issues, improving observability, automating operational processes, and driving long-term reliability improvements.
Required Skills & Experience
Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent professional experience.
10+ years of overall IT/SRE experience.
Strong production troubleshooting and observability/APM experience,
5+ years of exclusive, hands-on Oracle Transportation Management (OTM) experience.
Demonstrated experience supporting OTM in both on-premises and SaaS/cloud environments.
Strong knowledge of OTM architecture, configuration, integrations, interfaces, deployments, production support, and troubleshooting.
Hands-on SRE experience across:
Incident response
Monitoring and observability
Root cause analysis
Performance tuning
Automation
Capacity planning
Reliability engineering
Strong scripting and automation skills with Python, Bash, Shell, or similar technologies.
Strong troubleshooting knowledge across applications, infrastructure, networking, databases, APIs, and middleware.
Experience with production incident, problem, and change management within ITIL/SRE environments.
Ability to work effectively with distributed teams in a fast-paced production environment.
Experience with OTM SaaS migrations, upgrades, releases, patching, and environment transitions.
Nice to Have:
Experience supporting large-scale OTM implementations and mission-critical transportation or logistics platforms.
Experience with GCP and cloud-native technologies.
Knowledge of Docker,…