Summary
✨ AI‑Generated
A senior reliability engineering role focused on designing, operating, and improving large-scale enterprise application environments. The position requires deep experience with transportation management platforms, databases, programming, automation, monitoring, and incident resolution.
Highlights
Opportunity to engineer highly available enterprise systems, improve reliability, automate operations, and solve complex production challenges in a critical technology environment.
Description
About the Role
Design, implement, and maintain highly available, scalable, and reliable Oracle Transportation Management (OTM) environments across on premises and OTM SaaS platforms through SRE frameworks.
Responsibilities
Bring a minimum of 7-10 years of hands on, dedicated OTM experience, with strong expertise in OTM application Knowledge, production, integrations, deployments, and troubleshooting skills.Strong knowledge on Oracle database and Java programming.Perform SRE activities including monitoring, incident management, capacity planning, performance optimization, automation, problem management, and reliability engineering for OTM environments.Monitor application health, infrastructure performance, integrations, and critical business processes using observability and monitoring tools.Troubleshoot complex OTM application, middleware, database, integration, and infrastructure issues, driving incidents through resolution and minimizing business impact.Support OTM SaaS release, patching, configuration, deployment, and environment management activities while ensuring production stability and reliability.Work closely with application, development, infrastructure, database, integration, and cloud teams to troubleshoot issues and deliver reliable OTM solutions.Leverage Snowflake Observe and other observability capabilities to analyze application and platform health, identify trends, detect anomalies, and improve system performance.Implement and maintain Service Level Objectives (SLOs), Service Level Indicators (SLIs), alerting, dashboards, and reliability metrics for critical OTM services.Apply SRE best practices for automation, observability, security, compliance, capacity management, and disaster recovery across OTM environments.Support cloud based OTM SaaS solutions and contribute to cloud engineering initiatives, with Google Cloud Platform (GCP) experience considered an added advantage.
Qualifications
Bachelor’s degree in computer science, Engineering, Information Technology, or a related field, or equivalent professional experience.At least 7-10 years of overall IT/SRE experience, includingAt least 5 years of exclusive, hands on experience with Oracle Transportation Management (OTM).
Required Skills
Proven experience working with OTM both on premises and SaaS/cloud environments.Hands on experience performing SRE functions such as incident response, monitoring, observability, RCA, performance tuning, automation, capacity planning, and reliability improvement.Experience with cloud/SaaS platforms, preferably Oracle Cloud Infrastructure (OCI) and Oracle OTM SaaS.Experience with observability and monitoring platforms such as Snowflake Observe, OEM Monitoring or equivalent tools.Strong scripting and automation skills using Python, Bash, Shell, or similar technologies.Experience with Infrastructure as Code (IaC), configuration management, or automation tools such as Terraform, Puppet, or Ansible.Experience with OTM SaaS migrations, upgrades, releases, patching, and environment transitions.Experience with GCP, cloud native technologies, containers, Docker, or Kubernetes.Experience with Java and .NET applications and enterprise integration technologies.Experience with Kafka, REST/SOAP APIs, enterprise integrations, databases, and distributed systems.
Preferred Skills
Google Cloud Platform (GCP) experience considered an added advantage.