Tazapay - Staff DevOps Engineer

Tazapay โ€” United States ยท Posted ~2 days ago

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Description

Responsibilities (Not Exhaustive) Make architectural tradeoffs, design and deliver highly scalable, reliable, secure and fault tolerant cloud infrastructure.Be a role model for DevOps and platform engineers, mentor engineers.Demonstrate technical leadership and create impact across teams.Lead AI strategy and adoption for productivity, infrastructure automation and operational efficiency.Drive infrastructure and deployment practices while being secure and compliant (PCI-DSS, SOC2 and others).Participate in infrastructure, architecture and security design reviews to maintain our high engineering standards.Partner with engineering and product management teams to define and execute the platform roadmap.Translate business and engineering requirements into scalable and extensible infrastructure design.Proactively manage stakeholder communication related to deliverables, risks, changes and dependencies.Coordinate with cross functional teams (Backend, Frontend, Data, Security, QA etc.) on planning and execution.Continuously improve platform reliability, developer experience, deployment velocity and operational excellence.Engage in capacity and demand planning, system performance analysis, tuning and cost optimization.Lead production outages, incident response, post-mortems and drive SRE practices across the engineering organisation. The Ideal Candidate Education Degree in Computer Science or equivalent (B.E/B.Tech or higher) with 10+ years of experience in DevOps, SRE, Platform or Cloud Engineering roles for large distributed systems in reputed organizations. Requirements Must Have Hands-on experience in designing, building and operating cloud infrastructure for large scale production systems on AWS.Deep knowledge of Linux as a production environment.Strong knowledge of distributed systems, networking, systems internals and asynchronous architectures.Expert in at least 1 of the following languages for automation and tooling : go, python, shell scripting.Extensive experience with AWS services (ECS, EC2, VPC, IAM, RDS, Kinesis, Secrets Manager, SSM, WAF and more).Infrastructure as Code expertise with Terraform, CDK or CloudFormation.Strong experience with CI/CD tooling such as GitHub Actions, GitLab CI, Jenkins, and GitOps tools.Hands-on experience with observability stacks - Prometheus, Grafana, OpenTelemetry, distributed tracing and log aggregation.Experience with container infrastructure - Docker, container runtimes and image security.Experience administering and operating RDBMS/NoSQL systems at scale, such as Postgres, MongoDB and Redis.Ability to design and operate low latency services behind load balancers and API gateways.Strong understanding of system performance, scaling and reliability engineering.Possess excellent communication, sharp analytical abilities with proven design skills, able to think critically of the current platform in terms of growth and stability.Experience with microservice architecture and service-to-service communication patterns.Continuously refactor infrastructure and tooling to ensure high-quality design.Ability to plan, prioritize, estimate and execute platform releases with good degree of predictability.Ability to scope, review and refine user stories for technical completeness and to alleviate dependency risks.Passion for learning new things, solving challenging problems. Nice To Have Prior experience with fintech and payments.Expert level proficiency with Kubernetes in production.Prior experience operating infrastructure for stablecoins, blockchain or cryptocurrency platforms.AWS Cloud Certifications.Familiarity with security standards and compliance - PCI-DSS, SOC2, OWASP, static code analysis.Familiarity with data engineering services - Clickhouse, Dbt ETL.Experience working for a start-up. (ref:hirist.tech)