Senior SRE Engineer (Observability Focus)
Capital.com โ Poland ยท Posted ~1 day ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
We are a leading trading platform that is ambitiously expanding to the four corners of the globe.
Our top-rated products have won prestigious industry awards for their cutting-edge technology and seamless client experience.
We deliver only the best, so we are always in search of the best people to join our ever-growing talented team.
We're building out our observability practice and need a senior engineer who can own it end to end.
This is a hands-on role.
You'll design and operate the telemetry stack that gives our engineering teams real visibility into production โ across a hybrid AWS and on-premise environment, at scale.
Responsibilities:
Own the full observability stack: metrics (VictoriaMetrics), logs (OpenSearch), and traces (OpenTelemetry) โ from pipeline design to day-2 operationsArchitect and run VictoriaMetrics cluster topology (vmstorage/vminsert/vmselect), including vmagent scraping, remote write configuration, vmalert rules, and cardinality controlOperate OpenSearch clusters: index lifecycle management (ISM), hot-warm-cold architecture, shard tuning, and ingest pipelines via Data PrepperBuild and maintain OTEL Collector pipelines โ receivers, processors, exporters โ and instrument services across Java, Python, and JS/TS stacks (auto and manual)Run Kafka as the telemetry transport layer (OTEL Collector โ Kafka โ backends), including topic design, partition strategy, consumer group lag monitoring, and throughput tuning for high-volume telemetryManage log shipping infrastructure using Fluent Bit, Vector, or Fluentd; define structured logging standards and field normalization across servicesBuild Grafana dashboards and alerting that engineers actually use โ clear, actionable, with well-structured variables and thresholdsWork with platform and application teams to improve sampling strategies (head/tail), batching, and context propagation across distributed servicesContribute to incident response, post-mortems, and reliability improvements driven by observability signalsMentor engineers on observability practices, tooling, and structured logging standards
Requirements:
6+ years in a DevOps, SRE, or platform engineering role, with at least 2 years focused on observability tooling at production scaleDeep hands-on experience with VictoriaMetrics (or Prometheus) โ MetricsQL/PromQL, exporters, service discovery, remote write, downsampling, and retention managementSolid OpenSearch or Elasticsearch skills: cluster operations, Query DSL, ISM policies, and ingest pipeline designProduction experience with OpenTelemetry: Collector configuration, OTLP, context propagation, and instrumentation across multiple languagesStrong Kafka skills โ producer/consumer patterns, consumer group management, Kafka Connect, Schema Registry, and JMX-based monitoring.
Strimzi experience a plus if you've run Kafka on KubernetesProficiency with log shippers (Fluent Bit, Vector, Fluentd) and structured log parsing/normalizationWorking knowledge of Kubernetes (operators, Helm), Argo CD/GitOps, and Terraform/AnsibleComfortable in a hybrid AWS + on-prem environment; solid understanding of networking as it applies to scraping and shipping pipelinesScripting ability in Bash or Python for automation and toolingStrong communication skills โ you can explain observability tradeoffs clearly to engineers and non-engineers alikeEnglish proficiency
What you will get in return:
Competitive Salary: We believe great work deserves great pay! Your skills and talents will be rewarded with a salary that makes you feel valued and motivated Work-Life Harmony: Join a company that genuinely cares about you - because your life outside of work matters just as much as your time on the clock Generous Time Off: Need a breather? Our annual leave policy lets you recharge and enjoy life outside of work without a worry Employee Referral Program: Love working here? Share the love! Bring your talented friends on board and get rewarded for growing our awesome team Comprehensive Health & Pension Benefits: From medical insurance to pension plans, weโve got your back.
Plus, location-specific benefits and perks! Workation Wonderland: Live your digital nomad dreams with 30 extra days to work remotely from anywhere in the world (some restrictions apply).
Adventure awaits! Volunteer Days: Make a difference! Take two additional paid days each year to support causes you care about and give back to the community
Be a key player at the forefront of the digital assets movement, propelling your career to new heights! Join a dynamic and rapidly expanding company that values and rewards talent, initiative, and creativity.
Work alongside one of the most brilliant teams in the industry.
Our company has an Internal Reporting Procedure.
It is available from the Human Resources Department upon request hr@capital.com.
You may report a violation referred to in the Procedure under the terms specified therein.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses.
These tools assist our recruitment team but do not replace human judgment.
Final hiring decisions are ultimately made by humans.
If you would like more information about how your data is processed, please contact us.
We have 66,203 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 66,203 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume