Observability Architect
Job Summary
We are always looking for amazing talent who can contribute to our growth and deliver results! Geotab is seeking an SRE Observability Architect who will define the strategic vision technical architecture and engineering standards for observability across the organizations cloud platforms. The projects will vary in scope complexity and affected business area. If you love technology and are keen to join an industry leader we would love to hear from you!
As an SRE Observability Architect your key area of responsibility will be defining the foundational observability architecture that enables scalable cost-efficient and highly reliable insight into distributed systems while leading the design of next-generation observability platforms. You will need to work closely with SRE platform engineering and application development teams as well as align with security and compliance stakeholders.
To be successful in this role you will be a strong analytical and systems thinker with the ability to navigate complex ambiguous technical challenges and articulate technical architecture to executive audiences. In addition the successful candidate will have deep expertise in designing enterprise-scale observability platforms and the ability to influence and drive technical direction across multiple teams and organizational boundaries.
Define and own the enterprise-wide observability architecture establishing technical standards reference architectures and multi-year roadmaps.
Evaluate select and standardize observability tooling (e.g. Grafana Prometheus VictoriaMetrics Tempo Loki Elastic Stack OpenTelemetry) to reduce tool sprawl and optimize total cost of ownership.
Design scalable data pipelines and storage strategies capable of ingesting and querying petabyte-scale telemetry data across metrics traces logs and continuous profiling.
Design Terraform modules and Helm charts for declarative observability infrastructure provisioning across multi-cloud environments.
Establish and enforce instrumentation standards using the OpenTelemetry framework including SDK guidelines collector deployment patterns and semantic conventions.
Define and champion SLO/SLI/error-budget frameworks across engineering teams providing architectural guidance on service-level objective implementation.
Serve as a senior escalation point during critical incidents leveraging deep observability expertise to accelerate diagnosis and resolution.
Provide architectural mentorship and technical guidance to Observability Engineers and SRE team members.
5-8 years of experience in Observability Architecture Site Reliability Engineering (SRE) or Platform/Infrastructure Engineering.
Post-secondary Diploma/Degree in Engineering Computer Science or a related field.
Mastery of the OpenTelemetry ecosystem and expert-level knowledge of Prometheus-compatible metrics systems (VictoriaMetrics Thanos etc.).
Advanced experience with tracing systems (Grafana Tempo Jaeger) and log aggregation platforms (Loki Elasticsearch Google BigQuery).
Expert-level proficiency in cloud infrastructure (GCP strongly preferred) and Kubernetes architecture.
Strong software engineering skills in Go Python or similar languages for building cloud-native tooling.
Excellent communication skills with the ability to influence technical direction across organizational boundaries.
Preferred certifications: Google Cloud Professional Cloud Architect or Certified Kubernetes Administrator (CKA).
Flex working arrangements
Home office reimbursement program
Baby bonus & parental leave top up program
Online learning and networking opportunities
Electric vehicle purchase incentive program
Competitive medical and dental benefits
Retirement savings program
*The above are offered to full-time permanent employees only
The annual base salary for this position is the expected annual salary for this role and may be subject to change. Geotab offers various perks and benefits and other compensation components that an individual may be eligible for. The actual base salary for this position depends on a variety of factors such as but not limited to skills qualifications education and overall experience including the location the applicant lives while performing the job. This also includes equity with other team members and alignment with local market data. All offers of employment are contingent upon proof of eligibility to work and the individuals ability to pass a background check.
Hiring Range
$116200 - $155000 CAD
Required Experience:
Staff IC
About Company
Our GPS fleet tracking & management system equips thousands of fleets worldwide with technology to automate, track and manage a truly optimized operation.