Staff Observability Engineer @Tealium
Software Development
Salary 290,000 - 375,0..
Remote Location
Employment Type full-time
Posted 1wk ago

[Hiring] Staff Observability Engineer @Tealium

1wk ago - Tealium is hiring a remote Staff Observability Engineer. πŸ’Έ Salary: 290,000 - 375,000 pln per year πŸ“Location: Poland

Role Description

We are seeking a Senior or Staff Observability/SRE Engineer to help implement Tealium’s observability strategy across customer-facing products, AI features, platform services, and internal systems. This role requires a strong ownership mentality, strategic thinking, and data-driven decision-making to establish robust reliability practices, telemetry pipelines, and cost controls across conventional and AI-powered systems.

You will focus heavily on AI observability, ensuring model usage, agentic workflows, retrieval systems, and AI-assisted features produce trustworthy, timely, and economically sustainable outcomes. Working cross-functionally with SRE, MLOps, data engineering, security, and product teams, you will leverage modern tools and AI workflows to drive scalable, high-impact reliability solutions across the enterprise.

Your Day to Day

  • Lead end-to-end observability design, telemetry schemas, and OpenTelemetry pipeline architectures across Tealium products, platform services, data pipelines, and internal tools.
  • Partner cross-functionally with engineering and product teams to define service-level indicators, objectives, error budgets, and production-readiness requirements, embodying the Win Together and Trusted Outcomes WOWs.
  • Architect comprehensive AI observability solutions for Amazon Bedrock and agentic workflows, tracking model selection, latency, retries, token usage, tool execution, and cost attribution.
  • Establish end-to-end traceability across user requests, prompts, model calls, retrieved context, downstream services, and final responses while maintaining data privacy and security controls.
  • Develop automated dashboards, alerts, SLOs, runbooks, and operational views to drive rapid incident diagnosis, performance optimization, and economic sustainability.
  • Innovate continuously by defining measurable signals for response quality, groundedness, guardrail outcomes, and task completion across AI systems.
  • Participate in on-call rotation (approximately 20% of the time) and drive proactive failure injection, production testing, and capacity planning.

Qualifications

  • 6+ years in Site Reliability, Observability, or Platform Engineering supporting 24x7x365 production systems.
  • Deep experience with OpenTelemetry and platforms like Datadog, Sumo Logic, Prometheus, or Grafana for telemetry, logging, tracing, and alerting.
  • Experience designing observability for distributed systems, APIs, event-driven architectures, and asynchronous workflows.
  • Hands-on experience instrumenting AI/ML or GenAI systems, including model invocation, prompts, retrieval, tool use, and evaluation.
  • Proficiency with Amazon Bedrock or similar platforms to track model usage, latency, throttling, token metrics, and cost attribution.
  • Familiarity with agentic workflows, prompt engineering, vector databases (e.g., Neptune), RAG architectures, and frameworks like LangChain, LlamaIndex, or SageMaker.
  • Proficiency in Java, Python, or Go.
  • Strong AWS expertise, with an understanding of cloud networking, IAM, security, and service quotas.
  • Solid IaC, CI/CD, and container experience (Terraform, Kubernetes, Argo CD, Jenkins, or GitHub Actions).
  • Experience with data modeling, telemetry pipelines, cardinality management, retention, responsible AI, and data privacy controls.
  • Strong communication, mentoring, and cross-functional leadership skills across engineering, product, and non-technical stakeholders.

Wage Transparency

This position offers a base salary range of 290,000 - 375,000 PLN (Polish zloty) annually. The final offer is determined by job-related skills, experience, and qualifications. The role may also be eligible for a performance-based bonus and equity options.

Benefits

  • Tealium WOWs (Ways of Work), our award-winning culture.
  • Mosaic, our commitment to diversity, equity, and inclusion.
  • Tealium Cares, offering 15 hours of paid work time for volunteer activities annually.
  • Tealium Connects (remote-first working) with new hire stipends for home office support.
  • Tealium Ownership, new hire equity grants.
  • Tealium Time, flexible paid time-off policy and robust leave programs.
  • Healium, health and wellness programs for overall well-being.
  • Tealium LIFT (Learning is Facilitated at Tealium), offering over 6,000 courses for professional development.
  • Health and Related Benefits Programs, offering market competitive benefits.
Before You Apply
️
remote Be aware of the location restriction for this remote position: Poland
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Staff Observability Engineer @Tealium
Software Development
Salary 290,000 - 375,0..
Remote Location
Employment Type full-time
Posted 1wk ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
remote Be aware of the location restriction for this remote position: Poland
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 126,868+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later