Jobs Companies Focused Staff SRE - Observability

Über diese Staff SRE - Observability Stelle bei Focused

Focused · Vor Ort · Chicago, Illinois, United States

 

Who we are:

At Focused, we move quickly to deliver quality software that achieves client outcomes and meets their customer’s needs. We strategically partner with our clients to leverage our expertise in design and software, while our clients bring their own domain expertise. We work with a variety of clients from different industries, collaborating as we get new products to market, modernizing legacy systems, or helping teams learn the skills they need to be successful.   

Our values:

  • Listen first • We are experts in product practices but life long learners in the domain of our customers. We research, collaborate, and understand. 
  • Learn why • We ask questions and talk to users to understand problem spaces, objectives, and goals, which allows us to deeply invest and drive towards the outcomes of our clients. 
  • Love your craft • We love diving into a variety of domains and solving problems.  We take pride in delivering value, in communicating progress, and guiding our clients to success.

We are seeking an experienced Staff Observability Consultant with deep expertise in OpenTelemetry and strong Platform Engineering capabilities to help organizations implement, optimize, and scale their observability infrastructure. This role requires a seasoned consultant who can design comprehensive telemetry strategies, implement distributed tracing solutions, establish robust monitoring practices, and interface closely with clients on the observability journey.

Key Responsibilities:

OpenTelemetry & Observability

  • Design and implement end-to-end OpenTelemetry solutions across diverse technology stacks
  • Configure and deploy OpenTelemetry Collectors for efficient data collection, processing, sampling, and routing
  • Establish telemetry pipelines for metrics, traces, and logs across microservices architectures
  • Optimize collector configurations for performance, reliability, and cost-effectiveness

Platform Engineering & Infrastructure

  • Augment existing infrastructure with with integrated observability solutions
  • Implement Infrastructure as Code (IaC) solutions using Terraform, Pulumi, CloudFormation, etc.
  • Architect and manage Kubernetes clusters with comprehensive monitoring and logging
  • Build CI/CD pipelines with embedded observability and automated testing

Site Reliability Engineering (SRE)

  • Establish and maintain Service Level Indicators (SLIs), Objectives (SLOs), and Agreements (SLAs)
  • Implement error budgets, toil reduction strategies, and capacity planning
  • Support incident response procedures and post-mortem processes

Cloud & DevOps Engineering

  • Deploy and manage observability infrastructure across AWS, GCP, and Azure
  • Establish security, compliance, and governance frameworks for telemetry data
  • Experience automating Agent Evaluations in CI/CD pipelines and observability backends.

Required Qualifications:

Core Observability & OpenTelemetry

  • 3-7 years of experience in observability, monitoring, and distributed systems
  • Deep hands-on experience with OpenTelemetry ecosystem, including SDKs, APIs, and specifications
  • Proficiency with OpenTelemetry Collector configuration, processors, exporters, and receivers
  • Strong understanding of telemetry data models, semantic conventions, and instrumentation best practices

Platform Engineering & DevOps

  • 5+ years of Platform Engineering or DevOps experience with focus on site reliability, observability, and incident response
  • Proficiency with Infrastructure as Code tools (Terraform, Pulumi, CloudFormation, CDK)
  • Strong experience with CI/CD platforms (GitHub Actions, GitLab CI, Jenkins, ArgoCD)

Cloud & Infrastructure

  • Hands-on experience with major cloud providers (AWS, GCP, Azure) and their observability services
  • Experience with container technologies (Docker, Podman) and container registries
  • Knowledge of networking, security, load balancing, and distributed systems concepts

Site Reliability Engineering

  • Experience implementing SRE practices including error budgets and toil metrics
  • Proficiency in incident management, on-call procedures, and post-mortem culture
  • Experience with capacity planning, performance optimization, and scalability design

Programming & Automation

  • Proficiency in multiple programming languages preferred (Go, Python, Java, Node.js, Rust)
  • Strong scripting and automation skills (Bash, Python, PowerShell)
  • Understanding of software engineering best practices and testing methodologies

Preferred Qualifications (Exceptional Candidates)

AI & Agentic Frameworks

  • Understanding of Large Language Models (LLMs) and their application in DevOps
  • Knowledge of vector databases, embeddings, and retrieval-augmented generation (RAG)
  • Experience with AI/ML model deployment and monitoring in production environments

Leadership & Communication

  • Strong technical writing and documentation skills
  • Ability to present complex technical concepts to diverse stakeholders
  • A passion for knowledge sharing

Key Competencies

  • Systems thinking and ability to design holistic observability solutions
  • Strong analytical and troubleshooting skills for complex distributed systems
  • Curiosity about emerging technologies, particularly AI applications in operations
  • Adaptability to rapidly evolving cloud-native and observability technologies
  • Collaborative mindset with focus on enabling developer productivity and system reliability

What Sets Exceptional Candidates Apart:

  • Experience with Honeycomb
  • Contributions to open-source observability or AI framework projects
  • Track record of implementing platform engineering solutions that significantly improved developer experience
  • Experience scaling observability infrastructure to handle high event volume

What to know before you apply: 

  • This role will require being in the Chicago office three days per week and up to 20% travel within the United States.
  • Focused is unable to sponsor or take over sponsorship of the employment Visa process at this time.
  • The Chicago base salary range for this role is $160,000 - $200,000.
Bereit, sich bei Focused zu bewerben?
Bei Focused bewerben

Wie sich dieses Gehalt für SRE vergleicht

Diese Stelle zahlt $180,000/yrunter der üblichen Spanne für SRE Stellen.

$148,665 dem Median $180,000 $210,250

Übliche Spanne $180,000–$187,500/yr, aus 6 vergleichbaren SRE Anzeigen auf JobsRadar (Vergütung auf USD hochgerechnet). Gehaltseinblicke für SRE ansehen →

Über Focused

Careers with Focus

Hello, we're Focused

At Focused, we take a unique approach to developing high-quality, business-focused, software. We believe that digital products can and should be built to evolve with your business. Our approach is structured around delivering products to market fast, testing with real customers, and iterating based on their feedback.

We work with people who are the best at what they do and who care about making others the best at what they do too. We want to be great people to work with first—who just happen to be exceptional at building software.

Our values:

  • Listen first - Every decision we make is informed by deep listening. We hear every perspective, and we keep listening to find the right solution.
  • Learn why - We keep digging, keep asking questions, and learn new skills continuously to make ourselves and our team better. 
  • Love your craft - We believe in becoming the best at what we do, which means finding the best answers to the hardest problems—not the expected ones.

See yourself working here? Join our team by applying below!

Alle Jobs bei Focused ansehen →

Ähnliche Jobs

Ripple
Site Reliability Engineer, Observability
Ripple
⚡ Früh bewerben Chicago, Illinois, United Stat... Vor Ort $160,000–$200,000
● Neu 👁 Gesehen ✓ Beworben vor 21 Std.
Okta
Staff Site Reliability Engineer - Kubernetes
Okta
⚡ Früh bewerben Bellevue, Washington; Chicago,... Vor Ort $194,000–$267,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Okta
Senior Database Reliability Engineer (DBRE)
Okta
⚡ Früh bewerben Bellevue, Washington; Chicago,... Vor Ort $160,000–$220,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Ripple
Senior Site Reliability Engineer, Observability
Ripple
⚡ Früh bewerben Chicago, Illinois, United Stat... Vor Ort $160,000–$200,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Advanced Technology Services
Reliability Engineer - Industrial Maintenance
Advanced Technology Services
⚡ Früh bewerben United States- Chicago, Illino... Vor Ort $102,970–$131,690
● Neu 👁 Gesehen ✓ Beworben vor 2 Wo.
TransMarket Group
DevOps/SRE Intern
TransMarket Group
⚡ Früh bewerben Chicago, Illinois, United Stat...
● Neu 👁 Gesehen ✓ Beworben vor 1 Mon.
VBP
Site Reliability Engineer
VBP
⚡ Früh bewerben Cebu City, Cebu, Philippines Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
VBP
Site Reliability Specialist
VBP
⚡ Früh bewerben Cebu City, Cebu, Philippines Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Fluidstack
Principal Operations Engineer, Reliability
Fluidstack
⚡ Früh bewerben Austin, TX Vor Ort $242,000–$278,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Focused

Alle Jobs bei Focused ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos