Jobs Companies Bluecatnetworks Site Reliability Engineer – AI-first Platform

Über diese Site Reliability Engineer – AI-first Platform Stelle bei Bluecatnetworks

Bluecatnetworks · Hybrid · Belgrade
We're BlueCat—a Great Place to Work for good reason. Our team solves critical network challenges for some of the world's largest organizations. In simple terms, we manage the systems that keep networks running smoothly, securely, and reliably—the backbone infrastructure that powers digital transformation for enterprises globally. Our Intelligent Network Operations platform delivers and enables AI-driven agentic ops at scale, automating and simplifying how companies manage, secure, and optimize their networks.

But what makes us different is how we work: we believe great work happens in an environment where you're trusted, heard, and supported. With teams globally, we're building a workplace culture that values collaboration and integrity as much as innovation. If you're looking to advance your career with a company that invests in its people, this is it.

About the Team
The BlueCat Horizon team powers all BlueCat SaaS products. Our mission is to deliver enterprise-grade software on a reliable, fast, globally distributed, and cost-efficient cloud platform. We are building an AI-first platform, where agentic systems run as core production workloads on Kubernetes (EKS) and AWS AgentCore.

About the Role
We are looking for a Site Reliability Engineer  to help make our AI agent platform production-ready, secure, observable, and scalable. This role focuses on operating, automating, and hardening systems built by development teams, ensuring experimental AI capabilities are reliably delivered as enterprise-grade production services. You will work across platform engineering, infrastructure, and development teams to improve reliability, automation, and delivery speed.

What You’ll Do
  • Operate and scale Kubernetes (EKS) clusters running AI and cloud-native workloads
  • Support productionization of AWS AgentCore-based systems
  • Design, build, and maintain CI/CD pipelines using GitLab CI/CD for applications and platform services
  • Build and improve automated deployment workflows across environments
  • Implement and maintain observability systems (metrics, logs, traces)
  • Define and enforce SRE practices: SLIs, SLOs, alerting, incident response, postmortems
  • Partner with development teams to ensure services are production-ready and operable
  • Participate in on-call rotation and support incident resolution and root cause analysis
  • Build infrastructure using Terraform for repeatable deployments
  • Improve cost efficiency, performance, and reliability of the platform

What You Bring
  • 5+ years of experience in SRE, DevOps, or Platform Engineering
  • Strong hands-on experience with AWS and Kubernetes (EKS)
  • Experience operating or supporting AWS AgentCore or similar AI/agent platforms
  • Strong proficiency in Python for automation
  • Solid understanding of AWS identity and access management (IAM), including SSO/Identity Center, roles, policies, trust relationships, SCPs, and cross-account access patterns
  • Experience with Infrastructure as Code (Terraform preferred)
  • Strong understanding of CI/CD using GitLab CI/CD
  • Ability to quickly onboard into existing AWS environments and operate independently
  • Experience working with on-call and incident management in production systems

Nice to Have
  • Experience with AI/LLM or agent-based systems
  • Running stateful workloads on Kubernetes (PostgreSQL, Redis, etc.)
  • Knowledge of RAG, vector databases, or knowledge systems
  • Multi-region AWS architecture experience
  • Cost optimization and capacity planning experience

If you share our enthusiasm for the future of our company and are eager to contribute to our vibrant workplace, we look forward to receiving your application! Our comprehensive benefits encompass your health, financial well-being, and overall wellness, and we are committed to providing an exceptional work environment, enriching employee programs, and fostering a remarkable company culture. At our core, we champion values such as transparency, curiosity, respect, and above all, the pursuit of enjoyment.
 
In addition, we offer a range of appealing perks, including:
 
A Professional Development Budget
Dedicated Wellness Days and Wellness Week
A Lifestyle Spending Account
An Employee Recognition Program
 
Join us in shaping the future of our organization, where your talent and dedication can truly thrive. We invite you to apply and become a valuable member of our team!
 
BlueCat is an Equal Opportunity Employer that is committed to inclusion and diversity. We also take affirmative action to offer employment and advancement opportunities to all applicants, including minorities, women, protected veterans, and individuals with disabilities. BlueCat will not discriminate or retaliate against applicants who inquire about, disclose, or discuss their compensation or that of other applicants. 
Bereit, sich bei Bluecatnetworks zu bewerben?
Bei Bluecatnetworks bewerben

Ähnliche Jobs

Unlimit
Site Reliability Engineer (SRE)
Unlimit
⚡ Früh bewerben Belgrade
● Neu 👁 Gesehen ✓ Beworben vor 5 Mon.
Verisign
Site Reliability Engineer - IBM AIX
Verisign
⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Verisign
SRE - Linux
Verisign
⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Verisign
Site Reliability Engineer
Verisign
⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Anduril Industries
Staff Reliability Engineer, Aircraft Systems
Anduril Industries
⚡ Früh bewerben Costa Mesa, California, United... Vor Ort $191,000–$253,000
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Anduril Industries
Systems Reliability Engineer, Surveillance Tower
Anduril Industries
⚡ Früh bewerben Irvine, California, United Sta... Vor Ort $146,000–$194,000
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Anduril Industries
Systems Reliability Engineer, Sentry Tower
Anduril Industries
⚡ Früh bewerben Irvine, California, United Sta... Vor Ort $146,000–$194,000
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Scopely
Senior Platform Engineer (Reliability) - Unannounced Project
Scopely
⚡ Früh bewerben ES - Spain; GB - United Kingdo... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Roku
Senior Machine Learning Engineer, DevOps/SRE
Roku
⚡ Früh bewerben San Jose, California Vor Ort $148,750–$361,000
● Neu 👁 Gesehen ✓ Beworben vor 9 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Bluecatnetworks

Alle Jobs bei Bluecatnetworks ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos