[MLA] Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind

Kraków, Lesser Poland Voivodeship, Poland, Remote Remote Full-time 1 mo ago
Workplace
Remote
Location
Kraków, Lesser Poland Voivodeship, Poland
Who can apply
Remote for people based in Poland

Check your CV against this job

Free · No signup · A 0–100 ATS match score and the keywords you're missing.

Check my CV free

$14.99/month, cancel anytime. Already have an account? Log in

About the role

Job Description

Project – the aim you'll have

We are the AI Experience Framework team that builds the platform powering ServiceNow's AI-first user interfaces - an SSR runtime (karuna) built on Lit and server-rendered web components, running behind a multi-tier proxy/HTTP2 routing chain with sharded V8 isolate pools, paired with a ServiceNow Glide/Java platform layer (karuna-glide) that supplies metadata, ACLs, and service artifacts. This role owns production reliability for that stack end to end: Kubernetes deployment and operations, observability, and hands-on troubleshooting of both the Node.js and JVM sides of the system - not generalist infrastructure work.

Position – how you’ll contribute

* Support the deployment, operation, and reliability of production services running on Kubernetes. * Monitor service health and investigate production incidents across distributed applications. * Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements. * Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams. * Support CI/CD, GitOps-based deployments, observability, and production monitoring. * Work within a client-directed backlog and established priorities.

Qualifications

Expectations – the experience you need

* 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering, or a closely related role, including strong recent hands-on experience supporting Kubernetes-based production services. * 3+ years of hands-on production Kubernetes experience strongly preferred. Kubernetes production operations, including deployment, scaling, rollout / rollback, resource tuning, and service-to-service troubleshooting * Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene * Splunk experience for log aggregation, search, and production troubleshooting * Prometheus and Grafana experience, specifically building alert rules and dashboards, not only using existing dashboards * CI/CD and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux * Strong Linux and networking fundamentals, including DNS, load balancing, TCP / HTTP, HTTP/2, and Kubernetes networking * Production troubleshooting experience across Node.js and JVM/Java services, with strong depth in at least one runtime environment. Experience may include Node.js heap snapshots, CPU profiling, event-loop and memory analysis, as well as JVM GC log analysis, thread dumps, JVM tuning, and Java service latency investigation. * Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication * Very good spoken and written English.

Additional skills – the edge you have

* Web Components / Lit experience, to perform first-level debugging of UI-related issues * Server-side rendering or isomorphic runtime experience * Canary rollout / multi-version production operations * Distributed tracing and request-context correlation * KEDA or event-driven autoscaling * Experience with enterprise platform integration layers

Additional Information

Our offer – professional development, personal growth:

* Flexible employment and remote work   * International projects with leading global clients  * International business trips   * Non-corporate atmosphere  * Language classes  * Internal & external training  * Private healthcare and insurance   * Multisport card  * Well-being initiatives

Position at: Software Mind

Get matched & apply with FindAJobAI

Upload your resume once. We score every job against your profile, tailor your resume and cover letter, and autofill the application.