← Sottava, jobs the hour they open
1 mo agofound 2 h ago
[MLA] Senior Site Reliability Engineer (SRE) – Kubernetes
What the posting is about
Own production reliability for AI Experience Framework's SSR runtime and metadata platform. Support Kubernetes deployment, monitor service health, troubleshoot issues, and improve reliability. Collaborate with engineering teams and participate in on-call support.
Read out of the posting
Levelsenior
Experience asked5+ years
EmploymentFull time
LocationKraków, Lesser Poland Voivodeship
RemoteYes
Visa sponsorshipNot stated
SalaryNot published, and most postings do not
Posted2026-08-17
Found viasmartrecruiters, direct from their system
We saw it 2 months after it went up.
The posting, as the company wrote it
Employment: Full-time
Experience level: Mid-Senior Level
Job Description
Project – the aim you'll have 
We are the AI Experience Framework team that builds the platform powering ServiceNow's AI-first user interfaces - an SSR runtime (karuna) built on Lit and server-rendered web components, running behind a multi-tier proxy/HTTP2 routing chain with sharded V8 isolate pools, paired with a ServiceNow Glide/Java platform layer (karuna-glide) that supplies metadata, ACLs, and service artifacts. This role owns production reliability for that stack end to end: Kubernetes deployment and operations, observability, and hands-on troubleshooting of both the Node.js and JVM sides of the system - not generalist infrastructure work. 
Position – how you’ll contribute
Support the deployment, operation, and reliability of production services running on Kubernetes.
Monitor service health and investigate production incidents across distributed applications.
Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements.
Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams.
Support CI/CD, GitOps-based deployments, observability, and production monitoring.
Work within a client-directed backlog and established priorities.
Qualifications
Expectations – the experience you need
5+ years of experience in  Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering , or a closely related role, including strong recent hands-on experience supporting Kubernetes-based production services.
3+ years of hands-on production  Kubernetes  experience strongly preferred. Kubernetes production operations, including deployment, scaling, rollout / rollback, resource tuning, and service-to-service troubleshooting
Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene
Splunk  experience for log aggregation, search, and production troubleshooting
Prometheus and Grafana  experience, specifically building alert rules and dashboards, not only using existing dashboards
CI/CD  and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux
Strong Linux and networking fundamentals, including DNS, load balancing, TCP / HTTP, HTTP/2, and Kubernetes networking
Production troubleshooting experience across Node.js and JVM/Java services, with strong depth in at least one runtime environment.  Experience may include Node.js heap snapshots, CPU profiling, event-loop and memory analysis, as well as JVM GC log analysis, thread dumps, JVM tuning, and Java service latency investigation.
Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication
Very good spoken and written English. 
Additional skills – the edge you have
Web Components / Lit experience, to perform first-level debugging of UI-related issues
Server-side rendering or isomorphic runtime experience
Canary rollout / multi-version production operations
Distributed tracing and request-context correlation
KEDA or event-driven autoscaling
Experience with enterprise platform integration layers
Additional Information
Our offer – professional development, personal growth:
Flexible employment and remote work  
International projects with leading global clients 
International business trips  
Non-corporate atmosphere 
Language classes 
Internal & external training 
Private healthcare and insurance  
Multisport card 
Well-being initiatives 
Position at: Software Mind
Company Description
Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities – these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.
Also open at Software Mind
Why this page exists
We read companies’ own hiring systems every hour, 1,182 of them, and show a job the hour it opens instead of when a job board gets around to indexing it. We saw it 2 months after it went up.
The feed is free. No card, no trial to expire.