[MLA] Site Reliability Engineer (SRE)

  • Full-time

Company Description

Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities – these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.

Job Description

Project – the aim you'll have 

You will join a product engineering team developing an enterprise identity security platform that helps organizations understand and manage who has access to what data, applications, and resources across cloud, SaaS, on-premises, and custom environments. The team values ownership, collaboration, and engineering excellence, offering the opportunity to work on a business-critical service used by enterprise customers worldwide.

Position – how you’ll contribute

  • Support the deployment, operation, and ongoing maintenance of the Karuna service running on Kubernetes. 
  • Monitor production environments to ensure high availability, reliability, and performance. 
  • Investigate, troubleshoot, and resolve production incidents, performing root cause analysis where appropriate. 
  • Analyze application logs and debug production issues using Splunk. 
  • Perform first-level troubleshooting of UI-related issues involving Web Components, collaborating with front-end engineers when deeper investigation is required. 
  • Support deployment activities and contribute to maintaining and improving CI/CD pipelines. 
  • Identify opportunities for automation and operational improvements. 
  • Work closely with engineering teams and technical stakeholders in an international environment to improve operational processes, enhance service resilience, and optimize observability. 
  • Take ownership of operational tasks and proactively drive issues to resolution while working independently with minimal supervision. 

Qualifications

Expectations – the experience you need

  • Commercial experience as a Site Reliability Engineer, DevOps Engineer, Platform Engineer, or in a similar role. 
  • Hands-on experience supporting production services running on Kubernetes. 
  • Experience monitoring distributed applications and responding to production incidents. 
  • Practical knowledge of log analysis and troubleshooting using Splunk or similar monitoring tools. 
  • Understanding of cloud-native applications and modern operational practices. 
  • Familiarity with CI/CD pipelines and deployment processes. 
  • Basic understanding of Web Components and the ability to perform first-level UI troubleshooting. 
  • Strong analytical and problem-solving skills. 
  • Ability to work independently, prioritize tasks, and make sound technical decisions in ambiguous situations. 
  • Excellent communication skills and confidence collaborating with distributed engineering teams. 
  • Very good spoken and written English. 

Additional skills – the edge you have

  • Experience working with one or more major cloud platforms. 
  • Experience supporting enterprise SaaS or security-focused products.

Additional Information

Our offer – professional development, personal growth:

  • Flexible employment and remote work  
  • International projects with leading global clients 
  • International business trips  
  • Non-corporate atmosphere 
  • Language classes 
  • Internal & external training 
  • Private healthcare and insurance  
  • Multisport card 
  • Well-being initiatives 

Position at: Software Mind

By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply

Privacy Notice