pracaon.plpracaon.pl

Site Reliability Engineer (SRE) – Migration Operations IRC305279

Wrocław, Polska
6Tg
Gehalt nach Vereinbarung
Vollzeit • Remote • IT, Daten & KI

Wichtige Merkmale des Angebots

  • Backend: Java / .NET / Node / Python

  • Auf der Suche nach Experten – Senior/Experte

  • Vollzeit

  • Remote-Arbeit - kein Pendeln

Description

We are seeking a Senior SRE to join our Fleet Operations Team. The team ensures uninterrupted service reliability and system health during major hardware and software platform maintenance cycles. You will be responsible for operational readiness, performance monitoring, and incident management while migrating workloads to new clusters. Eager to know more? If you are proactive, bring new ideas and suggestions, don’t waste any second and apply!

Skills

  • Bash

  • DevOps & CI/CD

  • Infrastructure as Code (IaC)

  • Infrastructure Monitoring

  • Linux

  • Logging & Log Management

  • Networking

  • Python

  • Saltstack

About GlobalLogic

  • GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

What we offer

  • Empowering Projects: With 500+ clients spanning diverse industries and domains, we provide an exciting opportunity to contribute to groundbreaking projects that leverage cutting-edge technologies. As a team, we engineer digital products that positively impact people’s lives.

  • Empowering Growth: We foster a culture of continuous learning and professional development. Our dedication is to provide timely and comprehensive assistance for every consultant through our dedicated Learning & Development team, ensuring their continuous growth and success.

  • DE&I Matters: At GlobalLogic, we deeply value and embrace diversity. We are dedicated to providing equal opportunities for all individuals, fostering an inclusive and empowering work environment.

  • Career Development: Our corporate culture places a strong emphasis on career development, offering abundant opportunities for growth. Regular interactions with our teams ensure their engagement, motivation, and recognition. We empower our team members to pursue their career goals with confidence and enthusiasm.

  • Comprehensive Benefits: In addition to equitable compensation, we provide a comprehensive benefits package that prioritizes the overall well-being of our consultants. We genuinely care about their health and strive to create a positive work environment.

  • Flexible Opportunities: At GlobalLogic, we prioritize work-life balance by offering flexible opportunities tailored to your lifestyle. Explore relocation and rotation options for diverse cultural and professional experiences in different countries with our company.

Experience

  • 3-5 years

Requirements

  • Preferred qualifications & skills

  • Education & experience: Bachelor of Science in Computer Science, Systems Engineering, or a related field, plus heavy experience in production system administration and incident response.

  • System Operations: Expertise in Linux system administration, storage mounting, network routing, and process monitoring.

  • Monitoring & Tooling: Experience using Prometheus, Grafana, and log aggregators to monitor high-throughput data operations.

  • Automation: Proficiency in Python, Bash, and SaltStack.

  • Nice to have skills:.

  • Experience managing live data center maintenance windows and fleet migrations

Job responsibilities

  • Monitor and maintain host stability, network bandwidth, and storage health during active fleet migrations.

  • Develop monitoring dashboards and alert triggers to detect migration degradation or hardware failures.

  • Execute operational runbooks and coordinate technical mitigation during migration maintenance windows.

  • Automate infrastructure recovery steps for failed VM migration attempts.

Stichwörter / Fähigkeiten

Bash
DevOps & CI/CD
Infrastructure as Code (IaC)
Infrastructure Monitoring
Linux
Logging & Log Management
Networking
Python
Saltstack
Das Angebot wurde von einem externen Portal importiert.Anzeigenquelle

Weitere ähnliche Anzeigen