JustJoin.IT Praca zdalna Senior New

Senior Engineer - SRE & Infrastructure Services

EPAM Systems

⚲ Poznan

Do uzgodnienia

Wymagania

  • Kubernetes
  • CI/CD
  • AWS
  • flux
  • Jenkins
  • Prometheus
  • Grafana
  • Python

Opis stanowiska

We are seeking a Senior Engineer to join our SRE/DevOps organization. The successful candidate will design and implement automated provisioning, deployment, management, and monitoring solutions for a large-scale, rapidly evolving portfolio of SaaS services. Working closely with architecture and development teams, the Senior Engineer will contribute to engineering standards and best practices, drive CI/CD and observability improvements, and support the team in delivering reliable, scalable infrastructure.

Responsibilities
• Design and implement CI/CD pipelines leveraging Kubernetes, Flux, and related cloud-native technologies
• Implement and maintain monitoring and management solutions for cloud-based products using a combination of commercial off-the-shelf (COTS) and in-house tooling
• Collaborate with architects and development teams on standardized, scalable approaches for log management, service components, and infrastructure elements
• Evaluate and integrate AI-enabled tooling across observability, pipeline efficiency, and SRE troubleshooting workflows
• Develop DevOps tooling that reduces manual toil, strengthens security posture, and minimizes human error
• Build and maintain resilient, self-scaling systems that minimize customer impact while supporting a sustainable operational environment
• Participate in incident response, root cause analysis, and post-incident review processes

Requirements
• Minimum 7 years of professional experience in software development, DevOps, and/or Site Reliability Engineering
• Minimum 3 years of experience building and maintaining CI/CD pipelines and SRE automation within cloud environments at scale
• Experience with monitoring and alerting platforms (e.g., PagerDuty, Prometheus, Grafana)
• Hands-on experience deploying and managing cloud infrastructure
• Experience with Amazon Web Services (e.g., EC2, Elasticsearch, Lambda, CloudFormation)
• Working experience with at least one additional cloud provider (GCP, Azure, or OCI)
• Experience with CI/CD toolchains (Jenkins, Kubernetes, Flux)
• Proficiency in one or more of the following languages: Python, Go, Java, or C
• Minimum English language level of B1+

Nice to have
• Experience applying Generative AI or ML-based tooling within an operations context
• Experience with DevSecOps practices and security automation
• Background in architecting monitoring and management systems for enterprise SaaS products

We offer
• We gather like-minded people:
• Top tech minds driving innovation in AI, cloud and digital platform modernization
• Supportive team and agile, startup-like culture
• Hybrid by design mode and opportunity to work remotely within Poland
• Chance to work abroad for up to 60 days annually
• Business-driven relocation opportunities
• We provide growth opportunities:
• Career development programs
• Thought leadership, mentoring, soft skills and well-being programs
• Certification (Anthropic, Gemini, GCP, Azure, AWS)
• English classes
• We cover it all:
• Stable pay
• Participation in the Employee Stock Purchase Plan with a 15% discount
• Benefits package (health insurance, multisport, shopping vouchers)
• Referral bonuses up to $2,000
• Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and more
• Corporate, social and well-being events
• Please, note:
• Benefits listed above are available to employees only
• We are open for working with Contractors. Terms of B2B cooperation agreements are agreed individually
• We will reach out to selected candidates exclusively
EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.

🔍 Dekoder Ogłoszenia

🔴
large-scale, rapidly evolving portfolio of SaaS services
Może oznaczać zarówno ekscytujący rozwój, jak i potencjalnie chaotyczne zmiany i presję na szybkie dostarczanie.
🔴
contribute to engineering standards and best practices
Może oznaczać tworzenie nowych standardów od zera w nieustabilizowanym środowisku, a nie tylko ich przestrzeganie.
🔴
Evaluate and integrate AI-enabled tooling
Może oznaczać eksperymentowanie z nowymi, niedojrzałymi narzędziami, które wymagają dużo pracy wdrożeniowej i testowej.
🔴
Develop DevOps tooling that reduces manual toil
Sugestia, że obecne procesy są bardzo manualne i wymagają znaczącej automatyzacji, co może być czasochłonne.
🟡
sustainable operational environment
Może oznaczać równowagę między wydajnością a unikaniem wypalenia, ale równie dobrze może być eufemizmem dla utrzymania stabilności przy ograniczonych zasobach.