Lead DevOps Engineer
⚲ Warszawa, Trójmiasto, Poznań, Wrocław, Kraków
Do uzgodnienia
Wymagania
- GCP
- GKE
- Terraform
- GitHub Actions
- CI/CD
- Grafana
- Loki
- Networking
- Troubleshooting
Opis stanowiska
Principal MLOps / Platform Engineer
Poland, Remote or Ukraine, Remote or Romania, Remote
Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions. Our vision is to become THE go-to resource for every cyber protection need individuals may face - today and in the future.
Join us for the ride!
We’re looking for a Lead DevOps Engineer to architect, automate, and lead our cloud infrastructure and continuous delivery strategy on Google Cloud. In this role, you will own the reliability, scalability, and security of our production environments, ensuring zero-downtime operations and seamless deployment pipelines.
You will drive our DevOps practices forward by codifying infrastructure, automating release management, and enforcing strict security and compliance standards. We are a high-trust, outcome-focused team that moves quickly to solve complex, large-scale infrastructure and operational challenges on GCP.
Core Responsibilities:
•
GCP Infrastructure & Architecture: Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
• Infrastructure as Code (IaC): Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
• CI/CD Pipeline Automation: Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.
• Networking & Traffic Management: Manage enterprise networking perimeters, service mesh architectures (Istio mTLS, VirtualServices), and API Gateways (Kong) for secure ingress, egress, and microservices traffic.
• System Observability & Reliability: Implement comprehensive telemetry, log aggregation, and alerting to ensure high availability, optimal resource utilization, and fast incident response.
• Security, Compliance & Operations: Enforce strict security postures (VPC-SC, IAM, Org Policies), manage identity and secrets storage, maintain compliance standards (such as SOC2), and participate in the operational on-call rotation.
What you bring to the table:
•
Deep GCP & Kubernetes Mastery: Senior/Lead level experience operating production workloads on GCP and GKE (cluster topology, Helm, Kustomize, and ArgoCD).
• Infrastructure as Code (IaC): Expert-level Terraform skills with extensive experience running GitOps workflows at scale (specifically via Atlantis).
• CI/CD Expertise: Strong track record of designing, building, and maintaining multi-stage automated deployment pipelines.
• Networking & Service Mesh: High proficiency with Istio (VirtualServices, mTLS, sidecar injection) and API Gateways (specifically Kong).
• Enterprise Security & Identity: Hands-on experience enforcing enterprise-grade security and secret management (Auth0, Dex, ESO, SOPS, VPC-SC, IAM).
• Database & Datastore Familiarity: Solid operational comfort with Cloud SQL (PostgreSQL), BigQuery, and in-cluster datastores like Elasticsearch or ClickHouse.
• At least an upper-intermediate level of spoken and written English.
It would be great if you also had:
•
Automation Savvy: Experience with Ansible for host configuration, cluster bootstrapping, and disaster recovery.
• Advanced Certifications: GCP Professional Cloud Architect, GCP Professional Security Engineer, or Kubernetes certifications (CKA / CKS).
• Telemetry & Monitoring: Practical experience with Loki, Grafana, or ClickHouse for centralized logging and metrics.
Why Join Us?
• Professional Growth: Enhance your skills by collaborating closely with talented engineers and architects.
• Meaningful Impact: Help create software that protects millions of users from cybersecurity threats.
• Supportive Environment: Join a team dedicated to innovation, mentorship, and continuous learning.
Poland, Remote or Ukraine, Remote or Romania, Remote
Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions. Our vision is to become THE go-to resource for every cyber protection need individuals may face - today and in the future.
Join us for the ride!
We’re looking for a Lead DevOps Engineer to architect, automate, and lead our cloud infrastructure and continuous delivery strategy on Google Cloud. In this role, you will own the reliability, scalability, and security of our production environments, ensuring zero-downtime operations and seamless deployment pipelines.
You will drive our DevOps practices forward by codifying infrastructure, automating release management, and enforcing strict security and compliance standards. We are a high-trust, outcome-focused team that moves quickly to solve complex, large-scale infrastructure and operational challenges on GCP.
Core Responsibilities:
•
GCP Infrastructure & Architecture: Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
• Infrastructure as Code (IaC): Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
• CI/CD Pipeline Automation: Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.
• Networking & Traffic Management: Manage enterprise networking perimeters, service mesh architectures (Istio mTLS, VirtualServices), and API Gateways (Kong) for secure ingress, egress, and microservices traffic.
• System Observability & Reliability: Implement comprehensive telemetry, log aggregation, and alerting to ensure high availability, optimal resource utilization, and fast incident response.
• Security, Compliance & Operations: Enforce strict security postures (VPC-SC, IAM, Org Policies), manage identity and secrets storage, maintain compliance standards (such as SOC2), and participate in the operational on-call rotation.
What you bring to the table:
•
Deep GCP & Kubernetes Mastery: Senior/Lead level experience operating production workloads on GCP and GKE (cluster topology, Helm, Kustomize, and ArgoCD).
• Infrastructure as Code (IaC): Expert-level Terraform skills with extensive experience running GitOps workflows at scale (specifically via Atlantis).
• CI/CD Expertise: Strong track record of designing, building, and maintaining multi-stage automated deployment pipelines.
• Networking & Service Mesh: High proficiency with Istio (VirtualServices, mTLS, sidecar injection) and API Gateways (specifically Kong).
• Enterprise Security & Identity: Hands-on experience enforcing enterprise-grade security and secret management (Auth0, Dex, ESO, SOPS, VPC-SC, IAM).
• Database & Datastore Familiarity: Solid operational comfort with Cloud SQL (PostgreSQL), BigQuery, and in-cluster datastores like Elasticsearch or ClickHouse.
• At least an upper-intermediate level of spoken and written English.
It would be great if you also had:
•
Automation Savvy: Experience with Ansible for host configuration, cluster bootstrapping, and disaster recovery.
• Advanced Certifications: GCP Professional Cloud Architect, GCP Professional Security Engineer, or Kubernetes certifications (CKA / CKS).
• Telemetry & Monitoring: Practical experience with Loki, Grafana, or ClickHouse for centralized logging and metrics.
Why Join Us?
• Professional Growth: Enhance your skills by collaborating closely with talented engineers and architects.
• Meaningful Impact: Help create software that protects millions of users from cybersecurity threats.
• Supportive Environment: Join a team dedicated to innovation, mentorship, and continuous learning.
🔍 Dekoder Ogłoszenia
🔴
Lead DevOps Engineer
Chociaż tytuł sugeruje rolę przywódczą, opis bardziej skupia się na technicznym aspekcie inżynierii platformy i MLOps, co może oznaczać mniejszy nacisk na zarządzanie zespołem.
🔴
architect, automate, and lead our cloud infrastructure and continuous delivery strategy on Google Cloud
Słowo 'lead' może sugerować odpowiedzialność za strategię i kierunek, ale bez jasnego określenia zarządzania zespołem, może oznaczać głównie techniczne przewodnictwo i inicjatywę.
🔴
own the reliability, scalability, and security of our production environments
Posiadanie odpowiedzialności za kluczowe aspekty środowiska produkcyjnego może oznaczać dużą presję i konieczność bycia dostępnym w sytuacjach kryzysowych.
🔴
high-trust, outcome-focused team that moves quickly
Określenie 'moves quickly' w połączeniu z 'outcome-focused' może sugerować szybkie tempo pracy i nacisk na dostarczanie wyników, co czasem może oznaczać mniej czasu na dogłębne planowanie lub procesy.
🟡
Join us for the ride!
Jest to typowe hasło marketingowe, które może sugerować ekscytującą podróż rozwoju firmy, ale nie dostarcza konkretnych informacji o codziennej pracy czy wyzwaniach.