Rapyd logo

NOC Team Leader

Rapyd

Tel Aviv-Yafo, Tel Aviv District, IsraelHybridFull-timeNo compensation foundPosted 1mo agoVerified open 3 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
No compensation found
Location
Tel Aviv-Yafo, Tel Aviv District, IsraelHybrid
Schedule
Full-time
Work Authorization
Not specified

Job overview

Rapyd is hiring a NOC Team Leader. The NOC Team Lead at Rapyd is responsible for ensuring the high availability, reliability, and performance of the company's production systems, infrastructure, and applications. This hybrid role combines traditional reactive monitoring with proactive, automated, and software-centric engineering practices, bridging the gap between Network Operations Center (NOC) and Site Reliability Engineering (SRE). The role involves leading global NOC teams, implementing monitoring tools, automating tasks, and managing incidents.

Key focus areas include Direct 24/7 global NOC teams, managing incident response, service uptime, and operational excellence, Implement monitoring, alerting, and observability tools (e.g., Prometheus, Grafana, Datadog) to track system health, and Lead efforts to automate repetitive operational tasks and manual troubleshooting to improve system reliability.

Successful candidates bring Managing Technical Teams. Important skills include Datadog, AWS, Azure, GCP, Kubernetes, and Linux/Unix. Preferred (not required): Prometheus and Grafana.

Skills & qualifications

RequiredNice to have

Skills

DatadogAWSAzureGCPKubernetesLinux/UnixPythonBashGoCommunicationLeadershipCrisis ManagementPrometheusGrafanaJenkinsGitLab

Qualifications

3+ Years Managing Technical Teams

Full job description

Description Rapyd has unified payments, payouts and fintech on one worldwide platform, and we’re assembling the world’s best team to liberate global commerce. With offices in Tel Aviv, Amsterdam, Singapore, Iceland, London, Dubai, Hong Kong, and the U.S., the opportunities at Rapyd are limitless.

We believe in straight talk, quick decisions, strong execution and elegant solutions. Rapyd is where hard work pays off and careers take off. Join us and let’s build the future of fintech together.

Get the tools to grow globally at www.rapyd.net . Follow: Blog , Insta , LinkedIn , Twitter

The NOC Team Lead is responsible for ensuring the high availability, reliability, and performance of a company's production systems, infrastructure, and applications. This hybrid role bridges traditional reactive monitoring (Network Operations Center - NOC) with proactive, automated, and software-centric engineering practices (Site Reliability Engineering - SRE).

Key Responsibilities

  • Operational Leadership: Directs 24/7 global NOC teams, managing incident response, service uptime, and operational excellence.

  • Proactive Monitoring & Observability: Implements monitoring, alerting, and observability tools (e.g., Prometheus, Grafana, Datadog) to track system health via golden signals (latency, traffic, errors, saturation).

  • Automation and Toil Reduction: Leads efforts to automate repetitive operational tasks and manual troubleshooting to improve system reliability and reduce human error.

  • Incident Management & Root Cause Analysis (RCA): Oversees the management of critical incidents, ensures timely communication, and performs post-mortem analysis to prevent recurrence.

  • Team Management & Development: Coaches, mentors, and develops NOC engineers and SREs, fostering a culture of high performance and continuous improvement.

  • Stakeholder Collaboration: Partners with engineering, development, and IT teams to align system performance with business goals and SLA requirements. Key Differences in Focus

  • NOC Focus: Primarily monitoring, detecting, and responding to incidents, ensuring connectivity and managing alerts.

  • SRE Focus: Focuses on engineering improvements, reducing technical debt, automation, and system resilience.

  • The Transformation: Modern managers are transforming traditional, manual NOCs into automated, SRE-driven environments.

Requirements

  • Experience: 3+ years in managing technical teams within NOC, SRE, or infrastructure domains.
  • Technical Skills: Proficiency in cloud platforms (AWS, Azure, GCP), Kubernetes, Linux/Unix, and scripting languages (Python, Bash, Golang).
  • Tools: Experience with monitoring platforms (Datadog) and CI/CD tools (e.g., Jenkins, GitLab).
  • Soft Skills: Strong communication, leadership, and crisis management skills. Job Candidate Privacy Policy - https://www.rapyd.net/candidate-privacy-policy

You've read the whole posting — now see how you match it.