Ayasdi logo

Lead Platform Engineer

Ayasdi

Remote · location unlistedJobPosted 2 days agoStill listed 2 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
No compensation found
Location
Remote · location unlisted
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Job overview

The Lead Platform Engineer will manage multiple installations of the IRIS Smart Manufacturing platform across cloud and on-premise environments. The role includes operating Azure and AWS infrastructure, leading production operations, automating deployments, developing CI/CD pipelines, and maintaining Kubernetes applications and supporting systems. The engineer will also conduct monthly audits, address security vulnerabilities and compliance requirements, and support multiple product teams and customers.

Skills & qualifications

RequiredNice to have

Skills

AWSAzureLinuxKubernetesDockerCI/CD PipelinesAzure DevOpsTerraformAnsibleCloud PlatformsSecurity Best PracticesVulnerability ManagementSecurity Testing ToolsProduction TroubleshootingRoot Cause AnalysisAgilePostgresElasticsearchRedisNginxGrafanaPrometheusEvent HubKafkaRabbitMQ

Qualifications

4+ Years AWS or Azure Experience3+ Years Linux Experience4+ Years Kubernetes Experience2+ Years Docker Experience4+ Years CI/CD Pipeline Experience2+ Years Terraform or Ansible Experience2+ Years Cloud Experience

Full job description

Introduction

Overview:

Our solutions hosted on our Iris Smart Manufacturing platform combine equipment and process domain expertise in Mining & Metals, Oil & Gas, Chemicals & Petrochemicals with the state-of-the-art in data sciences, machine learning, and process optimization. The IRIS platform can work in hybrid mode and is built using microservices architecture. The applications are containerised and run on Azure Kubernetes service and use Azure blob storage, Data Lake, IoT Hub, Event Hub, Event Grid, and Queues.

The development team embraces DevOps culture and delivers products using Continuous Delivery principles. You will be key in managing multiple installations of the platform across Cloud and on-premise environments, fixing security vulnerabilities on a monthly basis, conducting audits, and contributing to automation of installations. You will have a great opportunity to work with a world-class team and latest technologies and be able to learn and contribute.

Job Description

Responsibilities:

  • Contribute to the IRIS Platform operations roadmap and execute planned research and development
  • Deploy and manage cloud infrastructure on Azure and AWS using best practices and governance standards
  • Lead production operations, including incident management, troubleshooting, and root cause analysis
  • Automate so that IRIS Platform is deployable easily across multiple Cloud providers such as Azure and AWS, and on-premise such as K3s and RKE2
  • Work with customers and Services team to support multiple installations of the IRIS platform
  • Conduct monthly audits including vulnerabilities, DAST tests, and review of backups, cost and access reviews
  • Develop CICD pipelines for projects built microservices that would be deployable across multiple environments
  • Implement security best practices in cloud and on-premise environments to address vulnerabilities and compliance requirements
  • Develop and maintain Helm charts for packaging and deploying Kubernetes applications
  • Manage operations of the databases and messaging systems Postgres, Elastic, Redis, and Kafka including configuration, scalability, and backups
  • Implement auto-scaling of various platform services within the Kubernetes environment
  • Automate infrastructure provisioning with Terraform, Ansible, or similar IaC tools
  • Operate Kubernetes clusters, ensuring scalability, reliability, and security
  • Implement monitoring, logging, and alerting to maintain SLAs and availability targets
  • Apply security best practices and enforce compliance standards across cloud and deployment pipelines
  • Be the operations contact point for multiple product teams

Required Skills & Qualifications:

  • 4+ years of hands-on experience with AWS and/or Azure
  • 3+ years of experience working with Linux systems
  • 4+ years of commercial experience with Kubernetes
  • 2+ years of experience working with Docker
  • 4+ years of experience setting up CICD pipelines (AzureDevOps or similar)
  • 2+ years of experience with automation tools such as Terraform and Ansible
  • 2+ years of experience with Cloud such as Azure, AWS or GCP
  • Good knowledge of best security practices, vulnerability management, and testing tools such as Acunetix, Snyk, CheckMarx or Trivy
  • Troubleshooting problems in Production and comfortable doing root cause analysis
  • Able to work in an Agile environment

Preferred Skills & Qualifications:

  • Good working or operational knowledge of databases (Postgres, Elastic search, Redis or similar)
  • Exposure to configuring web servers such as Nginx
  • Working knowledge of monitoring tools such as Grafana and Prometheus
  • Working knowledge of a messaging framework such as Event Hub, Kafka, RabbitMQ or similar

Diversity & Inclusion Statement: We are committed to building a diverse and inclusive team and encourage candidates from all backgrounds to apply.

About Us

SymphonyAI is building the leading enterprise AI SaaS company for digital transformation across the most critical and resilient growth industries, including retail, consumer packaged goods, financial crime prevention, manufacturing, media, and IT service management. Since its founding in 2017, SymphonyAI today serves 1500+ Enterprise customers globally and has grown to 3,000 talented leaders, data scientists, and other professionals across over 30 countries.

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.