Apply Edge Start your job search

Product Lead and Director of DevOps

Arrowfoot.AI Ventures · Princeton, NJ

Apply & track with Apply Edge

Company OverviewPerfoarm Technologies Inc. is an AgTech venture and Arrowfoot.AI Ventures portfolio company with multiple products that provide monitoring, analytics, and operational intelligence platform for poultry farming. CyclOps Live Ops is a Computerized Maintenance Management System (CMMS) (web and app for Android and IoS) that tracks and optimizes asset uptime, schedules labor, tracks repair and maintenance workflows and assists with financial reporting. Another product combines IoT-enabled barn controllers, edge processing, Azure cloud services, Databricks Lakehouse architecture, MLflow, LLM-enabled workflows, and real-time telemetry to improve farm efficiency.Job DescriptionThe Product Lead will build, automate, secure, and operate the cloud and edge infrastructure behind CyclOps and be responsible for feature upgrades and support. For the farm efficiency product, the the role will focus on supporting Engineers in a Director of DevOps capacity (Azure infrastructure, CI/CD, infrastructure-as-code, observability, IoT device operations, Databricks platform reliability, and secure production deployments). The ideal candidate is hands-on, automation-first, and comfortable supporting distributed systems that connect on-site environments to cloud services.Product Lead - CylOps Live Ops CMMS SystemAssume full responsibility for technical aspects of CylOps (web and apps)Transform Beta version into an enterprise-ready platformDirector of DevOps - PerfoarmAzure Infrastructure & Platform OperationsDesign, deploy, and maintain Azure infrastructure for farm optimization including IoT Hub, Event Hubs, Blob Storage, Functions, Key Vault, App Services, networking, and monitoring services.Manage development, test, staging, and production environments with repeatable, secure, and well-documented configuration.Support secure connectivity between farm edge devices, Azure cloud services, data pipelines, APIs, and operational dashboards.Implement resilient patterns for reliability, fault tolerance, backup, recovery, and disaster readiness.CI/CD & Infrastructure-as-CodeBuild and maintain CI/CD pipelines using Azure DevOps, GitHub Actions, or similar tooling.Automate infrastructure provisioning using Terraform, Bicep, ARM templates, or equivalent IaC frameworks.Standardize build, test, release, rollback, and environment promotion processes across cloud, edge, data, and application components.Improve deployment visibility, release governance, and operational readiness for production systems.IoT, Edge & Databricks ReliabilitySupport Azure IoT Hub device provisioning, secure messaging, telemetry ingestion, and cloud-to-device communication.Help monitor edge connectivity, device health, telemetry gaps, message delays, and farm-site operational issues.Support Databricks workspace operations, pipeline monitoring, job scheduling, permissions, secrets, and platform reliability.Collaborate with data engineers to improve Databricks workflow stability, observability, and cost efficiency.Security, Observability & AI-Accelerated DeliveryImplement monitoring, logging, alerting, dashboards, and incident response workflows using Azure Monitor, Log Analytics, Grafana, Datadog, or similar tools.Manage secrets, certificates, service principals, managed identities, RBAC, and secure access patterns using Azure Key Vault and Azure AD.Use Claude Code or Codex as part of the standard DevOps workflow for scripting, automation, IaC, documentation, troubleshooting, and code reviews.Work closely with IoT, full stack, data, AI/ML, and product teams to support reliable production delivery.Required Qualifications5+ years of experience in DevOps, Cloud Engineering, Platform Engineering, or Site Reliability Engineering.Strong hands-on experience with Microsoft Azure production environments.Experience with Azure DevOps or GitHub Actions for CI/CD automation.Experience with Terraform, Bicep, ARM templates, or similar infrastructure-as-code tools.Experience with Docker, Kubernetes, cloud networking, security, monitoring, and production support.Strong scripting ability using Python, Bash, PowerShell, or similar languages.Experience with observability, incident response, logging, alerting, and operational runbooks.Experience using Claude Code or Codex as a required part of day-to-day engineering workflow.US-based candidates only.Preferred / Nice to HaveExperience with Azure IoT Hub, edge devices, gateways, MQTT, or industrial IoT deployments.Experience supporting Azure Databricks, Spark jobs, MLflow infrastructure, or data platform operations.Experience with Grafana, Prometheus, Datadog, OpenTelemetry, or advanced Azure Monitor implementations.Experience with secure remote infrastructure in distributed physical locations such as farms, industrial sites, or manufacturing facilities.Experience with SOC 2 and ISO 27001 compliance readiness, controls implementation, audit support, security documentation, or evidence collection is a big plus.Familiarity with AI/ML platform operations, LLM-enabled tooling, model deployment, or MLOps practices.