Location: 

(

Philadelphia

,

PA

)

Salary: 

$

175k

 - $

285k

Our client, a leading portfolio of globally recognized lifestyle brands, is seeking an accomplished Staff Platform Engineer to join the development of AI-powered digital experiences. This is a unique opportunity for an experienced platform professional to integrate algorithmic solutions with creative tooling across a rapidly evolving digital ecosystem.

In this role, the individual will leverage deep platform engineering expertise to architect, deploy, and maintain the foundational infrastructure that powers AI-driven applications. Working at the intersection of cloud-native architecture, security, and developer productivity, this engineer will collaborate closely with a talented cross-functional team of UX designers, data scientists, product managers, and domain experts to deliver significant business impact.

The ideal candidate is a seasoned platform engineer who combines deep cloud expertise with a security-first mindset and a passion for automation. They thrive in a culture of rapid technological change, embrace AI-augmented workflows, and are committed to mentoring and elevating those around them while ensuring infrastructure remains secure, scalable, and maintainable.

Role Responsibilities

  1. Infrastructure as Code: Design, deploy, and maintain modular infrastructure using OpenTofu and Terraform, enforcing a strict DRY architectural approach that ensures reproducible environments across development and production with zero configuration drift and secure defaults.
  2. Cloud Platform Engineering: Architect and manage Google Cloud Platform (GCP) services including Cloud Run, BigQuery, AlloyDB, Pub/Sub, Cloud Storage, and IAM to support scalable, high-performance AI workloads.
  3. Security-First Design: Champion security across the platform by implementing zero-trust architectures, least-privilege IAM policies, GCP Secret Manager integrations, and robust authentication layers including Identity-Aware Proxy and Workload Identity Federation.
  4. CI/CD Automation: Build and manage continuous integration and delivery pipelines using GitHub Actions and TACoS (Terraform Automation and Collaboration Software) solutions, with a strong emphasis on automated testing, linting, and reproducible builds.
  5. Observability & Reliability: Enhance the GCP-native observability stack—including Cloud Monitoring, SLOs, and Alert Policies—to ensure FastAPI services and Cloud Run instances maintain excellent uptime and performance, with clear correlation tracking across request traces and spans.
  6. AI-Driven Engineering: Embrace and help shape agentic AI workflows, recognizing that leveraging AI assistants is an essential practice for modern engineering teams. Apply an experienced guiding hand to ensure AI-generated code meets rigorous standards for uptime, security, and maintainability.
  7. Python Ecosystem Stewardship: Maintain and contribute to core tooling, orchestration frameworks (Airflow, LangGraph), and shared infrastructure libraries managed via uv and Pydantic, with a focus on human oversight and code quality.
  8. Mentorship & Collaboration: Act as a mentor to less experienced team members, actively seeking to elevate those around them through knowledge sharing, code reviews, and respectful communication across both deep architectural discussions and remote collaboration channels.

Role Qualifications

Must-Have

  1. Experience: A proven track record in Platform Engineering, DevOps, or Site Reliability Engineering, with demonstrated success in building and maintaining production infrastructure at scale.
  2. Cloud Expertise: Deep, hands-on knowledge of Google Cloud Platform services including Cloud Run, BigQuery, AlloyDB, Pub/Sub, Cloud Storage, and IAM.
  3. IaC Mastery: Extensive experience with Terraform and/or OpenTofu. Familiarity with TACoS (Terraform Automation and Collaboration Software) solutions is a strong plus.
  4. CI/CD Advocate: Strong experience building and managing CI/CD pipelines using GitHub Actions and related automation tools.
  5. Development Skills: Proficiency in Python, with the ability to navigate complex microservices architectures and data pipelines within a monorepo environment.
  6. Security-First Mindset: A demonstrated commitment to security, with experience implementing zero-trust networks and designing secure-by-default infrastructure.
  7. Cultural Fit: A respectful, kind, and collaborative professional who communicates effectively across remote and in-person environments and thrives in a culture that embraces rapid technological change.

Nice-to-Have

  1. AI/ML Infrastructure: Experience managing AI/ML infrastructure, specifically with Vertex AI or similar platforms.
  2. Stack Familiarity: Familiarity with the specific technology stack, including FastAPI, Pydantic, dbt, Airflow, and LangGraph.
  3. Developer Experience: Experience building CLI tooling and leveraging task runners like just to craft an exceptional developer experience for engineering teams.

The Perks

Our client offers a comprehensive suite of Perks & Benefits to all eligible employees. Availability and eligibility may vary based on employment status and location, but generally include competitive medical, dental, and vision coverage, generous PTO, substantial employee discounts, and robust retirement savings plans.

Ready to grow your career?

Let's get started