Role Overview
We are seeking a Senior Data Warehouse Platform Engineer with deep expertise in Snowflake, AWS, and Infrastructure as Code (IaC) to architect, automate, and maintain our scalable and reliable data platform infrastructure.
In this role, you will bridge the gap between Data Engineering, Site Reliability Engineering (SRE), and Cloud Infrastructure. You will be responsible for automating Snowflake provisioning, managing Cloud and Kubernetes orchestration layer assets, and ensuring high reliability, performance, and security across our data ecosystem.
Key Responsibilities
Snowflake Administration & Automation: Oversee end-to-end Snowflake database administration (account management, dynamic scaling, RBAC security models, external stages, and schema management) using programmatic automation.
Infrastructure as Code (IaC): Design, build, and maintain production-grade infrastructure on AWS using Terraform, Python, and/or Go.
Compute & Orchestration: Manage containerized workloads and services using Kubernetes (AWS EKS), optimizing network connectivity and cloud resource deployment.
Observability & SRE Practice: Implement robust monitoring, tracing, and alerting (using tools like OpenTelemetry) to maximize system uptime, optimize costs, and respond effectively to operational incidents.
Platform Security & Compliance: Enforce cloud networking best practices, secure access controls, and automated compliance policies across AWS and Snowflake environments.
Cross-Functional Collaboration: Partner closely with Data Engineers, Security teams, and Software Engineers to build self-service platform tools that simplify data pipelines.
Core Qualifications
1. Snowflake & Data Infrastructure
Strong hands-on experience as a Snowflake DBA/Platform Engineer.
Proven track record of automating Snowflake resource provisioning (schemas, virtual warehouses, roles, storage integrations, external stages) via code.
2. Cloud & Infrastructure as Code (IaC)
Expertise in AWS Cloud Platform core services (IAM, S3, VPC, PrivateLink, EC2).
Advanced proficiency in Terraform for enterprise infrastructure provisioning.
Strong programming skills in Python and/or Go for building custom platform tooling and automation scripts.
3. Kubernetes & Orchestration
Hands-on experience operating and networking AWS EKS (Amazon Elastic Kubernetes Service).
Solid understanding of cloud networking principles (VPC peering, CNI, ingress control, service mesh, DNS).
4. Observability & Reliability
Experience with modern telemetry standards (e.g., OpenTelemetry) and observability suites (e.g., Datadog, Prometheus, Grafana).
Strong Site Reliability Engineering (SRE) mindset—focused on proactive resilience, automated failovers, performance tuning, and incident management.
Nice-to-Have Skills
Hands-on experience with modern Kubernetes ecosystem tools: Helm, ArgoCD, ACK (AWS Controllers for Kubernetes), CRDs (Custom Resource Definitions), or Kro.
Experience with FinOps principles for optimizing Snowflake warehouse credit consumption and AWS compute costs.
Knowledge of CI/CD pipeline automation (GitHub Actions, GitLab CI, or Jenkins).
What You’ll Bring
Automation-First Mindset: A drive to eliminate toil by replacing manual configuration with robust code and pipelines.
Operational Excellence: Ability to balance aggressive feature delivery with stability, security, and performance.
Collaborative Problem Solving: Excellent communication skills with a track record of solving complex system challenges alongside multi-disciplinary teams.