Post Job Free
Sign in

Lead Data Engineer & Cloud Architect

Location:
Los Angeles, CA
Salary:
140000
Posted:
September 02, 2026

Contact this candidate

Resume:

Mario Tee

*********@*******.*** 281-***-**** linkedin.com/in/mario-tee

PROFILE

Lead engineer and data architect with 10+ years of experience delivering production software, data platforms, and cloud solutions across healthcare AI, global workforce management, and regulated insurance. Combined strong Python, SQL, data engineering, and cloud architecture with hands-on GCP, BigQuery, Dataflow, Dataproc, Cloud Composer, Pub/Sub, dbt, Airflow, Terraform, and Kubernetes experience to move analytical ideas from discovery through production deployment and support. Built scalable data pipelines, developed self-serve data platforms, established data governance and security practices, and translated complex analysis into decisions that product, operations, clinical, and engineering stakeholders could act on. Worked as a consultative technical lead across solution discovery, estimation, architecture, implementation, testing, documentation, deployment, and production support, balancing platform performance with maintainability, governance, cost, reliability, and business value. Known for pragmatic problem solving, production-quality engineering, clear communication with technical and non-technical audiences, and mentoring teams on repeatable data architecture and cloud delivery practices. SKILLS

• Cloud & Data Platforms: Google Cloud Platform (GCP), BigQuery, BigLake, Google Cloud Storage, Dataflow

(Apache Beam), Dataproc (Spark/Hadoop), Cloud Composer (Airflow), Pub/Sub, Confluent/Kafka, Looker, Vertex AI, BigQuery ML, Dataplex, IAM, VPC Service Controls, GKE/Kubernetes, AWS

• Data Engineering & Warehousing: dbt, Airflow, ETL/ELT, Data Modeling, Data Warehousing, Data Migration, PB-scale Migration, Self-Serve Data Platforms, Data Governance, Data Security, Data Masking, Data Encryption

• Programming & Infrastructure: Python, SQL, Terraform, Pulumi, Infrastructure as Code (IaC), Docker, Kubernetes, CI/CD, Agile/DevOps

• Leadership & Strategy: Data Architecture Roadmaps, North Star Data Strategy, Cross-Functional Team Leadership, Executive Stakeholder Management, Technical Risk Management, ROI Articulation, Mentoring PROFESSIONAL EXPERIENCE

Lead Engineer / Principal Data Architect, Tempus AI, Chicago, IL, Mar 2023 - Present

• Designed and led the implementation of a GCP-based data ecosystem supporting AI-powered healthcare products, establishing BigQuery as the central warehouse for clinical and operational data across precision medicine workflows.

• Built self-serve data platform capabilities that empowered decentralized product and clinical teams to access governed datasets while maintaining central standards for data governance and security.

• Architected streaming and batch ingestion pipelines using Dataflow (Apache Beam) and Pub/Sub to process real-time clinical event data and integrate with downstream analytics systems.

• Led the migration of legacy on-premise and cloud data workloads to Google Cloud Platform, including PB-scale clinical datasets, ensuring minimal disruption to production healthcare applications.

• Implemented Cloud Composer (Airflow) orchestration for complex multi-step data workflows, improving pipeline reliability and enabling automated retries, alerting, and monitoring across the platform.

• Deployed Dataproc (Spark/Hadoop) clusters for large-scale data transformation and feature engineering tasks supporting machine learning model training and evaluation.

• Established dbt-based transformation layers for analytics engineering, enabling version-controlled, tested, and documented data models consumed by product and clinical stakeholders.

• Enforced IAM, VPC Service Controls, and data masking/encryption policies to protect sensitive healthcare information and meet regulatory compliance requirements across the data platform.

• Automated infrastructure provisioning using Terraform and Kubernetes (GKE), ensuring repeatable, scalable, and version-controlled deployments for data services and applications.

• Integrated Looker for business intelligence and self-serve analytics, enabling clinicians and product managers to explore data without engineering intervention while maintaining governance guardrails.

• Introduced BigQuery ML and Vertex AI capabilities for predictive analytics and machine learning model deployment, supporting clinical decision-making and operational optimization.

• Acted as a trusted technical advisor to senior leadership, articulating the ROI of data platform initiatives and managing technical risk across cross-functional engineering, product, and clinical teams.

• Drove Agile/DevOps practices across data engineering teams, establishing CI/CD pipelines, code review standards, and automated testing to accelerate delivery while maintaining production stability. Senior Software Engineer / Senior Data Engineer, Deel, San Francisco, CA, May 2020 - Feb 2023

• Built and maintained data pipelines on GCP to support global payroll, compliance, and workforce management operations, processing large volumes of international employee and contractor data.

• Designed BigQuery data models and dbt transformations that enabled finance, operations, and product teams to analyze payroll accuracy, compliance status, and operational efficiency across multiple countries.

• Implemented Cloud Composer (Airflow) workflows to orchestrate complex ETL/ELT processes, automating manual data movement and reducing processing time for monthly payroll cycles.

• Used Dataflow (Apache Beam) for streaming data ingestion from event-driven systems, enabling real-time visibility into workforce transactions and compliance events.

• Integrated Pub/Sub and Confluent/Kafka for reliable message delivery between microservices and data platforms, ensuring consistent data flow across distributed systems.

• Enforced data governance and data security controls including IAM policies, data masking, and encryption to protect sensitive financial and personal information across international jurisdictions.

• Automated infrastructure with Terraform and containerized data services using Kubernetes, improving deployment consistency and reducing operational overhead for the data platform.

• Collaborated with cross-functional teams in an Agile/DevOps environment to deliver data products that supported rapid company expansion into new markets and regulatory environments.

• Contributed to the design of self-serve analytics capabilities, enabling business teams to access governed datasets through Looker while maintaining centralized data standards. Software Engineer, Kemper, Chicago, IL, May 2015 - Apr 2020

• Developed and maintained enterprise data pipelines supporting policy management, claims processing, and customer servicing for a large insurance platform.

• Built ETL/ELT workflows using Python and SQL to move insurance data across operational systems, data warehouses, and reporting environments, improving data reliability and accessibility.

• Modernized legacy data systems by migrating workloads to cloud infrastructure, including Google Cloud Platform services such as BigQuery and Cloud Storage, reducing on-premise operational costs.

• Implemented Airflow orchestration for scheduled data processing jobs, enabling automated execution, monitoring, and alerting for critical insurance workflows.

• Designed data models and data warehousing solutions that supported risk analysis, fraud detection, and operational decision-making across the organization.

• Applied data governance practices including data quality checks, lineage tracking, and access controls to ensure consistency and compliance across insurance data assets.

• Containerized data services using Docker and Kubernetes, improving deployment speed and environment consistency for development and production systems.

• Collaborated with data scientists and analysts to build feature pipelines and analytical datasets that powered predictive models for risk assessment and fraud detection.

• Participated in Agile ceremonies and DevOps initiatives, contributing to CI/CD pipeline development and automated testing for data engineering projects. EDUCATION

Bachelor of Computer Science

The University of Dallas, 2010 - 2014



Contact this candidate