Post Job Free
Sign in

Senior Data Engineer (ETL, Streaming, Cloud)

Location:
Dallas, TX
Salary:
60000
Posted:
July 24, 2026

Contact this candidate

Resume:

SRAVANTHI POCHAMPALLY

Dallas, TX +1-469-***-**** *********************@*****.*** H4 EAD

SUMMARY

Data Engineer with 6+ years of experience building ETL workflows, streaming pipelines, and cloud-native data platforms using PySpark, Databricks, Apache Spark, Kafka, Airflow, Snowflake, Redshift, AWS, Azure, Python, SQL, and dbt, delivering scalable data solutions that support 25+ reporting systems, dashboards, and analytics workflows while enabling reliable data integration, near real-time processing, and high-quality data delivery. PROFESSIONAL EXPERIENCE

Data Engineer BCBS Sep 2025 - Present USA

● Built claims and enrollment ingestion pipelines using PySpark and Databricks, integrating data from 7 source systems to deliver consistent datasets for recurring claims and membership reporting.

● Developed event-driven integrations using Apache Kafka, enabling near real-time data exchange across 4 care management applications and shortening reporting cycles by 1 business day.

● Automated dependency-driven workflows using Apache Airflow and Snowflake, executing 75+ scheduled jobs and streamlining data loads for 15 reporting and analytics datasets.

● Designed a reporting data mart using Amazon Redshift and SQL, optimizing access to operational and regulatory datasets and reducing dashboard refresh time from 15 minutes to under 5 minutes.

● Implemented validation frameworks using Python and Great Expectations, detecting 350+ claims and member data discrepancies each quarter before they affected downstream reporting.

● Standardized clinical and member data mappings using FHIR and HL7, enabling EHR/EMR data interoperability across 3 healthcare platforms and supporting 6 production releases without data exchange issues. Data Engineer Epsilon Jan 2023 - Aug 2025 USA

● Engineered customer data ingestion pipelines using AWS Glue and Amazon S3, consolidating campaign, web, and CRM data from 8 source systems and providing curated datasets for 18+ audience segmentation workflows.

● Constructed event-driven pipelines using Amazon Kinesis, enabling near real-time processing of engagement events across 4 customer engagement applications and delivering campaign metrics within 30 minutes of ingestion.

● Orchestrated incremental transformations using dbt, Snowflake, and Git-based CI/CD pipelines, managing 65+ production models and supplying trusted datasets for 18 campaign performance dashboards.

● Implemented dependency-based workflows using Apache Airflow, coordinating 50+ scheduled pipelines and ensuring timely data delivery for 20+ analytics datasets.

● Optimized campaign analytics workloads using Amazon Redshift and SQL, enabling performance reporting for 30+ active campaigns and producing 12 recurring client deliverables.

● Integrated customer identity and engagement datasets using Python and REST APIs, synchronizing data across 4 advertising and CRM platforms and supporting 8 production deployments without data synchronization issues. Data Engineer HCL Technologies Sep 2019 - Aug 2022 IND

● Developed ETL workflows using Informatica PowerCenter and Oracle SQL, integrating data from 5 business applications and creating reporting extracts for 12 operational and compliance reports.

● Designed batch processing jobs using Apache Spark and HDFS, transforming 1.8M+ transaction records monthly and generating 8 recurring data feeds for reporting and reconciliation processes.

● Migrated legacy ETL workloads to Azure Data Factory, consolidating 30+ workflows and simplifying data movement across 6 source-to-target interfaces supporting enterprise reporting systems.

● Configured dimensional models using SQL Server and SSIS, populating reporting layers for 10 departmental dashboards and organizing historical data across 6 business subject areas.

● Coordinated production migrations with data analysts, application teams, and QA engineers using Git and JIRA, implementing changes across 4 enterprise modules through 10 approved change requests. TECHNICAL SKILLS

● Programming Languages: Python, SQL, Oracle SQL, Unix Shell Scripting

● Data Engineering & Integration: ETL Development, ELT Development, Data Ingestion, Data Integration, Batch Processing, Data Transformation, Workflow Orchestration,Data Validation, Data Migration, Data Modeling, Dimensional Modeling, Data Profiling, Metadata Management, Data Quality Management, Data Partitioning

● Frameworks & Technologies: PySpark, Apache Spark, dbt, Apache Kafka, Amazon Kinesis, Apache Airflow, Informatica PowerCenter, Control-M

● Databases & Storage: Snowflake, Amazon Redshift, SQL Server, Oracle Database, Amazon S3, HDFS, Delta Lake, Data Warehousing, Data Lakes, Relational Databases, Star Schema, Snowflake Schema

● Cloud Platforms: Amazon Web Services (Glue, IAM, Cloudwatch), Microsoft Azure, Azure Data Factory (ADF), AWS

● APIs & Data Standards: REST APIs, JSON, XML, FHIR, HL7

● Data Governance & Compliance: HIPAA, Data Governance, Data Lineage, Data Security, Data Privacy, Access Control

● DevOps & Collaboration: Git, CI/CD Pipelines, JIRA, Agile, Scrum, SDLC EDUCATION

Bachelor of Technology in Electronics and Communications CVR College of Engineering, India Apr 2019



Contact this candidate