Dallas, TX
... SKILLS • Programming & Scripting: Python, SQL, Scala, Java, Shell Scripting, REST APIs • Distributed Processing & Big Data: Apache Spark, Hadoop Ecosystem (Hive, Pig, HDFS, MapReduce), Kafka, Flink, Presto, Delta Lake, Druid, Storm • Analytics & ...
- Jul 22
St. Louis, MO
... Technical Skills Category Skills Programming Languages Python, SQL, Scala, Java, VSAM, File-AID, REXX, COBOL Big Data Technologies Apache Spark, PySpark, Spark SQL, Hadoop, MapReduce, Hive Streaming Technologies Azure Event Hubs, Spark Structured ...
- Jul 22
Visakhapatnam, Andhra Pradesh, India
... CORE COMPETENCIES Analytics Engineering & Data: SQL, Python (Pandas, NumPy, Scikit-learn), Airflow, Hive/Hadoop, ETL pipeline design, dimensional modeling (star/snowflake schema), data warehousing Cloud Platforms: Google Cloud Platform, BigQuery, ...
- Jul 22
Los Angeles, CA
... Glue), Snowflake, Azure Data Factory, SQL Server, Oracle, MySQL, dbt ETL & Data Engineering: Alteryx, Informatica, PySpark, Hadoop, Power Query, ETL Development, Star Schema, Dimensional Modeling, Data Warehousing Business Intelligence & Reporting: ...
- Jul 22
New York City, NY
... Key Projects Enterprise Lakehouse Modernization • Migrated legacy Hadoop workloads to a Databricks Lakehouse using Delta Lake, Unity Catalog, and Terraform, reducing operational costs by 40%. Real-Time Fraud Detection Platform • Developed Kafka and ...
- Jul 22
East of 101, CA, 94080
... TECHNICAL SKILLS Cloud & Big Data Platforms - AWS, Azure, GCP, Hadoop, Spark, S3, Databricks, cloud-based analytics infrastructure, cloud automation Data Analytics & BI - SQL, Power BI, Tableau, Excel, Power Query, Power Automate, Power Apps, Looker ...
- Jul 22
College Park, MD
... (Fortune 500) -- PNC Bank Hyderabad, India • Built ETL pipelines ingesting data from SQL Server, Oracle, and MongoDB into Hadoop, transforming 10TB daily with Spark SQL to serve 30 downstream business teams • Rewrote 150 Spark SQL transforms using ...
- Jul 22
McKinney, TX
... Cedar Gate Technologies (Healthcare) Data Engineer Kathmandu, Nepal Jun 2020 – Aug 2021 Built and operated healthcare data pipelines processing claims, eligibility, enrollment, and provider datasets using AWS Glue, PySpark, Hadoop, Hive, and EMR, ...
- Jul 22
Virginia
... Spark SQL, Bash Processing & Orchestration: Apache Spark, Apache Airflow, dbt (tests, docs, incremental models), Trino, Apache Kafka, ETL/ELT design Data Platforms & Modeling: Databricks, Snowflake, Amazon Redshift, Hive/Hadoop, PostgreSQL, MongoDB; ...
- Jul 22
Cincinnati, OH
... ETL Pipelines, Pandas, PySpark Databases Oracle, Redshift, PostgreSQL, MySQL Big Data Technologies: Apache Spark, Hadoop Ecosystem, Hive, Spark SQL Transformation Tools dbt (Data Build Tool), PySpark, Pandas Visualization & BI Power BI, ...
- Jul 21