DARSHAN THIPPASHETTY
Email : *******.************@*****.*** Phone : +1-704-***-****
LinkedIn: https://www.linkedin.com/in/darshan-thippashetty Location: Charlotte, NC -28262
SUMMARY:
Over 5 years of IT experience with emphasis on Data Warehousing Technologies, development, design, implementation & support of ETL processes for large scale data warehouses using Informatica Power Center 9.5.1
Experienced in multiple domains like Banking, Financial and Insurance.
Extensive hands-on expertise in ETL and data integration using Informatica Power Center 9.x(9.5.1,9.1) tools like Designer (Source Analyzer, Warehouse designer, Mapping designer, Mapplet Designer, Transformation Developer, Target designer), Repository Manager, Workflow Manager & Workflow Monitor
Database experience using Oracle 11g/10g, SQL * Developer 2008.
Extensively worked on data extraction, transformation and loading data from various heterogeneous source systems like Oracle, flat files, XML and Salesforce.com files into EDW, ODS and Data marts using Power Center.
Strong experience in debugging, performance tuning and error handling of Informatica mappings to optimize session performance.
Experience in implementing the complex business rules by creating transformation, re-usable transformations and developing complex Mapplets, Mappings, SQL, performance Optimization and Audit logging.
Experience in the Implementation of full lifecycle in Data warehouse, Operational Data Store (ODS) and Business Data marts with Dimensional modeling
Understanding of Ralph Kimball and Bill Inmon Data Warehouse Methodologies and knowledge of the various data design models like Star Schema and Snowflake models.
Strong understanding of OLAP and OLTP Concepts.
Strong skills in Data Analysis, Data Requirement Analysis and Data Mapping for ETL processes.
Proficient in Data Profiling, Data Cleansing and Data validation of operational sources to ensure data quality.
Experience with multiple Databases like Oracle, SQL Server, and Salesforce.com as sources and targets along with flat files and xml.
Good experience in writing Unix Shell Scripting with respect to Business Requirement and developed Several Automations [Re-usable Components] to reduce manual interventions.
Involved in Technical Documentation, Unit test, Integration test and writing the test plan.
Extensive knowledge on Big data, Data Analysis and Data mining technologies.
Hands on experience on Big data tools like, Microsoft Azure, HDInsight, HIVE, SQOOP
Complete hands on experience in all phases of SDLC including requirement elicitation, creation of Functional and Technical design documents, development and support.
Extensive knowledge and experience in AGILE (Scrum) Methodology. Involved in sprint planning, product backlog creation and acted in the capacity of Scrum Master
Excellent communication and interpersonal skills, good analytical reasoning and high adaptability to new technologies and tools.
TECHNICAL SKILLS:
ETL Tool : Informatica Power Center 9.5/9.1
Databases : Oracle 11g, SQL Server, MySQL
Big Data & Analytics : HADOOP, HIVE, SQOOP, R, Tableau
Scripting : PL/SQL, UNIX Shell Scripting
Languages : Python, JAVA, Servlets, JSP, HTML5, CSS
EDUCATION:
University of North Carolina at Charlotte Masters’ in Information Technology Aug 2015–Dec 2016
J.N.N College of Engineering, Shimoga, India Bachelors’ in Computer Science Sep 2007–May 2011
AWARDS AND CERTIFICATIONS:
LIVE WIRE Award for excellent performance in AAA club project.
Certified SQL Developer
Ongoing certification on PYTHON Scripting by Coursera
WORK HISTORY:
DATA WAREHOUSE TEACHING ASSISTANT Sept 2016 – Dec 2016
University of North Carolina at Charlotte [USA]
Worked as a Teaching Assistant for the course Data Warehouse and Analytics
As part of the Data Warehousing course, worked with students on a major project to implement and explain the 3 tier architecture of data warehouse, data modeling and ETL operations.
Assisted my Professor in designing the data model and explaining Star / Snow flake schema
Designed and developed ETL Mappings using Informatica Power Center 9.5.1 to extract data from flat files and Oracle, and to load the data into the target database.
Responsible for SQL programming
Implemented SCD Type1 & Type2 mappings to update slowly changing dimensions to maintain full historical claims
Used Mapplets and Reusable Transformations to prevent redundancy of transformation usage and maintainability.
ETL / HADOOP Consultant May 2016 – Aug 2016 Saltmines Group LLC, Florida [USA]
Devised a POC to demonstrate how HADOOP can be integrated with existing data warehouse to improve the performance.
Deployed HDInsight clusters and developed SQOOP Scripts to import / export data between SQL Server and clusters.
Created HIVE tables on HDFS data and wrote HIVE queries to access them
Implemented Java Web application, which queries HIVE DB and then compared the results with traditional relation DB.
Actively involved in analyzing the existing database system and mapping tables, relationships, and columns to SQL Server 2008.
Worked on MS SQL Server to create tables and query the data as per the business requirements
Implemented FACT & Dimensions tables, Physical & logical data modeling.
Experienced on working with complex queries involving group functions, aggregators, joiner and sub queries
Worked on identifying the performance bottlenecks and performed query tuning
Involved in technical design and specification documentation based on Business user interactions.
RDBMS TEACHING ASSISTANT Sept 2015 – Apr 2016
University of North Carolina at Charlotte [USA]
Worked as a Teaching Assistant for the course Relational Database Management System
Used MySQL, Oracle 11g to teach database concepts and explain the database objects manipulation
Assisted students in understanding fundamental concepts of ER modeling
Worked with students to help them in writing SQL scripts which involves DDL, DML and TCL operations
Created triggers, indexes, stored procedures and views
Interacted directly with students to understand their difficulties and work with them to solve their issues
Helped students with to complete their academic projects successfully
ETL INFORMATICA Developer Jan 2013 - August 2015
HCL Technologies Ltd
Environment: Informatica Power center 9.5.1, UNIX, Oracle 11g, SFDC
Description:
AAA (pronounced "Triple A"), formerly known as the American Automobile Association, provides services to its members such as travel, automotive, insurance, financial, and discounts. The scope of this module was to develop a data layer for marketing team to communicate to customers. This Project deals with Data from the new IVANS policy data exchange is parsed and used to incrementally update the Club’s new Policy Operational Data Model. Salesforce.com was one of the major sources which provide incremental business transactions data.
Responsibilities:
Requirement Gathering and Business Analysis. Performed Business Discovery to define functional specifications.
Data Modeling using Ralph Kimball approach and populating the business rules using mappings into the target tables.
Developed complex ETL Mappings with relational / flat files / XML sources and targets
Implemented SCD Type 2, truncate and incremental load logic using Informatica mapping.
Involved in identification of facts, measures, dimensions and hierarchies for OLAP models.
Create complex Informatica mappings with extensive use of aggregator, union, filter, router, normalizer, joiner, sequence generator, stored procedure and lookup transformation.
Work in team to identify issues and improve performance across various aspects (targets, source, transformation and sessions) of the Informatica process.
Experience on implementation of Web Service providers and consumer transformation
Worked on importing WSDL, SOAP request, XML files
Performed data integration between systems using Informatica Power Center
Extensively worked on Informatica scheduler for scheduling Informatica workflows.
UNIX shell scripts - to create dynamic parameter files, indirect files, send customized mails, pre and post load auditing, etc.
Implemented parallelism and reusability by configuring reusable transformations & the concurrent run property in the workflow
Scheduling and monitoring the jobs. Worked on debugger to analyze and debug the run time errors
Experienced in Data Modeling - Logical/Physical, Star/Snowflake schema, FACT & Dimension tables
Working with an Agile, Scrum methodology to ensure delivery of high quality work with every sprint.
Involved in migrating the code from development to production and making sure the migration is successful with no errors, validating the mappings, sessions, connectivity and work-flows.
Perform the post deployment (QA/UAT/prod) health checks and monitor for any issues during the initial job execution.
Provided Knowledge Transfer to end users and created extensive documentation on the design, development, implementation, daily loads and process flow of the mappings.
ETL Developer June 2011- Dec 2012
HCL Technologies Ltd
Environment: Informatica Power center 9.5.1, UNIX, Oracle 11g
Description:
The scope of this project is to allow the club to operate independently of the infrastructure of the IE. The mappings are pointed to flat files that are being generated for the club by the IE. The flat files are fixed width and are of the same format as that of the oracle tables. The IE will be providing 2 sets of files for each of the tables that CDC is provided. The main file will continue all the CDC activity while a second audit file contains a count of the number of records in the associated CDC file.
Responsibilities:
Analyzing the Requirements and developing the code using ETL tool Informatica.
Created Informatica mappings to process the flat files as data sources
Used Mapplets and Reusable Transformations to prevent redundancy of transformation usage and maintainability.
Used session partitions, dynamic cache, and index cache to improve the performance Wrote UNIX scripts for creating, dropping, analyzing tables, pre/post-session shell scripts to download/upload files from shared server, to generate control files, check status (failure/success) after completion of all batches
Developing T-SQL stored procedures, functions, triggers, views.
Tuned and tested the mapping using different logic’s to provide maximum efficiency.
Debug the sessions by utilizing the logs of the sessions.
Create test cases to perform unit testing. Monitored Workflows and sessions using Workflow Monitor.
Involved in Error handling (Ignore, rejecting bad records to a flat file, loading the records and reviewing them)
Address production issues by performing root cause analysis, implementing and deploying the fix
Performance tuning of SQL queries in production with utilities like Explain and Trace.
Worked with Informatica production support team to find and resolve post production failures
Scheduled jobs, observe the run time session properties and record the statistics of load to analyze the bottlenecks I had written a SQL script to run on Informatica repository server to extract the metadata and run statistics
Created technical design documents for development of mapping and implementing business logic through transformations in the mappings.