We provide IT Staff Augmentation Services!

Senior Big Data Developer / Architect Resume

0/5 (Submit Your Rating)

Alpharetta, GA

SUMMARY

  • More than 8 years of total experience in analysis, design, development testing / validation, application support of ETL, Client/Server, System/data Conversion, data warehousing, analytic reporting and big data based projects.
  • Extensive experience in Hadoop, Hbase, Hive, Pig, Sqoop, Flume, Oozie, Mahout, Impala, Ruby, Python, Java, Scala, Netezza, SQL, PL/SQL, C/C++, Pro*C, Oracle, Unix Shell Scripting, Perl, FCP, XML, Java Script, ESP Cybermation, SAP Business Objects, Power pivot, JSON.
  • Excellent leadership and organization skills including but not limited to task planning, coordination, delegation, scheduling, tracking and progress reporting to concerned stakeholders.
  • Understanding of various software development models such as Waterfall, Agile, RAD, and Prototype models and experience in their application from project conception to project deployment, support and maintenance.
  • Knowledge of functions and responsibilities of project lead, environment manager, application support staff member, onsite /offshore coordinator, analyst programmer, and technical business analyst roles.
  • Excellent interpersonal, communication, troubleshooting and analytical skills and demonstrates pro - activeness and quick learning.
  • Subject matter expert in telecom domain functional areas such as Order Management, Provisioning, Pricing, Billing, Payments/Adjustments, Fraud and Risk Management, Collections, Reporting, Divestiture and System Data conversion.
  • Experience in linear and digital media domain analytics such as subscription management and reporting, click through (web) analytics, social media analytics and data integration of disparate data sources and their reporting.
  • Completed CFA (Chartered Financial Analyst) Level I exam successfully in CFA program.

TECHNICAL SKILLS

Roles: Project Lead, Team Coordinator, Analyst Developer, Application -Support MemberEnvironment Manager, Technical Business Analyst

Operating Systems: Solaris 5.9, Windows XP, NT, 2000, 98, MS DOS, HP-UX, Ubuntu 12.04, RHEL 5.6

Languages: Hadoop, Hive, Pig, Sqoop, Flume, Ruby, Python, Java, Perl, SQL, PLSQL, FCP, XMLJSON, C/C++, Pro*C, Netezza, Oracle, JavaScript, ESP Cybermation, Business Objects

Databases: Oracle (7.x, 8.x, 9.x, 10g, 11g), MS Access, Netezza, SQL Server

SQL Utilities: SQL*Loader, Export, Import, External Tables, NZLOAD

Version Control: Dimensions, PVCS, CVS, SVN

IDE: PLSQL Developer, TOAD, SQL Developer, VISIO, IntelliJIdea, Rubymine, Eclipse, FCPNetbeans

PROFESSIONAL EXPERIENCE

Confidential

Senior Big Data developer / Architect

Responsibilities:

  • Designing, developing and testing ETL for legacy analytics framework by employing data processing methods proposed by Kimball, Inmon and Data Vault initiatives.
  • Researching, prototyping, designing, developing production ready open source based big data technology solutions.
  • Collaborating with product managers and business analyst to gather business requirements and provide solutions for the same.
  • Designing, developing and implementing data mapping and physical schema designs.
  • Designing, developing and implementing, process and data flows for parallel processing.
  • Develop Subject Matter Expertize for big data systems and act as single point of for big data platform.
  • Design, development and deployment of gaming site’s social badge tracking and analytical reporting solution for enhancing user experience. Ruby scripts extracted JSON data from Mongo DB and un-nested to relation format for further ETL processing to derive metrics about badge activity, user activity, scores and time spent on badges.
  • Design, development and deployment of subscription tracking and its reporting for subscription and product management and its analytics. JRuby scripts extracted data from third party API and staged it into database staging area. The staged data was processed using subscription rules to derive subscription metrics such as new product subscription, new product upgrades/downgrades, voluntary and involuntary churn for NBA and NASCAR league passes.
  • Design, development and deployment of ETL/ELT/ETLT processing framework to automate ETL processing activities from feed reception, cleansing, staging, aggregation and summarization. Ruby and JRuby scripts derive their processing from metadata based flow. These scripts tag incoming data with metadata for parallel processing of large click stream data with its associated ad flight/impression data.
  • Designed and developed multiple summaries for reporting click through data metrics to assist better decision making. The developed metric functional areas included SEO/SEM reporting, ad revenue monetization, web analytics, and social alert tracking for multiple media brands. Flume was used to ingest the data at regular intervals. Pig was used for processing complex transformation rules. Hive was mostly used for atomic data aggregation to create detailed granular summaries. At the end of hadoop processing, these summaries were fed through sqoop at regular intervals in Netezza, which acted as data warehouse/mart. The Netezza then processed these granular summaries to create summaries by functional area, business unit, line of business for direct reporting by BI tools.
  • SEO/SEM reporting included developing metrics by identifying keywords as branded or organic and reporting to enhance search engine results. In addition, developing automated integration between Google webmaster and our analytics platform. The automation was done using python for extracting daily keyword and page based google webmaster data. This extracted was presented in single report thus reducing no. of reports and no. of schedules.
  • Web analytics reporting included developing page impressions, unique visitors, page clicks by page, keywords, referrals, section, business unit etc. Hive, Pig and NZSQL were used to develop scripts and Netezza procedures to create these metrics. The reporting was done through PowerPivot enabled excel sheet for adhoc analysis and BO for canned reporting.
  • Monetization summary development included calculation of ad revenue by matching ad impressions with ad flights and click rate to assist tracking and monitoring of traffic and their behavior on digital media. The data from separate sources (web log, ad server) were integrated using hive and NZSQL based on rules to calculate click through rate, and monetized value by integrating them on page view id (surrogate key).
  • Social brand alerting and reporting development included capturing and classifying social media messages to track and monitor brands in real time. Ruby/Python/Scala scripts were deployed to extract twitter and Facebook feeds to constantly monitor social space. This data was fed into hdfs and batch process evaluated each tweet, status update to watch for possible brand management alerts.
  • Architected and lead the initial Hdfs/Hadoop deployment for AIR project from scratch and made it fully operational to move the processing from Netezza platform on HDFS within one quarter. The effort involved setting up first POC hadoop infrastructure and designing data organization, data flow, processing flow artifacts and their implementation. The effort also involved modifying current system architecture to include Hadoop, which acts as data hub. In this modified architecture Netezza act as data warehouse /mart. The oracle summary layer was scaled back to just allow metadata based processing and retain business critical legacy analytics.

Environment: Hadoop, HDFS, Netezza, Ruby, Python, Perl, Scala, Java, Netezza, Business Objects, SAP Data Services, C/C++, Oracle, SQL Server, SQL, NZSQL, PLSQL, UNIX, Shell Scripting, Linux, Ubuntu, XML, Mongo DB, JSON.

Confidential, Alpharetta, GA

Senior Technical Associate / Team Lead

Responsibilities:

  • Lead for entire SDLC including but not limited to requirement capture with business, functional (FD) and technical design (TD) document specifications, code construction, unit test, system test support, deployment and warranty support.
  • Demonstrate knowledge of task planning, coordination, delegation, scheduling, tracking and its progress reporting to development project owners.
  • Create Database Objects, SQL / PLSQL, NZSQL.
  • Design and develop Procedures, Functions, Triggers and Packages to convert Business logic to code.
  • Design and develop Front End GUI applications using FCP a CASE tool.
  • Design and develop back end services specification using Client server architecture with one to one GUI design mapping.
  • Design and develop asynchronous batch jobs in Pro*C, UNIX, Perl, Java for pricing, provisioning, fraud detection / prevention, billing and reporting.
  • Design and develop J2EE web services for web based order management system.
  • Provide solutions and take responsibility of requirement capture, impact analysis and using business knowledge to deliver optimal solutions.
  • Design and develop ETL Jobs and System Data conversion using SQL Loader, Business Objects Data Integrator, SQL / PLSQL, NZSQL, Import, Export and OCI.
  • Provide application support for asynchronous batch jobs from incident inception to closure with SLA adherence.
  • Develop knowledge and train team members on Usage of Perl, AWK, Unix Shell commands, shell scripts and Java based Applications built on Web-logic and Tomcat.
  • Use professional structured methods, standards and techniques in the development of data requirements, data models, and data dictionaries per CMMi - Level 5 standards.
  • Reverse-engineer existing databases in order to create logical data models as part of system design and re-design.
  • Design and develop in-detail ER Models and Maps in development phase.

Environment: - C, C++, OCI, Pro*C, Oracle 9i/ 10g, SQL, PL/SQL, UNIX (scripting), Perl, Core Java, Apache tomcat, XML, JDBC, JAXP, J2EE, FCP (Foundation Cooperative Processing), Weblogic 10.3, Business Objects, AWK, Windows 2000, Sun Solaris, HPUX, Remedy, PLSQL Developer.

Confidential

Technical Associate / Coordinator

Responsibilities:

  • Strong knowledge of Creating Database Objects, SQL / PLSQL.
  • Created Procedures, Functions, Triggers and Packages to convert Business logic to code.
  • Created Front End GUI applications using FCP a CASE tool.
  • Created back end services in Pro*C using Client server architecture corresponding to each GUI.
  • Created asynchronous batch jobs in Pro*C, Unix, Perl, Java for pricing, provisioning, fraud detection / prevention, billing and reporting.
  • Responsibilities included requirement capture, impact analysis and using business knowledge to deliver optimal solutions.
  • In-depth knowledge on usage of Perl, AWK, Unix Shell commands and shell scripts.
  • Used professional structured methods, standards and techniques in the development of data requirements, data models, and data dictionaries.
  • Delivered artifacts and assignments included functional and technical design document specifications, code, unit testing, assembly/ system test support and its deployment followed by warranty support.

Environment: - C, C++, OCI, Pro*C, Oracle 9i/ 10g, SQL, PL/SQL, UNIX (scripting), Perl, Core Java, Apache tomcat, XML, JDBC, JAXP, FCP (Foundation Cooperative Processing), Weblogic 10.3, AWK, Windows 2000, Sun Solaris, HPUX, Remedy, PLSQL Developer.

We'd love your feedback!