Job Location: Hyderabad/Secunderabad
Role and responsibilities :
– Creating Project Technical Documentation
– Designing Solution architecture, and work on Data Ingestion, Preparation and Transformation. Debugging the production failures and identifying the solution.
– Developing efficient frameworks for development and testing using (AWS Dynamo DB, EKS, Kafka, KinesisSparkStreamingPython) to enable seamless data ingestion process on to the Hadoop platform.
– Enabling Data Governance and Data Discovery on Hadoop Platform
– Building data processing framework using Spark, HQL
– Exposure of Security Framework with Kerberos, Ranger, Atlas
– Exposure of Data Pipeline Automation using DevOps tools
– Exposure of Job Monitoring framework along validations automation
– Exposure of handling structured, Un Structured and Streaming data
Technical skills requirements :
– The candidate must demonstrate proficiency in,
– Solid hands-on and Solution Architecting experience in Big-Data Technologies (AWS preferred)
– Hands on experience in: AWS Dynamo DB, EKS, Kafka, Kinesis, Glue PySpark, EMR PySpark
– Hands-on experience of programming language like Python, Scala with Spark.
– Good command and working experience on HadoopMap Reduce, HDFS, Hive, HBase, and No-SQL Databases
– Hands on working experience on any of the data engineeringanalytics platform (HortonworksCloudera MapR AWS), AWS preferred
– Hands-on experience on Data Ingestion Apache Nifi, Apache Airflow, Sqoop, and Oozie
– Hands on working experience of data processing at scale with event driven systems, message queues (Kafka FlinkSpark Streaming)
– Hands on working Experience with AWS Services like EMR, Kinesis, S3, CloudFormation, Glue, API Gateway, Lake Foundation
– Hands on working Experience with AWS Athena
– Data Warehouse exposure on Apache Nifi, Apache Airflow, Kylo
– Operationalization of ML models on AWS (e.g. deployment, scheduling, model monitoring etc.)
– Feature EngineeringData Processing to be used for Model development
– Experience gathering and processing raw data at scale (including writing scripts, web scraping, calling APIs, write SQL queries, etc.)
– Experience building data pipelines for structuredunstructured, real-timebatch, eventssynchronous asynchronous using MQ, Kafka, Steam processing
– Hands-on working experience in analyzing source system data and data flows, working with structured and unstructured data
– Must be very strong in writing SQL queries
– Strengthen the Data engineering team with Big Data solutions
– Strong technical, analytical, and problem-solving skills
– Strong organizational skills, with the ability to work autonomously as well as in a team-based environment
– Pleasant Personality, Strong Communication & Interpersonal Skills
Submit CV To All Data Science Job Consultants Across India For Free

