Job Location: Hyderabad/Secunderabad
Roles and Responsibilities
Responsibilities:
1. Develop and deploy batch and streaming data pipelines in cloud ecosystem.
2. Automation of manual processes and performance tuning of existing pipelines.
3. Data loading and processing from multiple source locations into Data lake, Datamart and Datawarehouse while keeping cost, performance and security in mind.
4. Automate and develop analytics tools and occasionally involve in visualization set up processes.
5. Develop processes for migrating on-premise data to cloud environment
Qualifications :
Mandatory Skills :
1. 2+ years of experience in IT programming, application/product development.
2. Experience working in any of the Bigdata cloud ecosystems like AWS, GCP, PCF, Azure etc.
3. Strong in SQL/RDBMS and any of the programming languages like Java/Python/Scala.
4. Experience with one or more of the big data tools like Hadoop, Kafka, Spark, Beam etc.,
5. Good experience with AWS services like EC2, EMR, Redshift or equivalent GCP services like Compute engine, Big query, Dataflow etc.
6. Good knowledge in Data structures.
7. Experience working with multiple OS like Windows, Linux and Unix and good scripting knowledge including Shell, Bash.
8. Basic knowledge in any of the web and server side frameworks like AngularJS, ReactJS, NodeJS, Django
9. Good knowledge in NOSQL DB concepts.
10. Very good communication and team player skills.
Requirements
Preferable Skills :
1. Working knowledge on MQs, publisher-consumer models, streaming data processing
2. Experience with Devops tools like Github, Jenkins, Monitoring systems etc.
3. Good knowledge in pipeline orchestration tools like Airflow, Composer or equivalent services like AWS Glue.,
Desired Candidate Profile
Perks and Benefits
Submit CV To All Data Science Job Consultants Across India For Free

