Job Location: Bangalore
Job Description and Requirements
You will be responsible for developing Data Analytics workflows driving Diagnostic, Prescriptive, and Predictive analytics. Workflows must be tuned for large-scale data processing in a performant and scalable manner. The ideal candidate must be an independent problem solver with excellent communication skills. This role involves collaborating with domain experts to understand analytics objectives and identifying which process will be most efficient in extracting the insight.
The candidate must have knowledge of the underlying mathematical foundations of statistics, machine learning, and analytics. The ability to transform a data analytics objective into a mathematical process is a must-have skill for this role.
- Extensive previous experience with Python 3.
- Python libraries used for data (such as Pandas, NumPy, SciPy).
- Spark query tuning and performance optimization.
- Hands-on experience with PySpark, Spark RDD, MLlib and GraphX.
- Experience working with S3 and SQL database integration (Postgres and/or MySQL)
- Orchestration framework knowledge (eg: Airflow or Perfect or Luigi) is good to have.
- Deep understanding of distributed systems (e.g. CAP theorem, partitioning, replication, consistency, and consensus).
- Exposure to stream-processing systems like Spark-Streaming, Kafka Streams, etc
- Experience in cloud
- Sharp analytical and problem-solving skills.
B.E./B.Tech./M.Tech./MS degree in the relevant domain (CS, Statistics, ML/AI, Data Science) with 2 to 5 years of experience is desirable.
Job Category
Country
Job Subcategory
Hire Type
Submit CV To All Data Science Job Consultants Across India For Free

