Job Location: India
- Requirements:Over 2-5 years of experience in designing scalable and secure pipelines for data needs
- Experience with data replication tools such as Fivetran, Stitch, Meltano or Airbyte.
- Strong analytical skills related to working with semi-structured datasets such as json.
- Experience with AWS cloud services such as ECS, EC2, EMR, S3, IAM.
- Experience working with a modern data warehouse (Snowflake, Redshift, BigQuery, or similar).
- Experience creating orchestration architecture for MLOps, Data Ingestion, Data Transforms, Reverse ETL using Airflow/Prefect/Dagster.
- Experience in data-related programming languages such as Python/Scala/Java or similar.
- Experience with devops tools (Jenkins, Github Actions, Gitlab CI) will be preferred.
- Experience in Spark (PySpark, SparkSQL) would be a plus.
- Familiarity with Kafka or similar streaming tools would be a plus.
- Responsible:Build the infrastructure required for optimal ELT pipelines from a wide variety of data sources.
- Work with stakeholders including the Executive, Product, Data and Design teams to assist with data-related technical issues and support their data infrastructure needs.
- Apply software engineering best practices like version control, continuous integration/deployment, release management and peer reviews.
- Create pipelines as a service for data analyst and data scientist team members to assist them in building and optimizing products into an innovative industry leader.
- Keep data separated and secured across various regions in AWS.
- Understanding of various factors impacting cost optimization for AWS services.
- Governance / Best Practices – adhere and contribute to enterprise data governance standards. Also, educate and support colleagues in best practices to ensure that data is used appropriately.
- Good understanding of data warehousing and modelling concepts.
- Strong project management and organizational skills.
Submit CV To All Data Science Job Consultants Across India For Free

