Job Location: Remote
Dremio’s development leaders ensure that Dremio Cloud & our Data Lake value-add for the industry is enhanced with scalable, resilient solutions with uptime & performance that matches SLAs. Dremio is growing quickly and building cloud infrastructure, SaaS & services that enable developer velocity will have an immediate and visible impact on Dremio’s success. You will be enabling data-driven decision making and customer engagement by creating a self-service semantic layer for the product and sales teams to leverage.
What you’ll be doing
- Creating and maintaining a data lake of customer and product usage metrics, which will be used to derive insights to drive product and growth strategies.
- Design and implement workflows for ingestion and transformation for various data sources (S3, GCS, Google Analytics).
- Optimize the retrieval of structured and unstructured data to make it actionable in real time.
- Help develop a strategy for a long term data architecture, which will allow Dremio to make effective data-driven decisions to optimize the customer’s experience.
- Develop and maintain scalable and reliable data pipelines to support gradual increases in data volume and complexity.
- Collaborate with the Product Management and Engineering teams to incorporate new use cases and sources of data
What we’re looking for
- 5+ years of experience as a data engineer in a SaaS environment
- Strong logical and analytical skills
- Deep understanding of data lakes and relational databases
- Knowledge of data formats such as JSON and Parquet
- Experience with Apache Spark, Python libraries such as Pandas for data manipulation
- Experience with AWS and GCP
- Experience with ETL/ELT tools
- Experience working with data projects and ensuring the highest levels of data integrity and quality.
- You can scope, schedule, and resource complex projects in collaboration with other partners such as Product and Engineering.
- Experience working with CI/CD pipelines, DevOps and delivering quality in a fast paced environment.
- Familiarity with BI and data science tools such as Tableau, Superset, Jupyter.
- Excellent communication skills with both technical and non-technical audiences.
Bonus points if you have
- Experience with Apache Iceberg
Submit CV To All Data Science Job Consultants Across India For Free

