Architect, design, automate and maintain optimal data pipeline architecture.
Developing and implementing an overall organizational data strategy that is in line with business processes. The strategy includes data model designs, database development standards, implementation and management of?data warehouses?and?data analytics?systems.
Identifying data sources, both internal and external, and working out a plan for data management that is aligned with organizational data strategy.
Managing end-to-end data architecture, from selecting the platform, designing the technical architecture, and developing the application to finally testing and implementing the proposed solution.
Integrating technical functionality, ensuring data accessibility, accuracy, and security.
Planning and execution of big data solutions using Hadoop, Spark and HDFS environments. Entail the complete lifecycle management of Hadoop and data management and automation on Spark.
Architect and build the infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources using SQL and AWS ?big data? technologies
Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability.
Conducting a continuous audit of data management system performance, refine whenever required, and report immediately any breach or loopholes to the stakeholders.
Coordinating and collaborating with cross-functional teams, stakeholders, and vendors for the smooth functioning of the enterprise data system.?
Participate in any other initiatives running under the umbrella of Engineering like training, talks, estimates in Data Engineering domain.
Influence data engineering best practices within the team.
Mentor and guide data and information engineers within the company.