0% found this document useful (0 votes)
14 views2 pages

Snowflake Data Engineer

The job description outlines a Data Engineer position within the Data Analytics and AI organization, focusing on developing data solutions for various operations in the pharmaceutical industry. Key responsibilities include building and maintaining data pipelines, ensuring data integrity, and collaborating with data scientists and stakeholders. Essential skills required include experience in ETL/ELT processes, cloud services, data visualization, and strong communication abilities.

Uploaded by

mistryboychirag
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
14 views2 pages

Snowflake Data Engineer

The job description outlines a Data Engineer position within the Data Analytics and AI organization, focusing on developing data solutions for various operations in the pharmaceutical industry. Key responsibilities include building and maintaining data pipelines, ensuring data integrity, and collaborating with data scientists and stakeholders. Essential skills required include experience in ETL/ELT processes, cloud services, data visualization, and strong communication abilities.

Uploaded by

mistryboychirag
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

Job Description

Data Engineer – Data Analytics and AI

You will be joining the Data Analytics and AI (DA&AI) organization within Operations IT, responsible for
delivering state-of-the-art Data Analytics and AI solutions for Operations capability areas like
Pharmaceutical Technology Development, Manufacturing & Global Engineering, Quality Control,
Sustainability, Supply Chain, Logistics and Global External Sourcing and Procurement. Here our work has a
direct impact on patients – transforming our ability to develop life-changing medicines. We empower the
business to perform at its peak and lead a new way of working, combining cutting-edge science with
leading digital technology platforms and data.

As the data engineer, you will be responsible for designing, building, and maintaining the infrastructure
and systems necessary to collect, store, process, and analyze large volumes of data. They play a critical role
in ensuring that data is accessible, reliable, and ready for analysis by data scientists, analysts, and other
stakeholders. You will work closely with technical leads and cross-functional teams to implement data
solutions that meet business needs.

Key Responsibilities:
 Develop, Maintain, and Optimize the scalable data pipelines for data integration, transformation,
and analysis, ensuring high performance and reliability.
 Demonstrate proficiency in ETL/ELT processes, including writing, testing, and maintaining high-
quality code for data ingestion and transformation.
 Improve the efficiency and performance of data pipeline and workflows, utilizing advanced data
engineering techniques and best practices.
 Troubleshoot and resolve complex issues related to data pipelines and applications, using logs and
code reviews to ensure smooth data flow and system functionality.
 Work closely with data scientists, analysts, and stakeholders to understand requirements and
deliver effective, high-quality datasets and data solutions.
 Utilize and manage cloud-based services, particularly on AWS, for data storage and processing,
ensuring optimal use of resources.
 Ensure data accuracy and integrity by implementing data validation and cleansing techniques to
maintain consistency.
 Perform unit testing, system integration testing, regression testing, and assist with user acceptance
testing to ensure data solutions meet quality standards.
 Create and maintain data visualizations and dashboards to provide actionable insights.
 Implement and manage CI/CD pipelines, version control, and deployment processes.
 Stay updated with best practices and emerging technologies in data engineering. Contribute to and
learn from all phases of the software development lifecycle (SDLC) processes.
 Liaise with internal teams and third-party vendors to address application issues and project needs
effectively.
 Maintain clear documentation for Knowledge Base Articles (KBAs), data models, pipeline
documentation, and deployment release notes.
 Ensure that data pipelines and solutions adhere to the FAIR (Findable, Accessible, Interoperable, and
Reusable) principles to enhance data usability and sharing.

Essential Skills:
 Minimum 5+ years of experience in developing and delivering software engineering and data
engineering solutions.
 Extensive experience with ELT/ETL tools such as SnapLogic, FiveTran, or similar.
 Deep expertise in Snowflake, DBT (Data Build Tool), and similar data warehousing technologies.
 Proficient in designing and optimizing data models and transformations for large-scale data
systems.
 Strong knowledge of data pipeline principles, including dimensional modelling, schema design, and
data integration patterns.
 Familiarity with Data Mesh and Data Product concepts, including experience in delivering and
managing data products.
 Strong data orchestration skills to effectively manage and streamline data workflows and processes.
 Proficiency in data visualization technologies, with experience in advanced use of tools such as
Power BI or similar.
 Solid understanding of DevOps practices, including CI/CD pipelines, version control systems like
GitHub. Ability to implement and maintain automated deployment and integration processes.
 Automated testing frameworks (Unit Test, system integration testing, regression testing & data
testing).
 Excellent communication and interpersonal skills, with a proven ability of managing stakeholder
expectations, gathering requirements, and translating them into technical solutions.
 Experience working in Agile development environments, with a strong understanding of Agile
principles and practices. Ability to adapt to changing requirements and contribute to iterative
development cycles.
 Advanced SQL skills for data analysis. Expertise on problem-solving skills with a focus on finding
innovative solutions to complex data challenges. Ability to proactively identify and address
potential issues and bottlenecks.
 Fluent in multiple coding languages, such as Python or Java and able to learn new ones quickly
 Technical skills in metadata capturing and management using Collibra to document, organize, and
ensure the quality of metadata assets
 Strong Knowledge of cloud-based data, compute, and storage services, including AWS S3, EC2, RDS,
EBS, Lambda, Airflow (MWAA), Containerization Services (ECS, EKS) and so on.

Desired:
 Bachelor's or Master's degree in health sciences, Life Sciences, Data Management, Information
Technology, or a related field, or equivalent experience.
 Significant experience working in the pharmaceuticals industry, with a deep understanding of
industry-specific data requirements
 Demonstrated ability to manage and collaborate with a diverse range of stakeholders, ensuring
high levels of satisfaction and successful project delivery.
 Proven capability to work independently and thrive in a dynamic, fast-paced environment,
managing multiple tasks and adapting to evolving conditions.
 Experience working in large multinational organizations, especially within pharmaceutical or similar
environments, demonstrating familiarity with global data systems and processes.
 Certification in AWS Cloud or other relevant data engineering or software engineering certifications,
showcasing advanced knowledge and technical proficiency.

You might also like