Data Engineer Intern- Job Description
Duration: 2 months
Location: Gurgaon
Job Summary
• The Data Engineer – Intern role requires basic experience in any of the traditional
warehousing technologies (e.g. Teradata, Oracle, SQL Server) or modern database/data
warehouse technologies (e.g., AWS Redshift, Azure Synapse, Google Big Query,
Snowflake)
• As a Data Engineering intern, you will also collaborate with a multi-disciplinary team of
solution architects, visualization engineers, and data scientists on a wide range of
business problems.
• You will be working on large scale projects to provide value to our customers of mining,
QSR, healthcare, financial services, retail industry, etc.
Job Responsibilities
• Assist in developing high-performance distributed data warehouses, analytics systems,
and cloud-based architectures
• Support efforts to integrate data from multiple sources into databases and object
stores, and contribute ideas to improve data reliability, efficiency, and quality
• Support work using modern cloud data platforms (e.g., AWS Redshift, Azure Synapse,
Google Big Query, Snowflake)
• Learn and apply ETL/ELT tools and frameworks (e.g., SSIS, Azure Data Factory, AWS
Glue, Matillion, Talend)
• Support ensuring that data warehousing and big data systems meet business
requirements, including automation, security, performance, and logging/monitoring
best practices
• Work with both on-premise and cloud data environments
• Help transform and prepare data according to logical and physical data models and
overall data architecture
• Contribute to building data pipelines and dataflows that meet business requirements
• Support projects involving data lake design, data management, and cloud migration of
data warehouses
• Gain exposure to large-enterprise data engineering practices and standards
• Interacts with clients including in-person meetings, towards making business
recommendations with effective presentations of findings, including visual displays of
quantitative information
• Brings in new R&D ideas for learning
Education
• Bachelor of Technology/Engineering (Computer Science Engineering/Information
Technology Engineering)
Work Experience
• Experience with at least one traditional data warehousing technology (e.g., Teradata,
Oracle, SQL Server) and/or modern platforms (e.g., AWS, GCP, Snowflake, Databricks,
Microsoft Azure)
• Experience in one scripting language i.e. python, java or scala
• Experience in SQL scripting, tuning, indexing, partitioning, data access patterns, and
scaling strategies
GENERAL
Other Work Experience
• Basic Knowledge on any of the ETL tools and frameworks (e.g. SSIS, Azure Data Factory,
AWS Glue, Matillion, Talend) Is preferred, with a focus on how these technologies affect
business outcomes.
Knowledge, Skills and Abilities
• Ability to translate complex business problems into data-driven solutions
• Basic Working knowledge on any of reporting tools like Power BI , Tableau etc
• Ability to identify data quality issues that could affect business outcomes
• Flexibility in working across different database technologies and propensity to learn new
platforms on-the-fly
• Strong interpersonal skills
• Team player prepared to lead or support depending on situation
Travel Requirements
• Not Applicable
People Management
• It will be an Individual Contributor Role
Level of Autonomy
• Intern will be required to do basic investigation and analysis on their own and there will
be someone to assist in next stages of work before submission. Work will be reviewed
for overall accuracy and adequacy.
GENERAL