Role Overview
Build the data foundation behind modern AI. Every AI application, every dashboard, every decision is only as good as the data behind it. You'll build the pipes nobody sees — and everything depends on. This internship introduces emerging technologists to how data is collected, transformed, validated, stored, and prepared for applications, analytics, and AI.
Responsibilities
- Trace broken numbers through transformation steps to identify and fix data issues.
- Work with Python, SQL, and ETL/ELT processes.
- Handle data ingestion, REST APIs, and data transformation.
- Manage data quality and work with Relational databases, Cloud storage, and Data lakes.
- Utilize Spark, Databricks, and workflow orchestration for analytics pipelines.
Requirements
- Basic Python and/or SQL knowledge.
- Analytical thinking and comfort working with data.
- Attention to detail and interest in databases and cloud technologies.
- Willingness to investigate data discrepancies.
- Bachelor's degree in Engineering, Computer Science, Data Science, Statistics, Mathematics, Economics, or Finance.
Skills
- Python
- SQL
- Spark
- Databricks
- ETL
Benefits
- Practical data-engineering, Python, and SQL experience.
- Data pipeline and cloud platform exposure.
- Certification support for eligible participants.
- Project experience letter on successful completion.
- Potential professional reference based on performance.
- Commuter assistance, flexible schedule, and internet reimbursement.