Data Engineer Intern
Data · Remote · Internship
About the role
Selfinity AI is looking for a Data Engineer Intern to help build the data foundation for new SaaS and technology products.
The primary focus of this role is collecting raw data from internal systems, public sources, APIs, and approved web crawling workflows, then transforming that data into reliable and usable datasets. You will help design data ingestion systems, automated pipelines, and storage structures that allow the company to identify market trends, evaluate new opportunities, and support future product development.
This role is suitable for candidates with 0–3 years of experience who are interested in web data collection, data pipelines, databases, and scalable data infrastructure.
What you'll do
- Develop and maintain web crawlers, API integrations, and ingestion services for collecting structured and unstructured data
- Build automated ETL and ELT pipelines to clean, transform, validate, and load data into internal systems
- Design database schemas and data models that support analytics, product development, and future machine learning use cases
- Improve the reliability, scalability, and performance of data collection and processing workflows
- Collaborate with product, engineering, and business teams to identify useful data sources and translate business needs into data solutions
What we look for
- 0–3 years of experience in data engineering, software engineering, data analytics, or a related field
- Experience with Python, SQL, or similar programming languages
- Basic understanding of databases, APIs, ETL or ELT workflows, and relational data modeling
- Strong analytical, debugging, documentation, and communication skills
What we prefer
- Experience developing web crawlers, scrapers, or automated data-collection tools
- Familiarity with cloud platforms, data warehouses, data lakes, or pipeline orchestration tools
- Personal, academic, or professional projects involving data pipelines, databases, or large datasets