Data Scientist Intern

Expatiate CommunicationPasadena, CA
Onsite

About The Position

This unpaid internship is focused on providing educational and professional development aligned with academic and professional goals. It accommodates academic schedules, offers hands-on training, and complements rather than replaces paid employee work. The internship is limited to a maximum of three months and does not guarantee or imply entitlement to a paid job upon completion. Location: On-Site | Pasadena, CA Employment Type: Internship (Unpaid) Job Type: Unpaid Internship (there is no expectation of compensation) Duration: Up to Three (3) months. (The Internship is conducted without entitlement to a paid job after the consultation of the internship.) Internships are limited to the period in which the internship provides the intern with beneficial learning to the extent to which the internship is tied to the intern’s formal education program. Due to the nature of the work and access to educational student data, interns must pass a satisfactory Live Scan background check and provide TB clearance to proceed with the internship. (all associated costs for the Live Scan and TB clearance are the intern's responsibility.) We are looking for a motivated Data Scientist Intern with a strong foundation in Python, MS SQL Server, MongoDB, AI Engineering and experience using automation tools like Playwright or Selenium for web scraping. The role will involve extracting and processing data from various sources to support our data-driven decision-making and product development processes. Also help the Data Department with integrating AI to automate manual workflows.

Requirements

  • Currently pursuing or recently completed a degree in Computer Science, Data Science, or a related field.
  • Proficiency in Python, with experience using libraries like pandas, NumPy, and BeautifulSoup.
  • Familiarity with Selenium for web automation and data scraping.
  • Knowledge of data visualization tools like Matplotlib or Seaborn.
  • Basic understanding of machine learning concepts and frameworks.
  • Experience with databases (SQL or NoSQL) for storing and retrieving data.
  • Knowledge of (AWS, GCP, or Azure) for deploying solutions.
  • Knowledge of data visualization tools like Power BI.
  • Knowledge AI engineering tools/techniques like Langchain, Langgraph, VectorDB, RAG, Prompt Configuration, PII Management.
  • Familiarity with version control systems like GitHub.
  • Strong automation tools skills like playwright, selenium.

Responsibilities

  • Use Playwright, Selenium to automate data scraping and web interactions.
  • Develop and maintain robust web-scraping pipelines, capability to use AI with these automation tools.
  • Wrap the above automation pipelines into APIs which are capable of being deployable to production environments.
  • Clean, preprocess, and analyze large datasets using Python libraries such as pandas, NumPy.
  • Perform exploratory data analysis (EDA) to extract meaningful insights.
  • Assist in building APIs which integrate Automation tools and AI to make the daily data operations robust.
  • Test, validate, and setup guardrails for AI
  • Work closely with data scientists, engineers, and product teams to understand data needs and deliver actionable insights.
  • Document and present findings effectively to both technical and non-technical stakeholders.
  • Perform research and development to supports tasks/projects.

Benefits

  • This is an unpaid internship intended to provide educational and hands-on learning experience. Interns are not eligible for wages, employee benefits, or other forms of compensation unless otherwise required by applicable law or expressly stated in a written agreement. Eligibility for benefits may vary based on employment status, classification, and other factors. Not all benefits are available to all employees. Benefits are administered in compliance with applicable federal and state laws and regulations.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service