
Undergrad Intern – Machine Learning Engineer
Posted 14 hours ago

Posted 14 hours ago
This is a fully remote position, open to applicants in United States.
• Assist in the design, development, and testing of scalable data pipelines for the ingestion and transformation of multiple data sources into enterprise data lakes and warehouses.
• Collaborate with data engineers and scientists to clean, prepare, and analyze both structured and unstructured data.
• Support the development and automation of pipelines for model training, evaluation, and deployment.
• Investigate and analyze pharmaceutical commercial datasets to derive data-driven insights.
• Engage in the design of visualizations and dashboards for internal stakeholders.
• Acquire knowledge in hypothesis testing, regression, and classification techniques.
• Work together on the implementation and monitoring of machine learning models in production settings.
• Document technical processes, models, and tools systematically.
• Experiment with various data engineering and MLOps tools and methodologies.
• Take part in agile ceremonies, including sprint planning and retrospectives.
• Anticipated completion of at least one year of study at an accredited college or university prior to the start of the internship.
• Expected ongoing enrollment in an accredited college or university following the internship.
• The student must reside in the United States for the duration of the internship.
• Availability for a full-time work schedule is required.
• Must be willing to accept and commit to future full-time employment by July 2028, if offered.
• Candidates must be 18 years or older.
• Currently enrolled in a full-time Bachelor’s Degree program at an accredited institution.
• A minimum GPA of 3.0 or its equivalent is preferred.
• Degree concentration in Information Technology, Computer Science, Engineering, Business, or a related field is preferred.
• Experience in the biotechnology, pharmaceutical, or healthcare industry is a plus.
• Intermediate proficiency in Microsoft Word, Excel, and PowerPoint is preferred.
• Foundational knowledge in Python or R, including libraries such as pandas, NumPy, matplotlib, seaborn, scikit-learn, or XGBoost is preferred.
• Familiarity with cloud platforms like AWS, GCP, or Azure is preferred.
• Basic understanding of SQL and experience querying relational or big data sources is preferred.
• Interest or coursework in NLP, time-series analysis, or statistical modeling is preferred.
• Familiarity with Git and collaborative coding practices is preferred.
• Awareness of Databricks, Apache Spark, or Apache Airflow is advantageous.
• Coursework in business systems analysis, Lean, Agile/SCRUM, SDLC processes, software development, database modeling, web design and development, IT management, cloud-based application management, B2B collaboration, or enterprise systems is preferred.
• Must be authorized to work in the U.S. throughout the duration of the program.
• Sponsorship for future full-time positions is not assured.
• Competitive benefits package.
• A collaborative work culture.
• Support for professional and personal growth and well-being.
• Opportunities for executive and social networking events.
• Participation in community volunteer projects.
• An inclusive workplace environment.
• Reasonable accommodations for individuals with disabilities.
Sistema Fibra
Amgen
Narvar
phData
Get handpicked remote jobs straight to your inbox weekly.