Data and Machine Learning Intern
Posted 570 days ago, which is unusual. The employer's own board was still carrying it when we last read it, just now.
This is the employer's own posting, not a copy on a job board.
What we know
Is it still open?
Confirmed still open
Last checked 10h ago — checked against the employer's own applicant tracking system, which is the company answering directly.
We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.
How old is it?
Posted 570d ago
The date the source published, not the day we noticed it (2025-03-10). Last seen at its source just now.
We have tracked this listing since 22 Sep 2026 (9 days). The employer's own board has carried it every time we have read it, most recently just now.
Is it remote?
The listing says yes
The location field doesn't say remote, so our assessment is based on the title or the description. Read the listing before applying.
Who may apply?
Colombia
The description states no restriction of its own. This is the source's own tag.
Pay not stated
Similar roles pay $1,050–1,938/mo
Middle 50% of 18 listings that do state pay — Operations · all levels · Colombia · USD/month. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.
Skills named in the ad
Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.
Carried by 1 source
-
greenhouse employer's own board first seen 9d ago · last seen just now
The listing
Duration: Six months
Format: Full time (40 hrs/week), paid
In the last year at Loka, our teams launched almost 200 GenAI projects for companies of all kinds, including the world’s Number 1 GenAI reading tutor, a startup that transforms homes into batteries and a leading cancer-fighting laboratory. And we did it all while enjoying every other Friday off 😎
As a Data - Machine Learning Intern, you'll gain professional experience supporting Loka’s certified specialists, technical experts and PhDs while elevating your skillset, building a portfolio and launching projects you’re proud of.
The Role
- Assist in designing, developing and maintaining data pipelines to ensure clean, reliable and timely data.
- Collaborate with the team to implement and optimize ETL processes.
- Integrate data from various sources into warehouses, data lakes and lakehouses.
- Support data management tasks, including data cleaning, validation and transformation.
- Understand business objectives and develop models that help achieve them, plus metrics to track their progress.
- Implement ML systems using classical ML, DL and Foundation Models following best practices.
- Participate in client communications by helping gather requirements and communicate deliverables.
- Explore and visualize data with a careful eye for issues that require data cleaning as well as differences in data distribution that may affect performance after deployment.
- Identify and analyze model errors.
Required Hard Skills
- Last year of a bachelor’s degree in Computer Science or related
- Proficient in English
- Basic knowledge of Python, ML, and Data libraries
- Basic knowledge of Databases
- Understanding of statistical, ML and deep learning algorithms
- Experience visualizing and manipulating big datasets
- Problem solving
- Bonus: AWS knowledge, (Py)Spark, Airflow, Data Lakes and Data Warehouses
Required Soft Skills
- Curiosity: You’re ambitious to learn and grow in different industries utilizing a modern tech stack.
- Autonomy and positivity: We’re a fully remote, globally distributed team.
- Teamwork: Enjoy a collaborative approach.
- Adaptability: Operate with a startup mindset and move at a startup pace.
- Dependable: You can be trusted to deliver high-quality work.
Benefits
- Every other Friday off
- Health Bonus
- Remote and flexible
- Paid sick days and local holidays
Please submit your CV in English.