Data & Machine Learning Engineer

Bogotá remote Until 9/1/2026 8+ years exp H-1B sponsor history First posted July 3, 2026 Last posted July 3, 2026
Job description
This is a full-time  opportunity for a star Data/ML Engineer from LATAM.  In-person verification will be conducted.
 

IDT Corporation is a global communications company founded in 1990 and headquartered in Newark, New Jersey. We are industry leaders in prepaid communication and payment services and one of the largest international voice carriers. We are listed on the NYSE, employ over 1800 people across 20 countries, and have over $1.5 billion in revenues. 

IDT is not ”another big IT corporation”— we encourage and support in-house entrepreneurs in developing their ideas into business actions.

Our flagship brand, BOSS Revolution, offers Money Transfer, International Calling, and Mobile Top-Up services and supports IDT’s mission of enabling people to keep in touch and share resources with family and friends worldwide.

 
We are looking for a skilled Data/ML Engineer to join our BI team and take an active role in designing, building, and maintaining the end-to-end data pipeline, architecture and design that powers our warehouse, LLM-driven applications, and AI-based BI. If you're looking for a company that will give you the maximum flexibility in choosing a location to work, this opportunity is for you!

Responsibilities:

  • Design, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model-training workflows, and real-time inference services.
  • Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.
  • Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI-backed insights.
  • Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine-tuning routines to power conversational interfaces.
  • Oversee vector data stores and develop efficient indexing methodologies to support retrieval-augmented generation (RAG) workflows.
  • Partner with data stakeholders to gather requirements for language-model initiatives and translate into scalable solutions.
  • Create and maintain comprehensive documentation for all data processes, workflows and model deployment routines.
  • Should be willing to stay informed and learn emerging methodologies in data engineering, MLOps and LLM operations.
  •  

Requirements:

  • 8+ years of experience as a Data Engineer with 2+ years focused on MLOps.
  • Excellent English communication skills.
  • Effective oral and written communication skills with BI team and user community.
  • Demonstrated experience in utilizing python for data engineering tasks, including transformation, advanced data manipulation, and large-scale data processing.
  • Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows.
  • Skilled at integrating open-source LLM frameworks into data engineering workflows for end-to-end model training, customization, and scalable inference. 
  • Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments.
  • Hands-on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real-time data ingestion. 
  • Experience designing complex data pipelines extracting data from RDBMS, JSON, API and Flat file sources.
  • Demonstrated skills in SQL and PLSQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, along with hands-on experience in one or more relational database systems and cloud-based database services such as Snowflake/Redshift.
  • Understanding of software engineering principles and skills working on Unix/Linux/Windows Operating systems, and experience with Agile methodologies.
  • Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment.
  • Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data-driven decision-making and strategic insights. 
  •  

Pluses

  • Experience with vector databases such as DataStax AstraDB, and developing LLM-powered applications using popular open source frameworks like LangChain and LlamaIndex–including prompt engineering, retrieval-augmented generation (RAG), and orchestration of intelligent workflows. 
  • Familiarity with evaluating and integrating open-source LLM frameworks–such as Hugging Face Transformers/LLaMA-4 across end-to-end workflows, including fine-tuning and inference optimization.
  • Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.
  •  

What we offer:

  • Remote work opportunity!
  • B2B Employment ($, gross).
  • Stable job with long-term growth perspective with talented people around.
  • Really good hardware.
  • Great learning and growth opportunities.
  • Compensation for professional training, seminars, and conferences.
  • Referral program – get rewarded for helping us grow the team with talented people.
  • Company-supported English classes to enhance your professional growth.
Please attach CV in English.
The interview process will be conducted in English.
 
In-person verification will be conducted. Fake profiles will be reported.
 
Only accepting applicants from LATAM.
About this role

Summary

Design and maintain data pipelines, architecture, and ML workflows for AI-based BI.

Job title

Data & Machine Learning Engineer

Experience level

8+ years

Minimum experience

8+ years exp

Industry

communication

Location requirements

In-person verification required in Bogotá, remote work allowed

Salary

Not specified

Visa sponsorship

H-1B sponsor history

Management role

No

Skills & keywords

Required skills

pythonvector databasesllm frameworkscloud platformsbig data technologiessqlbusiness intelligencerelational databasessoftware engineering

Preferred skills

vector databases such as DataStax AstraDBllm frameworks like Hugging Face Transformersprompt engineeringretrieval-augmented generation (RAG)MLOps toolingCI/CD pipelines

Specializations

data pipelinesmlopsvector databasesllm frameworkssemantic retrieval
Locations

Structured locations inferred from the posting.

Bogotá, Bogota, Colombia

Hybrid City