Data & Machine Learning Engineer

IDT Corp
  • San Juan, PR
    30+ days ago

    Job Description

    Data & Machine Learning Engineer

    Bogotá / Santiago / Brasilia / Mexico / Lima / Asunción / Montevideo / Sucre / Guatemala City / Managua / San Salvador / Buenos Aires / Panama / puerto rico

    IDT Corporation - Technology DW /

    BizIntel /

    Remote

    apply for this job

    This is a full-time opportunity for a star Data/ML Engineer from LATAM. In-person verification will be conducted.

    IDT(www.idt.net) is an American telecommunications company founded in 1990 and headquartered in New Jersey. Today it is an industry leader in prepaid communication and payment services and one of the world's largest international voice carriers. We are listed on the NYSE, employ over 1300 people across 20+ countries, and have revenues in excess of $1.5 billion.

    We are looking for a skilled Data/ML Engineer to join our BI team and take an active role in designing, building, and maintaining the end-to-end data pipeline, architecture and design that powers our warehouse, LLM-driven applications, and AI-based BI. If you're looking for a company that will give you the maximum flexibility in choosing a location to work, this opportunity is for you!

    Responsibilities:

    • Design, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model-training workflows, and real-time inference services.
    • Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.
    • Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI-backed insights.
    • Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine-tuning routines to power conversational interfaces.
    • Oversee vector data stores and develop efficient indexing methodologies to support retrieval-augmented generation (RAG) workflows.
    • Partner with data stakeholders to gather requirements for language-model initiatives and translate into scalable solutions.
    • Create and maintain comprehensive documentation for all data processes, workflows and model deployment routines.
    • Should be willing to stay informed and learn emerging methodologies in data engineering, MLOps and LLM operations.

    Requirements:

    • 8+ years of experience as a Data Engineer with 2+ years focused on MLOps.
    • Excellent English communication skills.
    • Effective oral and written communication skills with BI team and user community.
    • Demonstrated experience in utilizing python for data engineering tasks, including transformation, advanced data manipulation, and large-scale data processing.
    • Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows.
    • Skilled at integrating open-source LLM frameworks into data engineering workflows for end-to-end model training, customization, and scalable inference.
    • Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments.
    • Hands-on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real-time data ingestion.
    • Experience designing complex data pipelines extracting data from RDBMS, JSON, API and Flat file sources.
    • Demonstrated skills in SQL and PLSQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, along with hands-on experience in one or more relational database systems and cloud-based database services such as Snowflake/Redshift.
    • Understanding of software engineering principles and skills working on Unix/Linux/Windows Operating systems, and experience with Agile methodologies.
    • Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment.
    • Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data-driven decision-making and strategic insights.

    Pluses

    • Experience with vector databases such as DataStax AstraDB, and developing LLM-powered applications using popular open source frameworks like LangChain and LlamaIndex-including prompt engineering, retrieval-augmented generation (RAG), and orchestration of intelligent workflows.
    • Familiarity with evaluating and integrating open-source LLM frameworks-such as Hugging Face Transformers/LLaMA-4 across end-to-end workflows, including fine-tuning and inference optimization.
    • Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.

    Please attach CV in English.

    The interview process will be conducted in English.

    In-person verification will be conducted. Fake profiles will be reported.

    Only accepting applicants from LATAM.

    We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

    apply for this job

    Numbers & Facts

    LocationSan Juan, PR

    Skills

    • Agile Programming Methodologiesunmatched
    • Amazon Web Services (AWS)unmatched
    • Apache Hadoopunmatched
    • Apache Kafkaunmatched
    • Apache Sparkunmatched
    • Application Programming Interface (API)unmatched
    • Artificial Intelligence (AI)unmatched
    • Big Dataunmatched
    • Business Intelligenceunmatched
    • Business Operationsunmatched
    • Cloud Computingunmatched
    • Communication Skillsunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Integrationunmatched
    • Data Collectionunmatched
    • Data Managementunmatched
    • Data Processingunmatched
    • Data Warehousingunmatched
    • Database Architectureunmatched
    • Database Technologyunmatched
    • Develop Methodologiesunmatched
    • Documentationunmatched
    • Engineering Managementunmatched
    • English Languageunmatched
    • JSONunmatched
    • Linux Operating Systemunmatched
    • Machine Learningunmatched
    • Machine Toolunmatched
    • Microsoft Windows Azureunmatched
    • Microsoft Windows Operating Systemunmatched
    • Modeling Languagesunmatched
    • Open Sourceunmatched
    • Open Source Application Frameworksunmatched
    • Oracle PL-SQLunmatched
    • Presentation/Verbal Skillsunmatched
    • Profit & Lossunmatched
    • Python Programming/Scripting Languageunmatched
    • Relational Databases (RDBMS)unmatched
    • Reporting Dashboardsunmatched
    • Requirements Managementunmatched
    • SQL (Structured Query Language)unmatched
    • Sales Pipelineunmatched
    • Scalable System Developmentunmatched
    • Snowflake Schemaunmatched
    • Software Engineeringunmatched
    • Source Code/Configuration Management (SCM)unmatched
    • Telecommunications Industryunmatched
    • Unix Operating Systemsunmatched
    • Unstructured Dataunmatched
    • Warehousingunmatched
    • Writing Skillsunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder