Senior Data Engineer

Idt · Bogotá (Remote) · posted Sep 22, 2026

Open to candidates in 26 countries

Argentina, Belize, Bolivia, Brazil, Chile, Colombia, Costa Rica, Cuba, Dominican Republic, Ecuador, El Salvador, Guatemala, Guyana, Haiti, Honduras, Jamaica, Mexico, Nicaragua, Panama, Paraguay, Peru, Puerto Rico, Suriname, Trinidad and Tobago, Uruguay, Venezuela

More remote jobs open to candidates in Brazil, Colombia, Mexico and Argentina

Full-timesenior
Can this job hire you?

We never charge to apply.

What this role actually asks for

Extracted by RemoteHunt

Must have

  • •8+ years of experience as a Data Engineer
  • •Python for data engineering
  • •Big data technologies (Spark, Hadoop, Kafka)
  • •Complex data pipeline design
  • •SQL and PLSQL mastery
  • •BI and data warehouse methodologies
  • •English communication skills

Nice to have

  • •Vector databases (DataStax AstraDB)
  • •LLM application development (LangChain, LlamaIndex)
  • •Prompt engineering, RAG, orchestration
  • •Open-source LLM frameworks (Hugging Face, LLaMA-4)
  • •MLOps tooling and CI/CD pipelines

Tools and technologies

PythonApache SparkHadoopKafkaSQLPLSQLSnowflakeRedshiftUnixLinuxWindowsGitDataStax AstraDBLangChainLlamaIndex

Worth checking before you apply

  • ⚠in-person verification

The full posting

This is a full-time opportunity for a Senior Data Engineer from LATAM.

Only accepting applicants from LATAM. In-person verification will be conducted.

We are looking for a senior Data Engineer to join our BI team and take an active role in designing, building, and maintaining the end-to-end data pipeline, architecture and design that powers our warehouse, LLM-driven applications, and AI-based BI.

Responsibilities::

  • Design, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model-training workflows, and real-time inference services.
  • Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.
  • Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI-backed insights.
  • Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine-tuning routines to power conversational interfaces.
  • Oversee vector data stores and develop efficient indexing methodologies to support retrieval-augmented generation (RAG) workflows.
  • Partner with data stakeholders to gather requirements for language-model initiatives and translate into scalable solutions.
  • Create and maintain comprehensive documentation for all data processes, workflows and model deployment routines.
  • Should be willing to stay informed and learn emerging methodologies in data engineering, MLOps and LLM operations.

Requirements::

  • 8+ years of experience as a Data Engineer.

  • Demonstrated experience in utilizing Python for data engineering tasks, including transformation, advanced data manipulation, and large-scale data processing.

  • Hands-on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real-time data ingestion.

  • Experience designing complex data pipelines extracting data from RDBMS, JSON, API, and Flat file sources.

  • Demonstrated skills in SQL and PLSQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, along with hands-on experience in one or more relational database systems and cloud-based database service s such as Snowflake/Redshift

  • Understanding of software engineering principles and skills working on Unix/Linux/Windows Operating systems, and experience with Agile methodologies.

  • Proficiency in version control systems, with experience in managing code repositories , branching, merging, and collaborating within a distributed development environmen t.

  • Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data-driven decision-making and strategic insights.

  • Effective oral and written English communication skills with BI team and user community.

Pluses:

  • Experience with vector databases such as DataStax AstraDB, and developing LLM-powered applications using popular open source frameworks like LangChain and LlamaIndex–including prompt engineering, retrieval-augmented generation (RAG), and orchestration of intelligent workflows.
  • Familiarity with evaluating and integrating open-source LLM frameworks–such as Hugging Face Transformers/LLaMA-4 across end-to-end workflows, including fine-tuning and inference optimization.
  • Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.

What we offer::

  • Remote work opportunity!
  • B2B Employment ($, gross).
  • Stable job with long-term growth perspective with talented people around.
  • Really good hardware.
  • Great learning and growth opportunities.
  • Compensation for professional training, seminars, and conferences.
  • Referral program – get rewarded for helping us grow the team with talented people.
  • Company-supported English classes to enhance your professional growth.
Send this job to a friend:TelegramWhatsApp

Similar remote jobs

Get new jobs like this by email

Once a week. Remote data engineer jobs that can hire you in your country.

Hiring in: your countryChange

Weekly, only when there are at least 3 new jobs. No account needed. Unsubscribe in one click. You never pay to apply. Privacy

Is this one actually worth your time?

RemoteHunt scores every remote job 0–100 against your own resume, so you apply to the handful that fit instead of the hundred that don't. Free plan, no card required.