Skip to content
ReferMeAJob
LiveRemoteFull-timeApply by 1 Nov 2026

Data Engineer (LLM, RAG)

Remote

Experience
5–15 years
Employment
Full-time
Work mode
Remote
Salary
Not disclosed
Deadline
Apply by 1 Nov 2026
Posted
2025-04-15

Required skills

SkillExperienceLevel
LLM API5+ yearsExpert
LLM integration5+ yearsExpert
Large Language Model (LLM)5+ yearsExpert
RAG5+ yearsExpert
Python5+ yearsNot specified
Apache-Kafka5+ yearsExpert
Apache Spark5+ yearsExpert
Hadoop5+ yearsExpert
Machine Learning5+ yearsExpert
NumPy5+ yearsExpert
Pandas5+ yearsExpert
Pytorch5+ yearsExpert
SciKit-Learn5+ yearsExpert

About the role

JOB OVERVIEW:

Job Title: Data Engineer (LLM, RAG)

Company: Trantor Inc

Experience: 5-10 years

Location: Remote Working

No. of Positions: 1


MANDATORY CRITERIA:

  • Strong Data Engineer Profile
  • 5+ Years of Experience in statistical machine learning, deep learning, data mining
  • 2+ Years of Experience working with large language models (LLMs)
  • Proficiency in Python + at least one other programming language
  • Prompt engineering and retrieval-augmented generation (RAG) techniques
  • Experience with frameworks and libraries like PyTorch, Numpy, Pandas, SciPy, Scikit-Learn, LangChain, and Hugging Face Transformers


JOB RESPONSIBILITIES:

  • Design, train, fine-tune, and deploy LLMs using prompt engineering and RAG techniques
  • Build scalable solutions using frameworks and libraries like PyTorch, Hugging Face Transformers, and LangChain
  • Collaborate with cross-functional teams to deliver data-driven solutions
  • Develop efficient data pipelines using big data technologies like Spark and Hadoop
  • Ensure compliance with data privacy and ethical AI practices


REQUIRED SKILLS:

  • Python programming with working knowledge of Java, Scala, or C++
  • Experience with LLMs, including training and deployment
  • Prompt engineering and RAG techniques
  • Frameworks and libraries like PyTorch, Numpy, Pandas, SciPy, Scikit-Learn, LangChain, and Hugging Face Transformers
  • SQL and big data technologies like Hadoop and Spark
  • Cloud platforms like AWS, Azure, or Google Cloud


IDEAL CANDIDATE:

  • Master’s or Ph.D. in Computer Science, Statistics, or related field
  • 5+ years of experience in machine learning, deep learning, and statistical modeling
  • Strong proficiency in Python and other programming languages
  • Expertise in PyTorch, Hugging Face, LangChain, Scikit-Learn, Pandas, NumPy, and SciPy
  • Strong analytical and problem-solving abilities
  • Effective communication skills and ability to work in cross-functional teams


WHAT TO EXPECT:

  • Remote working with flexible work arrangements
  • Opportunity to work with a large-scale/global company

Role categories

Similar roles

AM

Remote

NewRemoteFull-timePythonAgile MethodologyArtificial IntelligenceTensorFlowPytorch
Not disclosed
Apply Now
ME

Remote

NewRemoteFull-timePythonTensorFlow
Not disclosed
Apply Now
DA

Remote

NewRemoteFull-timeSQLPythonR
Not disclosed
Apply Now
FD

Remote

RemoteFull-timeNode.JsAngular (All Versions)
Not disclosed
Apply Now