Back to vacancies
Inferent · Artificial Intelligence / Research & Development
LLM Training & Inference Engineer (Python)
We are looking for an engineer specialized in training, evaluating, and deploying large language models (LLMs), with experience in Python, machine learning, and efficient inference systems.
México Full time · Mid-level
About the role
At Inferent, we are building infrastructure and products based on applied artificial intelligence.
We are looking for a technical profile specialized in large language models (LLMs), capable of working from experimentation and applied research through the deployment of intelligent systems in production.
Responsibilities:
- Design and run training and fine-tuning processes for LLMs.
- Build data preparation, cleaning, and evaluation pipelines.
- Implement efficient and scalable inference systems.
- Optimize model memory, latency, and performance.
- Experiment with modern artificial intelligence architectures.
- Integrate proprietary and external models into Inferent products.
- Build internal evaluation and benchmarking tools.
- Collaborate with engineering and product teams.
Responsibilities
- Design and run training and fine-tuning processes for LLMs.
- Build data preparation, cleaning, and evaluation pipelines.
- Implement efficient and scalable inference systems.
- Optimize model memory, latency, and performance.
- Experiment with modern artificial intelligence architectures.
- Integrate proprietary and external models into Inferent products.
- Build internal evaluation and benchmarking tools.
- Collaborate with engineering and product teams.
Requirements
- Advanced Python.
- PyTorch.
- TensorFlow.
- Transformers.
- Transformer architecture.
- Linux.
- Git.
- Machine learning.
Preferred
- • Hugging Face.
- • Llama.
- • Mistral.
- • Qwen.
- • Fine-tuning.
- • LoRA / QLoRA.
- • RAG.
- • Vector databases.
- • CUDA.
- • Docker.
- • Kubernetes.
- • C++.
- • Rust.
- • Go.