LLM & Generative AI Engineer
Posted 6 days 9 hours ago by GenixBit Labs Pvt. Ltd.
Permanent
Full Time
Other
London, United Kingdom
Job Description
About the Role 
Join our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks.
Key Responsibilities- Fine-tune open-source models (Llama, Mistral, Qwen) for specific domain functions
- Optimize model deployment pipelines for low latency and high throughput
- Build advanced context management and semantic search solutions
- Implement prompt evaluation frameworks and guardrail architectures
- 3+ years of experience focusing on Natural Language Processing and Generative AI
- Hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA)
- Experience deploying models with vLLM, Ollama, or Triton Inference Server
- Strong background in software engineering best practices and clean code