Senior AI Platform Engineer - LLM Inference Backend

Posted 4 hours 19 minutes ago by JPMorgan Chase & Co.

Permanent
Full Time
Other
London, United Kingdom
Job Description

JPMorgan Chase & Co. is hiring a Software Engineer III in London to build backend services for LLM inference and scalable production systems.

You will work on routing, batching, scheduling, streaming responses, and quota management while improving APIs, observability, and reliability across the platform. You will explore model architectures, tokenization costs, and GPU utilization, collaborating with product teams and using enterprise AI tooling to boost performance and security of critical