JPMorgan Chase is hiring a Software Engineer III for its Data Analytics Platform in London, onsite, full time. This role blends infrastructure and ML to scale LLM serving, building core backends for routing, batching, streaming, and quotas, shaping APIs and SDKs, and driving end-to-end performance and reliability with observability and autoscaling. You should be proficient in Go, Python, or TypeScript, with a strong CS foundation and curiosity about distributed systems and LLM fundamentals like batching, token costs, KV cache, and context-length tradeoffs. Preferred: Kubernetes, gRPC, Redis, SLOs, and AI-assisted development tooling; secure coding. Apply with concrete, data-driven results, showcase CI/CD and incident response, and highlight collaboration and safety practices.
We have an opportunity to impact your career and provide an adventure where you can push the limits of what's possible.
As a Software Engineer III at JPMorganChase within the Firmwide LLM Serving Platform team, you are an integral part of an agile team that designs, builds, and operates the services that make large language models usable at scale. This is an infrastructure-meets-ML role: you don't need to be an ML researcher, but you should be excited to learn how model architectures and inference constraints translate into real production systems. You will contribute to a living platform where we optimize performance — pushing down latency, increasing throughput, maximizing GPU utilization, and eliminating waste across the request lifecycle.
Job responsibilities
Build core backend services for LLM inference, including request routing, batching, scheduling, streaming responses, and quota/limits.
Implement and maintain APIs and SDKs used by product and application teams across the firm.
Profile and optimize performance end-to-end across CPU, memory, network, serialization, concurrency, GPU utilization, and caching.
Improve reliability and operability through health checks, graceful degradation, autoscaling behaviors, incident follow-ups, and runbooks.
Contribute to system design by breaking down ambiguous problems, proposing approaches, and making pragmatic tradeoffs.
Add observability with metrics, tracing, logging, dashboards, and actionable alerts tied to SLOs.
Support safe deployments through CI/CD improvements, canarying, feature flags, backward compatibility, and rollback plans.
Learn LLM serving fundamentals — tokenization costs, KV cache, quantization, context length tradeoffs, throughput vs. latency.
Required qualifications, capabilities, and skills
Formal training or certification on software engineering concepts and applied experience.
Bachelor's Degree in Computer Science or equivalent.
Solid programming fundamentals: data structures, concurrency basics, debugging, testing.
Comfort working in one or more of Go, Python, or TypeScript, with the ability to ramp up quickly on the others.
Interest in distributed systems and system design, even if you haven't built large systems yet.
Curiosity about LLMs and AI model architecture, with willingness to learn quickly.
A measurement-driven mindset: you like profiling, benchmarking, and proving improvements with data.
Preferred qualifications, capabilities, and skills
Experience with performance profiling tools such as pprof, flamegraphs, or distributed tracing systems.
Familiarity with containers and orchestration (Docker, Kubernetes) and service-to-service networking.
Understanding of inference concepts: batching, streaming tokens, GPU memory constraints, KV cache.
Experience with high-throughput APIs (gRPC/HTTP), eventing/queues, or caching layers such as Redis.
Exposure to reliability practices: SLOs/SLIs, on-call rotations, incident reviews.
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.