Back to search
Listing unavailableData & AnalyticsHybrid

AI Inference Internship

Perplexity · Hybrid

This listing is no longer available.

MeritLog keeps this source-backed description for reference. Availability is not verified, and there is no application link here.

Last seen by MeritLog August 17, 2026Source: AshbySource version: ashby-public-job-posting-v1

Source: the employer's Ashby job board. Open the original listing for current details. Availability is not verified for this retained page.

Job details

Work model
Hybrid
Salary
Not listed by source
Location
London

Job description

Perplexity is excited to announce the Internship Program for exceptional Master’s or PhD students studying Computer Science or Engineering in the UK, enrolled in the 2025-2026 academic year. This is an intensive program in which you will work directly with our AI Inference team. This program offers a unique opportunity to gain valuable experience in a rapidly growing AI startup. Outstanding performers might be offered a full time position at the end of the program. Our AI Inference team is responsible for running the models behind the Perplexity products. The team maintains the inference engine and deployments behind models ranging from single-node embeddings to distributed sparse Mixture-of-Experts models, maintaining large GPU clusters. With a keen focus on latency and throughput, the Inference team is responsible for the entire serving stack, from GPU kernels to networking and monitoring infrastructure.  Responsibilities - Work with the inference team to improve serving latency and throughput - Bring up support for new models and state-of-the art inference optimizations or quantization schemes - Optimize inference across the entire stack, from GPU kernels to serving endpoints Qualifications - Strong engineering track record with proven knowledge of fundamentals and programming languages (multi-threaded programming, networking, compilation, systems programming, etc) - Pursuing a Master's or PhD in Computer Science with a focus on performance-related subjects (HPC, Compilers, Distributed Systems) - Experience with ML frameworks (Torch, JAX)  - Experience with GPU programming (CUDA, Triton) - Experience with High-Performance Computing (OpenMPI) Schedule - Internship program: 13 weeks, full-time or part-time, in-person in London office (hybrid schedule: 3 days from the office, 2 days WFH) Interview Process - Fill out the application on Perplexity website - If selected, People Ops and technical interviews will be involved. - Offer. We’re impressed! We’d love to welcome you to our Internship program! - Start. We have a desk waiting for you in our London office!  FAQ Do you sponsor visas? Can I apply if I need a visa to work in the UK? → Unfortunately we are unable to sponsor visas What if I’m on a student visa? → You need to seek approval from your University (to determine if you are eligible to work full time or part time only)  How many internship spots are there? - We have spots for 2-3 interns in our 2026 class. Is housing provided? - Unfortunately we cannot provide housing.  Is health insurance provided? - Unfortunately we cannot provide health insurance for interns. Full time employees receive full health insurance and benefits. How many full time offers are available at the end of the residency? - There is no limit. All outstanding performers will be given a full time offer! At Perplexity, we've experienced tremendous growth and adoption since publicly launching the world's first fully functional conversational answer engine in 2022. We've grown from answering 2.5 million questions per day at the start of 2024 to around 20 million daily queries in December 2024. We also offer Perplexity Enterprise Pro, which counts leading companies like Nvidia, the Cleveland Cavaliers, Bridgewater, and Zoom as customers. To support our rapid expansion, we've raised significant funding from some of the most respected technology investors. Our investor base includes IVP, NEA, Jeff Bezos, NVIDIA, Databricks, Bessemer Venture Partners, Elad Gil, Nat Friedman, Daniel Gross, Naval Ravikant, Tobi Lutke, and many other visionary individuals. In 2024, our employee base grew nearly 300%, and we're just getting started. Final offer amounts are determined by multiple factors, including experience and expertise, and may vary from the amounts listed above.

Keep exploring

Available Data & Analytics roles

These current listings are available to explore now.

Search all jobs

Privacy choices

Analytics and advertising stay off unless you allow them. Private data stays out.

Read the privacy notice