About this role
Salary: £57,000 - 73,000 per year
Requirements: We are looking for an advanced degree (MS or PhD) in Computer Science, Machine Learning, or a related field, or an equivalent combination of education, training, and/or experience.We want strong software engineering skills with a proven track record of building complex systems.We require expertise in Python and experience with deep learning frameworks, preferably PyTorch.We value familiarity with large-scale machine learning, particularly in the context of language models.We look for the ability to balance research goals with practical engineering constraints.We need strong problem-solving skills and a results-oriented mindset.We expect excellent communication skills and the ability to work in a collaborative environment.We care about candidates who are thoughtful about the societal impacts of their work.Preferred experience includes working on high-performance, large-scale ML systems.Preferred experience includes familiarity with GPUs, Kubernetes, and OS internals.Preferred experience includes language modeling using transformer architectures.Preferred experience includes knowledge of reinforcement learning techniques.Preferred experience includes a background in large-scale ETL processes.We ask for a bachelors degree or equivalent, with a field of study relevant to the role demonstrated through coursework, training, or professional experience.Minimum years of experience will correlate with the internal job level requirements for the position. Responsibilities: We conduct research and implement solutions in model architecture, algorithms, data processing, and optimizer development.We independently lead small research projects while collaborating with team members on larger initiatives.We design, run, and analyze scientific experiments to advance our understanding of large language models.We optimize and scale our training infrastructure to improve efficiency and reliability.We develop and improve dev tooling to enhance team productivity.We contribute to the entire stack, from low-level optimizations to high-level model design.We may take on sample projects such as optimizing the throughput of novel attention mechanisms, comparing compute efficiency of Transformer variants, preparing large-scale datasets, scaling distributed training jobs, designing fault tolerance strategies, and creating interactive visualizations of model internals. Technologies: AIETLKubernetesMachine LearningPyTorchPythonSupport More:
We are Anthropic, a public benefit corporation headquartered in San Francisco, focused on creating reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society. We work as a single cohesive, highly collaborative team on a few large-scale research efforts, and we value impact, communication, and rigorous empirical science. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, a lovely office space, and a location-based hybrid policy with an expectation that staff be in our offices at least 25% of the time. We also sponsor visas where possible and encourage applications from candidates of all backgrounds, including underrepresented groups.
last updated 36 week of 2026