Research Engineer / Research Scientist
Pre-training, Large Language Models, Multimodal AI
Location
Zürich, Switzerland
Work type
Hybrid
Employment
Full Time
Experience
5-10 years
Compensation
Fr280K - Fr680K per year
Posted
1d ago
Summary and responsibilities
Role overview
Summary
As a Research Engineer/Scientist on the Pre-training team, you will contribute to developing the next generation of large language models with multimodal capabilities. This role involves conducting cutting-edge research, implementing solutions in areas like model architecture and algorithms, and optimizing training infrastructure for safe and steerable AI systems.
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About the team
We are seeking passionate Research Scientists and Engineers to join our growing Pre-training team in Zurich. We are involved in developing the next generation of large language models. The team primarily focuses on multimodal capabilities: giving LLMs the ability to understand and interact with modalities other than text.
In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.
Responsibilities
Conduct research and implement solutions in areas such as model architecture, algorithms, data processing, and optimizer development
Independently lead small research projects while collaborating with team members on larger initiatives
Design, run, and analyze scientific experiments to advance our understanding of large language models
Optimize and scale our training infrastructure to improve efficiency and reliability
Develop and improve dev tooling to enhance team productivity
Contribute to the entire stack, from low-level optimizations to high-level model design
Qualifications & Experience
Degree (BA required, MS or PhD preferred) in Computer Science, Machine Learning, or a related field
Strong software engineering skills with a proven track record of building complex systems
Expertise in Python and deep learning frameworks
Have worked on high-performance, large-scale ML systems, particularly in the context of language modeling
Familiarity with ML Accelerators, Kubernetes, and large-scale data processing
Strong problem-solving skills and a results-oriented mindset
Excellent communication skills and ability to work in a collaborative environment
You'll thrive in this role if you
Have significant software engineering experience
Are able to balance research goals with practical engineering constraints
Are happy to take on tasks outside your job description to support the team
Enjoy pair programming and collaborative work
Are eager to learn more about machine learning research
Are enthusiastic to work at an organization that functions as a single, cohesive team pursuing large-scale AI research projects
Have ambitious goals for AI safety and general progress in the next few years, and you’re excited to create the best outcomes over the long-term
Sample Projects
Optimizing the throughput of novel attention mechanisms
Proposing Transformer variants, and experimentally comparing their performance
Preparing large-scale datasets for model consumption
Scaling distributed training jobs to thousands of accelerators
Designing fault tolerance strategies for training infrastructure
Creating interactive visualizations of model internals, such as attention patterns
If you're excited about pushing the boundaries of AI while prioritizing safety and ethics, we want to hear from you!
Logistics
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
How we're different
We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.
The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Come work with us!
Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
Updated 23h ago
Candidate fit
Skills and qualifications
Additional skills
Experience
5-10 years
How this role is positioned
Role classification
Job domains
Industries
Employment
Full Time
Contract duration
Permanent
Hiring type
Direct
Global hiring
Location specific
Offer details
Compensation and benefits
Compensation
Fr280K - Fr680K per year
Benefits and perks
Location, schedule, and role shape
Work setup
Work conditions
Bandwidth profile
Context on the employer
Company snapshot
Company
Anthropic
Team size
Growing team
Location
Zürich, Switzerland
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
Popular Domains
Explore opportunities across specialized functional areas.
Trending Industries
Discover roles in the world's most innovative sectors.
Research Engineer / Research Scientist
Zürich, Switzerland • Full Time