Anthropic brand banner

Anthropic Fellows Program, Reinforcement Learning

AnthropicWorldwidePosted 15d agoonsite
via Arbeitnow
<div class="content-intro"><h2><strong>About Anthropic</strong></h2> <p>Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.</p></div><p><a href="https://bit.ly/afpsafety"><strong>Apply using this link</strong></a><strong>. </strong>Applications for the next cohort of Anthropic Fellows close at <strong>11:59pm PT on July 26</strong>. The cohort is expected to start <strong>November 2</strong>. In some circumstances, we can accommodate fellows starting outside the usual cohort timelines — please note in your application if the November start date doesn't work for you.</p> <p>This page is specific to one of the Anthropic Fellows Workstreams, see also the main&nbsp;<strong><a href="https://job-boards.greenhouse.io/anthropic/jobs/5023394008">Anthropic Fellows posting</a></strong>.</p> <h1><strong>Anthropic Fellows Program overview</strong></h1> <p>The Anthropic Fellows Program is designed to foster AI research and engineering talent. We provide funding and mentorship to promising technical talent - regardless of previous experience.</p> <p>Fellows will primarily use external infrastructure (e.g. open-source models, public APIs) to work on an empirical project aligned with our research priorities, with the goal of producing a <strong>public output</strong> (e.g. a paper submission). In one of our earlier cohorts, over 80% of fellows produced papers.&nbsp;We run multiple cohorts of Fellows each year and review applications on a rolling basis.</p> <h2><strong>What to expect</strong></h2> <ul> <li>4 months of full-time research&nbsp;</li> <li>Direct mentorship from Anthropic researchers&nbsp;</li> <li>Access to a shared workspace (in either Berkeley, California or London, UK)</li> <li>Connection to the broader AI safety and security research community</li> <li>Weekly stipend of 3,850 USD / 2,310 GBP / 4,300 CAD + benefits (these vary by country)</li> <li>Funding for compute (~$15k/month) and other research expenses</li> </ul> <h2><strong>Interview process</strong></h2> <p>The interview process will include an initial application &amp; reference check, technical assessments &amp; interviews, and a research discussion.&nbsp;</p> <p><strong>We encourage you to apply even if you do not believe you meet every single qualification.</strong> Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.</p> <h2><strong>Compensation</strong></h2> <p>The expected base stipend for this role is 3,850 USD / 2,310 GBP / 4,300 CAD per week, with an expectation of 40 hours per week for 4 months (with possible extension).</p> <h2><strong>Fellows workstreams</strong></h2> <p>Due to the success of the <a href="https://alignment.anthropic.com/2024/anthropic-fellows-program/">Anthropic Fellows for AI Safety Research</a> program, we are now expanding it across teams at Anthropic. We expect there to be significant overlap in the types of skills and responsibilities across the roles and will by default consider candidates for all the workstreams.</p> <p>Some of the workstreams may include unique assessment steps; we therefore ask you for workstream preferences in the<a href="https://bit.ly/afpsafety"> application</a>. You can see an overview of the current workstreams below:</p> <ol> <li>AI Safety Fellows</li> <li>AI Security Fellows</li> <li>ML Systems &amp; Performance Fellows</li> <li>Reinforcement Learning Fellows</li> <li>Economics &amp; Societal Impacts Fellows</li> </ol> <p>This page is specific to one of the Anthropic Fellows Workstreams, see also the main&nbsp;<strong><a href="https://job-boards.greenhouse.io/anthropic/jobs/5023394008">Anthropic Fellows posting</a></strong>.</p> <h2><strong>Across the workstreams, you may be a good fit if you:</strong></h2> <ul> <li>Are motivated by making sure AI is safe and beneficial for society as a whole</li> <li>Are excited to transition into empirical AI research and would be interested in a full-time role at Anthropic</li> <li>Have a strong technical background in computer science, mathematics, or physics</li> <li>Thrive in fast-paced, collaborative environments</li> <li>Can implement ideas quickly and communicate clearly</li> </ul> <h2><strong>Strong candidates may also have:</strong></h2> <ul> <li>Strong background in a discipline relevant to a specific Fellows workstream (e.g. economics, social sciences, or cybersecurity)</li> <li>Experience in areas of research or engineering related to their workstream</li> </ul> <h2><strong>Candidates must be:</strong></h2> <ul> <li>Fluent in Python programming</li> <li>Available to work full-time on the Fellows program</li> </ul> <h1><strong>Reinforcement Learning Fellows</strong></h1> <h2><strong>Mentors, research areas, &amp; past projects</strong></h2> <p>Fellows will undergo a project selection &amp; mentor matching process. Potential research areas and mentors include:</p> <ul> <li>Ruhua Jiang</li> <li>Kaidi Cao</li> <li>Sunny Duan</li> <li>David Brandfonbrener</li> <li>Colt Steele</li> <li>Dino Distefano</li> <li>Will Williams</li> </ul> <p>Projects in this workstream may include:</p> <ul> <li>Building model-based tools to better understand AI training data and improve training data quality</li> <li>A research project to better understand generalization</li> <li>Creating RL environments to improve Claude models at capabilities that are within your domain of expertise</li> <li>Building RL environments for safety-related tasks</li> <li>Conducting research and implementing solutions in areas such as RL algorithms</li> </ul> <h2><strong>Unique candidate criteria</strong></h2> <p>You might be a particularly great fit for this workstream if you:</p> <ul> <li>Have strong software engineering skills with experience building complex ML systems</li> <li>Can balance research exploration with engineering rigor and operational reliability</li> <li>Enjoy collaborating across research and engineering disciplines</li> <li>Are comfortable working with large-scale distributed systems and high-performance computing</li> <li>Have experience with training, fine-tuning, or evaluating large language models</li> <li>Are adept at analyzing and debugging model training processes</li> </ul> <h1><strong>Logistics</strong></h1> <p><strong>Logistics Requirements: </strong>To participate in the Fellows program, you must have work authorization in the US, UK, or Canada and be located in that country during the program.</p> <p><strong>Workspace Locations:</strong> We have designated shared workspaces in London and Berkeley where fellows will work from and mentors will visit. <strong>We are also open to remote fellows in the UK, US, or Canada</strong>. We will ask you about your availability to work from Berkeley or London (full- or part-time) during the program.</p> <p><strong>Visa Sponsorship: </strong>We are <strong>not</strong> currently able to sponsor visas for fellows. To participate in the Fellows program, you need to have or independently obtain full-time work authorization in the UK, the US, or Canada.</p> <p><strong>Program Duration: </strong>The program runs for 4 months, full-time. If you can't commit to the full duration, please still apply and note your constraints in the application. We review these requests on a case-by-case basis.</p> <p><strong>Please note:</strong> We do not guarantee that we will make any full-time offers to fellows. However, strong performance during the program may indicate that a Fellow would be a good fit for full-time roles at Anthropic. In previous cohorts, 25-50% of fellows received a full-time offer, and we’ve supported many more to go on to do great work on AI safety and security at other organizations.</p> <p>Applications and interviews are managed by <a class="underline underline underline-offset-2 decoration-1 decoration-current/40 hover:decoration-current focus:decoration-current" href="https://www.constellation.org/" target="_blank">Constellation</a>, our recruiting partner. Clicking "Apply here" will take you to their portal, and updates will come from a Constellation address. Constellation also runs the Berkeley workspace and provides program support for fellows working on AI safety and security; fellows on capabilities-focused projects are supported directly by Anthropic. All applicants currently use the same application portal but we are working to separate applications for safety/security and capabilities focused projects in future rounds.</p> <h1><a href="https://bit.ly/afpsafety"><strong>Apply here</strong></a></h1> <p><span style="text-decoration: underline;">The below are Anthropic's policies for full time roles. These do <strong>NOT</strong> apply to the Fellows Program.</span></p><div class="content-conclusion"><h2><strong>Logistics</strong></h2> <p><strong>Minimum education: </strong>Bachelor’s degree or an equivalent combination of education, training, and/or experience</p> <p><strong>Required field of study:&nbsp;</strong>A field relevant to the role as demonstrated through coursework, training, or professional experience</p> <p><strong>Minimum years of experience: </strong>Years of experience required will correlate with the internal job level requirements for the position</p> <p><strong>Location-based hybrid policy:</strong> Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.</p> <p><strong data-stringify-type="bold">Visa sponsorship:</strong>&nbsp;We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.</p> <p><strong>We encourage you to apply even if you do not believe you meet every single qualification.</strong> Not all strong candidates will meet every single qualification as listed.&nbsp; Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.<br><br><strong data-stringify-type="bold">Your safety matters to us.</strong> To protect yourself from potential scams, remember that Anthropic recruiters only contact you ----- addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit&nbsp;<u data-stringify-type="underline"><a class="c-link c-link--underline" href="http://anthropic.com/careers" target="_blank" data-stringify-link="http://anthropic.com/careers" data-sk="tooltip_parent" data-remove-tab-index="true">anthropic.com/careers</a></u>&nbsp;directly for confirmed position openings.</p> <h2><strong>How we're different</strong></h2> <p>We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.</p> <p>The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI &amp; Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.</p> <h2><strong>Come work with us!</strong></h2> <p>Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. <strong data-stringify-type="bold">Guidance on Candidates' AI Usage:</strong>&nbsp;Learn about&nbsp;<a class="c-link" href="https://www.anthropic.com/candidate-ai-guidance" target="_blank" data-stringify-link="https://www.anthropic.com/candidate-ai-guidance" data-sk="tooltip_parent">our policy</a> for using AI in our application process.</p></div>

Find Jobs in United Kingdom on Arbeitnow

Skills

  • AI Research & Engineering

Similar roles

  • PlanetScale logo

    Developer Educator

    PlanetScale·Worldwide

    PlanetScale is growing rapidly and reinventing the database space. Our database platform offers developers the fastest and most reliable Postgres and Vitess+MySQL databases, enabling businesses to efficiently handle data and workloads of all scales. We also take DX seriously. Our platform enables developers with deep performance insights, AI integration, powerful cluster management, and schema migration tooling. Our customers trust us…

    • Remote
    • Full-time
  • Postscript logo

    Business Development Representative

    Postscript·Worldwide

    Postscript is the AI messaging platform trusted by 20,000+ Shopify brands — including Brooklinen, Ruggable, True Classic, and Dr. Squatch. With a mission to make SMS your number one revenue driving channel, we built the best-in-class SMS marketing platform and launched revenue-driving AI features not available on any other platform. And we're just getting started. Backed by Greylock and Y…

    • Remote
    • Full-time
  • Mixpanel logo

    Revenue Strategy &amp; Operations Manager

    Mixpanel·Worldwide

    About Mixpanel Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how people use the products they build. By combining powerful analytics with AI that knows your business, Mixpanel helps teams see what’s working, diagnose what’s not, and decide what to build next. Learn more at mixpanel.com. About the Team The…

    • Remote
    • Full-time
  • Postscript logo

    Director, Email Deliverability

    Postscript·Worldwide

    Postscript is the AI messaging platform trusted by 20,000+ Shopify brands — including Brooklinen, Ruggable, True Classic, and Dr. Squatch. With a mission to make SMS your number one revenue driving channel, we built the best-in-class SMS marketing platform and launched revenue-driving AI features not available on any other platform. And we're just getting started. Backed by Greylock and Y…

    • Remote
    • Full-time
  • Pipedrive logo

    Principal AI/ML Scientist & Engineer

    Pipedrive·Berlin, Germany

    We believe it takes great people to create a great product. That’s why our team lives our company values, and we hire based on them, too. Since 2010, Pipedrive has been on a mission to support sales and marketing teams with easy-to-use, powerful tools that make everyday work faster and easier. Today, our cloud-based software is trusted by over 100,000…

    • On-site
    • Full-time
    • Blue Cardtied; settle 21–33mo
  • Freshworks logo

    Lead - Solution Engineer(Dutch speaking)

    Freshworks·London, United Kingdom

    Freshworks is looking for a self-motivated Lead Solutions Engineer to complement our Pre-Sales team. You will play a critical role by being the technical expert that helps the sales team with new business opportunities as well as consulting or selling to the most strategic prospects and accounts at Freshworks. Lead Solutions Engineers are responsible for articulating our value proposition, architecting…

    • Remote
    • Full-time
  • Filtronic logo

    Lead RF Engineer

    Filtronic·Sedgefield, United Kingdom

    Filtronic is looking for a Lead RF Engineer to join our team, based at our new state-of-the-art facility in County Durham UK, to work on a range of exciting products from concept through to production. With a wide range of design tools, experienced engineers and in-house RF/mmW manufacturing capabilities, you will play a key role in the design and delivery…

    • Hybrid
    • Full-time
  • Harvey Water Softeners logo

    Plumbing and Maintenance Service Engineer

    Harvey Water Softeners·London, United Kingdom

    Culligan Harvey Water Softeners are looking for a Service Engineer to join our growing team. Everyone at Culligan Harvey’s is responsible for delivering an amazing customer experience and living the company’s values, dare to care, do the right thing, keep it simple, yes I can and make it fun. You will be supporting the business by attending service calls in…

    • Full-time