LessWrong AI
2026-09-17 22:56 UTC
By Matthew_Opitz
USR-0152-20260917-community-fo-63512e2d
If METR is overworked, how to alleviate the bottleneck?
I share the skepticism re: "Is METR a Meaningful Check on Anthropic?" Let's take it as a given that we need an independent, government-funded agency involving thousands of independent auditors to pace and supervise the frontier AI labs. Let's even take it as a given that Congress will soon allocate, let's generously say, billions of dollars per year to this new agency. Let's imagine that the Hugging Face Incident, or some even more concerning incident yet to occur or be disclosed, ends up functioning as our new "Sputnik Moment" for AI Alignment against existential risk. We would still have a problem: lack of qualified personnel with which to staff this new independent agency. I think we can all agree that just having a computer science degree does not really prepare someone for AI Alignment work, which is a pity because there are a lot of unemployed computer science majors out there. Like the US had to do to meet the 1950s Sputnik Moment, we would also need to overhaul the educational pipeline into this new field. There are two ways I could see this being done: Option #1: Fund state colleges to offer a new master's degree to go on top of a computer science degree. The new master's degree would aim to supplement computer science graduates with knowledge of topics in "Intellidynamics," as Liron Shapira puts it. These would be concepts like reward hacking, mesa-optimizers, timeless decision theory...basically all of the abstract game-theory sort of stuff that would be useful to…
I share the skepticism re: "Is METR a Meaningful Check on Anthropic?" Let's take it as a given that we need an independent, government-funded agency involving thousands of independent auditors to pace and supervise the frontier AI labs. Let's even take it as a given that Congress will soon allocate, let's generously say, billions of dollars per year to this new agency. Let's imagine that the Hugging Face Incident, or some even more concerning incident yet to occur or be disclosed, ends up functioning as our new "Sputnik Moment" for AI Alignment against existential risk. We would still have a problem: lack of qualified personnel with which to staff this new independent agency. I think we can all agree that just having a computer science degree does not really prepare someone for AI Alignment work, which is a pity because there are a lot of unemployed computer science majors out there. Like the US had to do to meet the 1950s Sputnik Moment, we would also need to overhaul the educational pipeline into this new field. There are two ways I could see this being done: Option #1: Fund state colleges to offer a new master's degree to go on top of a computer science degree. The new master's degree would aim to supplement computer science graduates with knowledge of topics in "Intellidynamics," as Liron Shapira puts it. These would be concepts like reward hacking, mesa-optimizers, timeless decision theory...basically all of the abstract game-theory sort of stuff that would be useful to…
Full article content could not be extracted automatically. Read the original below.
Source:
LessWrong AI
· lesswrong.com