Lizka

I'm the Content Specialist at the Centre for Effective Altruism; I run the non-engineering side of the EA Forum. I have a more detailed bio here.

Please feel free to reach out!

Wikitag Contributions

Comments

Sorted by

FYI: the paper is now out. 

See also the LW linkpost: METR: Measuring AI Ability to Complete Long Tasks, and a summary on Twitter

(IMO this is a really cool paper — very grateful to @Thomas Kwa et al. I'm looking forward to digging into the details.)