RSSAmplifier

Blog

Foxy Scout

Foxy Scout

foxy-scout.comRSS feed ↗15 posts

Latest posts

My 2025 AI predictions and 2024 evaluations

Below, I evaluate my 2024 AI forecasts then register my 2025 forecasts . Evaluating 2024 forecasts GPT-4.5 and end of 2024 capability forecasts In Feb 2024, I made some forecasts about GPT-4.5 capabilities. Unfortunately GPT-4.5 hasn’t been released, so theoretically this hasn’

What do you mean by "overconfident"?

The word "overconfident" seems overloaded. Here are some things I think that people sometimes mean when they say someone is overconfident: They gave a binary probability that is too far from 50% (I believe this is the original one) They overestimated a binary probability (e.g. they said

A template for GPT-4.5 Forecasting (and my forecasts)

I made a spreadsheet for forecasting the 10th/50th/90th percentile for how you think GPT-4.5 will do on various benchmarks (given 6 months after the release to allow for actually being applied to the benchmark, and post-training enhancements). Copy it here to register your forecasts . If

Brief thoughts on forecasting/epistemics interventions

All views are my own rather than those of any organizations/groups that I’m affiliated with. Trying to share my current views relatively bluntly. Note that I am often cynical about things I’m involved in. Thanks to Adam Binks for feedback. Crossposted to EA Forum comment

Scenario Forecasting Workshop: Materials and Learnings

(cross-posting from Alignment Forum , co-authored with Charlie Griffin) Disclaimer: While some participants and organizers of this exercise work in industry, none of the six scenarios presented here were written by industry employees, no proprietary info was used to inform these scenarios, and they represent the views of their

Post on AI Strategy Nearcasting

I just co-authored a post on AI Strategy Nearcasting : "trying to answer key strategic questions about transformative AI, under the assumption that key events will happen in a world that is otherwise relatively similar to today's". In particular, we wrote up our thoughts on how

My review of "Is power-seeking AI an existential risk?"

I've written a review of Joe Carlsmith's report Is Power-Seeking AI an Existential Risk? I highly recommend the report and previous reviews for those interested in getting a better understanding of considerations around AI x-risk. I'll excerpt a few portions below. Thinking

Samotsvety's AI risk forecasts

Introduction In my review of What We Owe The Future (WWOTF), I wrote: Finally, I’ve updated some based on my experience with Samotsvety forecasters when discussing AI risk… When we discussed the report on power-seeking AI, I expected tons of skepticism but in fact almost all

My take on What We Owe The Future

Overview What We Owe The Future (WWOTF) by Will MacAskill has recently been released with much fanfare . While I strongly agree that future people matter morally and we should act based on this, I think the book isn’t clear enough about MacAskill’s views on longtermist priorities,

Discussion on utilizing AI for alignment

Summarized/re-worded from a discussion I had with John Wentworth . John notes that everything he said was off-the-cuff, and he doesn’t always endorse such things on reflection. Eli: I have a question about your post on Godzilla Strategies . For background, I disagree with it; I&

Prioritizing x-risks may require caring about future people

Introduction Several recent popular posts ( here , here , and here ) have made the case that existential risks (x-risks) should be introduced without appealing to longtermism or the idea that future people have moral value. They tend to argue or imply that x-risks would still be justified as a priority

Reasons I’ve been hesitant about high levels of near-ish AI risk

I’ve been interested in AI risk for a while and my confidence in its seriousness has increased over time, but I’ve generally harbored some hesitation about believing some combination of short-ish AI timelines [1] and high risk levels [2] . In this post I’ll

Personal forecasting retrospective: 2020-2022

Overview I’ve been forecasting with some level of activity for over 2 years now, so I’m overdue for a retrospective. [1] I discuss: My overall track record on each platform : My track record is generally strong, though my tails on continuous questions are systematically not fat

Personal update: EA entrepreneurship, mental health, and what's next

tl;dr In the last 6 months I started a forecasting org, got fairly depressed and decided it was best to step down indefinitely, and am now figuring out what to do next. I note some lessons I’m taking away and my future plans . The High In December

Matrix Moments

This article describes a "Matrix Moment" as part of the development of NBA players (see also this podcast ): It’s special to watch a player when they come to the realization that they’re unstoppable. And that’s exactly what I witnessed as Barrett poured