
@jackclarkSF
@AnthropicAI, ONEAI OECD, co-chair @indexingai, writer @ https://t.co/3vmtHYkIJ2 Past: @openai, @business @theregister. Neural nets, distributed systems, weird futures
Ever since co-founding Anthropic I've been obsessed with the need for third-party measurement of AI systems. (It's in my tweet announcing Anthropic: t.co/eoyuAnAcQ3). Now, we're piloting a way for outside researchers and organizations to study what's happening on AI platforms like Claude while protecting user privacy. This matters because AI systems are giant, complex sociotechnical things with innumerable properties. To think AI labs are going to be able to figure out all the appropriate ways to measure and assess these systems is hubristic and just obviously wrong. So we need to figure out ways to externalize not only the properties of the systems, but also data about how these systems are interacting with the people in the world, which is platform telemetry. What we're trying to do here is prototype the next form of transparency that may be necessary, which is platform transparency while preserving user privacy. Along with this, I want to find ways to empower researchers outside Anthropic to perform their own studies on the kinds of data that companies uniquely have access to, as I think this is a good way to create a more vibrant external research ecosystem. This is another step within a larger project of The Anthropic Institute about measuring and externalizing more details about the properties of our AI systems and the platforms we deploy them on, and sits alongside existing efforts from our Economics team via the Anthropic Economic Index. We are now soliciting research ideas from others for our next wave of projects - please apply!
Anthropic@AnthropicAI·For the first time, we’ve given external researchers a way to study AI’s impacts using real, privacy-preserved Claude usage data. To date, this work has only been possible within AI labs. We can’t tell the whole story alone, so we opened up our tools. anthropic.com/research/enabl…
Ever since co-founding Anthropic I've been obsessed with the need for third-party measurement of AI systems. (It's in my tweet announcing Anthropic: t.co/eoyuAnAcQ3). Now, we're piloting a way for outside researchers and organizations to study what's happening on AI platforms like Claude while protecting user privacy. This matters because AI systems are giant, complex sociotechnical things with innumerable properties. To think AI labs are going to be able to figure out all the appropriate ways to measure and assess these systems is hubristic and just obviously wrong. So we need to figure out ways to externalize not only the properties of the systems, but also data about how these systems are interacting with the people in the world, which is platform telemetry. What we're trying to do here is prototype the next form of transparency that may be necessary, which is platform transparency while preserving user privacy. Along with this, I want to find ways to empower researchers outside Anthropic to perform their own studies on the kinds of data that companies uniquely have access to, as I think this is a good way to create a more vibrant external research ecosystem. This is another step the larger project for The Anthropic Institute about measuring and externalizing more details about the properties of our AI systems and the platforms we deploy them on, and sits alongside existing efforts from our Economics team via the Anthropic Economic Index. We are now soliciting research ideas from others for our next wave of projects - please apply!
Anthropic@AnthropicAI·For the first time, we’ve given external researchers a way to study AI’s impacts using real, privacy-preserved Claude usage data. To date, this work has only been possible within AI labs. We can’t tell the whole story alone, so we opened up our tools. anthropic.com/research/enabl…
This morning, I hiked through dust and into mossy forests, past coyotes and wild turkeys, climbed to find myself atop the world, able to peek at land islands garlanded with fog beneath and scalloped clouds above. #VOTENATURE2026
@MikeIsaac i should state we could afford new tins, but were/are also kind of a miserly/penny-pinching family. there was a period where we used to go to ASDA after it had been bought by walmart to gleefully buy the 5p (about 10 cent) ramen packs
@rachelclif it does, however, then look like vomit, so I thought I'd save the timeline from that
@zkevinbai (I did subsequently add some hot sauce, just to liven things up)
Sometimes I have to celebrate my English heritage by making oddly unappetizing seemingly underseasoned food that is, nonetheless, absolute scran.
Consider: baked potato.
one of my colleagues put this in our 1-1 doc for today
@RishiBommasani To some extent, TAI wasn't really a re-org, but rather a formalization of an ad hoc grouping of research teams which had already existed and worked with me for years, though since announcing it we've added more teams (e.g, @mattbotvinick 's rule of law crew)
@RishiBommasani @mattbotvinick To the extent we fund external entities, it's not in a typical grantmaking mode, but rather things like targeted funding to create evidence to compound insights of our own research teams (e.g, the econ team are helping to fund RCTs related to our broader eco research agenda)
Rough eras of recent AI progress in terms of what research community is collectively hillclimbing on: 2018-2022: Basic capabilities (summarizing, coding, etc) 2022-2026: Norm & time coherence (rlhf/CAI, longer context, agents) 2026 - ?2028?: scientific intuition / independence
the future of the world is available to anyone who takes the time to read arxiv papers every week
This morning I crested a hill to see the morning light tearing into a thick low-hanging cloud, causing yellow light to rip at the edges and cascade down as angelic golden rain. #VOTENATURE2026
The beautiful and tragic thing about having young children is you do stuff with them like blowing bubbles on your front porch with your extended family and you think "this is the best time of my life" and you're right.
7pm: Thinking to myself after closing the door on my toddler who I've just tucked into bed: They are an angel. I am so blessed. I love them infinitely. 8pm (after they've come out of their room 7 or 8 times): What if I bricked up their door at night?
Nice analysis from Peter here about one of the puzzles in AI - given all the changes AI is causing within the economy, why hasn't it had measurable impacts on unemployment? x.com/PeterMcCrory/s…
Props to OpenAI for publishing this post on some safety and alignment issues observed in internal deployments - there are many counter-incentives to publishing stuff like this, but by making it public we all get better info about safety at the frontier. openai.com/index/safety-a…
writing this up for next issue of Import AI
Very interesting personal milestone - Fable just had a good suggestion for some text of a fictional story to delete. I find AI writing broadly excruciating, AI sub-editing pretty helpful, and feels like 'AI editing' is now coming online. Story coming in Import AI tomorrow!
Legendary game.
At this point, everyone at the frontier of AI agrees that third-parties should test out AI systems and use these to develop standards to feed into policy - excellent to see @demishassabis laying out a framework to do this! x.com/demishassabis/…
Import AI is skipping this week due to the two forces that currently heavily influence my life - demands of toddlers, and England's world cup journey.
While hiking today I encountered the gnarled base of a tree like some disembodied wooden hand clutching at grass and stone, and behind it the sun loomed so bright that it distorted the picture I took, giving itself an unearthly vast corona. #VOTENATURE2026
though many experts disagree on the nature of the singularity, all agree that the demarcation for the beginning of the "true foothills" lies on June 30th 2026, the day arXiv went from red to black.
@rSanti97 @maxbodach by flourish I mean crash. my typos are merely an indication of my great enthusiasm
@deredleritt3r put another way, I feel like either some weird stuff is happening in eco and we're talking about that, or we have a load of data that says weird stuff is _likely imminent_ and we're talking about that.
Today at 6pm ET I'm in NYC with @sckimbriel for an @aspeninstitute conversation on AI and the future of society. We'll be talking about RSI, the futures implied by AI progress, and the choices we'll need to consider as the technology advances. Watch live: youtube.com/live/iP9wk0pkC… x.com/sckimbriel/sta…
Hiked in a cool redwood valley on a scorching day, the sun so bright it made parts of the canopy shimmer luminous. I walked with my head craned back and mouth agog marveling at the beauty, accompanied by the whispering sussuration of a million leaves. #VOTENATURE2026
Niche tweet: If you have an excitable toddler, get sick of listening to Bluey and Raffi, and need the occasional bit of non-brainslop kid TV to let you do basic tasks like unload the dishwasher, you might like this fan-made video for the Geese song Cobra. youtube.com/watch?v=R6UcL3…
It's sunday, so getting a little loose with Import AI. Wrote up a fun recent paper from @akorinek about measuring AI in the economy, will be covered tomorrow in issue 459.
Wrote a fictional story for the next issue of Import AI that is pretty different to my previous ones. I think there are beautiful things ahead for humanity as AI gets more powerful and I've tried to capture some of that here. Issue should come out on Tuesday.
@headinajaro Note that the above is more about, to use your term, "helping the monkeys see the god". I agree that "controlling" the things requires different tools. Though it's hard to imagine that the control piece doesn't depend in part on legibility and technical assessment
Many people act like AI policy is some mystery where the right solution demands some kind of Policy Einstein who invents general relativity for tech regulation. This isn't true at all! There are many sensible ideas we could do today. All we need to do is choose to do them. x.com/gabriel_weil/s…
The best part is the majority of these ideas also have the property of continually generating information and building state capacity around advanced technology, so they start "paying out" to society in the form of more well-informed governance almost immediately.
Gotta hand it to him - Juergen Schmidhuber had some amazing papers in the 2010s - early stuff on handwriting recognition (arxiv.org/abs/1003.0358), computer vision (arxiv.org/abs/1102.0183), suggesting intelligence be benchmarked via video games (arxiv.org/abs/1109.1314), etc.
Now of course Schmidhuber is divisive and does spend a lot of time trying to appropriately assign credit in the history of AI and through this has become something of a punchline. But you have to admit that he has demonstrated incredible vision and foresight.