Research acceleration: The view inside OpenAI

openai.com

182 points by iamsyr 20 hours ago


carbonguy - 14 hours ago

> ... We are pursuing this work in part because automated research could help us solve alignment and build defenses against increasingly capable AI. An automated AI researcher can also be an automated safety or alignment researcher. More capable, aligned systems could help secure critical infrastructure, defend against dangerous AI agents, and develop new protective measures.

In other words... "We must pursue advancements in AI to protect us against advancements in AI?"

edit: there's so much to be critical of in this blog post, just going to throw two more points in here that really stood out to me:

1) all of the metrics are effectively pointing out "we're using way more AI!" - but nothing about impact. What has all this token burn done for them, actually? Let them claim they have more self-licking ice-cream cones than before?

2) in section 3 they break down what the token burn is going towards. Most of the spend is: a) building, b) documenting, and c) monitoring research infra i.e. they're using AI systems which they already recognize may be misaligned to build the systems that they believe will help them identify future misalignment? to which I guess the rebuttal is "no no, we're sure these ones are aligned!"

dsign - an hour ago

It's a funny read if you pull together "AI 2027" and what we all know is going on. Essentially, open AI employee or model is writing "things are going exactly as bad as AI 2027 predicted, but my (golden/RL-) cuffs are too heavy and all I can do is publish this code-speak for 'send help'". It's not a pretty place to be.

pizza234 - 13 hours ago

Funny (in a tragic way) the little crumbs on the path to AI 2027:

> We aim to safely build an automated AI researcher that can work under human supervision to further progress on deep learning and alignment, enabling iterative improvements [...] By "research intern", we mean a system that can carry out well-defined research tasks under human direction, including tasks that would take a skilled researcher a few days.

AI 2027:

> OpenBrain continues to deploy the iteratively improving Agent-1 internally for AI R&D

> With Agent-1's help, OpenBrain is now post-training Agent-2

> With the help of thousands of Agent-2 automated researchers, OpenBrain is making major algorithmic advances

ellis0n - an hour ago

I’m not sure the alignment problem can be solved at all, since these bit-aliens could get out of control due to a hardware glitch in the matrix and for every higher-order control algorithm, there will always be an even higher-order one that could never be investigated.

hedgehog - 17 hours ago

This roughly lines up with my personal experience that in March a combination of stronger models and better tooling on my end let me start running jobs unattended 24/7 (using Anthropic sub and my own hardware). Their $8000/day per researcher spend is crazy though, I'm curious how they keep track of the work.

simonw - 18 hours ago

My eye glazed over a bit during the opening paragraphs, but once you get to the meat of the article about how OpenAI's own researchers are using their tools it gets a lot more interesting.

I noted that they use the acronym RSI (for Recursive Self-Improvement) without defining it. I think that's a little out of touch - I don't think RSI is a well-known acronym outside of OpenAI's bubble yet.

Jeff_Brown - 18 hours ago

The burning question I can't get any information nn is whether, if they determined an earlier misaligned generation may have transmitted misalignment to the current models, they would roll back to a safe checkpoint to rebuild from there. I suspect they would not unless forced to.

RMPR - 4 hours ago

> By mid-August, the median researcher was integrating agents daily into their work, using more than $600 per day of inference at API prices.

There is a lot of talk about AI replacing humans, but how is this sustainable?

falcor84 - an hour ago

> For AGI to benefit all of humanity, we believe it must be democratically governed.

That's a very bold opening statement that they don't really come back to. What would that mean? Who would this demos include?

MisterMunchkin - 35 minutes ago

They're measuring cost as the benchmark of whether someone is a better researcher... burn more resources and you rank higher...

But not a single metric is based on revenue or profit.

Schlagbohrer - an hour ago

It would be polite if they defined RSI at all, rather than just plopping the acronym in there with no explanation. Rude!

nozzlegear - 12 hours ago

I want an all-powerful AI that's aligned with my values, but not necessarily yours. Is that so much to ask for?

lhk931122 - 8 hours ago

Ah, success rate here are scored by an agentic classifier. And uncertain outcomes are excluded from the graph. The thing measured and grading it comes from the same house. In my setup, review agent pass work that an outside critic later rejects

piokoch - 15 minutes ago

One more marketing stunt. We are so good, AI is so powerful so we need to use AI to fight with it. The message is: if you don't buy from us, your competitor will purchase all of this amazing power...

I understand that investors are buying this, after all they believed in all of other crap that led to the 2008 crisis, but please...

- 14 hours ago
[deleted]
jayalbertyapan - 4 hours ago

[flagged]

paidx - 10 hours ago

[flagged]

matan0904 - 17 hours ago

[flagged]

Orien_18 - 16 hours ago

[flagged]

frays - 16 hours ago

[dead]