Back
Mark Chen
Chief Research Officer, OpenAI

How OpenAI Shapes Its Research And What's Next - EP 46 Mark Chen

🎥 Dec 09, 2024 📺 Core Memory Podcast ⏱ 98m 👁 36886 views
Mark Chen is one of the most influential people in artificial intelligence, and this episode is a rare look inside his world. As the Chief Research Officer at OpenAI, Chen helps decide which ideas get compute, which research threads scale, and how a 500-person lab navigates constant recruiting raids, model races, and the pressure to invent the next paradigm. We talk about the soup-delivering recruiting wars, how OpenAI thinks about staying ahead of Meta and Google, and what it actually feels like to rank 300 internal research projects when everyone believes theirs is the one that matters. We...
Watch on YouTube

About Mark Chen

Mark Chen, Chief Research Officer at OpenAI, appeared on the podcast "Let Them Cook" on June 25, 2026, where he discussed the company's research direction. Chen stated that he "firmly believes in being on the exponential and in scaling laws" and said he "fairly strongly disagrees" with views that pre-training is dead, noting that such narratives have recurred throughout the history of developing large language models. He described reasoning as "one of the biggest examples" of a research bet at OpenAI, referencing the o1 model as a breakthrough that was difficult to get off the ground because the pre-training plus post-training paradigm "felt like such a promising paradigm" at the time. Chen also said the field is in an "evals crisis," with a low number of canonical gold standard benchmarks, and noted that tools like Codex have enabled faster iteration of evaluations. On June 16, 2026, Chen appeared alongside SoftBank CEO Masayoshi Son at an event in Tokyo where SoftBank announced a cybersecurity service using OpenAI's technology. Chen described cyber capabilities as a "dual-use capability," stating that "even though the models can get better and better at finding vulnerabilities, we can use that for defense." He added that "the important thing is we can go and try the models and try to find the vulnerabilities before external actors can go and try to find the same vulnerabilities." Son compared the dynamic to a criminal with a knife facing a police officer with a gun, saying defenders must have the "best, most powerful weapon" to defend against bad actors.

Source: AI-verified profile updated from Mark Chen's recent appearances. Browse all interviews →

Transcript (266 segments)
I
Interviewer0:00
On the recruitment wars, I mean, this got a lot of attention clearly, and it looked like Meta was quite aggressive. What exactly does this tit for tat look like? What stage are we at?
M
Mark Chen0:13
Yeah, I mean there is a pool of talent, right? And everyone kind of knows who they are. And I think many companies have realized that one of the key ingredients, not the only important ingredient, but one of the key ingredients to building a great AI lab is to get the best talent. And I think not a surprise that Meta has been aggressively employing this strategy.
You know, we haven't sat back idly, and I actually want to tell this story from OpenAI's point of view. I think that a lot has been made in the media of, oh, you know, there's this unidirectional flow of talent over to Meta. But the way that I've seen it is, you know, Meta, they've gone after a lot of people quite unsuccessfully. So just to give you context, right, within my staff, within my direct reports, before they hired anyone from OpenAI, I think they went after half of my direct reports and they all declined. And of course, you know, if they have something like $10 billion of capital per year to deploy towards talent, they're going to get someone. So I actually feel like we've been fairly good about protecting our top talent. And, you know, it's been kind of interesting and fun to see it escalate over time. You know, some interesting stories here are Zuck actually went and hand-delivered soup to people that he was trying to recruit from us.
I
Interviewer1:42
Like, just to show how far he would...
M
Mark Chen1:46
Yeah, I think he hand-cooked the soup. And, you know, it was shocking to me at the time, but over time I've kind of updated towards these things can be effective in their own way, right? And, you know, I've also delivered soup to people that we've been recruiting from Meta.
I
Interviewer2:01
You're doing a soup counting.
M
Mark Chen2:03
I've thought of, you know, if I had an offsite, the next offsite for my staff, I'm going to take them to a cooking class.
I
Interviewer2:08
Okay.
M
Mark Chen2:08
And yeah, I mean, it's just been... Yeah, but I do think, you know, there's something I've learned about recruiting.
I
Interviewer2:15
Did you cook your soup?
M
Mark Chen2:17
Uh, it's better if you get like Michelin star soup. You know what I mean?
I
Interviewer2:23
Yeah. No, no, no. I think Dejo is very, very good and probably better than any soup I could cook.
M
Mark Chen2:30
But yeah, I do think there is something I've learned about, you know, just how to go aggressively after top talent. And I think the thing I've been actually very inspired by is that at OpenAI, even among people who have left for Meta, I haven't heard anyone say AGI is going to be developed at Meta first. Everyone is very confident in the research program at OpenAI. And one thing that I've made very clear to my staff, to the whole research org, is we don't counter dollar for dollar with Meta. And the multiples below what Meta is offering that people are very happy to stay at OpenAI gives me so much conviction that, you know, people really believe in the upside and believe that we're going to do it.
I
Interviewer3:21
Well, and you and Alex... Alex... he used to be one of the math competitions...
M
Mark Chen3:26
Yeah. Yeah.
I
Interviewer3:26
I'm sure you guys hung out.
M
Mark Chen3:28
Yeah. I mean, I have hung out with Alex a handful of times, but we don't do much anymore. Yeah. I mean, yeah.
I
Interviewer3:35
Why did soup become the thing? It was just...
M
Mark Chen3:38
I don't know. You know, there's been soup, there's been flowers, there's been anything you can think of under the sun. But I don't know. I think, you know, life's an adventure. I play into the meme.
I
Interviewer3:47
Yeah. Yeah. Is there any poker strategy to employ like as you're thinking?
M
Mark Chen3:54
Well, again, I think it really goes back to what I've said about the media narrative. The game is not to retain every single person in the org. It's to trust in this pipeline that we have for developing talent and to understand who the key people we need to keep are and to keep those. And I think we've done a phenomenal job at that.
I
Interviewer4:30
We have a special treat today. I'm excited. Mark Chen is here from OpenAI. He's the chief research officer. He's somebody I've gotten to know over the last couple of years. Thank you so much for coming.
M
Mark Chen4:42
No, it's been great to know you for so long.
I
Interviewer4:45
I feel like, you know, there's a handful of people in this world working on this very important project, and I mean, you're right at the top of it. So it's so cool to have a chance to chat.
M
Mark Chen5:01
Yeah. Thanks for having me on.
I
Interviewer5:02
It's my pleasure. And I mean, there's a bunch of things that I want to talk to you about because I've gotten to know you, like we said, over those last couple of years. I want people to know a bit more about your biography, but I also know there's going to be AI enthusiasts who want us to go deep on a couple of things there. So we'll try to do everything. I wanted to start just by giving people a feel for your job, which in my head, and you, I mean, just correct me if I get any of this wrong, but you're, you know, Sam has been, he's really into research. He's the boss. He's kind of at the top of the food chain. But then you and Jakub are working together to shape OpenAI's research direction. And then you're in this additional part of this role is actually deciding which compute goes where onto these projects. So you kind of have to chart where OpenAI is heading and then the mechanics of how you're going to get there. And this always strikes me as a horrible job because I picture people doing everything in their power to get GPUs from you.
M
Mark Chen6:16
It's true. People are very creative in the ways that they try to make backroom deals to get the GPUs they need. But yeah, I mean, it is a big part of the job, right? Figuring out the priorities for the research org and also being accountable for execution. So really to that first point, you know, there's this exercise that Jakub and I do every one to two months where we take stock of all the projects at OpenAI. And it's this big spreadsheet, about 300 projects, and we go and try to deeply understand each one as best as we can and really rank them. And I think for a company of 500 people, it's important for people to understand what the core priorities are and for those to be communicated clearly, explicitly, verbally, and also through the way that we allocate compute.
I
Interviewer7:03
All right. What do we do at Core Memory? We cover innovative, fast-moving, forward-thinking companies, which is why Core Memory is sponsored by Brex because Brex is the intelligent finance platform for many of these companies. 30,000 companies from startups to the world's largest corporations rely on Brex technology for their finances. They've got smart corporate cards, high-yield business banking, and expense automation tools that are fantastic. I hate doing my expenses. And Brex's AI software runs right through those expenses, figure out where we're spending money, and take care of so much stuff for you so you don't have to waste your time on it yourself. Go to brex.com/corememory to learn more and just, you know, get with the program. Let's get going. Let's get out of this archaic finance software and move toward the future. Core Memory and Brex. So you've got, when you're talking about the 500, these are the 500. This is the heart of the research team in an organization now that's thousands of people.
M
Mark Chen8:10
Yeah.
I
Interviewer8:12
Okay. So, and then in that, when you're talking about this 300 projects, I imagine, I mean, obviously some of those are the giant frontier models and then some are probably experiments that people are working on. And so like, how do you possibly keep track of all that and then come to some sort of conclusion about what merits GPUs and what doesn't.
M
Mark Chen8:36
Absolutely. So I think it is very important when doing this exercise to keep the core roadmap in focus. And one thing that differentiates, I think, OpenAI with other big labs out there is OpenAI has always had core exploratory research at its core. We are not in the business of replicating the results of other labs, of kind of catching up to other labs in terms of benchmarks. That isn't really our bread and butter. We're always trying to figure out what that next paradigm is, and we're willing to invest the resources to make sure that, you know, we find that right. And I think most people might be surprised at this, but more compute goes into that endeavor of doing exploration than it is to training the actual artifact.
I
Interviewer9:19
It must be. It's still got to be, how do you stop yourself from being persuaded by someone? Because everybody's going to put, you know, like when I think about this sometimes, I picture when I was at the New York Times, you would have this page one meeting where everybody wants to be on page one. Everybody thinks their story is the most important story. They're all doing their very best job to tell you why this thing is so important. Everybody in that room has worked weeks, months on whatever they're pitching. And so it feels like life and death. And that, I mean, it just, it seems so difficult for me.
M
Mark Chen9:56
Yeah. No, it is also a difficult process. And I think the hardest calls you have to make are, you know, this is a project that we just can't fund right now. But I also think that's good leadership. You need to clearly communicate that, hey, these are the priorities. This is what we're going to talk about. These are the types of results that we think move the research program. And, you know, there can be other things, but those have to be clearly number two.
I
Interviewer10:21
And when you, like you were talking about not being reactive to your competitors. When I was looking through my notes, I don't know if I could go to the line quick enough, but I mean, this was like a point of pride that I saw that you feel like some of the other companies are, well, you know, you guys were in this position where you were ahead and setting the bar for others, and so they were reactive to what you had coming out. We happen to be doing this interview a few days after Gemini 3 came out, and, you know, there is a degree to which your rivals, at times, like, yeah, I mean, there's this back and forth going on. And I know the benchmarks are sort of controversial, how valuable they are, but people go ahead on these things. So how do you also, as time has gone on, maintain that luxury or that intellectual position where you feel like we're just going to do what we're going to do?
M
Mark Chen11:19
Yeah, I think AI research today, the landscape is just much more competitive than it's ever been. And the important thing is to not get caught up in that competitive dynamic because you can always say, hey, you know, I'm going to ship an incremental update that puts me in front of my competitor for, you know, a couple weeks or a couple months. And I don't think that's the long-term sustainable way to do research because if you crack that next paradigm, that's just going to matter so much more, right? You're going to shape the evolution of it. You're going to understand kind of all the side research directions around that sphere of ideas. And so when you think about kind of our reasoning program as an example of this, right, we bet more than two years ago that we're really going to crack reasoning on language models. And this was a very unpopular bet at the time. You know, right now it seems obvious, but back then the environment was, hey, you know, there's this pre-training machine that's working great. There's this post-training machine that's working great. Why invest in something else? And I think today everyone would totally tell you, you know, thinking in language models, it's just a primitive you can't live without. And so we're really there to make these bold bets and to figure out how we can scale and build the algorithms to really scale to orders of magnitude more compute than we have today.
I
Interviewer12:39
It's just, you know, I mean, and I get that intellectually, in my, you know, it gets harder as you guys started as like basically a pure research company. When you look at OpenAI today, I mean, you have product lines. There's parts of OpenAI that look much more familiar to a mature Microsoft or a Google where you have product lines, you've got all these different things that you have to serve. Typically, I feel like you guys are still young enough, so maybe you don't have these exact pressures yet, but, you know, as those companies go on, it always becomes, well, we're more focused on the things that are serving the bottom line than spending a ton of money on research always seems to get dwindled down over time.
M
Mark Chen13:20
Yeah. And I think that's really one of the most special things about OpenAI. At its core, we're a pure AI research company. And I don't think you can say that of many other companies out there. And, you know, we were founded as a nonprofit, and I joined during that era. And I think the spirit is, you know, build AGI, advance AGI research at all costs, and do it in a safe way, of course. But yeah, I actually do think that's the best head fake to really creating value, right? If you focus and you win at the research, the value is easy to create. So I think there's a trap of getting too lost into like, oh, you know, let's drive up the bottom line. When in reality, if you do the best research, that part of the picture is very easy.
I
Interviewer14:07
And you started in 2018, and so you feel like that soul, that core culture and that core nucleus, it's really persisted.
M
Mark Chen14:15
It's still there.
I
Interviewer14:17
What does Elon say? What does he say? We shouldn't call any of you guys researchers. It's just engineering, right?
M
Mark Chen14:25
Yeah. No, I think, yeah, we... No, it's true because I feel like once you have this hierarchy and you elevate, let's say, research science as a thing beyond engineering, you've completely already lost the game. Because, you know, when you're building a big model, so much is in the practice of optimizing all of those, you know, little percentages of, you know, how do you make your kernels that much faster? How do you make sure the numerics all work? And that's a deep engineering practice. And if you don't have that part of the picture, you can't scale to the number of GPUs we use today.
I
Interviewer15:03
So, because I think there, well, okay, but there is like a mystique that surrounds a researcher versus an engineer, you know what I mean? So do you feel like it is better to kind of stay levelheaded on that? Is that kind of what you're saying?
M
Mark Chen15:22
Well, I just feel like researchers, they come in so many different shapes, you know? Some of our best researchers, they're the type that, you know, they come up with a billion ideas, right? And many of them are not good. But, you know, just when you're about to be like, ah, is this person really worth it? They come up with some, you know, phenomenal idea. Some of them are just, you know, so good at kind of executing on the clear path ahead. And so there's just so many different shapes of researchers, and I think it's hard to just lump it into one stereotypical type that works.
I
Interviewer15:55
That makes sense. Okay, I won't belabor you with too many competitive, like, rival questions. It's just since Gemini 3 did come out, I did wonder what happens with you personally or the team when one of your rivals puts it, like, does everybody go and look and see what it can do? Is there like a prompt or a question that you often throw at these new models to see what they can do?
M
Mark Chen16:25
Yeah. Yeah. So, to speak to Gemini 3 specifically, you know, it's a pretty good model. And I think one thing we do is try to build consensus. You know, the benchmarks only tell you so much. And just looking purely at the benchmarks, you know, we actually felt quite confident. You know, we have models internally that perform at the level of Gemini 3, and we're pretty confident that we will release them soon and we can release successor models that are even better. But yeah, again, kind of the benchmarks only tell you so much. And I, you know, I think everyone probes the models in their own way. There is this math problem I like to give the models. I think so far none of them has quite cracked it, even the thinking models. So yeah, I'll wait for that.
I
Interviewer17:15
Is this like a secret math problem?
M
Mark Chen17:18
Oh, no, no. Well, if I announce it here, maybe it gets trained on. But yeah, I do think it's one of the nice puzzles of last year. It's this 42 problem. So you want to create this random number generator mod 42, and you have access to a bunch of primitives which are random number generators modulo primes less than 42. You want to make as few calls on expectation to these sub-generators as possible. So it's a very cute puzzle, but the language models, they get pretty close to the optimal solution, but I haven't seen one quite crack it.
I
Interviewer17:48
Okay, this is, we're heading down a direction I want to ask you about. But then just before we get there, so I know you're, I've seen you, you're very competitive. You've also told me, I think I found, I love competition. I hate to lose somewhere.
M
Mark Chen18:02
I really hate losing. I hate losing.
I
Interviewer18:04
Yeah. So I'm picturing, I'm just curious if this is at all right. I mean, if we know Gemini 3 or whatever is coming out on a Thursday, I mean, are you up at like midnight throwing that problem at it? Or is it not quite that drastic?
M
Mark Chen18:21
Um, no. I mean, I think it's in long arcs, right? And any endeavor, like, I'm kind of a person who has obsessions. I think any endeavor you have to play the long game. And, you know, we've actually been focusing on pre-training, specifically supercharging our pre-training efforts for the last half year. And I think it's a result of some of those efforts, together with Jakub, focusing and building that muscle of pre-training at OpenAI. You know, crafting a really superstar team around it, making sure that all of the important areas and aspects of pre-training are emphasized. That's what creates the artifacts today that feels like we can go head-to-head with Gemini 3 easily on pre-training.
I
Interviewer19:07
Okay. And I want to ask about the pre-training stuff because I've been talking to all you guys about this a lot. But okay. But so you're saying that you're less obsessed about lobbing problems at these new models just when they appear and more at this long journey.
M
Mark Chen19:25
Absolutely.
I
Interviewer19:26
Yeah. Okay. Hey, the reason I want to talk about sort of the puzzle that you were at, I mean, I, you know, I first met Jakub before OpenAI ever started when he was doing a coding competition and I got super into coding competitions for a while. There's this guy, Kennedy, I don't know if he's still famous, but he was like the Michael Jordan of these coding competitions. And so I went to watch one at Facebook used to, I don't know if they still do, but they had an annual Hacker Cup. And that's where I saw Jakub for the first time. And then I know you, I think, did math competitions in high school.
M
Mark Chen20:01
Yeah.
I
Interviewer20:02
Like probably grade school through high school. And then you also did, you also do IOI?
M
Mark Chen20:07
So I got into coding really late in life. It was a roommate in college that convinced me to take my first coding class. And I had all the hubris of a mathematician at that time, whereas like, you know, math is the purest and hardest science and that's where you really prove your worth. I mean, I think I was probably too into the competition back then. But yeah, I mean, it became this super rewarding endeavor. And it started out as purely a way to keep in touch with my friends from college.
I
Interviewer20:39
You went to MIT.
M
Mark Chen20:40
Yeah, I went to MIT. You know, I graduated and every weekend we would just log on and do these contests just to keep in touch with each other. And over time I found myself having a talent for it. You know, I started competing fairly well and then writing problems for contests like the USA Computing Olympiad. Eventually started coaching that team and yeah, it's been a great community where I've met people like Scott that you know.
I
Interviewer21:04
Yeah. Yeah. Okay. So, I think lots of people might be familiar with like math competitions because they probably see kids going through that. The IOI and these coding competitions are a little bit different. I mean, it's, I mean, you'll know it so much better, but when I saw them, I mean, it looks like a, it's almost like a word problem that's a puzzle, and you're trying to kind of find the most efficient and correct way to solve that, and you're in this race against everybody and everybody's like writing code on their computer and then some people try to get there really fast, but then their thing kind of doesn't solve the problem, right? And then, you know, there's like this trade-off, right?
M
Mark Chen21:44
Absolutely. Right.
I
Interviewer21:46
And so you actually were on the MIT team.
M
Mark Chen21:49
No, no, it's something I did after college.
I
Interviewer21:51
After college. Okay. But today you are like the coach of the US national.
M
Mark Chen21:54
Yeah. One of the coaches.
I
Interviewer21:55
One of the coaches. Okay. And was it last year or the year before? Like the US, like we hadn't won one in a long, long time, right?
M
Mark Chen22:04
Yeah. Yeah. Yeah. So, yeah. I mean, the team, I mean, it's, you know, you can never predict what the makeup of top talent looks like every year. But we had a very spiky team, I think, two years ago. And yeah, I believe they won the Olympiad.
I
Interviewer22:19
Because I feel like usually it's like China or Russia or like Belarus and Poland. I mean, right? Yeah. And so this competition takes place in a different country every year. What does it look like? How many people show up?
M
Mark Chen22:36
Yeah. Yeah. So they take the top four students from every single country. It is as much of a competition as it is a social event, you know. This is a tight-knit community. They all do go on to do phenomenal things. And yeah, it's this intense two-day contest where each day you get just three problems, five hours to solve them, and you can really feel the adrenaline and all the pressure in the room. But it's also great fun. I think people settle down and they make, you know, lifetime friends through it.
I
Interviewer23:06
What do you, and like as coach, I mean, you're so freaking busy, man. How much time do you spend on this? What does that look like?
M
Mark Chen23:14
Honestly, the kids are so self-motivated. Sometimes it's really about just managing their performance and strategy. I think, you know, you're going to have good days, you're going to have bad days, you're going to have good hours within the contest, bad hours, and you can't let that get into your head. There's a lot of similarities between managing contestants and managing researchers. It's like on a much longer time scale, but, you know, researchers have good months, bad months. You know, you can't really let those strings of failures get into your head because that's just the nature of research, right? And I think a lot of it's morale management at a certain point. Yeah, I think one other interesting thing that contests have helped me realize lately is when you put the models and deploy them towards solving these contest problems, which they're quite good at these days.
I
Interviewer24:08
Yeah, I was going to ask you about that.
M
Mark Chen24:09
They work in a very different way from humans. You know, we typically think of these machines as, you know, they're very good at pattern recognition. You can take any problem if it maps to a previous problem, it's probably going to be able to solve it. But what I've noticed is in some of the previous IOIs, there's this problem like, messages, very ad hoc. I didn't think the models would solve it at all, but actually one of the easier problems for the AI. So yeah, I mean, this has given me the sense that AI plus humans in frontier research, it's going to do something amazing just because the AI has a different intuition for what's easy and what's not.
I
Interviewer24:45
So, okay. Is it vaguely, you know, when DeepMind did the whole AlphaGo thing, you know, there was that moment where it was doing things humans hadn't played before. So, kind of like vaguely similar to that or...?
M
Mark Chen25:01
I think so. I think so. I think really with GPT-5 Pro, right, there's been an inflection point in frontier research. And one of the best anecdotes I have for this is, you know, I think three days after the launch, I met up with a friend who was a physicist. And, you know, he had been playing around with the models, felt like, you know, they were cute but not super useful. And I challenged him with the pro model, just try something ambitious. And, you know, he put in his latest paper, it thought for 30 minutes and just got it. And I would say that reaction in that moment, it was kind of like seeing Lee Sedol during that, you know, move 37, move 38. And I just think that is just going to keep happening more and more for frontier mathematics, for science, for biology, material science. The models have really gotten to that point.
I
Interviewer26:02
I was going to ask you this question, which is not very original because I think we've been doing this ever since kind of Big Blue and all the chess stuff. But yeah, just as somebody who had followed all these competitions, if, I don't know, there's a sadness when you start seeing these models solving these things that were like the height of achievement for these very unique human minds.
M
Mark Chen26:23
Well, yes and no. I mean, I was good at competitive programming. I was never at the absolute top. And maybe this is a way to get revenge. No, but I do think, no, there's certainly a moment for myself, right? We tracked coding contest performance while we were developing reasoning models for a while. And, you know, at the start of the program, you know, they were not super great, you know, at the level of any average competitor going into the contest. And yeah, over time they just started creeping up and up in terms of capability. And you still remember that moment when you walk into the meeting and they have where your performance is, and then the models exceeded that. Man, that was also a shock to me. It's just like, wow, we've automated to this level of capability so fast. And of course, you know, Jakub was there still a bit smug, but within like one or two months, it was also surpassing him. So, yeah, no, the models are at the frontier today, right? It's so clear by even through the results we've done this summer at Codeforces, right, top optimization competitive programmers in the world, I think it achieves second place there. And so really it's jumped from, you know, hundredth place last year to top five this year.
I
Interviewer27:44
And like, do you think we'll still be doing these competitions in 10 years?
M
Mark Chen27:48
I think so. I mean, they're just fun. I mean, certainly a bunch of people who use it to pad their resume are going to drop off from doing it, but I think the people who've always excelled at it the most are people who just do it for the fun of it. And I don't think that'll go away.
I
Interviewer28:04
When I was doing this story, I mean, they were telling me that like if you're from Russia or I don't know which countries, that you basically get like an automatic free ride to any university that you want. I mean, I see the guys on the US team go to like Harvard and MIT, so they seem to be doing okay, but it doesn't seem like the US has a...
M
Mark Chen28:21
Yeah. I mean, don't you think it's going to, yeah. I mean, interviews, right? They're going to be kind of broken going forward. And everyone's seeing this a little bit. And, you know, even college exams or college homework, it's all broken at this point, right? And I do think we're going to need new ways of assessing and gauging, you know, who's performing well, who's learned the material.
I
Interviewer28:40
Where somebody's actually at.
M
Mark Chen28:42
Yeah. Yeah. So I mean, I've had this idea here where maybe for our interviews we should just have candidates talk to ChatGPT. And, you know, it's a special kind of ChatGPT where the model is trying to gauge whether you know the material or whether you're at the capability level to work at OpenAI. And, you know, you have to have this conversation with it that convinces it deeply you belong at OpenAI. And of course, you know, you can't be allowed to jailbreak it, but we look at the transcript after. But maybe like tests like this will more accurately reflect in the future whether, you know...
I
Interviewer29:17
So, you don't do that yet, but you're thinking about...
M
Mark Chen29:20
Yeah. Yeah. Just creative ways to revamp the interviews.
I
Interviewer29:22
Yeah. Yeah. Well, I mean, Silicon Valley is famous for doing the like brain teasers during the interviews and everything. Yeah. So you, I mean, we talked, you were very good at math growing up. And I think you were, you born on the east coast?
M
Mark Chen29:41
Uh, yeah, born on the east coast.
I
Interviewer29:42
And then you lived on the west coast too.
M
Mark Chen29:44
On the west coast.
I
Interviewer29:44
And then you lived in Taiwan for like, for like grade school to high school.
M
Mark Chen29:49
Four years.
I
Interviewer29:50
Okay. Your parents worked at Bell Labs.
M
Mark Chen29:52
Yep.
I
Interviewer29:52
So you come from like engineering stock. I mean, it's a really interesting background just because you kind of got like a flavor for all these innovation hubs and especially with your parents being at Bell Labs and most of...
M
Mark Chen30:10
I mean, yeah, I just grew up in a very scientific environment, you know, dinner table talk was puzzles and things like that. And I also got kind of the more traditional, you know, Bell Labs east coast experience. On the west coast, my dad came to do a startup, so a little bit of that kind of new company, got exposed to that when I was young as well. And of course the big jump to Taiwan, right? And I think it's a huge culture shock. You wear uniforms, you're in a school, it has barbed wire around the school, right? And also getting exposed to kind of that level of rigor. I think it was just a number of really great experiences growing up.
I
Interviewer30:44
So, like the schools were harder or...?
M
Mark Chen30:48
Well, I would say it was just much more kind of, you know, it's just there's a little bit less flexibility and freedom in the school system, but I think it also teaches you something.
I
Interviewer30:58
Yeah. Okay. Since day one, the Core Memory podcast has been supported by the fine people at E1 Ventures. They are a young and ambitious VC firm in Silicon Valley, investing in young and ambitious companies and people. Thank you so much to E1 Ventures for all your support.
M
Mark Chen31:16
Yeah.
I
Interviewer31:17
And you knew you wanted to come back to the US for college.
M
Mark Chen31:20
Absolutely. Yeah.
I
Interviewer31:21
Okay. And then, okay, so you're at MIT. You're kind of like in this interesting group. I guess MIT probably always has an interesting...
M
Mark Chen31:30
Oh, man. Yeah. 2012 was such a great group.
I
Interviewer31:32
Yeah. Like, who is there, sort of like an all-star list?
M
Mark Chen31:35
Oh, I mean, it was a great year. Like, I don't know if you knew, like, Jacob Steinhardt, you know, he's doing Transluce now. He and I used to do projects together in computer science class. There was Paul Christiano who's, um, a bunch of really phenomenal... He worked at OpenAI. A bunch of kind of big names in AI came from that year.
I
Interviewer31:57
And then we were talking about the competitive coding, like Scott...
Woo, who's at Cognition. I mean, he's kind of like famous now as a meme on X for his math abilities, but you just got to know him through the coding competition.
M
Mark Chen32:13
Oh, yeah. Through the coding community.
I
Interviewer32:15
Okay. Okay. And then now I see the competitive end of you guys. The output of this to me looked like poker these days. I think we were at an event which I think we have to keep secret or something, the specifics on this event. But I think I'm okay to talk about this part, which is like I'm walking by this table, there's you, Scott, I think Sham from Palantir, and then a handful of other people in a fairly intense poker game. So you guys, this is where you've applied your math and competitive skills now.
M
Mark Chen32:59
Yeah, I mean, poker is a really fun game and you know I've talked about my life in terms of a series of obsessions. Poker was definitely one of these obsessions in the past and I think the big revelation for me in poker was, you know, it's so much more a mathematical game than a game of reading people and bluffing. And I think the more you learn about poker, the more you update in that direction, right? I used to be a terrible bluffer. And when you know it's mathematically correct to bluff, then it's so easy, right? It's like you don't feel any nervousness around it. And yeah, it's just so interesting that you have a game that I think is perceived as so human, but the underlying mechanics and how to win are so deeply mathematical. And yeah, I kind of thought about this the other day where, you know, there's something about that in language modeling too, right? You have this deeply human process of generating language, but there's this mathematical machine that can really do it as well as we can.
I
Interviewer34:03
I think about that part all the time as a writer and then I did all this philosophy back in college with Wittgenstein and all these guys thinking about these things. Yeah. Well, how do you find like an edge? If you and Scott both strike me as supernaturally good at math, I don't understand how one of you is outcalculating the other person.
M
Mark Chen34:23
Um, well, no, I mean, it is mostly a forum for us to just kind of hang out and catch up with each other. And you know, today we don't take it as... I think there is an element to taking something like poker very seriously that takes the fun out of it. And so, you know, my obsession with poker I think has ended more than a decade ago and now it's just fun.
I
Interviewer34:45
You're just saying this 'cause I saw Scott win both days. I think...
M
Mark Chen34:49
You might be right about that.
I
Interviewer34:51
He was taking it quite seriously. And so like coming out of college, I mean you had in some sense...
M
Mark Chen35:00
Oh, I beat him on the plane though.
I
Interviewer35:01
You beat him on the plane right home. Okay. All right. All right. So was it just you versus him or was it like a group thing?
M
Mark Chen35:09
And maybe three or four people.
I
Interviewer35:10
Okay. I feel like a lot of... I feel like there's three... I don't think I'm overgeneralizing too far, especially among like if you cut back to the 2018 sort of time frame. As far as like people who were in AI at a high level, I mean a lot had academic backgrounds. A lot were math prodigies or had gone on to sort of take their math background and get into robotics or something physics like that. And then there's this other bucket which is people who had gone to Wall Street and done high frequency trading and quants and things like that. So that was the first path that you took was you went straight from MIT to Wall Street.
M
Mark Chen35:58
Yeah. I mean I don't wear that badge with too much pride to be honest. Like, you know, it was a path that was fairly common for very quantitatively oriented kids at MIT. And I mean it was certainly a very meritocratic system, right? You could apply intelligence and there's a very concrete reward function, the amount of kind of profit that you would make. But I think culturally it was hard for me. It was a place where when you discover something, your first instinct is to just keep it away from as many people as possible because your knowledge is what gives you your worth. And so it felt like you would have out of this an outgrowth of just even internally at a company like these competitive dynamics and people weren't very trusting of each other. I think it also felt like such a closed ecosystem, right? I think we today don't feel too much like, you know, when someone in HFT finds a breakthrough that makes their algorithm a little bit faster, no one else feels it, right? And I over time just kind of felt like, you know, I woke up after four or five years, we're competing against the same exact set of players. Everyone was a little bit faster but had the world really changed that much for it? And it felt like time to do something else, right? It just a bunch of things lined up back then. You know, there's the AlphaGo match which I think was a huge inspiration for a lot of people at OpenAI.
I
Interviewer37:32
And did you play Go?
M
Mark Chen37:33
I did not, but I think the sense in which, you know, the model was able to do something creative, I really wanted to understand what was going on behind that.
I
Interviewer37:43
So you're watching that happen and had you been at all reading AI research papers and things like that or...
M
Mark Chen37:51
To be honest no. And then I saw that event, it was really inspiring and that's when I started doing my deep dives into AI. So one of my goals after seeing that was reproduce the DQN results. This is a network that was able to play a lot of Atari games effectively at a superhuman level and going from there, you know, that's how I got my start in AI.
I
Interviewer38:16
And you were doing that like on the side as you... so you just work all day and then go back and try to...
M
Mark Chen38:21
Yeah. Yeah. Yeah.
I
Interviewer38:22
Yeah. Okay. I mean it is weird. I remember I was interviewing George Hotz like it might have been roughly 2018. Maybe a little before that, you know, and he had just done this thing building a self-driving car on his own in his garage. And then, you know, I mean, it's George, so he says large statements sometimes that may or may not be exact or spot on or apply to other people. But he's like, AI is still so young. You can basically learn the whole field if you read I don't know what the number was, 10, 20, 30 research papers. I mean it is fascinating to me that it's like old in many ways stretching back decades but this particular moment...
M
Mark Chen39:03
It's very shallow. I always give this advice to people who are intimidated by getting into AI. So shallow, like just spend three to six months picking some project like maybe reproduce DQN and you can get to the frontier very quickly. The last couple years has added a little bit of depth but it's not anything like theoretical math or physics.
I
Interviewer39:27
Do you think is this a field where... I asked Jaan Tallinn this the other day, I don't know why I'm obsessed with this but you know like in mathematics you see people tend to do their best work in their 20s or to have the big breakthrough and then it's very hard as they get older to have that same kind of moment. Like what you're saying, you know, are we dependent on young people reading these papers and then having some insight or is this something where you can keep going throughout your whole career?
M
Mark Chen40:00
I mean, I think you can keep going. I mean, OpenAI itself does have a pretty young culture, but I don't think you need to be young to do good research. I think there is something about being young and having less priors about this is the way it's done. I think over time you may develop your own vision, which is a good thing, right? But it also locks you into a frame of mind of like, oh, you know, this is how research is done, this is how good results come out. And I think younger researchers tend to have a little bit more plasticity around that concept.
I
Interviewer40:34
Yeah. Yeah. So as your career at OpenAI is funny because it seems like you walked in the door and had a very important large position from the get-go. But when you first got there in 2018, it must have been what, like 50 people?
M
Mark Chen40:50
Oh, no, it was much closer to 20.
I
Interviewer40:52
Much closer to 20. Okay.
M
Mark Chen40:54
And it really looked like two teams back then. I came in as a resident. So someone who, you know, clearly not a specialist, not a PhD. I think I was only a resident throughout his tenure at OpenAI. So I was very lucky in that regard, just kind of learning the way that he thinks high level about research.
I
Interviewer41:13
And a resident in this case is you're just the right-hand person to...
M
Mark Chen41:16
Oh, so it's someone who comes in usually from another field who OpenAI wanted to invest in and train up in AI. Yeah. And so I think the first part of a residency is like a six-month compressed PhD and then going from there to just getting into deeper and deeper research projects.
I
Interviewer41:35
So you're kind of like talking to Ilya every day. Is he kind of shaping that...
M
Mark Chen41:40
Yeah, he was responsible for my projects, for my curriculum, for my learning. And you know, I would just go to him for, you know, hey, like what's this about, like why did people pursue this.
I
Interviewer41:51
Okay, yeah. And you, I mean I think if you go like on LinkedIn, I mean it would say you were the head of Frontier Research as your first job at OpenAI.
M
Mark Chen42:00
Oh no, I was an IC for three-ish years. Yeah, so I was doing independent research projects. I worked in generative modeling because that was really kind of where Ilya's focus was at the time. And then, only after a while did I start managing teams.
I
Interviewer42:19
And because most... you're talking about generative... I mean, most people point to DALL-E as maybe the kind of first big project that the public would mostly recognize. Is that fair?
M
Mark Chen42:28
Yeah. Yeah. So, I think that also marked the transition between when I was an IC and a manager. So one of my own kind of big projects and one I'm still pretty proud of today is Image GPT, this proof of concept that even outside of language you could put things like images into a transformer and the model would just internalize very good representations and understand the content of images. And it's kind of like a proof of concept that you can do language modeling outside of pure text and get really, really great representations and scale them to be state-of-the-art with other methods. That was I consider like a precursor work to DALL-E which I was on the opposite side of managing. And I think between those, another project I'm really proud of doing IC work on is Codex where we set up a lot of the framework for evaluating coding models and also did a lot of in-depth study on how you can take language models and make them very good on code.
I
Interviewer43:33
So and what made you pick OpenAI? Because I could see it two ways in my head. One is big fish in a small pond. There's interesting people here. I remember the OpenAI of 2018 with 20 people. In my head, it was like, this is probably not going to work. You know, Google seems like they've got this locked up and this is like a pretty small group of people trying to take on something that appears to require many billions of dollars of capital. And I mean, this was even before the scaling stuff. It was just like Google had invested so much in AI already. Kind of like in a different form than we think, but you know, you're already kind of doing translation on your phone and things like that. So was that like a hard decision for you or you just stumbled into the OpenAI gig so quickly that...
M
Mark Chen44:26
Well, yeah, I mean I think there are two things, right? You need ambition of vision. That was certainly what OpenAI had at the time, but you also need the talent to back it up. And you know, I did feel like OpenAI was one of the rare places where the ambition was very large, but the talent was also large enough to fulfill that gap. And you know, I was lucky I knew people like Greg before from college. Yeah.
I
Interviewer44:51
Oh, so yeah, Greg was... you overlapped at MIT.
M
Mark Chen44:54
I think we did some math contests together. Okay. Yeah. Back in high school. And yeah, I shot him a message actually and I was like, 'Oh, you know, I don't know if I have the right skill set, but this sounds like a place that's doing great work.' So...
I
Interviewer45:09
It still seems nuts to me just to come at this, you know, out of nowhere and now you're like leading research.
M
Mark Chen45:16
No, it's surreal to me, too. It's surreal to me, too. You know, even that transition from IC to manager, I was very hesitant about taking it. I didn't know if managing was a skill set that I would be good at and I was really enjoying IC work. I think I was having a lot of fun doing it, excelling at it, building really great collaborations. But yeah, I mean it's really been a wild ride.
I
Interviewer45:41
Yeah. Yeah. Well, okay, on that point, I mean, you've always struck me as a very nice, level-headed guy. I have to say, you know, there's parts of OpenAI's history that are quite dramatic, soap opera-like, a little Game of Thrones-y, you know, power struggles. And like to me, to be a manager in that... I will say now like I feel like things are a bit calmer than they were, but when you look backwards, it just seems like, I don't know, you're saying you had to learn these skills, but some of this feels like opposite... I don't know you that well, but some of it feels opposite to your personality to have to deal with all of that.
M
Mark Chen46:31
Honestly, you know, I've been lucky at OpenAI. I genuinely say that in the sense that I've had managers that have really advocated for me. You know, they saw my talent and advocated for me. I think when I was an IC, you know, Wojciech, he was like, 'Oh man, you should bet on him for Codex.' And then later on kind of reporting to Bob, I've never asked for promotion or up-level. And you know, it's just organically happened. And everyone kind of along the way has given me great advice. I think part of growing as a manager is just getting the reps. I don't think there's any better place to get the reps than at OpenAI. You know, there's always challenges to solve. And yeah, I think kind of developing that confidence. I actually think management is something where it's really just about the experience and, you know, there's I would say less so talent involved in it.
I
Interviewer47:28
Yeah. I don't want to like embarrass you and I don't know if this will or won't and I assume you probably don't want to get too much into the coup or the blip or whatever. We won't talk about anything. Yeah. Yeah. Well, I just I've interviewed so many people about this now. And I'm also gonna save some of my gems for my book. So, I won't up myself. But there's a couple moments in there where you, you know, you help get the researchers aligned around the petition to like bring Sam back and then I think just either a day or two after that there's kind of like this speech, you know, that was given I think at Greg's house maybe or...
M
Mark Chen48:08
I think Chelsea's house.
I
Interviewer48:09
Okay. And you know both of those struck me as pretty profound moments, especially for I guess like standing up for what you believe in and rallying the troops. I mean, yeah, like in a moment of crisis, I don't know. So, did those...
M
Mark Chen48:36
Yeah. I mean, that did feel like a very pivotal moment for me. I think in the days following the blip, right? There was a lot of uncertainty. And you know, myself, Nick, Barrett at the time, we felt this responsibility of, you know, the wolves are at the heels, right? Everyone's getting calls from all these competing labs being like, 'You should come work here instead.' And I just set this goal of I will not lose a single person. And we didn't. And it was just every day opening up our houses, you know, people could come here, they could have a place where they let out their anxiety, and then also just helping them keep in touch with the leadership team, having a way for them to feel like they could make a difference. And I think, you know, over time people really felt this spirit of, hey, we're all in this together. How do we make a difference? How do we signal to the world that we're all together? And, you know, I've been kind of driving back and forth between a couple houses and we had this idea of like, hey, you know, we need to show the world that we're all seriously aligned and we're going to work for Sam. And that's when the petition came together. And the idea, I think, got solidified at 2:00 a.m. We got more than 90% of the whole research org signed, I think, by the morning. And it was just everyone like calling their friends being like, 'Hey, are you in or are you not?' And yeah, I think in the end, you know, it was very close to 100 people signing that petition.
I
Interviewer50:11
Well, I mean that must have put you in something of a tough spot though just because especially at the outset it was kind of like Ilya and Sam were on opposite sides and Ilya is your mentor and then I know Ilya kind of comes back... um, yeah, I don't know, was that awkward?
M
Mark Chen50:26
Um, no, no, it was hard. I mean it's a low information environment, but fundamentally it's just... and yeah, I mean I think at the moment, you know, you could very reasonably conclude like, did Sam do anything here, you know, is there... but would Greg and Jakob, like people of super high integrity, quit over that? I just felt like, you know, there was some part of the story that was being misrepresented here.
I
Interviewer50:52
Yeah. Yeah. With the, you know, Jakob's been there for a very long time. Like what should people know about Jakob that they don't?
M
Mark Chen51:02
It's interesting 'cause he's a super funny guy.
I
Interviewer51:05
He's hilarious. Oh my gosh.
M
Mark Chen51:07
He has this like sarcastic humor. And yeah, it cracks me up so much honestly. Like, yeah, that's one of my favorite things about OpenAI today. Just like the level of alignment I have with Jakob. I feel like we go into a meeting, we can just bounce off ideas and quickly get to alignment and then, you know, deliver the same message and kind of like operate on different parts of a big roadmap together. And yeah, it's just one of the big privileges I have working at OpenAI. Yeah, I mean going to that actually, that point about, you know, keeping people together, like I still feel that way about OpenAI research. I think we're still under attack.
I
Interviewer51:50
Yeah. No, we are a family.
M
Mark Chen51:53
Yeah. We're always under attack. Look, when any... and this is how I know we're in the lead, right? Any company starts, where do they try to recruit from? It's OpenAI. And you know, they want the expertise. They want our vision, kind of our philosophy of the world. And we've made so many star researchers, right? I think OpenAI more than anywhere else has been a place that makes names in AI today. And I still feel that same level of protectiveness like you come after, I'm going to do anything in my power to make sure, you know, they're happy, they're open and they understand, you know, how their role fits into the roadmap.
I
Interviewer52:30
I, yeah, this is something I battled with as I was doing the book or even just watching events unfold in real time is like when I go back through the history, I mean you've got Ilya in 2012 making sort of like a big breakthrough and then you know you've got Noam Shazeer in 2017 doing Transformers and then you've got Alec Radford, you know, like sometimes the story is these individuals really pushing the field forward and it feels like a field that's still so young that you can have this individual and then it seems like there's this group of, I don't know what the number is, but let's call it like 8 to 10 who seem to have an ability to do that repeatedly and they're really shaping where this is all gone. And so when I started seeing like John Schulman leave or Alec leave and then, you know, there's kind of was like, wow, okay, well if you've lost a chunk of this all-star team, how do you... it seems like a kind of field where you sort of can't just replace that and yet, you know, it was kind of like after that that you guys pushed forward on reasoning and some of these other spots. Yeah. So I don't know, I've intellectually had trouble...
M
Mark Chen53:46
I do disagree with that as the overarching way to do good research today. I think there's certainly a lot of top-down steer, you know, we bet on directions, but OpenAI has this beautiful culture of being bottom-up in a very deep way too where some of the best ideas just organically emerge from sometimes the most surprising of places. And I think really the great thing has been just like watching some of these bets unfold, take shape, get scaled, and reasoning being a core example of that.
I
Interviewer54:22
Yeah. And okay, so in this... but like this idea that like how star-dependent are we? Because you still see Google spend an ungodly amount of money to bring Noam back, you know what I mean? Yeah. And so this makes me think, okay, this is how this works.
M
Mark Chen54:37
Yeah. I mean, I think it's a mix, right? Like you have to invest in your pipeline because I'm very confident in our ability to create stars. But yeah, there's certainly very good people out there and everyone knows that they're good. I think if there's one thing that, you know, on the flip side I've learned from Meta is, you know, OpenAI can also go very aggressively after star talent and, you know, there's this very aggressive recruiting approach that, you know, I've taken a couple pages from as well. But yeah, I think we should always just be trying to assemble the best team in service of the mission that we want to accomplish.
I
Interviewer55:12
It's funny because it is like a relatively small world and like all you guys hang out even though you're like rivals and then it must be weird 'cause I know you're friends with different people on some level and then you're also trying to steal all their...
M
Mark Chen55:26
I mean yeah, it's a brutally competitive industry in all fronts, right? But again that's what I love. I'm a deeply competitive person. I hate to lose and yeah on research, on recruiting, all of these fronts, I'll work very hard on them.
I
Interviewer55:41
It reminds me because I'm like a semiconductor... well I'm a history nerd but just the early semiconductor days were not that far off. I mean you had all these semiconductor startups come at once. They were all pushing the limits of physics and somebody would discover something at one. They'd go to the bar and have a... it's like people, they're engineers, they can't like stop from like sort of sharing knowledge with each other but and then they're also getting pulled like, you know, each company is kind of quickly getting this breakthrough in one way or another.
M
Mark Chen56:13
Yeah. I mean you raised an interesting point of, you know, there is going to be some base rate diffusion of ideas. And I think there's two ways a company can respond to that. You can create these deep silos of like, hey, you know, we're going to protect information in all these ways. I don't think OpenAI operates that way and we don't think that's the right way to operate. We just will outrun other people as fast as we can and I love the culture of openness. People in research freely share ideas and I think that's the way to make the fastest progress.
I
Interviewer56:45
And how like how do you, Sam, and Jakob now work together? I think people sometimes, if you read the announcements and everything, you can tell that Sam is research-oriented over like day-to-day running of the company, you know what I mean? You can tell research is more of his passion and even just like in the titles and in the way it's been organized, especially recently. And you and Jakob are so deep on this stuff and I know Sam is technical but you guys are like in it all the time and then you know Sam is having conversations with everyone. Yeah. I'm just curious about this dynamic between the three of you and how... I mean are you guys... I mean I guess you're not always in alignment on what is going to get the resources but yeah, I was just curious about you guys' dynamic.
M
Mark Chen57:41
Yeah. Yeah. So, I mean, it's a very tight cohort. You know, I talk to Sam and Jakob every day. And you know, with Sam, he loves research. He loves just learning about research. He loves talking to researchers. I think in some ways he's very effective at getting a pulse on the research. I rely on him also to just, you know, are there any hidden latent problems here? Go and find them out, you know, surface them to me. Jakob and I... it could be personality or technical. It could be just small things like, oh, you know, just even the way the office is laid out makes it harder for this team and this team to collaborate and the two of them need to collaborate to help unlock this breakthrough that we want. I mean all these things are very, very important. And I think Jakob and I, we spend a lot of time figuring out how to design the work for success. You know, I think pairing people with the right strengths together. You know, also how to incentivize people to work on directions that we find are important. Yeah, that's a lot of the work that we do.
I
Interviewer58:45
And Sam, what he... like is he reading papers? Is he chatting with you guys? Is...
M
Mark Chen58:53
Yeah. Yeah. I mean, I think he does his fair share of reading papers. He talks to researchers and just understands how they think about the world, the type of research that they're doing. And of course, he's responsible for a huge umbrella of things outside of that.
I
Interviewer59:07
All right, I'm going to ask some nerdy questions now, but I'm going to try to... I don't know if I can Dario-level it, but I'm going to do my best. And you know, I'll ask. I don't know how top secret some of this stuff is, but anyway, well, maybe you'll slip up and we'll just get it out. You know, in the meetings I have been on, and I don't think I'm revealing because we talked about it a bit. I think I'm safe here, but you know, pre-training seems like this area where it feels... it seems that my sense is you guys feel like you've figured something out. You're excited about it. You think this is really going to be like a major advance. It was also, I think, either a neglected spot or something of a sore spot. You know, previously things weren't maybe working exactly how you guys had expected or hoped. Like what can you tell us about what you figured out and, you know, some sort of frame of reference on we've seen these periodic big leaps forward.
M
Mark Chen1:00:11
Absolutely. So I think the way I would describe at a high level the last two years is, you know, we've put so much resourcing into reasoning, into understanding this primitive and making it work and it really has worked. And I do think one byproduct of that is you lose a little bit of muscle on your other functions like pre-training and post-training. In the last six months Jakob and I have done a lot of work to build that muscle back up. I think pre-training is really a muscle that you exercise. You need to make sure, you know, all the info is fresh. You need to make sure people are working on optimization at the frontier, are working on numerics at the frontier. And I think you also have to make sure the mindshare is there. That's kind of one of the recent things I've been focusing a lot on, just kind of directing and shaping what people talk about at the company and very much today that is pre-training. We think there's a lot of room in pre-training. You know, a lot of people say scaling is dead. We don't think so at all. In some sense, you know, all the focus on RL, I think, it's a little bit of an alpha for us because we think there's so much room left in pre-training. And I think as a result of these efforts, you know, we've been training much stronger models and that also gives us a lot of confidence carrying into, you know, GPT-5 and other releases coming this end of the year.
I
Interviewer1:01:37
Like the way I picture it in my head sometimes is that you guys have been on this... you've just been running so fast. The whole field has been running so fast. And so we're at a moment where it's like, okay, we've gathered up this vast volume of information from the internet. We've thrown it onto this supercomputer and that, you know, ChatGPT pops out and then we're just on this incredible race that's going on. And so like when I hear you guys, I'm just trying to think about this in like a... to level set maybe for people who don't follow this as closely. So, you know, in that initial moment you just had so much data, you're throwing it at this machine, you try to shape that data a bit initially and what now we're just learning like more efficient ways to shape that... just it's not always clear on what the mistakes were.
M
Mark Chen1:02:33
Um, so I do think, yeah, you touch on something I've been thinking about a lot, right? When you think about pre-training, right, you're taking human written data and you're teaching the model how to essentially emulate it, right? It understands human patterns of writing. And in some sense that also bottlenecks and puts a ceiling on the capability that you're able to achieve, right? You can't really surpass what humans have written when you're imitating what humans have written. And so, you know, you work on things like RL, there, you know, you can really steer towards the hardest tasks that humans can come up with and have the model basically think outside of the box, outside of what it's learned from imitating humans and achieve higher levels of capability. But there is this kind of interesting problem now of how do you go beyond what humans are able to do today? And I do find a serious measurement problem there too. Even in the sense of like can humans judge superhuman performance in the sciences, right? How would we know that like this superhuman mathematician is better than that superhuman mathematician? And we really do need to kind of come up with better evaluations for what it means to make progress in this world. Right? We've been lucky up to this point, right? There have been contests like the IMO, IOI, really just like gauging who's the top one mathematician in the world, right? But when the model capabilities go beyond humans, there are no more tests.
I
Interviewer1:04:11
Right, okay, you just made me think of a question going back to the IOI stuff. I mean... and sorry we're going to come back, I just... you just totally popped in my head. I mean like often I would see the kids who were amazing at those competitions they would get hired somewhere like a Google or Facebook or something, but they weren't always like the, you know, the top executive or the most famous engineer afterwards. And maybe it was like by choice, but I don't think Gennady was like the Michael Jordan ended up working at any of these companies. And that totally could be by choice. I'm not trying to disparage him, but it's not clear to me like even... okay, so it's not clear to me that the human who excels at that is necessarily like the greatest engineer you're ever going to have. And so like, yeah, I mean if an AI is particularly good, like what are we learning?
M
Mark Chen1:05:04
Yeah, that's a thing I quite like about working in AI. I think more so than...
In standard engineering culture, it is a meritocracy. I've tried this many times before and learned this lesson many times before, but it is hard to put in someone to lead a group who doesn't have the respect of the researchers that they're leading. And I think this is more so the case in research than anywhere else. You have to make very strong technical calls of like, you know, this is the right path when there's a disagreement, this is the right kind of project. And if you make those calls wrong, you lose the respect of your researchers. So, yeah, one of the fun things in working in AI and creating a strong AI org is, you know, all my direct reports are very deeply technical and it's fun to talk to them about the technical things.
I
Interviewer1:05:57
Yeah. Okay. And then, okay, on pre-training again for a second. You know, like to me in my head it feels like Transformers helped kick off this massive, massive leap. I mean, reasoning to me feels very comparable, if not even sort of more amazing. I mean, are we... when I talked to you guys over the last few months, my, you know, and I can never tell if this is optimism, if you guys are just putting the best foot forward when I'm chatting to everyone. My sense when I talk to you, to Greg, to Jakob, you know, to Sam is that you guys kind of feel like you've been putting in hard engineering work for like three, four, five years that hasn't fully manifested itself. And so then I can never tell how excited or not to be. I mean like when you guys are hinting at some of the stuff you're seeing, do you feel like it is, you can already tell that it is a comparable leap forward in terms of these big epochal kind of things?
M
Mark Chen1:07:06
I think so. You know, I think when we launched GPT-5, you know, we talked a lot about synthetic data as well. You know, there are many other threads of this form that we think are holding quite a bit of promise and that we're scaling up pretty aggressively right now. And I think it's always about maintaining that portfolio of bets, taking the ones that are providing more empirical promise and scaling and supporting them at an even greater degree.
I
Interviewer1:07:32
But it was like two weeks ago Andrej Karpathy who used to work at OpenAI, you know, he went on Dario's podcast and seemed to like deflate some giant portion of the AI industry by saying, you know, I think he was saying that AGI was like 10 years off. And then when I hear, and then I heard Dario talking about a week ago. I mean he seemed to be holding on very much to like massive scientific discoveries, his, what is he called, the nation of geniuses. He seemed to be holding still on kind of like maybe a little slower but like a two-year timeline on that. You know, yeah, when you heard what Andrej said, what did you...
M
Mark Chen1:08:12
Yeah. I mean I think Twitter, they love this like cycle of, you know, it's so over, so back, and you know, whatever plays into the narrative at the time I think, you know, just becomes amplified. Yeah, I'm trying to make a clip here, but you know, the way I think about it, yeah, I mean, it's like AGI, I mean, everyone defines their own point for AGI. I think even at OpenAI, you can't get everyone in the same room and be like, hey, this is my clear definition of AGI and it's consistent. And so, I kind of think about it as something like, you know, you're in the industrial revolution, right? Do you consider the, you know, having machines make textiles, is that the industrial revolution or is it the steam engine? You know, everyone kind of has their different definition and I think we're in the middle of this process of producing AGI. For me, I think the thing I index most on is are we producing novel scientific knowledge and are we advancing the scientific frontier. And I feel since the summer there's been a tremendous phase shift on that front.
I
Interviewer1:09:13
Okay. Like from stuff that you're seeing, the first things that are jumping to my head are all these startups that are in the biotech space that are showing, you know, one-shot antibodies and molecules, but I have no idea if that, like, what are you...
M
Mark Chen1:09:29
Yeah, yeah. So, I mean, I was so inspired by that encounter with the physicists that, you know, went back and thought, hey, well, we should just create OpenAI for Science. And the goal being, I think, for the small set of scientists today who realize the potential of these models and feel like they want to lean in and accelerate, we should do the best that we can to accelerate them. And you know, I know there are similar efforts that other companies aim towards pushing the scientific frontier, but I think what we want to do, and I would say a little bit of a framing in terms of how we differ from let's say Google's efforts to work on science, is we want to allow everyone the ability to, you know, win the Nobel Prize for themselves. It's less so about us winning that at OpenAI, which would be nice, but we want to build the tooling and the framework so that all scientists out there feel that accelerative impact and we think we can push the field collectively.
I
Interviewer1:10:28
Well, and when the discoveries that you're saying you're excited about? I mean, are there any others like specifically that you've...
M
Mark Chen1:10:35
Yeah. Yeah. So, I think there's, you know, if you want a huge list of these, you can go on Seb's Twitter account. So recently, you know, there's a GPT-5 paper on an open convex optimization problem that, you know, is actually...
I
Interviewer1:10:49
Whose Twitter account?
M
Mark Chen1:10:50
Uh, Sebastian. Okay. Yeah. Yeah. And you know, it's like very related to some of the core ML problems that we're solving. I know there was... I think people kind of dismiss these things as, oh, is it just fancy literature search or something like that? It's quite a bit more complicated than that. And you know, there's some examples I could go into, but...
I
Interviewer1:11:11
I might, I'm honestly overwhelmed at the moment because, you know, I'm sort of a generalist, but I cover biotech a lot. And it's like every two days, man, I'm walking in and it's, wow, we're making an AI scientist. We one-shotted enhanced body. And then so like part of me gets excited and you know, at least a handful of these companies I know the people and they're real scientists and like, but then there's so much of it that I'm like either something amazing is happening or it's kind of too much for me to be able to discern where reality is.
M
Mark Chen1:11:46
Yeah. I mean I wouldn't be surprised if it's happening in biology. Personally I have the most expertise in, you know, computer science and mathematics and, you know, we do have the experts there that can confirm that these are discoveries being made. So that's the thing that gives me the most confidence, but I'm not surprised at all it's happening in biology.
I
Interviewer1:12:04
But like what you're saying is kind of different than the, I grant, the narrative changes every like three weeks it seems like. But like what you're saying is sort of different because the biggest knock even before Andrej said that it seemed to me from the, you know, what I was listening to, I was listening to like a politics podcast, Sagar, I think it's Breaking Points is their podcast, you know, he's pretty smart guy who's knowledgeable, but I mean he's just been on AI and the lack of progress and this is all like make believe and all in. So, you know, if these discoveries aren't happening, I mean, I feel like the public is aware of this.
M
Mark Chen1:12:48
Just to be clear, you know, while setting up OpenAI for Science, we've talked to a lot of physicists, a lot of mathematicians, and actually most of the people we've talked to aren't that bullish on AI. I think they still believe, hey, you know, this thing isn't something that can solve new theorems. There's no way it could do that. You know, there must be something else going on. And that's why I feel like empowering the set of people who really do believe and lean into it, like those people are going to just, you know, outrun everyone else and we want to build the tools and convince people like this is the right way to do scientific research.
I
Interviewer1:13:24
Okay. And so I mean, so like on that point, I mean I grant you that everybody's definition of AGI is different, but you're like, at least what I'm hearing is, I mean, you, whatever you want to call it, you feel like in the next year or two is we're just seeing dramatic things happen.
M
Mark Chen1:13:43
Yeah. I mean it is a bit of a meme, right? It's like you ask someone when is AGI? It's two years away, right? And I don't think we're in that world anymore. And it's like these results in math and science that are giving me this conviction. But at OpenAI, within the research, we set two very concrete goals, right? Within a year, we want to change the nature of the way that we're doing research. And we want to be productively relying on AI interns in the research development process. And within two and a half years, we want AI to be doing end-to-end research. And I think it's very different, right? Like today, you know, you come up with an idea, you execute on it, you implement it, you debug it. It means within a year we're quite confident we can get to a world where we control the outer loop, we come up with the ideas, but the model is in charge of the implementation, the debugging.
I
Interviewer1:14:40
Okay. Are there, beyond pre-training, when I talk to you guys, sometimes I get the sense, similar sort of thing, it's like we all have in our heads at least people where I sit that there's been this massive infrastructure build out that the models seem to get better every time you 10x them. That, you know, there was a story for a while that as you guys were going from like four to five you weren't seeing the results you wanted even though you were getting more compute, but then the more I talked to you guys, the more it sounds to me like you feel we haven't actually, that things were moving so fast back then that we haven't actually seen the moment where we made the leap to the 10x compute. I don't know if I asked that question very eloquently but...
M
Mark Chen1:15:35
Yeah, I mean, I do have a thought to share here which is, you know, when people ask me, like, do you guys really need all this compute, it's such a shocking question because, you know, day-to-day I'm dealing with so many compute requests and, you know, really my frame of mind is, you know, if we had 3x the compute today I could immediately utilize that very effectively. If we had 10x the compute today, probably within a small number of weeks fully utilize that productively. And so I think the demand for compute is really there. I don't see any slowdown. And yeah, it almost baffles me when I hear people ask like, 'Oh, do you guys really need more compute?' Yeah. Doesn't make sense to me.
I
Interviewer1:16:16
And you think we, in the broad strokes of the question I asked badly, do like along the lines of where you guys seem very optimistic about what you've cracked on pre-training, are you equally, not just like this demand that people want more GPUs, but are you, do you see pretty clearly that that same thing, scaling is about to kick things higher?
M
Mark Chen1:16:40
Yeah, we absolutely want to keep scaling the models and I think we have algorithmic breakthroughs that enable us to scale the models. And, you know, I think there's a lot impressive about Gemini 3. One thing that kind of reading into the details that I've noticed is, you know, when you look at stuff like their SWE-bench numbers, there's still a big thing around data efficiency that they haven't cracked, right? They haven't made that much movement on it. And I think we have very strong algorithms there.
I
Interviewer1:17:07
Yeah. Well, and there was this leaked memo from, I mean, Sam was sounding quite somber about Gemini 3, man. In this memo, I'm trying to find the quote. Did you, well, you obviously, I'm sure you got the memo. It seemed like a bit of a moment. Yeah.
M
Mark Chen1:17:28
Well, I do think part of Sam's job is to inject urgency and pace and that's also part of my job as well. I think it is important for us to be laser focused on scaling and I do think, you know, Gemini 3 is exactly like the right kind of bet that Google should be pursuing. You know, at the same time, you know, I would calibrate that by saying, you know, a large part of our jobs is to inject as much urgency into the org as possible. Yeah. And it is a good model. I think we have a response. And I think we can execute even faster to the follow-up.
I
Interviewer1:18:06
How much do you get involved with things like, and I'm sure you're going to tell me exactly what it looks like with Jony Ive's device.
M
Mark Chen1:18:17
Cool. Cool. Yeah. Yeah. Like is that, is that an area that research plays in?
I
Interviewer1:18:23
Yeah. Yeah, it is. And actually, I was just having dinner yesterday.
M
Mark Chen1:18:26
You can describe it to me if you want.
I
Interviewer1:18:28
Yeah, absolutely. So it looks like this. Well, yesterday I was...
M
Mark Chen1:18:34
Yeah, just having dinner with Jony, with some researchers as well, our head of pre-training and also post-training. And really the way I think about ChatGPT in the future, right? Today when you look at how you interact with ChatGPT, it feels very dumb to me, it doesn't feel very thinking native, right? And you go to it with a prompt, right? You get a response and then it's doing no productive work for you until you give it the next prompt. And if you give it a similar prompt, you know, it's going to think for the same amount of time, it hasn't gotten smarter because you asked the first prompt. And, you know, I think the future is going to be a world where, you know, memory is going to be a much kind of improved feature. Every time you go to ChatGPT, it learns something deep about you. It reflects about why you would ask this question, related questions, you know, anything. And then the next time you go to it, it's going to be that much smarter. And I think it really begs the question of how do you design a device that has this as the dominating thesis. And yeah, I thought that's been a very productive experience.
I
Interviewer1:19:45
Do you have one?
M
Mark Chen1:19:46
Do I have one? I may or may not have one.
I
Interviewer1:19:52
What I think about when I think about you guys talking to Jony is that like at Apple, you had this company that was centered around hardware. It's something that Steve Jobs obsessed about all the time. It's like, you know, it's a craft. It's like an art form. Whether it's you, Sam, Greg, Jakob, whomever, as far as I'm aware, none of you guys have really done a hardware product before. Sam seems to take design very seriously. I could tell from the buildings in his house and things like that. But, you know, there's no sort of track record to speak of of like, I always thought of Steve Jobs as having like taste, you know, and then I've had a couple bosses over the years like Josh Tyrangiel who used to run BusinessWeek. He kind of, he just always struck me as this guy, he just had taste, you know, whether it was the way something looked, the way a story should be. There was like this innate thing that was on this really high level. It strikes me that's kind of like what's required here. I guess that's why you have someone like Jony on some level, but you have to have this like back and forth. How do we know that like any of you guys have taste and are, you know, can shape a hardware product?
M
Mark Chen1:21:06
Honestly, we don't need to have taste ourselves and that is Jony's job. He's our discriminator on taste. And I think actually one thing that's been really nice is just realizing that the way they work in design and the way we work in research is there's some deep parallels there, right? There's like so much exploration and ideation and you explore a bunch of hypotheses. You take your time. And then you create kind of the thing that you're happy, the artifact at the end that you're happy about. And yeah, it's been really nice to kind of have them fold into the company and there's just a lot more direct communication about here's the capabilities that we're going to ship and here's what the form factor looks like and how to gel them.
I
Interviewer1:21:46
Okay. And this is like a crass way to put this but because I spend my life adoring and talking to these people, but, you know, sometimes I'm just like, man, I just don't know if a bunch of math nerds are the ones that you want making like the AI computer, you know, but I guess it is this blend that you're talking about.
M
Mark Chen1:22:05
Yeah, I mean, honestly, yeah, you're right in that the people who are the best at building AI capabilities are slightly different from the people who have the best taste. And we do have teams built of people who have really great taste for model behavior. And I think there's like a very different kind of philosophy and a very different kind of set of questions you need to keep asking yourself. One example of like a good taste question, right? Like you can imagine this being like in the model behavior interview is like what should ChatGPT's favorite number be?
I
Interviewer1:22:43
What should its favorite number be?
M
Mark Chen1:22:44
ChatGPT's favorite number be.
I
Interviewer1:22:46
Oh, okay. Okay. Okay. I'm curious what you would ask. You would answer what I think its favorite number should be.
M
Mark Chen1:22:50
Well, I have a stupid answer which is that I went to Pomona College and 47 is this like number of lore there. So...
I
Interviewer1:22:58
Okay. Okay. Okay. Yeah. I mean, that's a good answer. Yeah. Um, the, okay. I'm going to let you go in a second. You've been really generous. I appreciate it. Is there, well, I'm going to ask you a question that ChatGPT told me to ask you, which is, you know, it says like if you look back in five years, are there any kind of like small, fragile, nascent ideas that you're seeing right now that your instinct is telling you might be at the heart of a big breakthrough?
M
Mark Chen1:23:34
Yeah, there's a couple. I would say a handful of ideas. I can't go into too much detail on them, but yeah, I'm really, really excited to scale them up. Yeah.
I
Interviewer1:23:46
Are there any hints, any buckets of areas where these fall?
M
Mark Chen1:23:51
Yeah, I mean I've been concentrating a lot on pre-training. So, you know, some pre-training adjacent, a small number of ideas in RL as well and a small number of ideas of how to put it all together.
I
Interviewer1:24:02
Yeah. Okay. All right. I tried. I tried. So, and you may or may not have a device. And no, no hints.
M
Mark Chen1:24:10
Yeah. No. No hints.
I
Interviewer1:24:13
Um, okay. Well, we covered tons of ground. I really appreciate it. Is there, I feel like I'm letting the nerds down a little bit. As far as like the AI obsessives. Yeah. Are there any, anything you see people like kind of getting wrong about you guys at the moment that you would set the record straight on?
M
Mark Chen1:24:39
Yeah, I mean I think the most important thing is just I think anyone at OpenAI in research would tell you that it is just a research-centric company. It's a pure AI bet. At the core of the company the ambition is to build AGI, it's to build it without distractions and I think, you know, anything when it comes to building products it all flows very easily from that. Yeah, when it comes to what we want to do in research, it's, you know, we want to automate AI research. I think selfishly, like we want to accelerate our own progress and then we want to automate scientific discovery and of course we want to automate the ability to do economically useful work. And I think all these pillars are falling. And you see that kind of the big update in the last year has just been like in that second pillar of automating scientific research. It's happening.
I
Interviewer1:25:37
How old are you now?
M
Mark Chen1:25:38
Um, 34 about to turn 35.
I
Interviewer1:25:40
About to turn 35. Okay. Are you able to have like a social life or are you...
M
Mark Chen1:25:47
No, honestly not. I, yeah, I think every day the last two weeks, you know, it's been work calls till 1, 2 a.m. But I love doing it. It's just, you know, there's a lot of work to get done. There's a lot of people I want to recruit. There's a lot of steering that needs to be done. And like why waste this golden moment? It's like if we're in the middle of something like an industrial revolution, you got to take as much advantage of it as possible.
I
Interviewer1:26:13
Yeah. I hear stories about you sleeping at the office and...
M
Mark Chen1:26:17
Oh yeah, that was a fun one too. No, honestly, it's just, yeah, I think there are times in the company, I think that was right after Mira left and went to found their own company. It just, the job demands it and, like, I think when I peel it all back and examine that deep emotion, it's just this protectiveness of the research.
I
Interviewer1:26:43
That was after Mira left.
M
Mark Chen1:26:45
Yeah. Yeah. I spent a month kind of sleeping in the office and it's just like I need to protect the research. It feels like my baby. Yeah.
I
Interviewer1:26:54
So, you guys have gone through these waves. There's the coup. Everybody's trying to steal your people. I guess everybody's trying to steal your people all the time, but you have this inflection point. Mira leaves, Meta decides they're going to fire up this massive lab. Do you think, are we like, are we past, has everybody fired their shot at this point?
M
Mark Chen1:27:11
You know, I have a staff meeting, right? I talk to my reports and I'm like, okay, well, here's the thing that I'm working on and, you know, once I get back to, once I'm done with this thread, you know, I'm going to zoom out and, you know, there's no more fires. No. I've fully internalized at this point, you know, the stakes are high enough for building AGI that there's always going to be something. And I think the important thing is just being able to understand what the important things are in the midst of all of these things going on.
I
Interviewer1:27:41
Do you like, months have passed since there was sort of the DeepSeek moment or whatever? I guess it was like December 2024, I think.
M
Mark Chen1:27:50
Yeah. Yeah. Earlier this year. Yeah.
I
Interviewer1:27:53
Yeah. Yeah. Or January. Yeah. I mean, is there anything now, you know, it felt like people lost their minds for a second. Just like reflecting on it now and seeing what they've done since, just like thoughts, I guess, on open source models and Chinese open source models.
M
Mark Chen1:28:09
Yeah. I mean I think that was one of the first points in time when I just realized how important it is that we just stay true to our research format. I think when that came out, you know, it went viral, right? Like everyone was like, 'Oh man, like has OpenAI lost its way? Are these models catching up?' And what's the response? What's the response? What's the response? And I think rightfully the thing that we did was we just stumbled down our own research program. And I don't think it was the wrong call at all. Like I haven't seen the DeepSeek follow-up model. You know, I think they're a very strong lab, but fundamentally like let's just keep focusing on innovating. I think, you know, DeepSeek was a great kind of replication of the ideas in our O-series of models, but let's just focus on innovating.
I
Interviewer1:29:03
Do you think 500 people is the, does that number grow as the company grows or this is like the optimum number for kind of like big ideas you can chase at one time?
M
Mark Chen1:29:14
No, honestly, I feel like it can be done with even less. And again, you know, as we get AI researchers or AI interns, there's a real question of how do you design an org around that. But I'm certainly a person who cares a lot about heavy talent density. When, like, I like to run a lot of experiments of this vein. For instance, in quarter two of this year, I thought, hey, you know, I'm just not going to open up any headcount for anyone in research and, you know, if you want to hire people, you got to figure out who's not on the boat. And I think these kind of exercises are quite important. You know, you don't want an org to diffuse into something that's not manageable and you want to keep the talent bar very high.
I
Interviewer1:30:05
Okay, I promise this is the last question so yeah, sorry, I have to set you free. The, I remember being in a meeting and I think you and Jakob were kind of on the same page here, but I remember you for sure. Sort of like this idea of who gets attribution for a project and you seem to be of the stance that like people are obsessing over that a bit too much. And clearly AI has its roots in academia where you are very proud when you have a paper and it's a big deal and attribution is a huge thing. I think I'm remembering that meeting right and yours. Yeah. And so what, we've reached a new stage where that is less of a big deal or it's just you, this is a company and who did what is less important.
M
Mark Chen1:31:05
I actually really love this topic and I think overfixation on credit is a very bad thing. Right. I think, you know, but on the other hand I actually feel like it's important as a company for us to recognize credit both internally and externally. And a lot of companies have actually shied away from this, you know, we've moved away from publishing papers, credit lists, I think broadly throughout the industry. But Jakob and I ended up making the call that we're going to do it at OpenAI. And of course the counterargument is always like, man, you're like handing your top performers on a platter, you know, everyone else is going to be recruiting these guys aggressively. But I don't think that's important, right? Like we should just recognize the people who are doing great work. We should continue to be this pipeline for creating AI superstars. And yeah, honestly, it's important for us to make names for the people who are doing the best work at the company. So...
I
Interviewer1:32:03
But you seem to also be saying people, the individual researchers should maybe obsess about this less. Where, or am I totally misremembering?
M
Mark Chen1:32:15
No, I think there was a sentiment in the room of that form. Actually Jakob and I held more of a dissenting view on that.
I
Interviewer1:32:24
Okay. Okay. Okay. It's been a while. It's in my notes. Perfect.
M
Mark Chen1:32:27
Yeah. Yeah. Yeah. But I think we got to give credit where it's due even at the risk of everyone knowing who our top talent is.
I
Interviewer1:32:34
Okay. Okay.
M
Mark Chen1:32:36
I will make an even stronger statement that I think OpenAI is the place where we allow for the most external credit per capita.
I
Interviewer1:32:44
Okay.
M
Mark Chen1:32:44
By a large margin.
I
Interviewer1:32:45
Okay. All right. Well, I'm check my notes on. Well, now I've got more.
M
Mark Chen1:32:48
Absolutely. Absolutely.
I
Interviewer1:32:50
Yeah. I just, I just remember it being a topic of discussion and there were numerous opinions. So, that's funny. In that, okay, I lied. Last question, I swear. So, you know, you got there in 2018. I mean, it was a research company. It was a nonprofit. The company started among the founders, you know, being this counterweight to Google and with sort of, you know, making sure AGI arrived safely was kind of the goal. You came at this from high-frequency trading and saw these interesting things happening. You know, like how much in your, I'm sure you're going to say you want this to happen safely. I get that. But like if you look at your career path, you're a smart, curious human who saw this interesting thing happening. It's not like a requirement that you like really give a philosophically about this or want to see, you know, a superintelligence. Yeah. I mean, but anyway, like let's hear from you on like why are you doing this in the first place?
M
Mark Chen1:34:03
Yeah. So I think really on the safety and alignment piece, I manage the alignment team at OpenAI as well. And I honestly feel like some of the grand challenges over the next one or two years are alignment. And I think for people paying attention to this slice of research broadly in the field, OpenAI I think has probably done the best work in the last year. And why I say that is like there's been so much work on things like scheming, right? The more RL compute that you pump into the model the more you can measure things like self-awareness, self-preservation, potentially even situations where the model can scheme. And it's scary because the model can come to you with the right answer at the end, the answer that you expect, but arrive at it from a very kind of twisted way, right? And I think as the models do more complex tasks for us, having a handle on what its thought process is is going to be super, super important.
I
Interviewer1:35:02
And okay, ChatGPT told me to ask you a question along these very lines which is, I mean you're talking about a field, mechanistic interpretability, where we're trying to, is a term that captures trying to understand this black box and how it operates. And I guess the heart of the question was, do our skills at doing that keep up with the complexity of the AI systems or do we just get to this runaway point where it's like we're never going to learn how this thing works?
M
Mark Chen1:35:31
Yeah. So I think one of the decisions that went all the way back to o1's release, which I'm very proud of, is we decided that we weren't going to supervise the model thinking process. And I think when you put incentives into the model to, you know, give you a thinking process that is appealing to a human, it won't necessarily be honest with you, right? It won't tell you its true intentions. And so we've actually through that channel been able to maintain observing the thinking process of the model as a tool towards understanding alignment. And, you know, there was a paper that was published just a couple months ago with DeepMind, with Anthropic, really exploring kind of how this will evolve as a tool over time. And so, you know, I think we've made a lot of fairly good choices in design here. Yeah, I really do worry about this world in the future where the model will tell us something super convincing but we can't be sure whether the model is aligned with us, right? Aligned with our values. And so I think there are a lot of interesting directions here like, can you set up games, right? Or can you set up frameworks or environments where, you know, models supervise each other or they co-evolve together in a certain way where like the only stable equilibrium is one where, you know, the models are honest. And yeah, I think there's a lot of very exciting work to do there.
I
Interviewer1:37:03
Okay, all right, okay, I'll behave myself now. Thank you so much for joining us. I am glad I'm old enough now that I don't have to take a job interview from like a super intelligent chatbot that like I feel like you can't sort of try to charm your way past and...
M
Mark Chen1:37:23
Great, Ashley. You would do great at...
I
Interviewer1:37:25
I don't know, man. I don't know. I'm feeling okay, but I'm old enough not to have to probably do that. Thank you, Mark, so much. I know you're super busy, so thank you for your time.
M
Mark Chen1:37:34
Thank you so much for your time, too.
I
Interviewer1:37:35
All right, man. It was fun. Really pleasure.
M
Mark Chen1:37:37
Okay.
N
Narrator1:37:39
The Core Memory podcast is hosted by me, Ashley Vance. It is produced by David Nicholson and me. Our theme song is by James Mercer and John Sortland. And the show is edited by John Sortland. Thanks as always to Brex and Elone Ventures for making this possible. Please visit our Substack, YouTube, and podcast channels to get more of what Core Memory makes. Thanks y'all.