We can't find the internet
Attempting to reconnect
Something went wrong!
Attempting to reconnect
Open Source Wins, AGI Is Here, and Scorsese's AI Toolkit with CEOs of Cerebras & Black Forest Labs
All-In with Chamath, Jason, Sacks & Friedberg · 1:03:57 · 12d ago
Transcript
We are in the race for superintelligence, and Andrew Feldman is back. And obviously, CEO and founder of Cerebris, doing inference chips, pioneered the space, had a successful IPO. We've talked about this a couple of times. We got to see each other in January at Davos. IPO happens. The boys and I got to sit with you recently. That was fun. At Liquidity. That was really fun. Had a great discussion with the boys. But I wanted to deep dive with you about a couple of topics. The first one is the build out of AI. We've never seen a build out like this since the Great Wall of China. Right. Who knows? The pyramids. Right. I mean, it feels like the amount of capital, time and intelligent people on the planet dedicating themselves to the build out of something. I can't think of anything in our lifetimes, but perhaps before our lifetimes, the war effort. This is a mobilization in a scale that we read about, we hear about, but you're actually doing it. You have customers who are building data centers, and you're a key piece of that. App Lovin started with an $8 domain and no VC funding and became one of the largest ad platforms in the world. Now that same engine powers AppLovin ads for e-commerce. Your ads run inside mobile games, reaching over a billion people with full screen, distraction free attention. The platform finds buyers and optimizes for profit. You set the target, it does the rest. One cookware brand went from $4 million to $16 million, turned profitable and is on pace for $80 million this year. Visit AppLovin.com slash all in to launch your first campaign today. Maybe you could just enlighten us in 2026. What is Cerebras doing and what is happening with this build out out in Texas? These are some gigantic, gigantic efforts. The size and scope of what is being built, the physical size and scope, usually when we talk about software or we talk about hardware, we're talking about chips and boxes and they don't have the same sort of physical enormity. Right. Right, right. And what we're talking about now are data centers that are in the next several years going to use more power than the previous 50 years on Earth took. Wow. Right. We're talking about individual buildings besides the football fields that have more power coming into them than midsize cities. And they're being built. They're being built across the US. They're being built in Canada. They're being built throughout the Nordics. They're being built here in Paris and throughout France and Europe and the Middle East in nations that sort of weren't front and center in anybody's mind previously. Kazakhstan, Tajikistan are building out Georgia, building out data centers of size, Armenia. Everybody sort of focused. Every country and every state obviously in America feels they need to participate in this. and the people who are buying the capacity, the OpenAI, Anthropics, SpaceX AI, the Googles, they are insatiable right now. Yeah. And they're building how many years out? When you talk to them, they were ordering chips from Cerebrus before you were finished with the chips. They're putting orders in ahead of time. The irony is, unlike many sort of exciting times in technology, they're trying to capture yesterday's demand, right? The demand is way outstripping our ability to build data centers and to fill them with hardware. All right. And so, you know, we have a $25 billion backlog. $25 billion backlog. And we are not alone in that. that open AI, Anthropic, you go through this list of Google wants more data centers, Microsoft wants more data centers, AWS wants more data centers, right? All of these players are not chasing sort of if you build it, they will come. They're chasing the demand is booked. How do we keep them from leaving? Right. And that's extremely unusual. It's very unusual. And now we have people who are, you know, we have a term for a token maxing. Yeah. And there's a great debate. Is this actually creating value? I'm curious where you stand. You know, is it even possible that this much demand could be created if value did not exist? There is clearly massive value happening. Yeah. But there's also massive experimentation. Oh, for sure. I liken this to when we first started with AWS, and it was so good to get around your own IT organization that you told every engineer, yeah, go ahead, put on your credit card, sign up. And a lot of it was really useful, and some of it was like, God, I wish we didn't do that. And so for sure, there's experimentation. But it doesn't mean that the net value isn't enormous. It means some of it is going to go nowhere. And it was the same. I remember when Costco opened up in the Palo Alto area in 1988. And people used to shop Costco like they shop Safeway. They'd go down every aisle. Yes. That's a horrible way to shop Costco because you end up with four things you didn't need and each was $22. Right. And as people got more sort of accustomed to it, you go to the back, you get the chicken, 18 cupcakes for the kid's birthday party. bang you were strategic and it's exactly the same i think at first people opened up and said everybody as much tokens as you want and in enterprises there's no open loop we don't give sort of any resource unconstrained to people and now we're jumping on saying whoa all right these guys should have as much as they need they're enormously productive over here we can use maybe an open source model, maybe a cheaper model over here. And now we're sort of running like a business. And we're really seeing a certain type of person emerge who knows how to deploy this technology. Systems thinking, which developers kind of have innately. CEOs tend to be great strategists and understand systems. But the intelligence is getting so much better every step along the way. that I'm watching individuals, typically startup founders, but also venture capitalists and associates who work at my venture firm. They start playing with the tool and then the tool starts playing with them. They start to go, oh, I haven't clearly defined what my goal is. I don't understand what a system is. I've never heard about making a requirements document. And the software's like, do you have a requirements document? What's your goal? The AI starts telling people, you're token maxing and you need to get a little more focused here. One of my colleagues 20 years ago, a really smart, smart computer scientist said, computers are really dumb. They do exactly what you tell them. And at first, prompting was like that. You modified your prompt a little bit and it changed the answer. Dramatically. Dramatically. And increasingly, it's understanding what your intent was. Right. Right. And if you have a chance to play with Fable or 5.6 from OpenAI, increasingly you don't have to get the prompt just right. You don't have to be a prompt whisperer. Instead, you ask it and it says, well, here are some things. And by the way, maybe you wanted the chart to go two ways. You wanted a line and a bar. And it's like, well, that's exactly what I wanted. I didn't ask for it, but that is better. And so it's understanding intent. And that's a huge leap. Which, if we were sitting here two years ago, the idea, we would never have been able to predict in a short 24 months that it would go from being a great summarizer researcher of web results to actually understanding your intent and then providing a solution and abstracting it all from you. That's right. Which is a very weird thing. I don't know if you've played with the ERMES agent yet. Yeah. Have you played with it yet? I mean, I asked it just this morning, and I was given a secret BitTensor project that has the new ZAI model, 5.2. And they gave me – GLM 5.2. GLM 5.2. So somebody in that BitTensor – I think you understand BitTensor. You've heard of it, the distributed crypto project. And so they have all this extra capacity. A whisperer told me, probably some capacity in China that has free energy. Okay, fine. So they gave me unlimited capacity. So I started having to do some really crazy jobs where I was saying, like, every hour, I want you to tell me what the trends in the world are that nobody else has identified yet. And you can do whatever you want to do that. But my goal is to be the smartest trend hunter in the world. And I watched what it was doing in the background. And it started debating itself on where it should find the things. It said, well, we should probably go to Hacker News and Reddit. And then it was like, yeah, but there's also social media and trends tend to manifest on Instagram. That's a reasoning model. You were watching a reasoning model work out. Yeah. Isn't that interesting? I mean, that's amazing. And it was collapsed. So as a civilian who doesn't hit the uncollapsed moment, and if you were using ChatGPT 3.5 or you were using 4.8, whatever it was, and you haven't used this new level of reasoning and inference and unlimited compute, essentially. It opened my eyes just this morning of what a world of unlimited tokens might look like. Because unlimited tokens, I believe, means unlimited reasoning. It does. What does that mean? Yeah. I mean, if you run these for 25 or 48 hours, you get amazing things now. And what if by using Cerebris we were 15 times faster and then you ran it for 24 hours, right? And you got weeks or months worth of thinking. And I mean it is extraordinary. And I think one of the things is people like Ilya and Sam in the early days were saying this was coming. Right. Right. And I think when you look back, you say to yourself, holy crap, those guys saw it. Yeah, they could see around the corner. That's right. And the rest of us were like, what? I'm not sure. When we had Sam on All In at one point, and he said, you know, I'd love to come on at some point. I said, sure, come on. And he was talking about it. He said, you know, I said, what's next? He said, reasoning. I said, unpack that. What does it mean? He's like, well, understanding what your intent was, just as you're saying, and then figuring out a strategy, and then maybe talking to other agents and other threads about like, is this the right thing to do and vetting each other's work? And I'm like, wow, we have come a long way from guess the next word. Right, right, right. Fill the sentence in, you know, summarize this PDF. Now, Cerebris is at the center of this because this reasoning is inference. This reasoning is inference and it's computationally intensive. Right, right. And so fast compute makes this sort of work fast and sort of tractable. It doesn't cripple it by taking a huge amount of time to get a good answer. And so it's exactly the fact that this reasoning consumes a huge amount of tokens internally that allows a blisteringly fast machine like ours. And I brought one because I'm never far without – when one costs half a billion to make, you bring it everywhere with you. We were tossing this back and forth at Davos. What's the model number of this one? this was in the first eight or ten got it so this has a special place this has a special place i mean my wife says it's like i'm a kid with a dirt bike for his eighth birthday he's in his bedroom at night i i carry him with me i mean when you have um you know your your next party at the house i highly recommend just a little hors d'oeuvre i think it'd be like a great fit it would be a great bet if you had some. That's right. But what we're looking at here is the ability to do that reasoning at scale. And what is Moore's law for inference and for cerebris? Do you have something internally you discuss as we're going to double this every X time period? So all chips prior to us in the processor world followed Moore's law. Got it. And we broke it. Doubling every 18 months. Doubling about every 18 months. Got it. And we crushed it with this chip. And we've carved out a whole new trajectory. And my view is in the next 18 months, we'll be way over 2X. Interesting. And so I think that early in an architecture, you have room to do much better than what was traditionally Moore's law. Now, if you've got a 20-year-old architecture, like the GPU, it's much harder. You have to rely on things like smaller geometry, going to the next fab node. But in a newer architecture, you have a huge amount of room still to learn about the work that is being presented and make optimizations that give you huge gains. How do you run the company? Just being the CEO now in the age of AI, you have $25 billion in demand. You have to deploy at just an incredible blistering pace. You have to hire people. You have to create a roadmap. I don't mean to give you a panic attack here. You have to keep up with somebody like OpenAI who's moving so unbelievably quickly. Yes. Right? And they're competitive. You've got to keep up. Right. Right? Your hardware, your software, your deployments have to keep up with some of the fastest moving organizations in history. They're demanding customers. They're not pushovers for sure. Yeah. And also potentially competitors down the road. Look, I think there is so much demand right now that there is no silicon that will go unused. But why is an open AI releasing jalapeno? Why is Amazon making their own chips? You see this reoccurring trend. Is it a way to let you know, to let Jensen and NVIDIA know, hey, we can do this too, so we need good pricing? Is it a little bit of a flex that way, or is that the future that they're going to be in your business? No, I think nobody likes being dependent. and i i think some of the lessons learned by the the hyperscalers of the x86 world is they were dependent on intel and uh some of the lessons learned by uh the gpu makers was they were dependent on a small number of hyperscalers yeah and they wanted more customers and so they set about to help fund these neoclouds. And so I think mostly it about an opportunity to control at least an important part of your destiny Got it And I think that a very reasonable thing I think you don have to sort of make the fastest chip You just can't be entirely dependent on other people's chips. And that dependency has become a hot topic. I'm not sure if you caught the episodes over the last two weeks, but we've been talking over the last year about open source. I've been championing that a lot just because I was early into OpenClaw and quickly started using Kimi and was like, wait a second. I'm blowing out my claw tokens, but this Kimi, I can't tell the difference. And then we started smart routing it, and suddenly this open source started to figure out reasoning, and the gap suddenly closed this year. you don't want to take your Ferrari to the grocery store there are times you want to drive your fun car and there are times you want to throw the kids in and don't worry if their Cheerios is on the floor minivan time and I think that as the sophistication of the user grows you're going to have hard problems and those are going to be frontier model problems they're going to be open AI problems There are going to be anthropic problems. There are going to be Gemini problems. And behind that, there are going to be a lot of ordinary problems. I mean, if you think about a company, you know how much time is spent cutting things out of workday and getting it in a different cell? Yeah. Think about – The cutting and pasting economy is real. That's right. And this doesn't need gold medal masks. No. What this needs is sort of rock-solid open-source capabilities. Yeah. And if you think about what, I mean, we've been thinking a lot about it in GNA, but a huge amount of GNA, all right, is not invention. Right. And you may not need sort of the most sophisticated agents for this. And another card that's turned over recently is some folks maybe have concerns with the ambition of the frontier models and maybe sharing their data, data leakage, and sovereignty of intelligence. And they're saying, hey, our company is going to choose. Maybe we're in a regulated industry, finance, healthcare, HIPAA, FINRA, all kinds of different regulations. We need to have this on-prem. We want to have domestically, and we'd like an open-source version where we have a little bit more control. Are you seeing that now? We are seeing that for sure. I think OpenAI made a good call releasing OSS-120B some months back. That was a good open-source model. But I think in the U.S., we need more domestic open-source models. We need to give the world a choice. right if they want to run open source right now it's oss 120b or chinese models nvidia has some nvidia has seen the same opportunity yes to push open source models i i think giving them more power might might be sort of well i was imagine that was you cut me off at the past like it my understanding was jensen was like hey we we don't even want to talk about these open source models we have because our customers, we're now going to be competing with Sam, Dario, Elon, Sergey. Do we want to be in that position? But we do need some more champions here, and it's open source so people can fork it. But that puts you in a more neutral position. That's right. We run today. We run GLM. We run Kimmy. We run the Quen set of models, and we run OpenAI's models, the closed source ones. We run models for, say, GlaxoSmithKline, which they wrote and developed. We run models for our partner in the UAE, G42 and MBZUAI, that are their models that they designed. So we have a wide variety. So sovereignty is a trend. A sovereignty is a trend and I think the government's actions with regard to stable and 5-6 where they said, oh, whoa, let's think and then we can act. I think sort of particularly here in Europe was a bit of a wake-up call. And when you saw this going down, there's a layer of partisanship in our country right now. It's pretty fervent. And Dario is pretty explicitly not part of this administration. They've been very adversarial. Both sides have admitted that. They're starting to work it out now. So it's hard, I think, for us not being in the room with these parties to understand what's partisanship, what's gamesmanship here. But do you believe that what they released was truly dangerous for cyber warfare, for cyber attacks? and that if you were to rate Dario's, not communication, because he's a very effervescent communicator, I think is a diplomatic way to say it, but to have a scheduled rolled out release. We'll put aside the government's control of it. But do you think that is a wise thing for us to do at this point? And do you think there was actually a major threat there? So what's interesting is I hadn't seen it before. Right. And I think if we just step back and say, is it reasonable? I don't know whether this was the right time, but at a time that a model is sufficiently creative in its thinking, that it poses a meaningful threat for the government to say, we'd like you to roll it out in steps. Yeah. This doesn't seem unreasonable to me. Not at all. Right. I mean, we do this with powerful pharmaceuticals, right? Sure. We'd like, I mean, we're certainly not encouraging seven years of trial and the amount of paperwork and all the garbage that has accrued to the FDA. But with a powerful new technology, it certainly doesn't seem unreasonable to say, hey, guys, let's at least do some red teaming at the government so we know our defenses can block this. Yeah, have we checked? Have we checked the infrastructure of the country? Of the NSA. Have we checked the infrastructure of them? Right. And can you give us two or three weeks to patch any obvious holes that are found? This doesn't seem to be an unreasonable thing for the government to ask. Right. But we, in this very polarized time, put on top of it, well, oh, my God, it's President Trump doing it. And then you have to think, well, what if it was President AOC or President anybody in between the two extremes? I think the polarization hurts a great deal. It hurts clear thinking. It hurts clear thinking. And both sides are going to do some dumb things and some really smart things. Right. Right. And in fact, what I found is that the people in the government are trying really hard. The rank and file. The rank and file are trying really hard. And this is moving fast. And I think that an ability to set aside some of the polarization and say, how do we do this in a reasonable manner? I mean, we want Dario and Sam competing like crazy. We want them. It's been awesome to watch. It's awesome. Yeah. Right. It's good for the technology. It's good for entrepreneurs to see even with thousands of people, this is what you can continue to achieve. Right. Right. This is a drive. Kick Google in the ass. That's right. get sharper, Amazon starting to wake up. Everybody got better because of that. We want that. And we certainly don't want to become sort of a region where the first thing we want to do is regulate it. Right. Right. But as it gets more powerful, and the industry really should do a better job of regulating itself perhaps. And it did seem like they were starting that process, but then the communication was lacking. Maybe it's, you know, I think not only are they racing hard, but they're inventing this as they go to. Yeah. Right. There's not a playbook. No. Right. They're inventing that we just put on guardrails. Well, they have to design the guardrails. Sure. Right. The guardrails have an impact. You know, one of the things that fast does is it makes the guardrails less painful. And so we discovered that in the last six weeks is that the very guardrails can add time and make it feel slower. And so fast ships like ours can really help that. But so they're racing against competition. They're racing against their own sense of greatness, which is maybe even the biggest driver here. And I think they're earnest trying to think about how to do the right thing. And all of those are mixed in this bucket. And sometimes you're on one side rather than the other. Yeah. And as you're saying, this is a first time, right? That's right. When 3.5 came out, it wasn't like when we were using ChatGPT 2.5, 3.5, it was taking down networks. But in talking to Nikesh from Palo Alto Networks, I asked him, like, hey, well, how would you grade this? And he said, we put it against our software. and we found bugs we were not aware of. Yes, it killed him. Yeah, he said, we had to stop everything we're doing and do patches for six weeks. Right, and that's when you know, right? I mean, Nikesh leads maybe the leading security software firm, right? And when it finds in an hour, right, tens of critical opens, you're like, whoa, this is a powerful tool and we need to think. And maybe you show it to a group first, right? Maybe you – I don't know what the right thing is. I mean, red teaming. Right. And we've always had – just when you were releasing the new version of an operating system, when you have your iPhone, you can say, I want to be part of the beta. That's right. Right. And there's like two other betas that you don't even get the chance to opt into as consumers. Right. And those ones are for security. Those ones are for making sure you don't lose your data or data that's a leak. That's right. Disappear or leak or corruption. Any number of these things. I think we can also know that there will be a massive data leak. Of course. We know this. Yeah. Right? And it's like Warren Buffett talked about the reinsurance industry that you know something bad is going to happen. You don't know when. Yeah. But you got to save up for it. Right? You put money away for it in reinsurance. Yeah. But there will be a tornado. There will be a massive earthquake. We know this and we can do our best to plan, but there'll be a massive breach and we have to steel ourselves in advance. And we have to think about it, think about the right response at the time and sort of prepare ourselves for a future that is in specific unknown. But in general, we're pretty sure something's going to happen. something will happen right yeah and yeah it's typically a black swan right that's right by definition it's going to be something we didn't consider or a question we didn't know to ask right but but even knowing that there's some unknown unknowns is a useful place to start yeah what are we not asking ourselves that's right with reasoning the ai is going to be able to tell us hey schmuck humans that's right by the way here's what you're not thinking about This is now my closing sentence when I do my prompting is I need you to make me a prompt that will help me do this trend scouting, for example. And then I always say at the end, please check your work. Right. And then tell me what I haven't considered in terms of my goals. And ask me some questions every time you run the job. And that has changed everything because it's like I checked my work. By the way, this was incorrect. Right. And I'm wondering, hey, would you like me to also do this? And some of the tools like Perplexi do that automatically. They give you your next three prompts. But if you give it explicit instructions, my Lord, is it good at that. So over the course of the last 10 years as I was raising money, I thought one of the smarter questions I got at the end of a conversation where someone asked, what was the smartest question you heard that wasn't covered by what I asked? It's incredible. Right? Now, that's somebody who's curious and thinking and humble and trying to sort of use this to get a picture of the space. And to the extent that you can ask the AI that and that it can sort of broaden your view, you know, maybe what questions should I have asked to be an expert in this? What would a PhD level questioner ask of this or a gold medal math? I mean, I think those are sort of questions that you know you don't even know how to ask. Which, you know, if we start thinking about AGI and superintelligence, you know, they're just definitions. But they're important definitions, I think, to kind of keep in mind because they're waypoints. That's right. And AGI, I think, I suspect you'll agree with me that we've hit it. We just haven't exactly deployed it fully. We have artificial general intelligence. Now, it feels like when we're talking about these reasoning moments and the ability for it to be as smart as any human. But let's talk about it. By any definition we had 20 years ago, we've hit it. Yes. Right. I mean, if you think about it, oh, there was a Turing test. Blew it away. Yes. I mean, you think about that any period of time, sort of 10, 15, 20, 30, 40, 50 years ago, any definition we would have previously put forward, we've blown past it. Which goes back to our previous point of like, do we know the questions asked? That's right. 20 years ago, science fiction authors had their say and we answered all their questions. Right. If they were to look at this today, they'd be like, well, I'm out of question. I'm out of question. Sorry. That's where sort of listening to people who sound sometimes like they're on the fringe. When Ilya was talking eight or ten years ago about the need for safety and you're like, what? Dead right. When Elon was talking about building rockets and driving the cost to near zero of a launch vehicle, you're like, what? There it is. And now you can see – and that's – I think that's why it's really fun to be a technologist now. Well, and with these tools specifically, we're talking about building all these tools, and then the tools are starting to build themselves in this recursive loop. That's right. We're kind of just starting to see people apply loops. in fact loop maxing became when i was doing my trend when i did my trend thing right it kept picking up loop looping and it kept picking up the maxing stuff and it created a buzzword for me loop maxing right and then it magically people started talking about loop maxing and i was like wow this is really weird it anticipated that this would other humans would come up with this word but talk a little bit about recursive and then the road to super intelligence and do you have a way andrew that you think about super intelligence and what it will mean for humanity and how we will define it and how we experience it yeah i i think let begin on on on loop maxing or sort of recursive learning I think what Sam and Ilya and then later Dario and Dennis saw six years ago or five years ago was that powerful recursive gains are exponential, right? You get better, you do it again. And if you continue to get gain, the slope of that curve is so steep. Yeah. And that we're just beginning to see that now. You ask it a question, you learn from the results, you ask it to do it again. The results get better and more information is added. Your answer gets better. You ask it to do again. It covers more material. And these sort of loops are producing sort of not a little bit better answers, but vastly better answers. Yeah. And that is enormously powerful because we don't quite know where it ends. Right. You keep throwing compute at it. I mean, how much better does the answer get? We run out of tokens or our budget, but holy cow. I mean, when does the exponential stop or does the answer keep going up and up and up to the right? Yeah. And that's sort of an enormously interesting intellectual question right now. Yeah, like when do we run out of problems to solve? Well, that's right. And when are the problems no longer sort of intellectual problems and they're now people problems? Yeah. Right, how to organize people to get done what the AI asked for. right i mean as you know and running your company a lot of your problems aren't hard intellectual problems they're people working together problems yeah right and motivation motivation you spend a lot of time as a leader spraying wd-40 on your team right right it just so friction is reduced and um how do we learn about those from ai right how do we get behavioral insight from AI. And I think that's some of the things the world models are going to bring us as they begin to watch human behavior. Yeah, we didn't even get to that. This is going to be for another interview. But when these things jump off the screens, right, and they're in the real world and the recursiveness starts, not trying to solve math problems and, you know, humanity's most difficult ones. But, hey, you know, there's an incredible world out here and here's the Palace of Versailles. Right. You're just like now we're like, make me a new version of Salesforce. And we're like, hey, you know what? I'd like the Palace of Versailles. I've got 100 acres somewhere out in Texas or Nevada. I'll just send 1,000 optimists out there. Make me the Palace of Versailles. Right. Sounds fantastical, but the Palace of Versailles would seem fantastical to people who lived 1,000 years before it. And it was fantastic, I think, to the people who built it. Right? Even to the builders. I think they were awed at it as they built it. Yeah, they're compounding recursive learning. That's right. And generations we talked about, you had really such a great insight of in building this place, you had generations of masons. Yeah. I think in all these large projects, often there were families who were specialists and you apprenticed on your father, your uncle. and when you had a project that took 50 or 70 or 100 years, you might have three or four generations of the same family, the same stonemason family working on the same structure. And passing on the learnings. New innovations, which is what we've modeled with this new models and what you're building in the infrastructure. It's pretty incredible when you think about it. Especially when we're sitting here. But the pace. And that's what – I mean, I think the problem with human learning is it often moves at the pace of a generation. And like elephants and other large mammals, we don't have generations but every 15 or 20 years. And if you want to move really quickly across generations, you want them happening more like drosophilae, like fruit fly. You want two a day. Yeah. Right. Then you see that in genetics. That's why we study them in genetics because learning encoded in the DNA, you can study over thousands of generations. I think that what we're getting is that equivalent in AI. We're getting sort of learning so quickly over the equivalent of thousands of generations. Yeah, Darwin would be in awe of this pace of evolution. That's exactly right. You think about it as – I remember when I was getting my psychology degree and they were teaching us about paradigms. And I was like trying to understand how the paradigms shifted. And the professor said to me, Jason, which you have to understand is paradigms don't die. They don't. People do. That's right. That's how Freud – He and Thomas Kuhn. That's right. Yeah, Freud and Skinner and Jung. It took them dying. That's right. It took the next generation to question it. And that was 20 years, sometimes 40 years as their students maintained positions of leadership until someone said, maybe we could do it differently. And I think what you're seeing is this iteration is a shortening of the intergeneration gap. And the learning is so fast. it's uh always so great to talk to you because uh one it's just intellectually um uh so your approach to it is so intellectually rigorous but also um with so much p doom in the world i feel so good that you're such an optimist about this technology and you're building it with such thoughtfulness and it i think for people who are hearing these horror stories about ai and job loss and everything. They need to understand there are people like yourself who are building this in an incredibly thoughtful way and this is going to be a net benefit for humanity that just is unimaginable. We have a shot with this technology. So not our children nor anyone they know dies of cancer. I mean, say it like that. There will be some dislocation in the economy. Sure. There will be. There was dislocation when cars came and it was a bad deal to be a guy who shooed horses or built carriages. But you got to also against that, make your tea of the cons and the pros. There's a shot that our children, none of them nor their people they love will die of cancer. And that's one thing that we can work on with this technology and we will have great purchase on. And I think you begin listing those and then it's a more thoughtful discussion. Yeah, unlimited energy, unlimited calories, unlimited knowledge, unlimited education, unlimited housing. And how we do it. Imagine sort of we know how to teach children and we don't do it, right? Aristotle was a tutor to Alexander the Great. Socrates was his tutor. We know that if you give a child a tutor and the tutor modifies the teaching for the child, they learn better. That's not how we teach in classes. No, factory farming. That's right. We teach to some sort of mid-level. Imagine if we built agents that taught children for their way of learning. Right. Right. And here's the way. We've been doing it the same way for a thousand years, and during that entire time, we knew how to do it better, and we chose not to. Yeah. And here's a way we can do it. Put that on the pro side. And so as long as we're sort of thoughtfully and fairly writing the good and the bad, I think it'll come out very well. You've got to get out there, Andrew. and keep communicating your version of the world because some people see around the corner and they get a little nervous and okay, fair enough. But I think the ledger, as you describe it, is heavily weighted towards abundance. I think it will create abundance for sure. Massive abundance. Andrew, I'll see you in six months for our checkup. That'll be great. I'm going all in. Industries, capital, and intelligence are converging into a single interconnected system, and the infrastructure behind it needs to evolve just as quickly. NASDAQ was built for this moment, powering more than 135 marketplaces and regulators globally and connecting capital to companies shaping the future. As the innovation economy accelerates, connectivity becomes the critical asset. NASDAQ is the leading technology platform that makes it possible and scalable. Learn more at nasdaq.com. I'm going all in. Robin Rombach is the co-founder and CEO of Black Forest Labs. You are based in Germany in Black Forest, which is a city in Germany. It's a mountain range, actually. A mountain range. Yes. Where you grew up. Where I grew up, yes. And you are working on open source image and video models. You worked at Stable Diffusion for a little bit. That's correct. Cut your teeth on that. And you're known for the open source model Flux and maybe also for some closed source models. Tell us about the business of Black Forest Labs. What is the business and what is the goal? 100%. One quick addition. We are based in the Black Forest. It's a town called Freiburg and in San Francisco. Oh, and in San Francisco, of course. You're splitting your time? I'm splitting my time to a certain degree. we started a company two years ago me and my co-forners as you said we've worked on stable diffusion in the past before that we invented an algorithm called latent diffusion which is basically the fundamental algorithm behind all of generative models that are being deployed for image generation, video generation even physical AI now it basically makes use of this principle that you can compress natural data such as images such as video, such as audio, into a much more efficient representation and then train a transformer model on that. And I mean, this is the stuff where JPEG, MP3, and all of that works. And we basically translated that into a neural algorithm a few years ago when we were still PhD students in Munich, actually. And then on top of that, we built stable diffusion And then on top of that, yeah, the generative models that we are developing today. And of course, like the technology has advanced, but we are now tackling, I would say, models that are really made for understanding like the whole world around us. Multimodal visual models, pre-trained on images, videos, audio data at the same time. and we are now like entering a new paradigm which is combining that with something that's called action prediction such that you can actually use the same model to make images to make videos to make audio and to predict actions which means you can ultimately deploy it on a robot in the real world wow so from the image to the video the audio and then eventually the real world with robotics and a real world model because if you can make the image you and you can train the model that means by default you understand the world in order to make a video of the world you have to understand the world yeah and the objects i think that's yeah i think that's like a really good like way to think about it uh it's like it's like an intuitive way uh to interact with the world, right? I would say there's these complementary forms of intelligence, ultimately. There's intuitive intelligence, and then there's a deep reasoning layer. Now, ultimately, you need for a complete form, you need both, and you need them to interact. And I think we've been approaching it more from the intuitive side. Images is a very natural way to approach this whole field because it's not as computationally intensive as, let's say, video, right? But now, yeah, I think, like, we're combining it. It's converging into, like, a multimodal model. And, yeah, we see, like, exactly, like, pre-training on videos gives, like, implicit understanding of the physics of interactions with the real world. And then you can get stuff like action prediction, like robotics out of the same model. and with these models and the training they're kind of um been a limitation in creating videos and creating images where the criticism of generative ai is it's a bit of a slot machine i give a prompt it gives me something back but how did it come up with that the training data But, you know, maybe I want a different style. Maybe I want a different color. Maybe I want a different, you know, aesthetic. How does that problem get solved? And do you actually understand what's happening when the image is being made under the hood? Yeah. Yeah, I think ultimately it's about exposing as many manipulation layers as possible to, I don't know, a user or developer that builds on top of this model, right? And I think we've seen that in the past. Like in the past, image models, they basically started from simple text to image systems. Right. Then they've expanded into text plus image to image systems, which means you could suddenly take an image, like a real image or a generated image, and iterate on that based on a text prompt, like edit it, modify it, right? And then this expanded into taking multiple images and a text prompt and combining them in a semantic way and producing new content. And the same principle now applies to video. And I think now it becomes actually even more interesting when all of these modalities are actually combined inputs and outputs of the same model. So let's talk about video. There's an announcement that you're working with the greatest director of all time, or living director Martin Scorsese. We'll talk about that in a second, yeah? Fantastic. But in a movie, this promise of being able to make a movie in which the camera angle, the sound could be something that a Martin Scorsese would be proud to release to his fans. How close are we? And maybe tell us a little bit about this partnership. The technology being able to make an actual movie like Goodfellas or a scene from Goodfellas versus where it is today where you can make interesting five or ten second clips and then maybe people struggle making 10 of them and then they use some post-editing software to put them together. But you immediately understand this is not that. It's not a movie. It's AI slop. It's kludgy. It doesn't pass the uncanny valley. Well, I think it's important. That's at least like the view that we have is that these AI models, they are a medium, right? we don't want to set any way of how they are supposed to be used. We don want to tell anyone especially not someone like Martin Scorsese how is he supposed to use his model He is like greatest filmmakers uh ever uh it was insane sitting in the same room with him multiple times and actually him seeing like exploring our models like as like one of the like core researchers behind it was like just an insane feeling right and at the same time i'm also like a big fan um so you sat in a room with martin scorsese and showed him your tools exactly yeah and what was his reaction What did he key off of? What was the thing that he found most inspiring or interesting? I think it was really this idea of like, yes, clearly a vision in his head of like a scene or a scenery where like maybe a new movie will be shot. And he's trying to explore that and kind of like we basically looked at the scenery of like a village in Eastern Europe somewhere. and he was describing it. We saw some outputs. We iterated on the outputs. And ultimately, I think, and that's what he said in the end, is getting the mental picture of something out of your head and communicating it in a visual way by making these images or the series of images is something that just makes it easier to communicate and convey an idea of what is actually in your head. And I think that's one of the very interesting and powerful ways to use this technology. And I think ultimately... Is to get the inspiration, to get the vision out of his head onto an image. Yeah, I mean, like, language ultimately is like a little bit of like a lossy communication medium, right? Yeah. It's also interpreted in different ways, but then visual information is so rich, so rich, like an image or video, there's so much signal in it, and it's just like another way of communicating. And I think that's like one of the beautiful things that this technology ultimately enables. And I think to your question of making full movies with, I don't know, a video generation model, for example, I'm not sure if that is the ultimate goal. Maybe it's interesting to plug this into some kind of agentic workflow and make a very long video. And I think that's really cool to explore. But I think ultimately, the real interesting use cases, they come when you have a human in the loop who iterates and uses it as a medium. And I think this is at least like a perspective that I take that makes it interesting. And this is most often when the most interesting outputs arrive or are actually being made. The brainstorming production level is so obviously a huge win. You can paralyze your brainstorming, basically. Yeah. And yeah, I like that. Paralyze your brainstorming. And they have an analogy for this. They do storyboards. And some of the great directors, Ridley Scott of Aliens and Gladiator, was known for making his own. I also believe Spielberg was also like to sketch Raiders of the Lost Ark and some of these. George Lucas was known for collaborating with many amazing artists, even making miniatures and making storyboards for the Star Wars franchise. He had those people on full time helping him with that. So that's the obvious place to start. But if we look at startups, startups always want to try to figure out how to do something cheaply. And people used to make a launch video for their startup for $100,000, $250,000. So they take their $10 million venture raise and spend $250,000 on a launch video. I've seen with a lot of the startups I'm investing in now, they'll just spend a week or two working with a director to make a launch video. You've probably seen this trend, yeah? And I'm sure people use Flux in some of your models for this. Have you seen this? Yeah, of course, yeah. Yeah, and what's your take on that? Because that feels like the early stage of storytelling. You're trying to communicate a product or service in a fun, engaging, punchy, 30-second, 90-second way, yeah? I mean, again, I think we support this exploration based on these tools, right? I think ultimately it's great to see all different kind of launch videos, products being built on top of the same base model or the same technology. And I think that's what's making it so interesting and also so powerful. Yeah. And what else are people using the technology for? I understand there's a Bitcoin movie coming out. Instead of using a green screen in this Bitcoin movie, I was talking to Gal Gadot, the actress who played Wonder Woman. I was talking to her at an event and she was telling me it was the Breakthrough Prize, Yuri Milner's event. And she was telling me she just did a Bitcoin movie and they did it on a soundstage without green screens. But all the actors just worked in a soundstage. and then all of the scenery behind them was being done by generative ai that's a real movie that's a 30 million dollar budget movie she said it would have cost 150 million if they had to build sets and the film would have never been greenlit are you starting to see people use that in production not just in the back end and the ideation phase but actually in production yet with your tools yeah um yeah we see some use cases like that in production i think like high-end film production is kind of like the one of the like most demanding use cases yes and i think i'm glad that it's being explored but i also yeah really want to um like it's i think it's important to see that this technology is like on a trajectory and it's improving it's improving rapidly i don't know if i look back at like where we started like a few years ago when i was doing my phd and this feels like the only thing that you could do was like images of 64 by 64 pixels how you can do multi-minute videos at a high resolution. But it's not going to stop there. It's going to continue to improve, and I think then it's going to unlock even more of these high-end use cases. But I think the main thing... How soon before we get to that? Hard to predict, I think. Hard to predict, and I think ultimately... A couple of years. Ultimately, I think you still want to have the tool that enables this human-in-the-loop kind of production workflow, right? But I think when I look at multimodal generative models as a whole, I think what really excites me is you can use the same kind of AI model to make a movie and deploy that as a brain on a robot. And I think this is so interesting. And I don't know, there's some thoughts around trying that in the digital world, which would be, for example, computer use. Remains to be seen if that is actually something that works or not. But I think the technology is so powerful and so versatile, and it's just moving into that. All the talk on world models, world action models, all of that, it's basically all the same. And I think that's what's making it so interesting and what I find most upsetting. So do you believe that the technology will be used to analyze or primarily to analyze real world? Like here's a video of somebody, you know, making a sandwich. Now we have the robot study it and make the sandwich. Or do you think there'll be a lot of synthetic data made that then the robots will just study the synthetic or they're going to just in some way innately know based on all this massive amounts of training data? I think it's a combination of prediction, right? and prediction is a way of, you can think about it as simulation, as generation. It's predicting actions, which is you have to understand the input, the visual inputs, in order to actually predict a reasonable next action. And it's about perception. It's like you can only do that if you understand, if you perceive the content. Then you can only, I don't know, like transform it into a new piece of content or predict an action or describe what you actually see And the combination of all of that is, I think, what's driving it. There's not a single one of them. It's a combination of these thoughts. And what's the best way to get that training data? Do you need to have people put on glasses, get a first-person perspective, have them put on gloves so you have that fidelity of understanding, hey, this glass is moving. I'm pouring this glass. I'm putting ice into it. you know and here's how that works and the splashing and the condensation water so i can pick it up and not drop it because it's wet on the outside or is it going to be just hey take the corpus of youtube videos and the robots know exactly what to do because they'll find a thousand videos of people pouring drinks i mean ultimately i think you would want to go to a place where you could like prompt a robot in context right as you can do with like a language model basically just tell it, hey, go and, I don't know, pick up this glass with the, I don't know, orange juice or whatever it is. Yeah, exactly. We're not there yet. But I think this is, like, one of the goals. And I think, like, how these models are deployed currently is there's, like, a lot of, like, different hardware, different robots that are running in factories that all have, like, some different kind of action representation that you need to kind of tune the models towards, right? So in practice, what you do is you have all this visual understanding in the models, and then you need only a very little bit of a few hours of fine-tuning data to adjust the model on that specific task. I think the goal would be to kind of move away from that towards as much in context as possible, but it is a little bit of a research problem. I think that open source has kind of having a moment right now. We've been discussing it on the podcast a whole bunch recently. And people are also talking about sovereignty. You have companies that own incredible IP libraries. I mentioned Star Wars before. Disney owns an incredible library. What should your advice? What would your advice be to a company like Disney? Should they take your open source software, train their own models or work with you to train their own models to control it and then hey this is our ip they've already made a point of working with chat gpt and saying hey you you can and cannot use certain characters in fact open air had a relationship with them that's for sora that's no longer happening but they officially licensed on the output some characters so how do you think about those major ip holders what's your advice to them are you in discussions with them we know about the martin spursese or tor deal but How do you think about content libraries? I think it is, look, I think, like, the most interesting use cases of this, like, if you think about, like, content creation is in generating something, making something that hasn't been there before, right? Like, that's a fundamental, like, interesting aspect of this technology. And then I think, like, yeah, when it comes to IP, what we implement, for example, on, like, our public-facing tools is you cannot generate certain IP with these models, right? And I think that's something that is a sensible approach. And then, yes, we do work with certain IP holders to develop models together with them, some of them based on our open source models, some of them based on our more powerful proprietary models. But I think that is a very attractive value proposition. What's the vision there? What do you think that will look like for consumers in another couple of years? What would potentially happen when you open up Disney Plus? I mean, that's a good question. I'm not in Disney, right? So it's up to them to decide that. But I think we want to enable them to build all kinds of stuff that they envision. And I think we can support them. We can support other companies in that space to, I don't know, integrate the technology in the best possible way. I think one of the very interesting angles of it is that it's becoming much faster. It's becoming more interactive. I can envision a whole bunch of very interesting interactive content creation tools that you could host on Disney Plus or elsewhere. I think the most interesting thing I've seen in this regard is fan films. Right. So there's a category before generative AI, fan fiction. People would write their own Star Wars story. Then there came fan films where people would dress up as Jedi Knights and record their own films. And George Lucas said, as long as you're not doing it commercially, you're not selling it, I give you permission to go make Jedi movies. And they even released how-to's on how to make a lightsaber or sound files of how to make the lightsaber sound. Now, people are taking the stories that haven't been told from the Star Wars universe, and they're recreating them using AI. And for the fans, they're becoming quite popular on YouTube. Star Wars Stories Untold is, I think, the biggest one. It's getting millions of views per video already. And I think that's really the future is letting the customer base pay a licensing fee or pay a fee, maybe rent software or maybe based on the output and let them be creative with the characters. Let them make their own stories. And you could be in a unique position to empower that. Well, 100%. I think if you find a model that works for the IP owners but then also can enable the super creative customization you schedule, I think that's great. I mean, for myself, when I read a book or whatever, watched a movie, I had so many ideas how it could be done differently or this could have happened. This is so nice that you can actually enable people to visualize these ideas. Yeah, it's going to be incredible. Continued success with it. You have an office in San Francisco. you're hiring people, yeah? We do, yeah. You've raised a bunch of money. We raised a bunch of money. We just crossed 100 people. We're hiring in Germany and in San Francisco. Fantastic. Who are you looking for? What's the right type of person, the right type of skill? Yeah. On the one hand, we are always looking for researchers who have experience in large-scale model training, experience in diffusion model training, flow matching training. We're looking for engineers who want to be working with the customers to develop these customized physical AI solutions or, for example, with an IP owner, develop these models jointly with them. We are looking for engineers who have experience in just large-scale compute infra, managing that and making sure that the training runs smoothly, that we maximize our MFU and all that. and we are looking for people who have interest in, you know, like getting the technology out there in the hands of people. The forward deployment of this, there's just so many great ideas and so many great partners for you. I think you're going to, with the open source specifically, you know, it seems like the corporates really want to have some additional level of control, but they also need the frontier models or your proprietary ones for some of those refined features. So I think you have a very bright future around it. 100%, exactly. All right, continued success. Thank you so much for doing the show. Thank you so much. A pleasure. Thank you so much. Thank you so much. I'm going all in. I'm going all in.