Podcasts about Alessio

  • 977PODCASTS
  • 2,759EPISODES
  • 43mAVG DURATION
  • 5WEEKLY NEW EPISODES
  • Jul 20, 2026LATEST

POPULARITY

20192020202120222023202420252026

Categories



Best podcasts about Alessio

Show all podcasts related to alessio

Latest podcast episodes about Alessio

Programas FM Milenium
Pablo y a la Bolsa: entrevista a Eduardo D´Alessio, presidente de D´Alessio IROL

Programas FM Milenium

Play Episode Listen Later Jul 20, 2026 13:13


Entrevista de Pablo Wende a Eduardo D´Alessio, presidente de D´Alessio IROL, a propósito del último informe de humor político y social.

ICF Singen/Villingen Audio
Hausrecht - eine offene Rechnung mit dem Feind | Alessio Passarella

ICF Singen/Villingen Audio

Play Episode Listen Later Jul 9, 2026 48:38


Fühlst du dich manchmal gefangen in wiederkehrenden Mustern, Ängsten oder Belastungen? In dieser Predigt erfährst du, wie die Bibel über geistliche Kämpfe, offene Türen und Gottes Schutz spricht und warum Jesus der einzige ist, der vollständige Freiheit schenkt. Lass dich ermutigen, alte Lasten loszulassen, Gottes Wahrheit anzunehmen und in der Autorität Jesu ein neues Kapitel zu beginnen.

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0
Why AI Infrastructure must evolve for Agent Experience — Akshat Bubna, Modal CTO

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

Play Episode Listen Later Jul 8, 2026 57:55


We've been running a bit of an Agent Cloud series surveying all the top inference/compute/cloud providers, from Databricks to Daytona to Railway and, even further back, E2B, but we're excited to conclude this series returning to Modal, which has just raised a monster $355M Series C.The cloud was built for developers. But agents are now changing that.The old infra stack was designed for a human who could read docs, reason through YAML, and understand dashboards to figure out what they need when something broke. While this was painful for developers, it worked since they could fill in missing context in their heads.However, agents don't have that luxury. Now in this new era of agents, everything has to be tighter.They need a place to write code, run it, inspect the output, change the environment, debug failures, and try again. Fast iteration and feedback loops with all the necessary context are crucial for agents to operate properly. Furthermore, sandboxes are a clear representation of this shift as agents can easily spin up isolated environments. This programmatic infra even extends to research:Two years ago, we were one of the first to cover Modal with CEO Erik Bernhardsson and Alessio designed our favorite LS thumbnail of all time:At the time, Modal was just a teeny little company with a $17M Series A.Today, fresh off their $355M Series C, Modal is one of the clearest examples of the agent cloud future being built in real time: a cloud platform moving past traditional web app assumptions toward the workloads AI actually creates such as elastic inference, sandboxes, GPU burst, post-training, background agents, and infrastructure that agents themselves can operate.In this episode, Modal CTO Akshat Bubna joins swyx and Vibhu to unpack why AI applications don't fit traditional cloud assumptions, why Kubernetes was never designed for bursty compute-heavy workloads, and why Modal is now shifting from developer experience to agent experience.We go deep on Modal's AI infra stack: serverless functions, decorator-based infrastructure, elastic inference for custom models, GPU snapshotting, DeFlash, speculative decoding, Auto Endpoints, sandboxes, persistent storage, networked containers, private IPv6, RDMA, multi-node training, and Modal's capacity pool across 17 cloud providers. Akshat also explains why RL rollouts can require 100,000 sandboxes, why production agents need hard guardrails, why observability may matter more than reading code, and why AI has made infrastructure exciting again.We discuss:* Why Kubernetes wasn't built for bursty AI workloads* How Modal started as a better runtime before becoming an AI cloud* Why Modal added GPUs before ChatGPT* The shift from developer experience to agent experience* Why observability matters when agents are writing the code* Elastic inference for custom models across audio, video, robotics, and comp bio* GPU snapshotting, cold starts, and why inference workloads are so bursty* Why RL rollouts can require 100,000 sandboxes* DeFlash, speculative decoding, and frontier-level inference performance* Auto Endpoints and making optimized inference easier to deploy* What Modal adds beyond vLLM, SGLang, and raw GPU rental* Modal's 17-cloud capacity pool and supercloud strategy* Networked sandboxes, sidecars, private IPv6, and RDMA* Serverless multi-node training for post-training and research workloads* Auto-research, model-guided sweeps, and agents launching GPU experiments* Compute strategy, capacity planning, and batch tiers* Why production agents need specialized sandboxes and hard guardrails* Modal's take on managed agents, CI, Gitpod/Ona, Python, TypeScript, and Modal BenchAkshat Bubna* LinkedIn: https://www.linkedin.com/in/akshat-bubna-188885103* X: https://x.com/akshat_bModal* Website: https://modal.comTimestamps00:00:00 Introduction00:00:39 Modal's origin and why Kubernetes wasn't enough00:04:32 Developer Experience → Agent Experience00:06:21 Modal's AI cloud primitives00:09:14 Sandboxes, agent loops, and proto-Cognition00:12:12 Elastic inference, GPU snapshotting, and 100,000 sandboxes00:15:24 DeFlash, speculative decoding, and Auto Endpoints00:19:59 Production-grade inference beyond raw GPUs00:22:00 Background agents, Ramp Inspect, and the agent lifecycle00:24:08 Modal's 17-cloud supercloud strategy00:26:40 Networked sandboxes, private IPv6, and RDMA00:32:48 Multi-node training, post-training, and auto research00:37:36 Compute strategy, capacity planning, and batch tiers00:40:55 Open models, real-time AI, and production agent infra00:43:06 Hard guardrails, managed agents, and specialized sandboxes00:46:06 Why AI made infrastructure exciting again00:48:30 Model APIs, differentiated products, and agentic video00:51:50 CI, coding-agent infra, SDKs, and Modal Bench00:57:28 Closing ThoughtsTranscriptIntroduction: Modal, Series C, and the Art PartySwyx [00:00:00]: We're here with Akshat, CTO of Modal, together with Vibhu. Congrats on your Series C.Akshat [00:00:10]: Thank you.Swyx [00:00:11]: Your party yesterday was amazing.Akshat [00:00:15]: Yeah.Swyx [00:00:15]: From all the photos and all the swag.Akshat [00:00:17]: We had a bunch of art installations, which was fun, seeing, like, our products on pedestals next to, like, Rodin.Swyx [00:00:25]: Very nice. Very nice. When you started, it was not the GPU inference company. Maybe it was in your mind. Take us back to the origin story.Modal's Origin: A New Runtime Beyond KubernetesAkshat [00:00:39]: I first met Eric, who's the CEO, through an investor. Back then Eric was already thinking about building, a new runtime, and he got there thinking through why are workflow orchestration products so hard to use. It's because you have to run them on Kubernetes. Kubernetes is hard to manage. It's not built for burstiness and, custom images,Swyx [00:01:03]: YeahAkshat [00:01:03]: It has a terrible developer experience.Swyx [00:01:05]: And I'll, I'll interjectAkshat [00:01:06]: YeahSwyx [00:01:07]: For listeners, who are new, we interviewed Eric two years ago, and there's a bit more of the story there from Spotify and all those things.Swyx [00:01:14]: And I came across Eric through Data Council because he did that talk on the serverless container stack that you guys did, which was like, that was my first like, “Okay, I need to take Modal very seriously” moment.Akshat [00:01:26]: Yeah.Swyx [00:01:26]: But it was still very unclear, like, do I need all this for just my data pipelines?Akshat [00:01:33]: Yeah. initially what we were thinking about was if we build a better runtime, it's a very useful primitive in itself. It's There's a lot of things that, get solved by serverless functions, like you can do, ETL stuff, you can do job queues, you can do all this, like, bursty processing, which it turns out every company had needs for. but then we also were thinking about this as like, this is a primitive that we can build a whole collection of products on, which are very verticalized. So perhaps data engineering would've been the first one, but we were thinking about inference. Back then it was more classical inference, like computer vision stuff and running XGBoosts and whatnot. But we added GPUs to the product a year before ChatGPT came out.From Serverless Containers to GPU WorkloadsSwyx [00:02:19]: Nice.Akshat [00:02:19]: We just didn't think it would be that big of a deal.Swyx [00:02:22]: Yeah, just like add A100.Vibhu [00:02:23]: Was there any, like, early key problem that really sparked off why you built it?Akshat [00:02:28]: Yeah. Primarily it's just, none of the tooling that was out there was built for, one, a really great developer experience, and also there's a general trend of, a lot of the workloads that we were seeing were very. I wish there was a better word for it, but compute-heavy. Like, they need, one, like, need a lot more resources, so you need to burst up and down a lot, versus like Kubernetes designed for, like, slow scaling and, more for, like, web server use cases. And also there's just a lot more specialization in, like, what kinds of environments these workloads run in. Like, we had sometimes they need accelerators, sometimes they need different kinds of images, and this is just like a consistent thing that we saw across a lot of companies. That would be the next step.Software-Defined Infrastructure and Decorator-Based DXSwyx [00:03:13]: Yeah. Yeah. Be nice. I don't know how much this factored into the early story, but I wrote a post when I was at Temporal about infrastructure, software-defined infrastructure or something like that.Akshat [00:03:22]: Yeah, the self-provisioningSwyx [00:03:23]: Self-provisioning.Akshat [00:03:24]: Yeah.Swyx [00:03:24]: Yeah. I can't even remember my own post.Swyx [00:03:26]: And then you put me on the landing page.Akshat [00:03:28]: Yeah. We really like, the term and so we stole it.Swyx [00:03:32]: Because you had the insight that everything can just be in decorators co-located with the code, right?Akshat [00:03:37]: Yeah.Swyx [00:03:37]: Was that a big part of the originalAkshat [00:03:39]: YesSwyx [00:03:39]: Story or it was just like a DX layer?Akshat [00:03:41]: That was, really important because we really didn't want people to spend, so much time, writing YAML, and it seemed like you could really condense the surface area of what you're doing, put it in code so you can operate on it just like you operate on other code, and like build stuff that's more expressive and dynamic. and so yeah, that was always a very important part.Swyx [00:04:04]: Then the pushback is this is a DSL.Akshat [00:04:07]: Yeah.Swyx [00:04:07]: It's you're closed source. I am locked into Modal.Akshat [00:04:11]: Yeah. We never really got pushback for that because the nice thing about Modal is you can bring whatever code you have, and sure, the DSL is at the configuration layer for, what hardware you're using, how you're scaling things up, but you still own the code.Akshat [00:04:27]: And that's, that's been an important, part of our story, even as we do inference now.Swyx [00:04:32]: Yeah.Vibhu [00:04:32]: How much of do you think still stays the same today? Like if you were to build something today, DevX very important, but I feel like, a lot of this has been changed with just hook it up to an agent, have Claude Code, have Codex implement a tool. there's very agent native primitives that are different than if I'm doing this myself, right?Developer Experience → Agent ExperienceAkshat [00:04:54]: We've changed our SDK team to think about agent experience instead of, developer experience and we think that the same benefits that apply for DX also apply for AX, which is why would you have an agent read through hundreds of Kubernetes files and like write YAML that's not even typed when it can make a couple of changes in a decorator and it gets this self-provisioning runtime of, being able to see its changes live in action? yeah, it just seems from the customers we talk to, they find Modal is much faster for agents to use versus operating on a different substrate.Swyx [00:05:34]: Yeah, because like you, again, you co-locate the infrastructure requirements to the code that runs it.Akshat [00:05:38]: Yeah.Swyx [00:05:38]: Well, the negative thesis now is that nobody's looking at their code anymore, so there's no point.Akshat [00:05:44]: Yeah, people aren't looking at code. one thing we still see is really important is observability.Swyx [00:05:51]: Yeah.Akshat [00:05:51]: Like how good is your dashboard? And of course, like we have, we push a lot of it to the CLI so the agents can do their own investigation, but you still need humans to go interpret what's going on and, make judgment calls and whatnot. and that's I feel like, Maybe more important now than looking at the code itself.Swyx [00:06:11]: Yes, because like, you can try to treat the code as a black box and then use, see the observable action that comes out of it, and then just prompt a change.What Modal Is For: AI Cloud PrimitivesAkshat [00:06:21]: Yeah.Swyx [00:06:22]: So I think it takes a bit of restraint to not specialize, to say, “I want to ship a new primitive,” and then just be general purpose.Swyx [00:06:31]: People ask you, “What are you for?” You're like, “ I don't know. We can do this, we can do that.”Vibhu [00:06:36]: Well, I'd be curious to see, like, okay, if we were to ask you, like, what is Modal for even at a high level? There's a lot you guys do, sandboxes, GPUs, everything. How do you answer?Akshat [00:06:46]: Modal is a cloud platform that's built for, where we've built the primitives from scratch for AI applications. and right now it covers, inference, training, batch processing, and sandbox workloads.Akshat [00:07:00]: But we're building a lot moreSwyx [00:07:02]: I noticed you didn't say web server, so there is still a role for, like, the always-on large-scale Kubernetes type things.Akshat [00:07:09]: Yeah, absolutely. We're, we're not trying to compete with the renders of the world, because yeah, we think the differentiator for us is the, are the workloads that need specialized compute, need to scale up and down a lot. yeah, they're, they're, they're just shaped differently.Working Alongside Frontier StartupsVibhu [00:07:26]: I think you're building a lot of it alongside the startups, right? They're innovating quite a bit, even in your, like, latest blog post. Like, even in the series C, the customers that you mention here, the cognitions, technical ones, ramps and whatnot, they're, they're innovating with you, right? And that's not something AWS is doing directly with.Akshat [00:07:45]: Yeah, absolutely. I think, this is again classic. We're a small team. We can move really fast. our engineers are working with our customers and figuring it out. Yeah.Swyx [00:07:54]: So my first week at Cognition, I walked in, there was someone wearing a Modal shirt. I was like, “What are you doing here?” They're like, “Yeah, I just. I am embedded inside of Cog.”Akshat [00:08:05]: Yeah, I think that was Peyton. We sent him overSwyx [00:08:07]: Yeah.Akshat [00:08:07]: Because, the latency of communication was too high otherwise.Swyx [00:08:12]: Yeah, distributed node, you have to - you have to place one and collocate.Vibhu [00:08:16]: Yeah.Swyx [00:08:16]: So I had a, I had direct personal experience, right? So I worked on smol developer three years ago. it was inspired by Claude 1. I think you onboarded me at some point, like, just before, and I was like, “Oh, like, I need some bursty compute. Like, I was just gonna try using Modal.” And it was a, it was a pretty pleasant experience. apparently, I showed up in the board meeting, like the analytics.smol developer, Sandboxes, and Proto-CognitionAkshat [00:08:39]: Yeah, you blew up on Hacker News and,Swyx [00:08:41]: YeahAkshat [00:08:41]: We got a big traffic spike. I. I think the way you used smol developer was Modal functions for running stuff, which was. Like, the, that was a good use case. but then, yeah.Swyx [00:08:53]: Yeah. That - So to me, that was proto-cognition.Akshat [00:08:55]: Right.Swyx [00:08:56]: If only I had, like, stuck to it.Swyx [00:08:58]: Like, that was like, if - did you say draw the tech treeAkshat [00:09:00]: AbsolutelySwyx [00:09:00]: You're just like, “Yeah, like, probably this will happen.”Akshat [00:09:02]: Yeah. Like, he was so close. You were just rebuilding upon usSwyx [00:09:04]: I just didn't realize.Akshat [00:09:05]: But the funny story there is at the same time, we were talking to a bunch of customers who needed something like sandboxing.Swyx [00:09:14]: Yeah.Akshat [00:09:14]: This is like twenty-three.Swyx [00:09:15]: Yeah.Akshat [00:09:16]: So we builtSwyx [00:09:17]: You introduced a new API right after that.Akshat [00:09:18]: Yeah.Swyx [00:09:19]: Yes.Akshat [00:09:19]: Like, we built sandboxes in May of twenty-three before anyone was even knew this was gonna be a thing. And the first example we published was, we took smol developerSwyx [00:09:28]: Smol developerAkshat [00:09:28]: And put it in a loop, so the agent can iterate on itself.Swyx [00:09:33]: Loops are hot these days.Vibhu [00:09:34]: It's the looper.Akshat [00:09:34]: Yeah.Vibhu [00:09:35]: Loops in. When was this, twenty-three?Akshat [00:09:38]: Yeah.Vibhu [00:09:39]: A small check.Akshat [00:09:39]: Yeah.Swyx [00:09:39]: It's like twenty-three. so the. the, those for listeners, like, the problem was the models are not built for any of this, right?Swyx [00:09:46]: Like, you're just trying to like. They're not post-training to understand, like, looping and, like, self-correction and tool calling was there, but, like, also not that great.Akshat [00:09:55]: Yeah.Akshat [00:09:55]: I don't remember if you used tool calling in this one, but yeah, the models would just diverge after like ten iterations and not produce anything meaningful.Swyx [00:10:03]: Yeah. But like, then. So okay, like now talking to myself three years ago, the answerVibhu [00:10:08]: Of course they will get betterSwyx [00:10:09]: Collect all the failures, build benchmark, and then collect all the, examples, build the RL environmentAkshat [00:10:15]: RightSwyx [00:10:15]: Sell it for like ten billion dollars to Meta.Swyx [00:10:17]: And then also train a model and then sell that for sixty billion dollars to Elon. And this isAkshat [00:10:23]: Yeah, of courseSwyx [00:10:23]: The funny machine. Like, it's like, it's about the hardware.Akshat [00:10:28]: It's hard to have that inherent conviction that the stuff will get that much better.Swyx [00:10:33]: In retrospect, it's so f*****g obvious.Akshat [00:10:36]: Fair enough.Swyx [00:10:37]: Like, what else were we doing back then? I don't know. anyway. Yeah. So this. That was the start of your sandboxing journey, right? I feel like it didn't blow up until, like, last year.Akshat [00:10:49]: Yeah.Swyx [00:10:50]: So there was like a couple years of quietness.Akshat [00:10:52]: Exactly, yeah. We wereVibhu [00:10:53]: I think very underrated product value. Like, my experience with Modal, Charles, before he had joined Modal, met this guy at a hackathon, and he really insisted we wanted to run some small model, not hosted anywhere, and he's like, “ there's this cool company, Modal. They'll like spin up a GPU sandbox, we can throw it on there. They'll take a Hugging Face link.” And like there's so much value just right there, right? Like instant hosting, spin it up, spin it down. It'll stay cold, but we run the demo a few days later, it'll come back up and like all this stuff in retrospect, like it's still what we needed like today.Akshat [00:11:27]: Yeah, it's still needed today. workload shapes have changed a lot as, we run stuff for people with really massive production scale and, there it's it's not about scaling from zero to one, but it's how do we scale really elastically, from like thousand to fifteen hundred GPUs very quickly in a given region. It's the same shape problem.Elastic Inference, GPU Autoscaling, and Custom ModelsVibhu [00:11:50]: Okay. So you look at, say, Cursor Composer, right?Akshat [00:11:53]: Yeah.Vibhu [00:11:53]: They had a. “We'll do RL on a model every couple hours.” you guys have a whole version of RL inference gym and whatnot.Vibhu [00:12:01]: When you look at workloads like that, you're doing train runs where you need to scale up, scale down every hour thousands of GPUs, right? That's the example for we do need it, right?Akshat [00:12:12]: Yeah. Well, so I'll, I'll take a step back and, maybe talk about like how people use Modal today. because our biggest use case is, elastic inference. And the thing we first found product market fit, with was inference for custom models. So we stayed away from the LLM space, and we were serving companies like Suno for audio, Runway for video, robotics, comp bio companies that train their own model elsewhere. But Modal is the best black box that for deployment, scaling to however many GPUs you need as your traffic pattern changes. And we saw all of them like have a very unpredict- predict- predictable, traffic pattern. it's like diurnal. It's Some days, like the company will do a launch and, they'll need like, way more. And it's not just one model that they deploy. They-- all these companies deploy, lots of different models in different regions, and so the autoscaling problem becomes even harder because then you have to scale within a certain region, and those cycles are offset. So different times you scale up in different regions.Akshat [00:13:20]: So that's like our sortVibhu [00:13:22]: And thatAkshat [00:13:22]: YeahVibhu [00:13:22]: That in and of itself is a huge category. There's a bunch of inference providers which, provide this fireworks, does this as a service together, whatnot, Base10. that's carved into its own niche for language models, at least right now.Akshat [00:13:36]: Yeah. the thing that we have specialized in is the autoscaling aspect.Vibhu [00:13:41]: Yeah.Akshat [00:13:41]: Because we found that it's not universally true that everyone else can autoscale, and we've gone deeper into it on the tech side by, we've incorporated GPU snapshotting into the product so we can take the GPU state, like your torch.compile model, snapshot it, and the next cold start is way faster. And so going back to your question, it's That's why you need a lot of burstiness for inference. But then people also do a lot of demand training, like for RL stuff, your rollouts are bursty, as you said. People also do a lot of batch jobs. So we'll see, a lot of companies, before they have a training run, they'll need thousands of GPUs to run encoding or something like that. And I think those things are much more bursty than. I agree that agents are not that bursty. sandboxes are, except when you're doing RL. RL is justRL, Batch Jobs, and 100,000 SandboxesVibhu [00:14:28]: Or commerceAkshat [00:14:28]: Insanely bursty.Vibhu [00:14:29]: Yeah.Akshat [00:14:30]: Yeah. Like when you're doing, rollouts, you sometimes need a hundred thousand sandboxes in your sandboxes.Vibhu [00:14:37]: Yeah. I'm curious if you've seen early sparks of continual learning. There are some people, like our friends, ngram, recently announced thisAkshat [00:14:45]: YeahVibhu [00:14:45]: They're, they're trying to do training. That also seems like a different workload, right? If you're doing training twenty-four/seven per se, there's a very weird dynamic of how you're using GPUs between people and whatnot, but seems like something you guys would work for.Akshat [00:15:00]: As you said, we're, we're fortunate to work with a number of, customers at the frontier and grab some of our customers. and they are taking the primitives we have, and trying to use them in very interesting ways, like continual learning. It's possible as the stuff gets better, some of that will be part of, our offering as well if, more people need it. but we're, we're just waiting to seeVibhu [00:15:23]: YeahAkshat [00:15:23]: How it shakes out.Vibhu [00:15:24]: Is there a primitive that you added after sandboxing that was the next step in the story?LLM Inference, DeFlash, and Speculative DecodingAkshat [00:15:32]: I guess we've been going much deeper into LLM inferenceVibhu [00:15:35]: YeahAkshat [00:15:35]: Because we realized that some of the advantages we have with like autoscaling, again, especially in different regions and whatnot, are, not present elsewhere. and the place where we had a gap was we weren't, working on the model layer itself. Like we were a black box. And, we realized that, we can get to frontier-level model performance, with, by having great people who work on this. And, we've been open sourcing a lot of our work, in terms of, Recently, we, shared our work on DeFlash, which is a block-based, speculator, and we've open sourced, all of it. So, you can - By using open source DeFlash, you can get the same performance as you would with one of the proprietary providers. And the next thing we're thinking about hereVibhu [00:16:23]: I thought this wasAkshat [00:16:24]: YeahVibhu [00:16:24]: An interesting blog post as well, right? Like, I think in here you make a claim that. Not a claim, just that how effective speculative deco-decoding really just get to.Akshat [00:16:33]: Yeah.Vibhu [00:16:33]: Anything you wanna point out from this around, what people should know?Akshat [00:16:39]: Yeah, absolutely. the high-level summary is, it would help to describe what speculative decoding is.Vibhu [00:16:44]: Yes.Akshat [00:16:44]: I will, yes.Vibhu [00:16:45]: I think, likeAkshat [00:16:46]: YeahVibhu [00:16:46]: So we've covered like Eagle and all thisAkshat [00:16:47]: YeahVibhu [00:16:47]: Like Hydra and all those things, but it was like two years ago.Akshat [00:16:51]: Yeah.Vibhu [00:16:51]: I think it doesn't hurt, right?Akshat [00:16:52]: Yeah. Speculative decoding is you have a smaller model, called a draft model, predict tokens ahead of the bigger model, and then you have the bigger model, verify all of this, all the tokens are predicted. And the reason it's faster is if you're predicting, one token at once, you're bound by memory bandwidth. But if you can batch the verification of, the draft model, then you're much more efficient using compute, and it's faster, and as long as your draft model is producing a lot of tokens that can get accepted, which is called the accept length, you can get a speed up that's, multiple times of, the original model speed. and well, that's what we highlight here. It's Like people talk a lot about we made these kernels faster and whatnot, but improving kernel will only give you like few percentage points of improvement, and, increasing accept length, literally is a multiplicative decreaseVibhu [00:17:47]: Like two to four X.Akshat [00:17:48]: Yeah, exactly.Vibhu [00:17:48]: Without much head-on performance.Akshat [00:17:50]: Yeah. I think it may - you are running a second model, right? So it may be something more expensive in the compute,Vibhu [00:17:57]: I meant quality performanceAkshat [00:17:58]: Probably not by muchVibhu [00:17:58]: But yeah. I thinkAkshat [00:17:59]: So there's no drop in quality performanceVibhu [00:18:01]: YeahAkshat [00:18:01]: Because you're always. You're never accepting a token that the big modelVibhu [00:18:04]: It's strictly betterAkshat [00:18:05]: YeahVibhu [00:18:05]: Or it's same.Akshat [00:18:06]: Exactly.Vibhu [00:18:07]: Right. Yeah.Akshat [00:18:08]: And so we've been working a bunch on DeFlash, which is a block-based speculator. so it's instead of predicting, one token at a time, it's predicting a block. And we've been open sourcing our work with it. The next thing for us here is for helping people train speculators and custom models. it's it's something that traditionally is very forward-deployed engineering driven, support deployed, engineer driven, like you work with customers and help them do that. And our vision for. This is why we launched Auto Endpoints, is we want to make frontier-level performance available to everyone. And so, we mentioned this in the announcement, we teased it. The next thing we're, we're launching is, as you run an auto endpoint, we shadow trafficAuto Endpoints and Frontier-Level PerformanceVibhu [00:18:54]: Do you want to explain what auto endpoints are?Akshat [00:18:57]: Yeah.Vibhu [00:18:57]: I lovely, yeah.Akshat [00:18:58]: Yeah. So, this is, I guess, going back to your Modal is you touch the code, but, sometimes people don't wanna touch the code, and they wanna get started with an endpoint that works and has all the great performance and, scalability that Modal has. So we've made that easier with, a way to create an endpoint from our UI, from the CLI, that has all of our optimizations that we talked about, like the DeFlash stuff already baked in, and there's full transparency. So we give you the code, you can go run it yourself, and if you want, you can eject out into the full Modal experience, which we see as people get sophisticated, they do wanna tweak the models, they wanna, fine-tune stuff. You can still do all of that. It's it's not a black box. And yeah, the next thing, as we teased later in the post, is how do we give you value even beyond this in terms of having your draft models evolve as your data distribution evolves, again, without having to talk to a person and, yeah.Vibhu [00:19:59]: I guess just to understand it directly, you have the GPUs, you have an endpoint that's compatible, you serve open model. If someone was to do this themselves, what's the delta that you guys provide? So you do a lot of open source great work on effective inference. how does it compare to, say, I take the same model, 5.2 FP8, take shelf inference engine, vLLM, SGLang, get compute of similar capacity, similar cost. What's the delta that plugging into something this, like this offers outside of the benefit of, scaling?Production Inference Beyond Raw GPUsAkshat [00:20:34]: It's interesting because we've taken the approach of open sourcing our contributions and upstreaming them. we work closely with the SGLang team. We want the improvements that our team, comes up with to be, there in open source for others to use, even outside of Modal. The benefit to us is we have a team that has significant expertise in terms of if you do have something that is not there, our team can help you get that performance, first. the other thing is with these endpoints, we are way more elastic, as you said, than, anyone else, and you have true scaling to zero. you have true, burstiness, and in practice, that matters a lot more to people than just finding, the GPU and, running Modal code on something.Vibhu [00:21:20]: Yeah. And I will say it's not that straightforward to just. like what I said is easier said than done, right?Akshat [00:21:26]: Yeah.Vibhu [00:21:27]: It's I think still for the average person, still hard to just gut check using different. There's, there's quite a bit of combinations you can make there. the trade-offs aren't really known at face value.Akshat [00:21:40]: Yeah. it's it's not just that. I think it's it's that running production-grade inference is a hard infer problem.Vibhu [00:21:49]: YeahAkshat [00:21:49]: Even if you subtract out the autoscalingVibhu [00:21:50]: YeahAkshat [00:21:51]: Is controlling things like tail latency and, making sure every, request is delivered at least once and whatnot.The Model and Agent LifecycleVibhu [00:22:00]: There's a lot of innovation that you can do here. I think, it's very interesting that you're starting to encroach on, like as you become a full cloud, you're starting to encroach on other people's turf.Vibhu [00:22:09]: What will you not do?Akshat [00:22:13]: Well, we wanna follow our users and, make sure they get like a platform that has everything that works well together. so right now we're focused on the model lifecycle and the agent, lifecycle. so both like going from data prep to training to inference, and then also if I want to deploy a background agent, let's say, sandbox, do persistent storage, a whole bunch of other stuff.Vibhu [00:22:38]: We talked to Cole, who did, OpenInspect. Yeah.Akshat [00:22:42]: Yeah.Vibhu [00:22:42]: And RealInspect also is on Modal.Akshat [00:22:44]: Yeah. So Ramp Inspect was a great example of a background agent that was really successful because they, were able to use some of the primitives like snapshotting and fast scaling to just have something that feels really reactive and works well.Ramp Inspect and Background AgentsVibhu [00:23:02]: Yeah. That's the new CTO of, Ramp right there.Akshat [00:23:05]: Yeah, Rahul.Vibhu [00:23:08]: It was really fun. yeah, okay, I think, all very bullish. Like, one of my reflections was also I did not originally. So when I met you guysThe Inference Inflection: CPU, GPU, and Co-LocationVibhu [00:23:19]: You weren't that much in the GPU game, and now you're all about, inference. And one of the points that I hinged on for Jensen's keynote at GTC this year was, what we're calling like the inference inflection, right? That let's say in AI workloads or machine learning workloads, it used to be like, let's call it eight to one GPU to CPU, and now it's more like one to one, which is like a interesting. Like, - because of how much agents are blocked or call out to this, to CPU heavy stuff the actual, like, limiting factor, like, swings back and forth from GPU to CPU a lot more than it used to be all GPU and then occasional CPU.Akshat [00:24:01]: Yeah.Vibhu [00:24:02]: GPU, CPU. And now it's like just constantly, and you just have to locate everything.Seventeen Clouds and the Supercloud StrategyAkshat [00:24:08]: Yeah. And that's one of the things that, again, we see as, something appealing about Modal, which is we've built this capacity pool that spans, 17 cloud providers, so we're, we're very good at Running on various kinds of cloud capacity across the worldSwyx [00:24:24]: You don't have your own data centers?Akshat [00:24:25]: We don't have our own data centers. We just run across a lot of neo cloudsSwyx [00:24:29]: Yeah. AreAkshat [00:24:30]: Metal providers.Swyx [00:24:30]: Yeah. Question mark.Swyx [00:24:31]: Yeah. You're, you're running the math, and you're like, “What's the cutover point where you're like.”Akshat [00:24:36]: Yeah, it's a good question. part of it is we see our differentiator in the software layer, and, being capital light and focusing on the software helps us move really fast. so far it's worked out well because there are so many other people building data centers that we're able to work effectively with them, and again, focus on what makes us, special.Swyx [00:24:55]: Yeah.Swyx [00:24:56]: 17 gets you into, like, the local providers sometimes. LikeAkshat [00:25:00]: The,Swyx [00:25:01]: Which was the most interesting one?Akshat [00:25:02]: There are a lot more neo clouds than you expect, and they all have various degrees of, various levels of reliability. And, that's why it's something we've invested a lot of time in, is building our own reliability layer on top. so if the GPU falls off the bus or something happens, we user workloads are not affected, and that lets us use a lot more capacity than,Swyx [00:25:30]: YeahAkshat [00:25:30]: You as a user would be able to.Swyx [00:25:32]: It's a useful thing to have because like now everyone knows, like, what layer you are and, like, you optimize for being the super cloud of all clouds.Akshat [00:25:41]: Yeah. That's, that's, that's the idea. and so I guess when you mentioned colocation, that's, that's another interesting thing where, one thing we've seen is people come to us when they want, very specifically located, CPUs or GPUs, like they wantSwyx [00:25:57]: Oh, they pin it in likeAkshat [00:25:58]: YeahSwyx [00:25:58]: EU?Akshat [00:25:59]: Exactly. Or EU, US.Swyx [00:26:01]: Right. Data resiliencyAkshat [00:26:02]: AustraliaSwyx [00:26:02]: Locality thing or performance or what?Akshat [00:26:04]: It's either data locality or latency, yeah.Swyx [00:26:07]: Yeah.Akshat [00:26:07]: Like, you want your. They're running sandboxes and model. They want them to be right next to aSwyx [00:26:10]: Yeah, it's easy thenAkshat [00:26:11]: YeahSwyx [00:26:12]: To. That is important in all those things. and so, like, you've accidentally, I don't know if it's accident, but, like, you've built the perfect primitive for agents to express themselves. And then, like, it's almost very funny how every extra development just involves more file system, just involves more CPU.Akshat [00:26:30]: Yeah.Swyx [00:26:31]: Just like the things that you already have. I don't know much about, if there's any, like, networking usages that are interesting, but you've also done some good work on networking.Networking, Sidecars, Private IPv6, and SandboxesAkshat [00:26:40]: Yeah, that's exactly right. Like, we're just taking compute storage and networking and building stuff on that layer, for, again, the stuff people need.Swyx [00:26:49]: YeahAkshat [00:26:50]: We see a few interesting networking things coming up. one is people want networked sandboxes. so we haveSwyx [00:26:57]: For like a Docker cluster type thing.Akshat [00:26:59]: Yeah.Swyx [00:26:59]: Sorry, Docker Swarm. Oh, f**k. What is it called?Akshat [00:27:02]: Compose.Swyx [00:27:03]: Compose type thing.Akshat [00:27:04]: Yeah. So if you want Docker Compose, our sandboxes now support, this thing called sidecars. So you can. A sandbox is a pod of containers, and you can run multiple containers in, a sandbox. also useful because, going back to networking, people want a lot of control over, outbound networking from a sandbox.Swyx [00:27:23]: Yeah.Akshat [00:27:23]: Like, they might wanna run a middle proxy for, like, maybe logging stuff for RL or, controlling how egress can happen to a domain, injecting credentials. and yeah. So we've, we've had to build a lot of that stuff ourselves.Swyx [00:27:38]: Yeah.Akshat [00:27:39]: But then also sometimes people want, sandboxes spanning multiple nodes to talk to each other, which is an emerging thing we're seeing. We have support for that for a different reason, and yeah, we'll see if that becomes stable.Swyx [00:27:52]: Like, just an open socket. It's a. This is directly like mTLS.Akshat [00:27:56]: We do support that, which is you can, expose a tunnel inside a sandbox.Swyx [00:28:01]: Yeah.Akshat [00:28:01]: And then you can either expose it to public internet or it can be, you can add like a HTTP, auth layer above it. But we have this thing called I6PN, which we haven't talked about, which is this, like, overlay network using IPv6 addresses. so if Modal containers, within the same workspace, when this is enabled, can address each other using this private IPv6 address, and no one else can.Akshat [00:28:28]: So it's like private networking, for containers. We built it because we needed it as a primitive for our distributed training product. so we have this other feature, which is you can add a decorator to a function, and you get a cluster of GPUs. and they have RDMA networking. so you can run a distributed training job, that's truly serverless. and we did the overlay network for that. But then we've seen that people are using it for other reasons, and, I'm intrigued to yeah, what would people do with it.Swyx [00:28:59]: Build primitives and let people figure it out, right?Akshat [00:29:01]: Yeah, exactly.Swyx [00:29:02]: You put out a pretty interestingAkshat [00:29:03]: They're like, they read the docs webpage. Let me use thatSwyx [00:29:06]: YeahAkshat [00:29:06]: Something they never intended to work. This is literally not even in our docs page. People somehow found it, and they're using it.RDMA, Memory Movement, and Distributed TrainingSwyx [00:29:12]: Huh.Swyx [00:29:14]: The way you portrayed it with, like, RDMA versus TCP, like, very well laid out, but just the transfer speed change at scale for RL, like yeah, you have it, you have it built in. I'm sure someone found it. It's found it to be a lot more efficient before you made a thing out of it, right?Akshat [00:29:32]: Yeah. And not to split hairs, I guess the overlay network is the TCP overlay network.Akshat [00:29:39]: The reason we have that is you need that to do the key exchange for RDMA before you set up the RDMA network on top of that. but then people found the TCP part.Swyx [00:29:48]: Can I tell you, this is like a big aha moment for me becauseAkshat [00:29:51]: YeahSwyx [00:29:51]: So I review 2,200 submissions for the World's Fair.Akshat [00:29:56]: Yeah.Swyx [00:29:57]: And then I got this from John OsterhoutAkshat [00:29:58]: HuhSwyx [00:29:59]: Who I don't know if. Do John Osterhout by name?Akshat [00:30:01]: The name sounds familiar.Swyx [00:30:02]: He published a. He's a well-known professor, published a lot of interesting software design books, and this is the talk he chose to submit, is on RDMA at Inference. And I'm like, you wouldn't think that this guy, who is like operating systems guy, would care about RDMA.Akshat [00:30:20]: I, it makes sense to me because I,Swyx [00:30:24]: This is the cloud, right? YeahAkshat [00:30:25]: Like, the way you move around your KV cache and how efficiently you can do it, how efficiently you move, your weights from your training GPUs to your inference GPUs in RL is there's a lot of degrees of freedom, and it is a systems problemSwyx [00:30:41]: YeahAkshat [00:30:41]: Moving memory aroundSwyx [00:30:42]: YeahAkshat [00:30:43]: Scheduling.Swyx [00:30:44]: This shows you how primitive my understanding of networking stuff is.Swyx [00:30:46]: Is this like the domain of WireGuard as well?Akshat [00:30:50]: Not quite.Swyx [00:30:51]: It's adjacent?Swyx [00:30:53]: Explain everything.Akshat [00:30:54]: Sure.Swyx [00:30:56]: How do we move memory around GPUs?Akshat [00:30:58]: Well, so sorry. Yeah, that is memory. Sorry, I was talking more, and maybe I was talking like five minutes back, about the private IPv6, addressing that you've set up.Swyx [00:31:09]: Yeah.Akshat [00:31:09]: Is it like it's a VPN?Swyx [00:31:10]: Yeah, it is like a VPN, and yeah, WireGuard is, yeah, you're right. It is,Akshat [00:31:16]: Right. Yeah, you already moved on to new topicsSwyx [00:31:17]: A similarAkshat [00:31:18]: OkaySwyx [00:31:19]: In the same space, WireGuard is, encrypted and this is,Akshat [00:31:23]: And you don't need encryption.Swyx [00:31:23]: Yeah.Akshat [00:31:24]: Yeah.Swyx [00:31:24]: This is not encrypted. that's the main difference. This is TCP and we have eBPF programs that will reject or allow the TCP connection based on whether you're allowed to do it.Akshat [00:31:35]: Used to involve a full sidecar, but now you have eBPF in the Linux kernel.Swyx [00:31:39]: Yeah.Akshat [00:31:40]: Yeah. I don't know if this is a natural follow-on to the topic of like my skepticism on distributed training is that while, like, people spend a lot of money on, like, cables to hook up GPUs, and even that is not, like, fast enough, and that's the bottleneck, is your networking fast enough?Swyx [00:31:59]: Yeah. So I guess you're talking about fully distributed training like, Dialog or something which is like cross data centerAkshat [00:32:06]: That would be, yes.Swyx [00:32:07]: That's the extreme.Akshat [00:32:08]: Yeah.Swyx [00:32:08]: You're in the middle, and then other people would have like the Mellanox cables up in, like, their actual data center.Akshat [00:32:14]: When you run multi-node training on Modal, RDMA, I think Mellanox, is, or InfiniBand is like a, is all seen as RDMA. but it's a way to bypass the TCP networking stack and, transfer, stuff much faster, between one node, to the other. And we have I think like 3 terabit per second, internal networkingSwyx [00:32:40]: OkayAkshat [00:32:40]: Which is the standard that's needed.Swyx [00:32:42]: Okay. So I misunderstood whatAkshat [00:32:43]: 50Swyx [00:32:43]: What part of the stack you wereAkshat [00:32:44]: 50 gigs overSwyx [00:32:45]: YeahAkshat [00:32:45]: If you wentSwyx [00:32:45]: YeahAkshat [00:32:46]: RDMA.Swyx [00:32:46]: Okay.Swyx [00:32:48]: Yeah. I, very impressive work.Multi-Node Training, Post-Training, and Auto ResearchSwyx [00:32:52]: So effectively you're extending like the model philosophy to the training cluster, like, yeah.Akshat [00:32:59]: Yeah. And we're, we're not going for like large scale training runs. the thing that we've built multi-node training for is, we see a lot of, smaller scale post-training. like, people are post-training like medium sized fund models, so they can, get higher quality on inference. this is a perfect fit, for something like that.Swyx [00:33:21]: Yeah. That is my impression of how a lot of these labs explore branches in post-training and then eventually merge whatever they find in.Akshat [00:33:31]: Yeah. The other use case we've seen for multi-node training is even if you have a big cluster, your researchers are still doing small runsSwyx [00:33:38]: YesAkshat [00:33:39]: Having elasticity thereSwyx [00:33:40]: Right, sureAkshat [00:33:40]: Matters a lot more.Swyx [00:33:41]: Yeah. the, like, this is like the current limiting factor for auto research, which is like you need to give your model some GPUs in order for it to completely run.Akshat [00:33:51]: We have a blog post on auto resource and model is,Swyx [00:33:55]: YeahAkshat [00:33:56]: Yeah, like, turns out to be pretty good substrate for that.Swyx [00:33:59]: So my impression is auto research means many things, likeAkshat [00:34:01]: YeahSwyx [00:34:01]: Anything that Andrej coins. Right now it's still science fair, right? Like not like, I don't know how many people are doing this.Akshat [00:34:08]: We're having a golf.Swyx [00:34:08]: Yeah.Akshat [00:34:09]: I thought the same thing.Swyx [00:34:11]: Yeah, you would know.Akshat [00:34:12]: We, like, our internal both training and inference teams use this the general shape of this quite a bit. like we have this one internal repo called auto inference, which essentially we've automated our own forward-deployed engineering efforts using, this harness, which is, the agent will just spin up a sweep of different things. It'll even run like, NVIDIA inside profiler and it'll like tweak configs and it'll arrive the right thing. it'll change your GPUs both from H200 to B200, and works really well.Swyx [00:34:47]: Nice.Akshat [00:34:47]: So yeah.Swyx [00:34:48]: By the way, I enjoy that your forward-deployed engineering is so technical that you have to do these things.Swyx [00:34:52]: It's very different from forward-deployed engineering from other people.Akshat [00:34:54]: Yeah. For our forward-deployed engineering team is, essentially they're like applied inference researchers or applied training researchers.Swyx [00:35:02]: Someone told me like they have to be able to build, but they also have to be able to sell. do they have to sell or are they like they're good, they're just like post-sale type of thing?Akshat [00:35:09]: It does, being able to talk to a customer and engage effectively with themSwyx [00:35:13]: YeahAkshat [00:35:13]: Matters a lot.Swyx [00:35:14]: They want the same thing.Akshat [00:35:15]: Yeah.Swyx [00:35:15]: ?Akshat [00:35:15]: But it's it's not really a sales, thing. We pair them with-- We have solution architects as well that are more on the sales side.Swyx [00:35:23]: Okay. Let's spend a bit more time on auto research. This is a big focus for for this year. Where does this go? like, have people explored enough? Like, there's all these beautiful charts of like improve and then level off a bit and then you find the next thing. Is this one abstraction up from normal training? Is that how we think about it, or do you think about it differently? Like model level training versus high, like driven hyperparameter search.Auto Inference and Modal BenchAkshat [00:35:51]: Yeah, like,Swyx [00:35:51]: Someone, some people call it like neural architecture search or whatever, right? Like.Akshat [00:35:54]: Yeah, - So the stuff I've seen people do with it is nowhere on the architecture level. It's pretty much tweaking parameters, but it's it's a hyperparameter sweep that's guided by some model intuition, so it's like much more efficient than, whatever other, sweep you would have.Swyx [00:36:12]: Yeah, it's just, it's just a question of where you want to spend your compute?Akshat [00:36:16]: Right.Swyx [00:36:16]: ‘Cause yeah, you can just throw infinite amounts of money on this and somehow you'll bang out Shakespeare?Akshat [00:36:22]: Yeah, infinite monkey.Swyx [00:36:24]: Yeah, so like the very good for model. and I think it's also very important that agents can spin up other agents, can spin up their infrastructure. Like very good for you. how good is our LLMs at generating model code? Like the benefit of existing LLMs is that you are in the data.Akshat [00:36:42]: Yeah. They're, they're surprisingly good. I think like pre Cloud 4 they were not, and then now they're able to shot, stuff out of the box. But we're playing around with releasing like a Modal Bench for like the harderSwyx [00:36:55]: YeahAkshat [00:36:55]: Things, that the LLMs cannot do yet and maybeSwyx [00:36:59]: What's an example of that?Akshat [00:37:01]: I think the things that- Sometimes agents struggle with, without right guidance and a skill is, how to, use the rest of our observability. Like how to. Something is failing, like how do you look at the logs and then update the right thing? It's reasoning about that. But they're able to shot, likeSwyx [00:37:23]: Yeah. You can just add a skill to it?Compute Strategy and Capacity PlanningAkshat [00:37:26]: Yeah. So we have a Modal skill now that. Which is why we built this Modal Bench. It's to find things like that, so we can address them in our tool.Swyx [00:37:35]: Tune a skill. Yeah.Akshat [00:37:36]: Yeah.Swyx [00:37:36]: No. it's it's good. are you facing any shortages? like we talk a lot about GPU shortages, but also CPU, also memory.Swyx [00:37:44]: Yeah.Akshat [00:37:45]: We have had a lot of growth, which means that, there's - we've had to be much better aboutSwyx [00:37:53]: PlanningAkshat [00:37:54]: Proactive capacity planning.Swyx [00:37:55]: Yeah.Akshat [00:37:55]: So we have,Swyx [00:37:57]: Which by the way, like it's like a MBA's like dreamAkshat [00:38:00]: YesSwyx [00:38:00]: Is like just planning this stuff. I think last time you and I talked about something maybe about this.Akshat [00:38:03]: Yeah. we have a really competent team of people that we call, The role is called compute strategy. so yeah, if anyone listening here or wants to work on thatSwyx [00:38:13]: Compute strategy?Akshat [00:38:13]: Yeah.Swyx [00:38:14]: I think,Akshat [00:38:14]: I feel like,Swyx [00:38:15]: I think the normies call it FP&A or something.Akshat [00:38:18]: Well, it's more It's it's not FP&A. It's it's There's a lot of interesting financial questions of like what is the blend between one year and three-year reservations? how do we forecast our own capacity? how do we. especially since our capacity is very fungible across different GPU types and different regions, like you have to model a lot of it. and you also have to have an opinion on how the supply chain is gonna evolve, and then you have to like, take bets,Swyx [00:38:49]: YeahAkshat [00:38:49]: Based on that.Swyx [00:38:50]: Tokenomics.Akshat [00:38:50]: Yeah.Swyx [00:38:51]: This is like probably a not a real point, but, I was trying to think about like what other industries. I was trying to think about like, we cannot be first to like these kinds of problems.Akshat [00:38:59]: Yeah.Swyx [00:39:00]: And what other industries have had this? And I was like, airlines with fuel and like they have to hedge their fuel and like, I think for a long time Southwest because they made like a hero fuel bet, they like were like super low cost becauseAkshat [00:39:12]: OhSwyx [00:39:12]: Compared to everyone else.Akshat [00:39:14]: Yeah. I hadn't thought about that.Vibhu [00:39:16]: We're at a fun time too?Akshat [00:39:18]: Yeah. It's. A lot of the compute business in general, for us is also about being very good about capacity management. That is how you have great unit, economics. but also over time it's how you can unlock more value for customers. Like, one of the things we're building now is like a way for customers to get, If they don't care about latency, like get much cheaper pricing and they'll get results back in like next 24 hours or something, like a batch tier essentially.Batch Tiers and Latency-Insensitive WorkloadsSwyx [00:39:47]: Yeah.Akshat [00:39:47]: And those are levers we have because we control the whole stack and scheduling and whatnot to give people a sufficientSwyx [00:39:53]: Yeah. I feel like they're not as popular. Like those, like the Frontier Labs have all those APIs. They're not as popular as they should be.Akshat [00:40:00]: The demand that we see for something like that is not for LLMs. although sometimes people wanna run evals andSwyx [00:40:08]: OkayAkshat [00:40:08]: Synthetic data prep and there it makes sense.Swyx [00:40:10]: Okay.Akshat [00:40:11]: But it's from a lot of LLM companies, like people who are doing computational bio, like they have to run really big batch jobs and they don't care about when they get it back.Swyx [00:40:22]: Yeah. And like they have a reasonable. It's it's also like a cousin to the stopping problem of like, will this finish in time?Akshat [00:40:30]: Yeah. You can bound it.Swyx [00:40:33]: Yeah.Akshat [00:40:33]: Like you can give peopleSwyx [00:40:34]: YeahAkshat [00:40:34]: SLAs on it.Swyx [00:40:35]: Yeah. I think what's, what's interesting is like the next phase of model.Swyx [00:40:38]: Like what, do people expect from you, now that you're established and you're like well-known compute player among all these leading companies. You had an inference launch week, and we talked a little bit about the launches. like what else? Like what else should people know?What Modal Builds NextAkshat [00:40:55]: We are building primitives that make our users' lives much easier. So, I think for example, with LLM inference, thousands more companies are gonna post-train their own models and, deploy open source models for inference. so we're thinking a lot about what is the best product shape for that. And, that involves everything from our training gym to, then, endpoints that get frontier-level performance. again, but I haven't talked to anyone. It looks somewhat different on other verticals. Like, we're also seeing a lot of real-time, audio-video stuff in there, which is why like, we're working on things like regional routing, with fallbacks. So you can get GPUs that are as close to users as possible. so you get like low latency for video streaming and whatnot. And then on the agent side, it's,Akshat [00:41:52]: We're still working very closely with our customers because stuff is changing so fast in terms of what they need. And, I think beyond sandboxes and persistent file systems, there's a lot of other things people will need from this agent stack as they build production agents. So yeah, we're thinking about those other things that fit in there.Swyx [00:42:13]: I want to ask what the other things are.Akshat [00:42:15]: Yeah. I probably should share right now.Swyx [00:42:17]: I think-- I think, okay, so, I do think a lot about the principal components of cloud, and you do talk about compute storage networking.Akshat [00:42:25]: Yeah.Swyx [00:42:25]: Because so far for me, it's fine. so far for the. the first couple generations of cloud, it's fine. What's different, qualitatively different about agents that you need some new permission level? Like a lot of people, okay, and I'll just kinda spew tokens at you until it like hopefully sparks something.Akshat [00:42:43]: Yeah.Swyx [00:42:44]: Like the new level now is whatever Claude Code does, which is dangerously scope permissions or like allow list by command or like whatever, right? And sometimes they're like, “Well, okay, we have like this adaptive thinking mode where like, just trust me, bro. I will make the calls for you.” Is that it? like mediated permissions.Hard Guardrails vs. LLM-Mediated PermissionsVibhu [00:43:03]: Now you're looping it with a goal and letting it roll.Akshat [00:43:06]: Yeah, I'm, I'm skeptical of LLM media permission for stuff that is at the sandbox level because you do want hard boundaries.Swyx [00:43:16]: Yeah.Akshat [00:43:16]: Otherwise, someone can exfiltrate stuff.Swyx [00:43:20]: But likeAkshat [00:43:20]: YeahSwyx [00:43:20]: Maybe that's old school thinking. Maybe we're the dinosaurs.Swyx [00:43:23]: Maybe the AI OS or the LLM OS is really the kernel is a goddamn LLM.Swyx [00:43:30]: Like it makes you feel uncomfortable.Akshat [00:43:31]: Yeah, I'm, I'm toldSwyx [00:43:32]: But that's what trusting the LLM is. Like imagine a spherical cow perfect LLM.Akshat [00:43:36]: Right.Swyx [00:43:37]: That it.Akshat [00:43:39]: Maybe.Swyx [00:43:41]: I wanna test the boundaries, right?Akshat [00:43:42]: Yeah.Swyx [00:43:42]: Like, and I don't believe that, but I wanna see where I'm wrong ‘cause that's, that's the consensus.Akshat [00:43:49]: Yeah. I think you always need hard guardrails when you want, And you can pair those with softer guardrails, right? And that's gonna be a lot of mediated.Managed Agents and Specialized SandboxesSwyx [00:44:00]: There. I'll also get you a end with a couple of your commentary on like the ecosystem outside of Modal. Manage agents. Everyone has one. Gemini, OpenAI, Claude, very useful for you, but also like it is their way of starting to edge into your space.Akshat [00:44:17]: Yeah.Swyx [00:44:17]: What's going on?Akshat [00:44:19]: Yeah, we're, very excited to partner with Anthropic and some of the other foundation labs, will not name who we're also working with. the way we see it is the manage agent thing is a great place to start if you're starting out building an agent and, But then when you get to, building something more production grade, like you're a company that's like Ramp that's building their own, Ramp also runs their accounting agent on us, so their external-facing agent. You need a lot more control over, your compute primitive on things like, what sort - how do you persist different files that the agent has access to, and how do you snapshot and restore? How do you control the networking? maybe you want GPUs. When you get to that point, you kinda want, a specialized sandbox provider, that gives you those things, and that's the role that we are trying to play.Swyx [00:45:15]: YeahAkshat [00:45:16]: We don't really have an opinion on the harness, whether it runs - it's a cloud-managed agent, and you hook it up to Model Sandbox, or you run the harness in Model Sandbox. We'll see where people converge with that.Swyx [00:45:26]: Yeah. Do you any opinions on like the meta harnesses, or just another layer on top of these things?Akshat [00:45:31]: You mean like the OpenPipeSwyx [00:45:33]: OpenPipe is one. I think Vercel had one, which I can't remember the name of right now. Fredshot had one. and then, to me, most recently was Data Databricks that had Omnigen. All these are meta harness. Like it's kinda pseudo agent cloud type things.Akshat [00:45:50]: I personally have not played around with them.Swyx [00:45:53]: Yeah.Akshat [00:45:53]: Build agents with them.Swyx [00:45:54]: Everything's bullish Modal, as long as it consumes more infra.Akshat [00:45:57]: That's why we're focusing on the infra layer. It's somewhere where our, relative competence is and, also it's a hard problem to solve.Swyx [00:46:06]: Yeah. I will say like just generally reflecting on that, I don't know if - if there's other topics on Modal, but like just generally reflecting as an infra person, not as intense as you, but in that field, this has like been the most exciting time in infra. Like it was boring for a while, and you couldn't really get people excited about data infrastructure. Like Eric would get on Data Console, everyone just watched the video and like say, “Look at how many sandboxes I can spin up,” and no one gave a crap.Why Infrastructure Became Exciting AgainAkshat [00:46:39]: Yeah.Swyx [00:46:40]: And like now everyone gives a crap.Akshat [00:46:42]: That's true. It is a very exciting time, and I think a lot of that's driven by just the amount of scale all of this stuff needs.Swyx [00:46:50]: I think the, like a lot of your initiatives or a lot of your like product directions make sense in retrospect, which is like the best kind, but I wouldn't necessarily have thought about it myself, which.Akshat [00:47:00]: We need the predictions.Swyx [00:47:02]: I think there's a lot that you just don't even see, right? Like you have the batch, you have the voice, you have the multimodal, but what else?Akshat [00:47:10]: What else is coming up for usSwyx [00:47:11]: Yeah. Where do you see things going?Akshat [00:47:13]: Yeah. I, in generalBiotech, Robotics, and Non-LLM AI WorkloadsAkshat [00:47:15]: It's it's clear that there's there's a huge shift happening. I think one thing that's not as obvious to people because LLM inference gets talked about so much and is also we work a lot of companies that are, doing things like drug discovery and computational bio, like the Chai Discoveries of the world. Big things are probably gonna happen there. we work a lot of robotics companies that are putting robots in like active deployments and getting good results out of them.Swyx [00:47:45]: Is there Air Gap Modal? Is there a version that is like prem air gapped whatever?Akshat [00:47:50]: No. We,Swyx [00:47:51]: You should cloud only.Akshat [00:47:51]: Yeah.Swyx [00:47:52]: Yeah. Okay. But yeah, so what you're saying is like because you're focused on primitives and they're good primitives, you find use cases in all these kinds of things.Akshat [00:48:01]: Yeah.Swyx [00:48:01]: Probably diversifies you a little bit away from LMS all the time.Akshat [00:48:05]: Yeah, absolutely. We're, we'- our goal isn't to only serve the LLM inference market.Swyx [00:48:10]: There are a lot just on the website, the audio,Akshat [00:48:12]: Yeah. We said both onSwyx [00:48:14]: Computational bio images. Yeah, there's a lot here. There's QTA TTS, customizing. Oh, Chatterbox. there was customizing Whisper.Akshat [00:48:24]: Okay. Yeah.Swyx [00:48:25]: This screen reminds me of a fallen competitor, which Replicate.Model APIs vs. Differentiated AI ProductsSwyx [00:48:31]: What's your postmortem on what happened?Akshat [00:48:34]: This is one thing we've stayed away from is providing an API for models because I think providing model APIs is some of it ends up serving like a really hobbyist market, which is much less sticky.Swyx [00:48:50]: Yeah.Akshat [00:48:50]: And we've always wanted to build for companies that are building products and need more flexibility that's not just an API.Swyx [00:48:57]: Which you can build an API for a model and this is clearly what it is. But you - but what you're saying, you can wrap it into a more fully functioning back end that you run.Akshat [00:49:06]: Yeah. So all of our examples, it's not that spin up this model, here's an API token, use it. They're all code.Swyx [00:49:13]: Okay.Akshat [00:49:13]: And so the point is that this is just an example.Swyx [00:49:16]: Starter code.Akshat [00:49:17]: Yeah. But you can tweak it however you want.Swyx [00:49:20]: Yeah.Akshat [00:49:21]: And if you're like a company building a product, like, computational bio whatnot, yeah.Swyx [00:49:26]: I guess I'm trying to tease out for listenersAkshat [00:49:28]: YeahSwyx [00:49:28]: When does it stop becoming, oh, you're just an API call and you're just a wrapper on API to becoming what you call a product, right?Swyx [00:49:36]: Like, what is that layer? Like what-- Like, more lines of code, but like beyond that, what is the substance that people add that qualifies it to be something more?Akshat [00:49:46]: I think there's a little bit of like a selection effect of like a lot of the companies who do wanna get deeper into that level are probably building something that's more differentiated. And, I think, an example is like - with LLM inference, originally we, worked with companies that were building their own post-training frameworks or they were, - Ramp early in the day was training their own tokenizer and like swapping out the tokenizer in Llama and whatnot. I'm not saying that's, that successful, in that case. But a better example is like, let's say Suno. because Suno, does not use Modal for training.Swyx [00:50:26]: Mikey on the pod. Yeah.Akshat [00:50:27]: But they use Modal for all their inference and that's because they have like a custom-- They have completely custom model architecture and that means that they have to be at the code level and tweak things that are not, just an API.Swyx [00:50:41]: It's interesting as well, like we had, Ethan, most recently on the xAI Groq team make a prediction that like the next tier in video gen is not a better video model, it's a better model or agent that orchestrates video models.Video Agents and Production WorkflowsAkshat [00:50:56]: Oh, interesting.Vibhu [00:50:56]: Language model backbone that can use toolsAkshat [00:50:58]: RightVibhu [00:50:59]: And write code.Akshat [00:51:00]: Like, yes, I can make my second video or my second video from Groq, but I want my minute video.Akshat [00:51:06]: And I'm not going there through normal video gen.Swyx [00:51:10]: Yeah, that's interesting. I - So we have GPU sandboxes and recently have seen a few companies doing agents that do video manipulation or,Akshat [00:51:22]: Yeah. Give it FFmpeg and just do it.Swyx [00:51:23]: Run FFmpeg. But likeAkshat [00:51:25]: That's not enough.Swyx [00:51:25]: Yeah.Akshat [00:51:26]: You need to give it Adobe.Swyx [00:51:27]: Yeah, I hadn't put it together with like it would be a video production thing. in my mind these things were going more towards editingAkshat [00:51:36]: Yeah.Vibhu [00:51:36]: Well, shout out Mantis.Akshat [00:51:37]: I think about this a lot.Swyx [00:51:38]: .Akshat [00:51:41]: Yeah. Sorry.Vibhu [00:51:41]: Luma. Luma Agent is a version of this for video production, but it's a off.Swyx [00:51:46]: I was gonna get your quick takes, on some other stuff that happensGitpod/Ona, CI, and Runtime SandboxesSwyx [00:51:50]: In recent news and just-just see if you have anything interesting. Gitpod, very li

Fluent Fiction - Italian
Tuscany Vineyards: Balancing Tradition & Innovation

Fluent Fiction - Italian

Play Episode Listen Later Jul 8, 2026 16:35 Transcription Available


Fluent Fiction - Italian: Tuscany Vineyards: Balancing Tradition & Innovation Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-07-08-22-34-02-it Story Transcript:It: Il sole estivo scaldava dolcemente i colli toscani.En: The summer sun gently warmed the Tuscan hills.It: Le vigne si estendevano come un mare verdeggiante sotto il cielo azzurro.En: The vineyards stretched out like a verdant sea under the blue sky.It: L'aria profumava di uva matura e fiori selvatici.En: The air was fragrant with ripe grapes and wildflowers.It: Alessio, con il volto leggermente segnato dalla fatica, sedeva pensieroso al tavolo sotto il pergolato.En: Alessio, his face slightly marked by fatigue, sat thoughtfully at the table under the pergola.It: Le foglie offrivano un po' di ombra, mentre l'odore di terra e vino li circondava.En: The leaves offered some shade, while the scent of earth and wine surrounded them.It: Gemma arrivò con passo leggero, portando con sé l'energia dell'arte e nuove idee.En: Gemma arrived with a light step, bringing with her the energy of art and new ideas.It: Aveva un sogno: trasformare parte del vigneto in un rifugio culturale.En: She had a dream: to transform part of the vineyard into a cultural retreat.It: Voleva integrare l'arte e la natura, un ponte tra tradizione e innovazione.En: She wanted to integrate art and nature, a bridge between tradition and innovation.It: Poco dopo, Luca si unì a loro.En: Soon after, Luca joined them.It: Il giovane appariva rilassato, ma dentro di sé lottava per trovare il suo ruolo tra le responsabilità familiari.En: The young man appeared relaxed but inside he struggled to find his role amidst the family responsibilities.It: Voleva fare la differenza, ma non sapeva come.En: He wanted to make a difference, but he didn't know how.It: "Allora, di cosa parliamo oggi?" chiese Luca, rompendo il silenzio.En: "So, what are we talking about today?" Luca asked, breaking the silence.It: "Tradizione," rispose Alessio fermo.En: "Tradition," Alessio replied firmly. "The vineyard must maintain its prestige.It: "Il vigneto deve mantenere il suo prestigio. Ogni anno ne va della nostra reputazione."En: Every year our reputation is at stake."It: Gemma si sporse in avanti, facendo tintinnare i bicchieri sul tavolo.En: Gemma leaned forward, making the glasses on the table clink.It: "Posso capirlo, ma immagina quante persone potremmo attrarre con un rifugio d'arte.En: "I can understand that, but imagine how many people we could attract with an art retreat.It: Porterebbe nuova vita qui."En: It would bring new life here."It: Alessio sospirò. "E se tutto questo fallisce? Rischiamo troppo."En: Alessio sighed. "And if all this fails? We risk too much."It: Luca guardò entrambi.En: Luca looked at both of them.It: Sentiva il peso delle loro aspettative.En: He felt the weight of their expectations.It: Poi, con un lampo di intuizione, intervenne.En: Then, with a flash of insight, he intervened.It: "E se provassimo con una parte piccola? Mantenendo il resto come sempre?"En: "What if we tried with a small part? Keeping the rest as it is?"It: Silenzio, poi un sorriso affiorò sul volto di Gemma.En: Silence, then a smile appeared on Gemma's face.It: "Una piccola parte... Potrei accettarlo."En: "A small part...I could accept that."It: Alessio esitò, poi annuì lentamente.En: Alessio hesitated, then nodded slowly.It: "D'accordo. Un esperimento. Vediamo come va."En: "Alright. An experiment. Let's see how it goes."It: Così, nacque un nuovo capitolo per la famiglia.En: Thus, a new chapter was born for the family.It: Alessio si sentiva più leggero condividendo il carico.En: Alessio felt lighter sharing the burden.It: Gemma aveva il suo spazio per creare.En: Gemma had her space to create.It: E Luca aveva finalmente trovato il suo scopo: gestire il progetto, un ponte tra il vecchio e il nuovo.En: And Luca had finally found his purpose: managing the project, a bridge between the old and the new.It: Il sole calava lentamente, tingendo il cielo di arancione e rosa.En: The sun slowly set, tinting the sky orange and pink.It: I tre fratelli brindarono sotto il pergolato, il vino riflettendo i colori di una nuova speranza.En: The three siblings toasted under the pergola, the wine reflecting the colors of newfound hope.It: La famiglia era unita, e la terra toscana accoglieva il loro sogno.En: The family was united, and the Tuscan land embraced their dream. Vocabulary Words:the summer: l'estatethe hills: i collithe vineyards: le vignethe fatigue: la faticathe table: il tavolothe pergola: il pergolatothe leaves: le fogliethe scent: l'odorethe wildflowers: i fiori selvaticithe dream: il sognothe retreat: il rifugiothe bridge: il pontethe tradition: la tradizionethe innovation: l'innovazionethe responsibilities: le responsabilitàthe difference: la differenzathe silence: il silenziothe prestige: il prestigiothe reputation: la reputazionethe glasses: i bicchierithe shade: l'ombrathe burden: il caricothe purpose: lo scopothe sunset: il tramontothe sky: il cielothe wine: il vinothe hope: la speranzathe family: la famigliathe land: la terrathe siblings: i fratelli

Unica Radio Podcast
Alessio Mura: voce del rap in lingua sarda tra Balentia e LIMBAS

Unica Radio Podcast

Play Episode Listen Later Jul 7, 2026 5:50


Alessio Mura racconta il nuovo EP dei Balentia, il valore del sardo nel rap contemporaneo e il festival LIMBAS che celebra identità, musica e cultura giovanile in Sardegna Il percorso di Alessio Mura nel rap sardo Nel panorama del rap sardo, la figura di Alessio Mura si distingue per continuità e coerenza artistica. Autore dei testi dei Balentia, Mura rappresenta una delle voci più solide di un movimento che da oltre trent'anni porta avanti la lingua sarda attraverso l'hip hop. Il suo lavoro si inserisce in una tradizione musicale che non è solo espressione artistica, ma anche strumento di identità culturale. Il rap sardo continua a evolversi e a trovare nuovi spazi grazie a progetti come il suo. Fueddu sardu: un EP tra lingua e identità Il nuovo EP Fueddu sardu segna un ulteriore passo nel percorso dei Balentia. Il progetto nasce dall'esigenza di raccontare la realtà contemporanea attraverso una lingua che conserva radici profonde. Nei quattro brani che compongono il lavoro emerge una forte attenzione ai temi sociali e culturali, con un linguaggio diretto e accessibile. Il rap sardo diventa così uno strumento narrativo che unisce passato e presente, tradizione e innovazione. Giovani e lingua sarda nella musica contemporanea Uno degli aspetti centrali del lavoro di Alessio Mura riguarda il rapporto tra giovani e lingua sarda. Negli ultimi anni si è registrato un rinnovato interesse verso le lingue locali, soprattutto attraverso la musica. Il rap sardo ha avuto un ruolo decisivo in questo processo, avvicinando nuove generazioni a una forma linguistica che rischiava di essere marginalizzata. La musica diventa quindi uno spazio di riappropriazione culturale e di espressione identitaria. LIMBAS: il primo festival del rap in lingua sarda Il 20 giugno 2026 a Sassari si terrà LIMBAS, il primo festival dedicato al rap in lingua sarda. L'evento rappresenta un punto di svolta per la scena musicale isolana, riunendo artisti e pubblico attorno a una stessa idea di cultura condivisa. Per Alessio Mura, protagonista dell'iniziativa, si tratta di un'occasione importante per dare visibilità a un movimento in crescita e per rafforzare la presenza del rap sardo nel panorama nazionale. Rap sardo e impegno sociale Nei testi dei Balentia emerge spesso una forte componente sociale. Il rap sardo viene utilizzato come mezzo di riflessione su temi attuali, dalla condizione giovanile alle trasformazioni culturali. La scrittura diventa così uno strumento di analisi e denuncia, ma anche di costruzione identitaria. Alessio Mura sottolinea come la musica possa ancora oggi avere un ruolo educativo e civile. Prospettive future e nuove generazioni Guardando al futuro, il lavoro di Alessio Mura e dei Balentia continua a orientarsi verso la valorizzazione della lingua sarda attraverso la musica. I consigli rivolti ai giovani artisti si concentrano sulla coerenza e sulla necessità di credere nella propria identità culturale. Il rap sardo, in questo senso, non è solo un genere musicale, ma un percorso di consapevolezza.

Fluent Fiction - Italian
Unraveling Tuscany's Mystery: A Summer of Secrets and Growth

Fluent Fiction - Italian

Play Episode Listen Later Jul 3, 2026 17:04 Transcription Available


Fluent Fiction - Italian: Unraveling Tuscany's Mystery: A Summer of Secrets and Growth Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-07-03-07-38-19-it Story Transcript:It: La villa storica in Toscana splendeva sotto il sole estivo.En: The historic villa in Tuscany shone under the summer sun.It: Era un luogo magico, con vigneti che si estendevano a perdita d'occhio e giardini rigogliosi.En: It was a magical place, with vineyards stretching as far as the eye could see and lush gardens.It: Le stanze erano luminose, con soffitti alti e mobili ornati.En: The rooms were bright, with high ceilings and ornate furniture.It: In un angolo della Villa c'era un piccolo segreto: l'antica biblioteca, un luogo perfetto per riflessioni e indagini.En: In one corner of the villa, there was a little secret: the ancient library, a perfect place for reflection and investigation.It: Sofia, giovane e acuta osservatrice, era lì per passare l'estate con i suoi cugini, Matteo e Alessio.En: Sofia, a young and keen observer, was there to spend the summer with her cousins, Matteo and Alessio.It: Matteo era spensierato e affascinante, sempre pronto a una risata, mentre Alessio preferiva i libri alla compagnia, riservato ma profondo nei suoi pensieri.En: Matteo was carefree and charming, always ready for a laugh, while Alessio preferred books to company, reserved but deep in his thoughts.It: Un giorno, un prezioso cimelio di famiglia scomparve misteriosamente.En: One day, a precious family heirloom mysteriously disappeared.It: Era un oggetto di grande valore, sia economico che affettivo.En: It was an object of great value, both economic and sentimental.It: Sofia sentì immediatamente il peso del mistero.En: Sofia immediately felt the weight of the mystery.It: Non voleva che i suoi cugini fossero accusati ingiustamente.En: She didn't want her cousins to be unjustly accused.It: Si mise al lavoro, fingendo una certa indifferenza mentre investigava.En: She set to work, pretending a certain indifference while she investigated.It: Osservava ogni dettaglio, ogni sguardo, raccogliendo piccoli indizi.En: She observed every detail, every glance, gathering small clues.It: Chiacchierava casualmente con Matteo e Alessio, cercando di capire cosa fosse davvero successo.En: She chatted casually with Matteo and Alessio, trying to understand what had really happened.It: La tensione cresceva nella villa.En: The tension in the villa was growing.It: Matteo si sentiva accusato a causa delle sue marachelle passate.En: Matteo felt accused because of his past pranks.It: Alessio temeva di essere colpevole per la sua distrazione.En: Alessio feared he might be guilty due to his distraction.It: Ma Sofia era determinata a trovare la verità.En: But Sofia was determined to find the truth.It: Dopo giorni di indagini silenziose, Sofia convocò i suoi cugini in biblioteca.En: After days of silent investigations, Sofia summoned her cousins to the library.It: C'era un'aria di suspense.En: There was an air of suspense.It: Con calma, rivelò ciò che aveva scoperto: Matteo aveva nascosto il cimelio come uno scherzo.En: Calmly, she revealed what she had discovered: Matteo had hidden the heirloom as a prank.It: Però lo scherzo era sfuggito di mano.En: However, the joke had gotten out of hand.It: In un attimo di dramma, Matteo capì l'importanza della responsabilità.En: In a moment of drama, Matteo realized the importance of responsibility.It: Si scusò con sincerità e restituì l'oggetto.En: He apologized sincerely and returned the object.It: I cugini si guardarono negli occhi, sentendo il peso delle parole non dette e delle tensioni sciogliersi come neve al sole.En: The cousins looked each other in the eye, feeling the weight of unspoken words and tensions dissolve like snow in the sun.It: La villa tornò serena.En: The villa returned to serenity.It: Con il cimelio al suo posto, i tre giovani godettero del resto dell'estate senza tensioni.En: With the heirloom back in place, the three young people enjoyed the rest of the summer without tensions.It: Sofia guadagnò fiducia nelle sue capacità; Matteo imparò il valore delle sue azioni; Alessio scoprì quanto il suo legame con i cugini fosse prezioso.En: Sofia gained confidence in her abilities; Matteo learned the value of his actions; Alessio discovered how precious his bond with his cousins was.It: Il sole tramontava sui vigneti, gettando lunghe ombre che nascondevano i segreti dal passato.En: The sun set over the vineyards, casting long shadows that hid secrets from the past.It: Ma per Sofia, Matteo e Alessio, il futuro era luminoso e pieno di nuove esperienze da condividere.En: But for Sofia, Matteo, and Alessio, the future was bright and full of new experiences to share.It: La villa risuonava di risate e complicità, un'estate in Toscana che nessuno di loro avrebbe mai dimenticato.En: The villa echoed with laughter and camaraderie, a summer in Tuscany that none of them would ever forget. Vocabulary Words:the villa: la villathe summer: l'estatethe vineyard: i vignetithe ceiling: il soffittoornate: ornatilush: rigogliosithe secret: il segretoreflection: riflessionethe investigation: l'indaginethe heirloom: il cimeliocarefree: spensieratocharming: affascinantereserved: riservatothe gaze: lo sguardothe tension: la tensioneto accuse: accusarethe prank: la marachellasincerely: con sinceritàto dissolve: sciogliersithe drama: il drammato apologize: scusarsithe responsibility: la responsabilitàthe bond: il legameprecious: preziosothe laughter: le risatethe camaraderie: la complicitàthe mystery: il misterothe shadow: l'ombrathe sunset: il tramontothe clue: l'indizio

Italiano ON-Air
Estate in Italia: guida a sagre, festival e feste popolari

Italiano ON-Air

Play Episode Listen Later Jul 1, 2026 8:36 Transcription Available


L'estate in Italia è sinonimo di piazze vive, musica all'aperto e, soprattutto, tanto buon cibo. Ma ti sei mai chiesto qual è la differenza esatta tra una sagra e un festival? O perché una manifestazione non è sempre una protesta?Nell'ultimo episodio della stagione 13 di

Il Cortocircuito
IL TRIO IN SPIAGGIA, mentre L'INDUSTRY AFFONDA!!!

Il Cortocircuito

Play Episode Listen Later Jun 30, 2026 165:00


Introduzione e Location Speciale Il trio composto da Pierpaolo, Alessio e Francesco si ritrova in una location insolita: uno stabilimento balneare a Passo Scuro, sulla costa laziale. La puntata speciale del "Cortocircuito" si svolge durante un pranzo a base di pesce, tra scherzi sull'audio basso e problemi tecnici di connessione legati allo streaming all'aperto.Aumento Prezzi Apple e Strategie Microsoft Il dibattito si accende sull'improvviso aumento dei prezzi dei prodotti Apple (MacBook, Mac Mini, HomePod) avvenuto senza preavviso. Si discute della legalità di cambiare i preventivi già emessi e si confronta questa mossa con gli aumenti di Microsoft per le console Xbox e le strategie di prezzo di Nintendo, analizzando l'impatto sul mercato PC e console.Il Futuro dell'Hardware e del Gaming Viene analizzato il mercato dell'hardware PC, con particolare attenzione ai prezzi delle RAM e alla crisi di componenti che non accenna a tornare ai livelli pre-2025. Si parla di Steam Machine, dell'ottimizzazione del software come alternativa alla potenza bruta e delle prospettive per le console di prossima generazione (PS6, Project Helix).Cinema, Serie TV e Reality Il trio condivide opinioni su film recenti come "Marty Supreme" (giudicato molto negativamente) e serie TV. Il discorso vira poi sui reality show italiani come Temptation Island e Amici, discutendo sulla loro veridicità, sul ruolo degli autori e sulla percezione sociale di questi programmi, includendo aneddoti personali su provini falliti.Intelligenza Artificiale e Sviluppo Videogiochi Un tema centrale è l'uso dell'IA generativa nell'industria del gaming. Si discute dell'etica delle "etichette AI" su Steam, della scomparsa dei ruoli "junior" sostituiti dall'automazione e di come l'IA cambierà radicalmente la produzione di asset (come alberi o texture) e il doppiaggio nei giochi futuri.Crisi degli Studi e Acquisizioni Si analizza la situazione critica di studi storici come Bungie (licenziamenti massicci e futuro incerto per Marathon), Double Fine e Obsidian. Il gruppo riflette sulla sostenibilità dei "Live Service" e sugli errori di gestione di Sony (era Jim Ryan) e Microsoft nel gestire i propri team interni.Consigli Finanziari e Chiusura In chiusura, Alessio e il gruppo analizzano l'andamento dei mercati finanziari e delle criptovalute. Si parla del crollo di titoli tech come ARM, delle prospettive di SpaceX sotto la soglia dei 150 dollari e dell'instabilità del Bitcoin, il tutto mentre la batteria del PC che trasmette la live sta per esaurirsi.

School Life Podcast
Blood in the Water - Boston, Alessio, Thomas, Henry and Archie - St Michaels Primary

School Life Podcast

Play Episode Listen Later Jun 26, 2026 4:10


CRASH! This is a story about four innocent brothers who crashed in the middle of the Bermuda Triangle, they did everything they could to get off of the island, along the way they encountered disgusting fish, sharks and a conveniently placed helicopter. Will they survive? Or will they perish in the depths of Armani cove! More episodes from St Michaels Primary here: https://www.archdradio.com/podcasts/slp/stmichaels-primary

RadioPNR
Alessio Schiavi ci presenta la "Festa di San Pietro in Antola"

RadioPNR

Play Episode Listen Later Jun 26, 2026 4:42


All'interno del programma di Radio PNR : City Life, condotto da Giampaolo Cacciatore, Alessio Schiavi, collaboratore Rifugio Monte Antola, ci presenta la "Festa di San Pietro in Antola" che si svolge a Carrega Ligure nelle giornate di sabato 27 e domenica 28 Giugno.

Italiano ON-Air
"Se mio nonno avesse le ruote...": l'arte italiana di rispondere alle ipotesi assurde

Italiano ON-Air

Play Episode Listen Later Jun 24, 2026 6:59 Transcription Available


Vi è mai capitato di discutere con qualcuno che continua a fare ipotesi impossibili su cose ormai passate? "Se avessi fatto così...", "Se fossi partito prima...". In Italia abbiamo un modo decisamente originale, ironico e un po' bizzarro per stroncare questi discorsi: tirare in ballo i nonni... e le ruote!In questa puntata di Italiano ON-Air, Katia e Alessio prendono spunto da una recentissima e virale conferenza stampa di Carlo Ancelotti ai Mondiali di calcio 2026 per fare un viaggio semiserio tra i proverbi regionali più divertenti d'Italia e la grammatica del quotidiano.

ICF Singen/Villingen Audio
Dämonen in der Kirche und was Jesus dagegen tut | Alessio Passarella

ICF Singen/Villingen Audio

Play Episode Listen Later Jun 24, 2026 69:18


SBS Italian - SBS in Italiano
Pizza napoletana, Simone e Alessio Zullo campioni del mondo per l'Australia

SBS Italian - SBS in Italiano

Play Episode Listen Later Jun 23, 2026 13:31


Simone Zullo, titolare con il fratello Alessio della pizzeria Fratelli Pulcinella di Parramatta, ha vinto la XXIII edizione del Campionato Mondiale del Pizzaiuolo – Caputo Cup, nella prestigiosa categoria Pizza Napoletana S.T.G. (Specialità Tradizionale Garantita).Seguici su Facebook e Instagram o abbonati ai nostri podcast cliccando qui. 

Fluent Fiction - Italian
Discovering Leonardo: A Journey of Art and Innovation

Fluent Fiction - Italian

Play Episode Listen Later Jun 21, 2026 16:12 Transcription Available


Fluent Fiction - Italian: Discovering Leonardo: A Journey of Art and Innovation Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-06-21-22-34-01-it Story Transcript:It: Alessio e Ginevra camminano fianco a fianco nel Museo della Scienza e della Tecnologia Leonardo da Vinci.En: Alessio and Ginevra walk side by side in the Museo della Scienza e della Tecnologia Leonardo da Vinci.It: È estate e il sole entra dalle grandi finestre, illuminando i modelli e le stazioni interattive.En: It's summer, and the sun streams in through the large windows, illuminating the models and interactive stations.It: Il museo è affollato di visitatori, affascinati dalle invenzioni geniali esposte.En: The museum is crowded with visitors, fascinated by the brilliant inventions on display.It: Alessio, con gli occhi scintillanti, è pieno di entusiasmo.En: Alessio, with sparkling eyes, is full of enthusiasm.It: "Guarda, Ginevra!En: "Look, Ginevra!It: Hanno un'intera sezione dedicata ai macchinari di Leonardo."En: They have an entire section dedicated to Leonardo's machinery."It: Ginevra sorride, osservando con calma.En: Ginevra smiles, observing calmly.It: "Sì, è incredibile.En: "Yes, it's incredible.It: Ammiravo sempre la sua capacità di fondere arte e scienza."En: I always admired his ability to merge art and science."It: Arrivano al negozio di souvenir.En: They arrive at the souvenir shop.It: È pieno di scaffali con libri e modelli intricati.En: It's filled with shelves of books and intricate models.It: Alessio si precipita subito verso una figura in movimento che riproduce un uomo vitruviano meccanico.En: Alessio immediately rushes towards a moving figure that reproduces a mechanical Vitruvian Man.It: "Ginevra, non è fantastico?"En: "Ginevra, isn't it fantastic?"It: chiede Alessio, impaziente di passare al prossimo oggetto.En: asks Alessio, impatient to move on to the next object.It: "L'idea mi piace, ma dobbiamo scegliere con cura.En: "I like the idea, but we must choose carefully.It: Ogni oggetto ha una storia," risponde Ginevra, con calma.En: Every item has a story," Ginevra responds calmly.It: Mentre lei esamina con attenzione una collezione di quaderni, Alessio si allontana, alla ricerca di qualcosa che lo colpisca immediatamente.En: While she carefully examines a collection of notebooks, Alessio wanders off, searching for something that immediately strikes him.It: La sua voglia di trovare un oggetto unico lo spinge a esplorare ogni angolo del negozio.En: His desire to find a unique object drives him to explore every corner of the store.It: Finalmente, raggiungono una sezione del negozio con in mostra un modello in edizione limitata della macchina volante di Leonardo.En: Finally, they reach a section of the shop displaying a limited edition model of Leonardo's flying machine.It: Alessio e Ginevra si fermano, entrambi attratti dalla meraviglia di quell'opera.En: Both Alessio and Ginevra stop, drawn by the wonder of that work.It: "Ginevra, è perfetto!En: "Ginevra, it's perfect!It: Rappresenta la sua innovazione e il suo splendore storico," esclama Alessio con entusiasmo condiviso.En: It represents his innovation and historical splendor," exclaims Alessio with shared enthusiasm.It: Ginevra annuisce, percependo l'unione ideale di tecnologia e storia.En: Ginevra nods, sensing the ideal union of technology and history.It: "Hai ragione, Alessio.En: "You're right, Alessio.It: Questo è qualcosa che rappresenta entrambi i nostri interessi."En: This is something that represents both of our interests."It: Acquistano il modello con un senso di soddisfazione.En: They purchase the model with a sense of satisfaction.It: Mentre escono dal museo, Alessio apprezza la riflessione di Ginevra sui dettagli, mentre Ginevra impara a godere dell'energia di Alessio verso l'innovazione.En: As they leave the museum, Alessio appreciates Ginevra's reflection on details, while Ginevra learns to enjoy Alessio's energy towards innovation.It: Insieme, escono nel luminoso sole estivo, contenti della scelta condivisa.En: Together, they step into the bright summer sun, pleased with their shared choice.It: Entrambi hanno imparato qualcosa di più l'uno dall'altro, lasciando alle spalle una giornata splendidamente memorabile nel cuore del genio di Leonardo da Vinci.En: Both have learned something more about each other, leaving behind a beautifully memorable day in the heart of the genius of Leonardo da Vinci. Vocabulary Words:the genius: il geniothe window: la finestrathe visitor: il visitatorethe invention: l'invenzionethe souvenir shop: il negozio di souvenirthe shelf: lo scaffaleintricate: intricatothe notebook: il quadernoenthusiasm: l'entusiasmoto merge: fonderebrilliant: genialesparkling: scintillanteto admire: ammirarethe reflection: la riflessionethe detail: il dettagliomechanical: meccanicomoving: in movimentoto examine: esaminareto explore: esplorareunique: unicothe corner: l'angolothe model: il modellolimited edition: edizione limitatathe flying machine: la macchina volantethe splendor: lo splendoreto purchase: acquistaresatisfaction: la soddisfazionehistorical: storicoto enjoy: goderebright: luminoso

Pillole di Storia
#792 - Alessio II, un bambino tra i complotti di Costantinopoli

Pillole di Storia

Play Episode Listen Later Jun 19, 2026 21:02


Adattamento audio: Matteo D'Alessandro - www.matteodalessandro.com Per approfondire gli argomenti della puntata: La nostra serie Imperatores, sugli imperatori romani : https://youtube.com/playlist?list=PLpMrMjMIcOkkIDocjNI3Q7gCk-4bOiVVO Le altre puntate sulla storia di Roma antica : https://youtube.com/playlist?list=PLpMrMjMIcOkkVlao9HeDl3jIHVKO3IcR_

Italiano ON-Air
Un museo in tasca: l'arte sulle monete italiane

Italiano ON-Air

Play Episode Listen Later Jun 17, 2026 8:05 Transcription Available


The Phone Hacks
Alessio Carducci-Ultimate Fap Champion

The Phone Hacks

Play Episode Listen Later Jun 16, 2026 61:37


Eddie and Carducci tell Capper about the UFC fight he missed. Also Capper compiles theworst warehousing job stories. Join the PATREON HERE - Just $7 (AUD) for bonus eps and content - get tons of behind the scenes hacks and pranks and help keep this podcast going! Go watch Capper's special Hold Me Closer Tiny Cancer HERE Follow CAPPER and ROHAN and PHONE HACKS on Instagram Subscribe where you're listening and leave a review to get the word out thereSee omnystudio.com/listener for privacy information.

The Watford FC Buzz Podcast
Who is new Watford Head Coach Alessio Dionisi?

The Watford FC Buzz Podcast

Play Episode Listen Later Jun 15, 2026 54:45


EP01: Who is the new Head Coach Alessio Dionisi? Hello and welcome to the Watford Buzz Podcast! The Home of your Watford FC chat, featuring journalist Tom Bodell (@TBBodell), analyst Jordan Wiemer (@JordanWeimer) and hosted by commentator and presenter Matt Mesiano (@MessyMesiano) We all have one thing in common, we're all huge Watford fans and we LOVE talking about the Hornets! On today's show, Matt, Tom and Jordan discussed:Pay our respects to Kenny JacketChat Watford Players at the World CupDiscuss the new head coach Alessio DionisiIf you want to get in touch you can do so really easily – just ping a message across on Twitter , BlueSky, OR send us an email to WatfordBuzzPodcast@gmail.com Hosted on Acast. See acast.com/privacy for more information.

Italiano ON-Air
Novecento

Italiano ON-Air

Play Episode Listen Later Jun 10, 2026 7:49 Transcription Available


Alessio e Katia ci portano a fare un viaggio emozionante nel tempo e attraverso l'Oceano, partendo dalle celebrazioni della Festa della Repubblica fino a toccare le storie dei milioni di italiani emigrati all'estero, grazie alle pagine di un grande capolavoro della letteratura italiana contemporanea: Novecento di Alessandro Baricco

Fluent Fiction - Italian
Unveiling Shadows: Alessio's Quest for Truth in Silence

Fluent Fiction - Italian

Play Episode Listen Later Jun 5, 2026 17:14 Transcription Available


Fluent Fiction - Italian: Unveiling Shadows: Alessio's Quest for Truth in Silence Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-06-05-22-34-01-it Story Transcript:It: Il sole del mattino filtrava attraverso le finestre del reparto psichiatrico, dipingendo ombre di luce sulle pareti bianche.En: The morning sun filtered through the windows of the psychiatric ward, painting shadows of light on the white walls.It: Alessio si svegliò lentamente, con la testa pesante e vuota.En: Alessio woke up slowly, with his head heavy and empty.It: Non sapeva come fosse arrivato lì.En: He didn't know how he had gotten there.It: Attorno a lui, il silenzio era rotto solo dal tenue rumore dei passi dei medici lungo i corridoi.En: Around him, the silence was broken only by the faint sound of doctors' footsteps along the corridors.It: Le pareti erano alte e fredde, un contrasto stridente con i fiori che fiorivano fuori, nel giardino della struttura.En: The walls were high and cold, a striking contrast with the flowers blooming outside, in the garden of the facility.It: Sapeva di dover capire cosa fosse successo, ma i ricordi erano come sabbia che scivola tra le dita.En: He knew he had to understand what had happened, but the memories were like sand slipping through fingers.It: "Buongiorno, Alessio."En: "Good morning, Alessio."It: La voce gentile di Giulia lo fece sobbalzare.En: The gentle voice of Giulia startled him.It: Era giovane, con occhi gentili e un sorriso che rassicurava.En: She was young, with kind eyes and a reassuring smile.It: "Come ti senti oggi?"En: "How are you feeling today?"It: "Confuso," rispose Alessio, cercando di mettere insieme i pezzi della sua memoria.En: "Confused," Alessio replied, trying to piece together fragments of his memory.It: "Non ricordo come sono finito qui."En: "I don't remember how I ended up here."It: Giulia annuì, registrando qualcosa sul suo taccuino.En: Giulia nodded, jotting something down in her notebook.It: "Ci lavoreremo insieme," disse, con una certa comprensione.En: "We'll work on it together," she said, with a certain understanding.It: "Nel frattempo, Marco vorrà parlarti più tardi."En: "In the meantime, Marco will want to talk to you later."It: Marco era un uomo all'apparenza cordiale, ma nei suoi occhi c'era qualcosa che metteva Alessio a disagio.En: Marco was an outwardly cordial man, but there was something in his eyes that made Alessio uneasy.It: Sembrava sapere più di quanto lasciasse intendere.En: He seemed to know more than he let on.It: Quella sera, mentre Alessio si rigirava nel letto, sentì un fruscio sotto il cuscino.En: That evening, as Alessio tossed and turned in bed, he felt a rustle under the pillow.It: Era un foglio di carta.En: It was a sheet of paper.It: Con mani tremanti, Alessio aprì la lettera.En: With trembling hands, Alessio opened the letter.It: Raccontava frammenti della sua vita prima del ricovero: una lite, un senso di tradimento, e nomi che non riusciva a collegare.En: It recounted fragments of his life before being admitted: an argument, a sense of betrayal, and names he couldn't connect.It: Alla fine, c'era un accenno a un coinvolgimento di Marco in qualcosa di più grande.En: At the end, there was a hint of Marco's involvement in something bigger.It: Il giorno seguente, Alessio affrontò Marco.En: The next day, Alessio confronted Marco.It: Gli occhi di Marco mutarono, passando dal fastidio all'accettazione.En: Marco's eyes changed, shifting from irritation to acceptance.It: "Non volevamo ferirti," disse Marco, con una sincerità dolente.En: "We didn't want to hurt you," Marco said, with a painful sincerity.It: "Era per proteggerti.En: "It was to protect you.It: C'era un pericolo per te."En: There was a danger for you."It: Giulia arrivò subito dopo, offrendo ad Alessio un'opzione.En: Giulia arrived shortly after, offering Alessio an option.It: "Possiamo aiutarti a riavere la tua vita, a cercare la verità nel modo giusto."En: "We can help you regain your life, to seek the truth in the right way."It: Con una nuova determinazione, Alessio accettò.En: With new determination, Alessio agreed.It: Capì che non era solo più una preda della sua mente, ma un uomo che poteva ancora lottare per la propria storia.En: He understood that he was no longer just a prey to his mind, but a man who could still fight for his own story.It: Fuori, la primavera continuava, con il suo profumo di rinascita e speranza.En: Outside, spring continued, with its scent of rebirth and hope.It: Alessio guardò il cielo azzurro oltre le finestre, sentendosi un po' più integro.En: Alessio looked at the blue sky beyond the windows, feeling a bit more whole.It: Retrovie di ricordi cominciavano a prendere forma, e con l'aiuto di Giulia e, in parte, di Marco, era pronto a scoprire la verità.En: Backdrops of memories began to take shape, and with the help of Giulia and, in part, Marco, he was ready to discover the truth. Vocabulary Words:psychiatric ward: il reparto psichiatricopaint: dipingereshadows: le ombreheavy: pesantecorridors: i corridoiblooming: fioriresand: la sabbiafingers: le ditastartle: sobbalzarereassuring: rassicurantejot down: registrareunderstanding: la comprensioneoutwardly: all'apparenzauneasy: a disagiorustle: il frusciotrembling: tremantearguement: la litebetrayal: il tradimentoinvolvement: il coinvolgimentoirritation: il fastidioacceptance: l'accettazionehurt: feriresincerity: la sinceritàdanger: il pericolodetermination: la determinazioneprey: la predarebirth: la rinascitabackdrops: le retrovietruth: la veritàwhole: integro

Fluent Fiction - Italian
Crafting Firenze: An Ice Cream Adventure

Fluent Fiction - Italian

Play Episode Listen Later Jun 2, 2026 18:14 Transcription Available


Fluent Fiction - Italian: Crafting Firenze: An Ice Cream Adventure Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-06-02-07-38-19-it Story Transcript:It: La piazza Santa Croce a Firenze era in fermento.En: La piazza Santa Croce in Firenze was buzzing with activity.It: Il calore estivo avvolgeva la città, e la musica dei violinisti di strada riempiva l'aria.En: The summer heat enveloped the city, and the music of street violinists filled the air.It: I turisti con gelati colorati in mano osservavano le celebrazioni della Festa della Repubblica, mentre l'aroma delle specialità fiorentine si mescolava a quello dei fiori nei balconi.En: Tourists with colorful ice creams in hand watched the celebrations of the Festa della Repubblica, as the aroma of Fiorentine specialties blended with that of the flowers on the balconies.It: Luca, con gli occhi brillanti di passione, guardava fuori dalla vetrina del suo piccolo negozio di gelato.En: Luca, with eyes bright with passion, looked out from the window of his small ice cream shop.It: Aveva sognato per settimane un nuovo sapore.En: For weeks, he had dreamed of a new flavor.It: Doveva catturare l'essenza di Firenze: la dolcezza dell'arte, la bellezza storica, e un pizzico di modernità.En: It had to capture the essence of Firenze: the sweetness of art, historical beauty, and a touch of modernity.It: Non era un compito facile.En: It wasn't an easy task.It: E c'era un motivo in più per la sua determinazione: un famoso critico gastronomico, Giovanna, era in città e lui voleva impressionarla.En: And there was an additional reason for his determination: a famous food critic, Giovanna, was in town, and he wanted to impress her.It: Ma la realtà del negozio lo portava con i piedi per terra.En: But the reality of the shop kept him grounded.It: La fila di clienti non finiva mai.En: The line of customers never ended.It: I turisti desideravano assaggiare il gelato artigianale di Luca.En: Tourists wanted to taste Luca's artisanal ice cream.It: I rumori delle risate e delle voci erano contagiosi, ma anche travolgenti.En: The sounds of laughter and voices were contagious but also overwhelming.It: Alessio, il suo fedele assistente, correva da una parte all'altra cercando di aiutare tutti.En: Alessio, his loyal assistant, ran back and forth trying to help everyone.It: Tuttavia, Luca sapeva che al ritmo attuale, non avrebbe mai trovato il tempo per sperimentare il suo nuovo sapore.En: However, Luca knew that at the current pace, he would never find the time to experiment with his new flavor.It: Inoltre, uno degli ingredienti chiave, una rara essenza di lavanda, era in ritardo a causa delle festività.En: Moreover, one of the key ingredients, a rare lavender essence, was delayed due to the holiday festivities.It: Luca dovette prendere una decisione difficile.En: Luca had to make a difficult decision.It: "Chiuderò il negozio per qualche ora," decise ad alta voce, sorprendente Alessio e i clienti all'interno.En: "I'll close the shop for a few hours," he decided aloud, surprising Alessio and the customers inside.It: Alcune persone mormorarono deluse, ma Luca era determinato.En: Some people murmured in disappointment, but Luca was determined.It: Sapeva che il rischio era grande, ma altrettanto lo era la possibile ricompensa.En: He knew the risk was great, but so too was the potential reward.It: Con Alessio che controllava la porta, Luca si mise al lavoro nel retrobottega.En: With Alessio managing the door, Luca got to work in the backroom.It: Tra scaffali pieni di ingredienti e pentole di acciaio, sperimentava sapori e consistenze.En: Among shelves filled with ingredients and stainless steel pots, he experimented with flavors and textures.It: Dopo ore di tentativi, trovò un sostituto per l'essenza di lavanda: il miele di acacia, dolce e profumato.En: After hours of trials, he found a substitute for the lavender essence: acacia honey, sweet and fragrant.It: Con il sole al tramonto, la porta del negozio si riaprì.En: With the sun setting, the shop door reopened.It: I clienti curiosi tornarono, attratti dalla promessa di un nuovo sapore.En: Curious customers returned, drawn by the promise of a new flavor.It: Giovanna, il critico, provò il gelato con una piccola esitazione, ma il suo viso si illuminò all'assaggio.En: Giovanna, the critic, tried the ice cream with a little hesitation, but her face lit up upon tasting it.It: "Delizioso!En: "Delicious!It: Riesco a sentire Firenze in ogni cucchiaio!"En: I can feel Firenze in every spoonful!"It: esclamò.En: she exclaimed.It: Le voci si diffusero rapidamente.En: Word spread quickly.It: Il gelato di Luca diventò il nuovo must della città.En: Luca's ice cream became the new must-have in the city.It: Il successo non solo portò nuovi clienti, ma anche un profondo insegnamento per Luca.En: The success not only brought new customers but also a profound lesson for Luca.It: Capì che a volte vale la pena fermarsi e seguire il cuore.En: He understood that sometimes it's worth stopping and following your heart.It: La piazza era ancora viva, piena di luci e musica, e Luca, con un sorriso sereno, si godette la vista del suo negozio vivace.En: The square was still alive, full of lights and music, and Luca, with a serene smile, enjoyed the sight of his lively shop.It: Sapeva di aver trovato il suo posto nel cuore di Firenze e nel cuore della sua gente.En: He knew he had found his place in the heart of Firenze and in the hearts of its people. Vocabulary Words:activity: l'attivitàheat: il calorearoma: l'aromaflower: il fiorepassion: la passionewindow: la vetrinaflavor: il saporeessence: l'essenzamodernity: la modernitàtask: il compitofoot: il piedeline: la filaartisan: artigianalelaughter: la risatavoice: la vocepace: il ritmoingredient: l'ingredienteholiday: la festivitàreward: la ricompensabackroom: il retrobottegashelf: lo scaffalesteel: l'acciaiohoney: il mielesunset: il tramontocurious: curiosospoonful: il cucchiaiomust-have: il mustlesson: l'insegnamentoheart: il cuoresight: la vista

Il Cortocircuito
"PIÙ GIOCHI PER TUTTI!" e IL SUPPORTO INFINITO di THE WITCHER 3!

Il Cortocircuito

Play Episode Listen Later May 31, 2026 149:45


L'Eredità di IO InteractiveSi discute del successo di 007 First Light, analizzando il milione e mezzo di copie vendute. Il focus si sposta sulla natura del gioco: un mix tra le meccaniche stealth di Hitman in chiave arcade e momenti hollywoodiani alla Uncharted. I conduttori lodano la scrittura e la capacità del team di rendere accessibile un genere spesso punitivo.Il Supporto ai Videogiochi: Tre Casi a ConfrontoIl cuore del dibattito ruota attorno a tre diverse filosofie di supporto post-lancio:Star Citizen: Raggiunto il miliardo di dollari di finanziamento, ci si interroga se sia ancora "crowdfunding" o una vendita diretta di asset digitali. Alessio lo definisce provocatoriamente una "truffa" o un "sogno" venduto a caro prezzo.CD Projekt Red & The Witcher 3: Si analizza l'annuncio di una nuova espansione/aggiornamento a 11 anni dal rilascio originale. Il dibattito si accende sulla distinzione tra supporto genuino e operazione di marketing per preparare il terreno a The Witcher 4.Bungie & Destiny 2: Viene discussa la gestione della comunicazione interna ed esterna riguardo alla fine del supporto e ai licenziamenti, con critiche alla trasparenza dell'azienda verso i dipendenti.Il Mercato Hardware e Steam DeckViene trattato il rincaro di Steam Deck e le polemiche scatenate dai tweet di Tim Sweeney (Epic). Si riflette sulla saturazione del mercato degli handheld PC e sulla strategia di Valve di alzare i prezzi per mantenere i margini, contrapponendola alla visione degli "aziendalisti" e dei "fan boy".Critica al Sistema delle Key e Conflitto di InteressiAlessio presenta una lunga analisi (19 slide) sul sistema di distribuzione delle chiavi per i recensori. Critica aspramente le proposte di automatizzare l'accesso alle key basandosi solo sui numeri di follower (populismo digitale), sottolineando l'importanza delle PR umane, della qualità dei contenuti e dei limiti tecnici imposti dalle piattaforme.Design e Innovazione: La Nuova Ferrari LuceL'ultimo grande tema è la presentazione della Ferrari Luce, la prima Ferrari elettrica a cinque posti. I conduttori commentano le reazioni negative di figure come Montezemolo e Briatore, discutendo del design curato da LoveFrom (Jonathan Ive) e del rischio di alienare l'identità storica del brand in favore di un nuovo target tecnologico e internazionale.POTETE TROVARE LE SLIDE DI PRESENTAZIONE ESPOSTE A SUPPORTO DEL DISCORSO DI ALESSIO QUI:https://docs.google.com/presentation/d/1_BDYQEPF9T-sUq4tkgLc5rcLUHndYI83T9OtVhJCk1M/edit?usp=sharing

Million Bazillion
What if there was only one currency in the whole world?

Million Bazillion

Play Episode Listen Later May 28, 2026 9:49


We've talked a lot about how money works and why countries have their own currencies here on “Million Bazillion.” But listener Alessio wants to know: Why DOESN'T the whole world use the same money? And could the world's nations all decide to just use one shared currency? In this bonus mini-episode, we'll get some answers!

Marketplace All-in-One
What if there was only one currency in the whole world?

Marketplace All-in-One

Play Episode Listen Later May 28, 2026 9:49


We've talked a lot about how money works and why countries have their own currencies here on “Million Bazillion.” But listener Alessio wants to know: Why DOESN'T the whole world use the same money? And could the world's nations all decide to just use one shared currency? In this bonus mini-episode, we'll get some answers!

Italiano ON-Air
Come sopravvivere all'incubo delle DOPPIE!

Italiano ON-Air

Play Episode Listen Later May 27, 2026 7:14 Transcription Available


In questo episodio, Katia e Alessio affrontano uno degli argomenti più temuti (e talvolta imbarazzanti) per chi studia la nostra lingua: le consonanti doppie. Se anche tu ti sei chiesto almeno una volta come riconoscerle e pronunciarle correttamente, questa è la puntata che fa per te!

Fluent Fiction - Italian
Unveiling Secrets: A Journalist's Quest on Costiera Amalfitana

Fluent Fiction - Italian

Play Episode Listen Later May 26, 2026 16:13 Transcription Available


Fluent Fiction - Italian: Unveiling Secrets: A Journalist's Quest on Costiera Amalfitana Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-05-26-22-34-02-it Story Transcript:It: Sulla Costiera Amalfitana, il sole brillava alto nel cielo primaverile.En: On the Costiera Amalfitana, the sun shone high in the spring sky.It: Il mare blu scintillava sotto di esso e le onde accarezzavano dolcemente la sabbia dorata della spiaggia.En: The blue sea sparkled beneath it, and the waves gently caressed the golden sand of the beach.It: I profumi dei fiori di bougainvillea e dei limoni riempivano l'aria, mentre le ville sulla scogliera si scaldavano al tepore del sole.En: The scents of bougainvillea flowers and lemons filled the air, while the cliffside villas basked in the warmth of the sun.It: Tra la sabbia, un relitto misterioso attirava l'attenzione di chi passeggiava sulla spiaggia, locale e visitatore allo stesso modo.En: Among the sand, a mysterious wreck drew the attention of those walking on the beach, both locals and visitors alike.It: Alessio, un giovane giornalista avventuroso, non poteva resistere al richiamo di quel mistero.En: Alessio, a young adventurous journalist, couldn't resist the allure of that mystery.It: Sperava di scrivere un articolo che avrebbe cambiato la sua carriera.En: He hoped to write an article that would change his career.It: Ma sapeva che per scoprire la verità aveva bisogno dell'aiuto di Bianca, una storica locale.En: But he knew that to uncover the truth, he needed the help of Bianca, a local historian.It: Bianca era conosciuta per il suo amore verso la storia del paese, ma anche per il suo desiderio di proteggere i suoi segreti.En: Bianca was known for her love of the town's history, but also for her desire to protect its secrets.It: Quando Alessio le chiese aiuto, Bianca esitò.En: When Alessio asked for her help, Bianca hesitated.It: Temeva che la storia avrebbe potuto sconvolgere la comunità.En: She feared the story might unsettle the community.It: "Bianca, voglio solo raccontare la verità," disse Alessio, il suo tono sincero.En: "Bianca, I just want to tell the truth," said Alessio, his tone sincere.It: "Credo che tutti dovrebbero conoscere il passato del nostro paese."En: "I believe everyone should know the past of our town."It: Alla fine, Bianca decise di fidarsi di Alessio.En: In the end, Bianca decided to trust Alessio.It: Insieme perlustrarono il relitto sulla spiaggia.En: Together, they explored the wreck on the beach.It: Lavorarono fianco a fianco, esaminando ogni pezzo di legno e ogni strano oggetto trovato.En: They worked side by side, examining every piece of wood and every strange object they found.It: Un giorno, scoprirono un piccolo compartimento nascosto.En: One day, they discovered a small hidden compartment.It: Dentro trovarono antichi manufatti che avrebbero potuto riscrivere la storia del paese.En: Inside, they found ancient artifacts that could rewrite the town's history.It: Bianca guardò Alessio.En: Bianca looked at Alessio.It: "Capisco l'importanza della tua storia, ma dobbiamo essere attenti," avvertì.En: "I understand the importance of your story, but we must be careful," she warned.It: Alessio annuì.En: Alessio nodded.It: "Scriverò un articolo che rispetta la storia e il tuo desiderio di protezione," promise.En: "I will write an article that respects history and your wish for protection," he promised.It: Con il suo aiuto, Alessio pubblicò una storia che non solo informava ma anche celebrava la ricca eredità culturale del paese.En: With her help, Alessio published a story that not only informed but also celebrated the town's rich cultural heritage.It: Il pubblico era incantato, e il paese trovò un nuovo interesse per il suo passato.En: The audience was enchanted, and the town found a renewed interest in its past.It: Alla fine, Alessio imparò l'importanza della collaborazione e del rispetto verso la storia locale.En: In the end, Alessio learned the importance of collaboration and respect for local history.It: Bianca, d'altra parte, riconobbe che rivelare alcune verità poteva portare a una maggiore comprensione e preservazione.En: Bianca, on the other hand, recognized that revealing some truths could lead to greater understanding and preservation.It: La spiaggia della Costiera Amalfitana continuava a brillare sotto il sole di primavera, ora casa non solo di un misterioso relitto ma anche di un nuovo capitolo nella storia del paese.En: The beach of the Costiera Amalfitana continued to shine under the spring sun, now home not only to a mysterious wreck but also to a new chapter in the town's history. Vocabulary Words:the coast: la costathe cliff: la scoglierathe wreck: il relittothe spring sky: il cielo primaverilethe warmth: il teporethe historian: lo storicothe artifact: l'artefattothe beach: la spiaggiathe scent: il profumothe sand: la sabbiathe compartment: il compartimentothe truth: la veritàthe past: il passatothe chapter: il capitolothe history: la storiathe heritage: l'ereditàthe villa: la villathe community: la comunitàthe article: l'articolothe career: la carrierathe visitor: il visitatorethe discovery: la scopertathe interest: l'interessethe presence: la presenzathe understanding: la comprensionethe collaboration: la collaborazionethe protection: la protezionethe preservation: la preservazionethe allure: il richiamothe mystery: il mistero

Fluent Fiction - Italian
Secret Bunker Revelations: Rediscovering Adventure and Friendship

Fluent Fiction - Italian

Play Episode Listen Later May 22, 2026 18:07 Transcription Available


Fluent Fiction - Italian: Secret Bunker Revelations: Rediscovering Adventure and Friendship Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-05-22-07-38-20-it Story Transcript:It: Lorenzo camminava lungo la strada con passo incerto.En: Lorenzo walked along the street with an uncertain step.It: Era primavera, ma l'aria era ancora fresca.En: It was spring, but the air was still fresh.It: Gli alberi erano pieni di foglie verdi e i fiori sbocciavano ovunque.En: The trees were full of green leaves and flowers bloomed everywhere.It: Aveva ricevuto un messaggio da Giulia.En: He had received a message from Giulia.It: "Vieni al bunker segreto," aveva scritto.En: "Come to the secret bunker," she had written.It: Il bunker era un luogo di cui tutti parlavano in città, ma pochi sapevano dov'era.En: The bunker was a place everyone in town talked about, but few knew where it was.It: Non era sicuro se fosse una buona idea, ma il desiderio di avventura lo spingeva.En: He wasn't sure if it was a good idea, but the desire for adventure drove him.It: Infine, Lorenzo si convinse.En: Finally, Lorenzo convinced himself.It: Si avviò verso il luogo indicato nel messaggio.En: He set off toward the location indicated in the message.It: Giulia lo aspettava all'ingresso nascosto da un pò di cespugli.En: Giulia was waiting for him at the entrance hidden by some bushes.It: "Ciao, Lorenzo!"En: "Hi, Lorenzo!"It: esclamò, energica come sempre.En: she exclaimed, energetic as always.It: Accanto a lei c'era Alessio, con un sorriso enigmatico.En: Next to her was Alessio, with an enigmatic smile.It: Lorenzo non vedeva Alessio da anni.En: Lorenzo hadn't seen Alessio in years.It: Era sempre stato misterioso, carismatico, capace di convincere chiunque a seguirlo.En: He had always been mysterious, charismatic, able to convince anyone to follow him.It: Il bunker era buio e illuminato solo da luci colorate appese qua e là.En: The bunker was dark and lit only by colored lights hanging here and there.It: Dentro, un gruppo di persone parlava animatamente o suonava strumenti strani.En: Inside, a group of people was talking animatedly or playing strange instruments.It: C'erano divani vecchi e sedie di ogni tipo.En: There were old sofas and chairs of every kind.It: Lorenzo si sentì fuori posto, ma anche affascinato.En: Lorenzo felt out of place, but also fascinated.It: "Perché siamo qui?"En: "Why are we here?"It: chiese Lorenzo, rivolgendosi ad Alessio.En: Lorenzo asked, turning to Alessio.It: "Voglio scoprire cos'hai in mente."En: "I want to find out what's on your mind."It: Alessio rise.En: Alessio laughed.It: "Non è nulla di pericoloso, te lo assicuro.En: "It's nothing dangerous, I assure you.It: Volevo solo riunirci tutti, come ai vecchi tempi.En: I just wanted to gather us all, like in the old days.It: Vedi, questo posto è speciale.En: You see, this place is special.It: Qui si può essere sé stessi."En: Here you can be yourself."It: Lorenzo guardò Giulia.En: Lorenzo looked at Giulia.It: Lei sembrava felicissima, completamente a suo agio.En: She seemed very happy, completely at ease.It: Forse era ora di smettere di preoccuparsi così tanto.En: Maybe it was time to stop worrying so much.It: Aveva sempre desiderato un po' di avventura, senza troppi rischi.En: He had always wanted a bit of adventure, without too many risks.It: Si sedette su un divano traballante.En: He sat on a wobbly sofa.It: Sentiva la tensione sciogliersi.En: He felt the tension melt away.It: Col passare delle ore, Lorenzo si unì alle conversazioni.En: As the hours passed, Lorenzo joined the conversations.It: Parlava con sconosciuti, rideva, ascoltava storie.En: He talked with strangers, laughed, listened to stories.It: Era una sensazione nuova e sorprendentemente piacevole.En: It was a new and surprisingly pleasant feeling.It: Iniziò a capire Alessio.En: He began to understand Alessio.It: Non c'erano pericoli nascosti qui, solo persone in cerca di nuovi legami.En: There were no hidden dangers here, only people seeking new connections.It: Alla fine della serata, Lorenzo si sentì cambiato.En: By the end of the evening, Lorenzo felt changed.It: Salutò Alessio e Giulia, grato per l'esperienza.En: He said goodbye to Alessio and Giulia, grateful for the experience.It: Tornò a casa con un sorriso.En: He went home with a smile.It: Aveva ritrovato il suo spirito d'avventura, senza perdere il controllo.En: He had rediscovered his adventurous spirit, without losing control.It: Alessio aveva le migliori intenzioni e, grazie a lui, Lorenzo aveva riscoperto un lato di sé che pensava perduto.En: Alessio had the best intentions and, thanks to him, Lorenzo had rediscovered a side of himself he thought lost.It: Da quel giorno in poi, Lorenzo trovò il giusto equilibrio tra cautela e avventura, pronto a dire "sì" alle opportunità che la vita gli offriva.En: From that day on, Lorenzo found the right balance between caution and adventure, ready to say "yes" to the opportunities life offered him.It: E sapeva che, con amici come Giulia e Alessio, non era mai solo nel suo viaggio.En: And he knew that, with friends like Giulia and Alessio, he was never alone in his journey. Vocabulary Words:the street: la stradathe step: il passothe spring: la primaverathe leaves: le foglieto bloom: sbocciarethe message: il messaggiothe bunker: il bunkerthe desire: il desideriothe adventure: l'avventurathe entrance: l'ingressothe bush: il cespuglioenergetic: energicathe smile: il sorrisoenigmatic: enigmaticothe sofa: il divanoto convince: convinceremysterious: misteriosocharismatic: carismaticoanimated: animatothe instrument: lo strumentowobbly: traballanteto exclaim: esclamaredangerous: pericolosoto assure: assicurarespecial: specialeto join: unirsithe stranger: lo sconosciutoto understand: capirethe connection: il legamethe opportunity: l'opportunità

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

Take the 2026 AI Engineering Survey and get >$2k in credits and AIE WF tickets!This was recorded before Railway suffered a major GCP outage on May 19, despite being a multi-AZ, multi-zone mesh ring, with HA fiber interconnects between their Metal GCP AWS, because workload discoverability was unintentionally still tied to GCP. All has been resolved with a post-mortem.Railway did not start as an AI infrastructure company.It was founded in 2020 years before agents became the default way people thought about deploying software. Jake Cooper, formerly at Bloomberg and Uber, started Railway with a simple obsession: the activation energy to ship something to production should be near zero. Push code, get a URL, iterate. No Docker files, no Kubernetes manifests, no Ansible scripts stacked on Ansible scripts.For years, this was a slow grind. Railway spent its first 18 months hand-acquiring its first 100 users with Jake personally greeting every Discord signup on a second monitor.Today, Railway has raised $124m and is growing very fast. A 35-person team supports 3 million users, adding roughly 100,000 signups a week. Their bare metal data centers have a 3-month payback period vs. renting in the cloud, with 70% margins funding aggressive cloud bursting when needed. The servers they own have actually appreciated in value as RAM prices have climbed basically meaning the value of their hardware now exceeds the capital they've raised.From rebuilding Railway's network overlay over a weekend to moving the vast majority of workloads onto its own bare metal data centers, Jake Cooper is trying to build a new cloud for an agent-native world. In this episode, Railway's founder and “conductor” joins swyx and Alessio to unpack why the next era of software infrastructure is not just “Heroku but newer,” what agents need that humans did not, and why the old deployment loop of Git, PRs, CI/CD, and static cloud resources may be heading for a rewrite.We go deep on Railway's infrastructure stack: own-metal data centers, three-month cloud payback periods, cloud bursting, data center debt, Railpack, Nixpacks, Temporal, feature flags, Central Station, content-addressable filesystems, agent-safe production forks, and why the CLI may become more important than the canvas in an agent world. Jake also shares the founder journey behind Railway, how the company survived losing $500K/month, why it now serves millions of users with only 35 people, and why he believes the pull request is dying.We discuss:* How Railway went from a slow six-year grind to adding 100,000 users a week* How Railway thinks about agents as the next dominant software species* Why agents need version control, observability, compute, storage, and orchestration at 1000x scale* The economics of Railway's own-metal data centers and three-month payback* How Railway uses cloud bursting while scaling its own infrastructure* Why data center debt can be a better tool than venture debt for infra startups* Central Station, Railway's internal system for clustering customer feedback and incidents* Why responsible disclosure and over-communication matter for platforms* Why feature flags, progressive rollouts, and shadow traffic are essential for agents* Temporal's strengths, pain points, and why workflows matter for agents* Railpack, Nixpacks, Nix, and lazy-loaded content-addressable filesystems* Why “cattle, not pets” may change if you can clone the pets* Why Railway is building a new cloud from scratch instead of copying hyperscalers* The solo founder path, focus, writing, and how Jake thinks about company buildingRailway:* Website: https://railway.com/* X: https://x.com/RailwayJake Cooper:* LinkedIn: https://www.linkedin.com/in/thejakecooper/* X: https://x.com/JustJakeTimestamps00:00:00 Introduction: What Is Railway?00:02:07 Jake's Path to Railway00:06:13 Railway's Six-Year Growth Story00:08:52 Rebuilding the Business After the Free Tier00:11:17 Agents as the Next Software Platform00:13:29 Railway's Infrastructure Philosophy00:15:42 Bare Metal, Cloud Economics, and the Compute Crunch00:17:22 Cloud Bursting and Five-Cloud Networking00:20:20 Data Center Debt and Infra Financing00:23:31 Data Centers in Space00:25:24 What Agents Need From Infrastructure00:28:24 CLIs, Canvas, and Agent-Native UX00:35:15 Central Station, Incidents, and Responsible Disclosure00:40:30 Safe Rollouts, SRE Agents, and Production Forks00:45:00 AI SRE, Specs, Code, and Tests00:48:24 Self-Replicating Infrastructure and the New Serverless00:53:18 Heroku, Temporal, and Workflow Engines01:04:07 Railpack, Nixpacks, and Lazy-Loaded Filesystems01:06:01 Coding Agents, Token Spend, and Roadmap Acceleration01:10:56 The Pull Request Is Dying01:12:28 Feature Flags and the Agent-Era SDLC01:16:15 Cattle, Pets, and Cloning Machines01:19:29 Solo Founder Lessons01:24:12 Focus, GPUs, and Building a New Cloud01:28:20 Closing ThoughtsTranscriptAlessio [00:00:00]: Hey, everyone. Welcome to the Latent Space Podcast. This is Alessio, founder of Kernel Labs, and I'm joined by Swyx, editor of Latent Space.Swyx [00:00:10]: Hey, hey, hey. Today we're in the studio with Jake Cooper of Railway.Alessio [00:00:14]: Conductor of Railway.Swyx [00:00:15]: Conductor at Railway. Yeah.Alessio [00:00:16]: Choo-choo.Swyx [00:00:17]: Do you actually have that anywhere, like on your business card?Jake [00:00:20]: We call some of our volunteer moderators conductors. I don't have a business card. We're not that big yet. At some point I will. I got handed a nice business card from the Supermicro folks, and I was like, “Damn, this is pretty official.”Swyx [00:00:30]: Business cards are coming back.Jake [00:00:32]: They're cool. They're hip. The conductor thing is good. We're trying to figure out what we want to call each other internally. Some people think it's super cringe and say, “You don't need a name for people internally.” Some people want to call each other something. We still don't have a really good one.Jake [00:00:55]: We've got New Railcrews, Trainiacs. Nothing has stuck yet.Swyx [00:01:00]: I like Trainiac. Trainiac sounds good. Railwayians. For those who don't know, what is Railway? Let's give people a crisp definition up front.Jake [00:01:09]: Railway is the easiest way to ship anything. You go to the canvas, or you talk with Claude, and you say, “Deploy a Postgres instance, deploy my GitHub repository, run this code,” and you're off to the races.Swyx [00:01:22]: You've got a nice animation on the landing page.Jake [00:01:24]: Thank you. None of my work, by the way. They don't let me touch the design stuff anymore.Jake [00:01:25]: We want to make it trivially easy not just to deploy things, but to evolve applications over time. Most tooling right now stacks entropy on top of entropy: Docker, Kubernetes, Ansible scripts, and all these other things. If we can version all of your software and keep track of all the changes, then we can make it trivial to clone environments, fork into a parallel universe, get copies of production data, get copies of any services, make changes, validate them, and collapse them back in without reproducing everything across a staging environment.The Railway Origin Story: From Uber Systems to a New CloudSwyx [00:02:07]: I was looking at your background: Bloomberg, Uber. Nothing immediately stands out as, “This guy is going to found the next great platform as a service.” What prepared you for Railway?Jake [00:02:21]: It was curiosity to keep going deeper. I started out on front-end stuff, working on Wolfram Mathematica and porting it over. Then I briefly moved to Bloomberg, then toward Uber and distributed systems, taking the Jump Bikes systems and moving them to a distributed system built on top of Cadence, the pre-Temporal Temporal.Swyx [00:02:44]: Which, by the way, I'm happy to talk about, pros and cons.Jake [00:02:48]: Totally.Swyx [00:02:51]: But let's do the Railway story.Jake [00:02:52]: It has been a continual step of wanting an experience. Whether it's walking up to a bike, unlocking it, and having it work frictionlessly, or something else, the depth required to make that happen follows from the experience. A lot of the work I do, and a lot of the team does, is in service of that experience. We fundamentally don't care how deep we have to go. We will swim to the bottom of the swimming pool to get the experience.Jake [00:03:17]: I don't have a physics PhD. I did an EECS degree. It has always been about figuring out the next step: how do we get there? That's what led to starting Railway for that experience and then moving all the way to bare metal data centers. I was adding patches to the kernel this week to get the experience there because I can see how much better it can be.Swyx [00:03:49]: Other patches to the Linux kernel this week?Jake [00:03:51]: Yeah. Not upstream. Our fork.Swyx [00:03:52]: That's a flex. Railpack? No, this is different. This is the OS on top of Railpack?Jake [00:03:57]: No, this is an actual kernel patch. It's always literally: what do we have to do to get that experience? Then figure it out. Anything is figureoutable.Swyx [00:04:10]: Would you send the patch upstream, or does it not fit other use cases?Jake [00:04:13]: Maybe. We have to work out the experience internally. It has to do with the storage layer we're building for some of the agentic stuff. Maybe it'll be useful upstream, but it's deeply useful for us internally.Open Source, Forks, and Non-Deterministic VersioningSwyx [00:04:29]: You mentioned open source before. How do you think about starting from open source, and then coding agents letting you do a lot more from forks of it?Jake [00:04:38]: GitHub's original sin is that it's almost a series of broken pointers. You have this thing, then you clone it, and now you've lost the whole upstream. How do we make it trivial for people to modify really small pieces of it?Jake [00:04:51]: We think of Git in a discrete sense: I've either made a change and merged upstream, or I haven't. What would it look like if it were percentage-based, a little more non-deterministic, or a stream of changes that users traverse as a percentage rolled out in general and then rolled all the way up?Jake [00:05:13]: We have the open-source kickback program and let you deploy templates because we want to make it trivial for people to version these shards over time. It solves a large problem around authentication, authorization, and security. NPM has a way to define, “Don't take any new packages.” The ideal end state is that you roll out progressively to users with the minimum impact zone and continue rolling up. JPMorgan should probably be the last one on the patch line, for all our sakes, because our money and livelihoods are there.Jake [00:05:53]: It's okay if Johnny Vibe Coder gets a broken patch because there's so much entropy in the system that the rubber has to meet the road at some point. You have to test at varying levels.The Long Grind: First Users, Free Tier, and Making the Business WorkSwyx [00:06:13]: I wanted to pull up this glorious chart, which is your usage or number of daily signups?Jake [00:06:22]: Daily signups, I think.Swyx [00:06:24]: You started six years ago. It was a slow grind, and now you're on a rocket ship. You say, “Don't doubt your fight and don't quit.” Maybe pick out certain points that were key inflections for the company.Jake [00:06:40]: At the start, it's about getting your first 100 users, hell or high water. We had a website and a support link. The support link was the Discord channel. I had notifications on with two monitors: the monitor I was working on and the other monitor with Discord. If anybody came in, I was immediately like, “Hey, how's it going?” It was rare, so getting those first 100 users to come back was the start.Jake [00:07:14]: Then you build a consultancy factory because users want all these things. You have to go back to the board and ask, “What is the actual product offering I want to build on top of this?”Jake [00:07:28]: VCs want charts that always go up and to the right, but in reality you don't necessarily want charts that look like that. For us, there have been periods of expansion where we add features to test use cases, and periods of compaction where we ask, “If the experience we have is good, how do we make it significantly better?” Maybe we strip out features that don't fit our ICP anymore.Jake [00:07:57]: The boom from 2022 to 2023 came from the free tier. Everybody under the sun was using it.Swyx [00:08:09]: A lot of Reddit bots and Discord bots.Jake [00:08:12]: And crypto miners. When you build an open product on the internet where anybody can sign up, the internet is a horrible place with so many things. You go through periods of asking, “How do I reach as many people as possible?” Then, “How do I fit the exact use case for the people who really matter and are really excited about this specific thing?”Jake [00:08:39]: Then there was a two-year period of making the actual business work. During the free-tier era, we were losing about half a million dollars a month.Swyx [00:08:59]: On a $20 million bank account.Jake [00:09:02]: On a $20 million bank account with maybe $50,000 a month in revenue. That's a horrible business. I don't know how anybody invested. But you have to go through it and say, “We have an experience people love, but the business has to work.”Jake [00:09:17]: There are two schools of thought. You can run the horrible business all the way up with bad margins, or you can go back and make it work. We've always wanted a super lean team. We're 35 people right now. It's very small.Swyx [00:09:36]: Supporting three million already?Jake [00:09:38]: Yeah. We're adding 100,000 users a week right now, so it's growing fast. We don't want to add headcount for the sake of headcount or throw bodies at problems. We want to build systems. It's hard to build systems during expansion because you're adding things to the system because people are asking for them or things are breaking.Jake [00:10:00]: We had to cut off the free users for a little while, rebuild the business, and make sure it worked. We want to reach as many people as possible because software is important. It's become difficult to create things in the physical world, so it's important to make it easy for people to build in the virtual world and have access to creation. But there are legs to that journey.Jake [00:10:30]: You can see divots in the charts. If you follow between 2025 and 2026, it's either summer or winter. People go on holiday with family.Swyx [00:10:50]: It affects that much?Jake [00:10:51]: Yeah. It's kind of B2C and kind of B2B. People are shipping constantly, then they stop. Our activation curve now shows more people activating on weekdays because we have more business users, so it smooths out over time.Agents as the New Interface to DeploymentSwyx [00:11:17]: Was there a point where you started prioritizing AI development or agent development?Jake [00:11:24]: We've prioritized agentic as a top-of-funnel thing. Over the last six months, we've deeply prioritized agentic as a mechanism to build and deploy things because we believe the curve is so steep and that is how people will build and deploy software.Jake [00:11:42]: It almost fundamentally doesn't matter whether this is dot-com or not because we're all on the internet anyway. If agents are going to deploy a bunch of things and we hit an inference wall at some point, we'll fix those problems. The dominant species over the next 10 years is that we've moved from assembly to C to C++ to JavaScript to words. You're going to need to close that loop.Swyx [00:12:13]: When you say this is dot-com, did you mean buying the domain, or the general case?Jake [00:12:17]: I mean the dot-com era, when companies had a huge run-up because people understood the internet was important. Then they hit bottlenecks, fundamental laws of physics, math didn't work, and everybody came back down to earth. But it didn't matter because the internet became so impactful. If you operate on a long enough time horizon, you should build these things anyway because you can see where it's going.Jake [00:12:45]: That's where I think a lot of agent stuff is. You get to a point where you're running thousands of agents in parallel. What is the inference cost? What is the compute cost? How do you make that efficient? How do you coordinate all this? We have issues coordinating humans; we don't even have good tooling for that. Now we have to figure out how to get agents to coordinate, safely version changes, and know when to raise their hand for someone to intervene. Otherwise it becomes an interrupt factory.Railway's Infrastructure Thesis: Network, Compute, Storage, and MetalSwyx [00:13:19]: Let's go right into the technical side. What are the core infrastructure or architectural beliefs of Railway that allow you to do what you do?Jake [00:13:29]: The primitives matter a lot for us. We need network, compute, storage, and orchestration around it. You need control over a lot of those things. We've talked a lot about how we don't really use Kubernetes because we want higher-order control to place workloads in very specific places.Jake [00:13:48]: The reason is that you have to be very efficient with agents: memory reuse and all these other things, or you're going to massively blow up your cost structure. Being able to rack and stack your own servers and build your own metal unlocks performance and cost. Experiences where you're running 1,000 agents in parallel are not massively cost prohibitive.Jake [00:14:13]: Token use and compute use are blowing up. Over time, those things have to get a lot more efficient. You can get a lot of margin to make those experiences solid by building your own metal. That's all in service of offering a differentiated experience to as many people as humanly possible.Swyx [00:14:51]: You have a data center in Singapore.Jake [00:14:53]: Yeah. We have two in every other region now. In Singapore, we're adding a second one in Q3.Swyx [00:14:58]: What's it like? I've never built a data center. Do you go to Equinix and say, “I want some slots?”Jake [00:15:05]: Yeah. Equinix. You basically go and say, “I want power and I want a cage.” They say, “Great, here's what it's going to be.” You rent the cage for a period of time, fill it with racks and servers, and hook up internet to it. That's all the pieces.Swyx [00:15:36]: Then you handle everything else.Jake [00:15:37]: You handle everything else.Swyx [00:15:39]: What's the math versus clouds doing it for you?Jake [00:15:43]: If we rented in the cloud, our payback period when we go to metal is about three months.Swyx [00:15:50]: Which is crazy.Jake [00:15:51]: It's nuts. That's four years of depreciated hardware. You're going to see a lot of this compute crunch because hyperscalers are buying up a lot of stuff. We're working directly with OEMs, resellers, and people building these machines: Supermicro, Dell, and others.Jake [00:16:11]: Upstream, there's a bunch of supply pressure. When we raised our last round, between deploying capital for servers and now, the amount of money we've raised is less than the amount of money we have in the bank plus the value of the servers because the servers have appreciated as RAM has gone up. It's nuts how valuable hardware has become.Jake [00:16:50]: If you look at hyperscalers, they deployed around $80 billion of capital expenditures this year, and next year will be more. That's a massive infrastructure build-out. You look at that and think it's crazy that they're spending way more than the Manhattan Project. But if every person is going to run dozens or hundreds of agents in parallel, you have no conceptual idea how much compute is required to make that experience happen, even if you're deeply efficient and sharing resources. And that doesn't even count inference.Swyx [00:17:22]: How do you plan the build-out? The growth chart is so vertical. Are you usually at 100% utilization as soon as racks are live? How far ahead are you planning?Jake [00:17:33]: We still maintain cloud presence for bursting. We work with AWS, GCP, and a few other clouds. We can rent, and then the moment we get space or power, we compact those workloads off the cloud. We started on the clouds, then built a system to migrate to our own metal. There's nothing that says you can't continually do that again, and that's exactly what we do. We never want to be compute constrained.Jake [00:18:09]: At the start of the year, we actually became compute constrained because one upstream provider wasn't able to give us quota at the rate we needed, and the hardware was slower. I spent a weekend rebuilding our entire network overlay so we could straddle five clouds: Oracle, AWS, ourselves, GCP, and one other one. We can do more than that now.Jake [00:18:38]: We got into a spot where we were trying to pack instances tight because we couldn't get enough compute. That led to a few reliability issues, which are now past us. I made a tweet pointing out that it's becoming harder and harder to acquire compute at the rate these models need to acquire compute. We got bit by it.Swyx [00:19:15]: How do you think about pricing knowing you might not have your own metal available at all times? Are you pricing assuming you need extra margin if you end up going into the cloud?Jake [00:19:26]: Because we've built out our metal data centers, our margins on metal are around 70%. We can deeply subsidize the cloud business if we want to scale at a reasonable rate. We have a few levers: metal, which makes the margins; cloud burst; debt to buy servers; and venture capital. It's an interesting operational problem: how much cash do we have, how much should we raise, how quickly can we deploy it, and can we scale revenue as quickly as we scale compute?Jake [00:20:05]: If we continue making it trivially easy for people to build and deploy, then the faster we close that loop and the more operationally excellent we are with capital, the faster the business can scale. It's almost a straight linear deployment rate.Financing Infrastructure: Hardware Debt, VC, and Operational LeverageSwyx [00:20:20]: I think infra startups raising debt is a tool people don't utilize enough or know enough about. What can you tell us about that? Is it secured against your CPUs?Jake [00:20:32]: It's secured against our hardware.Swyx [00:20:37]: What rates do you get? Who are the lenders?Jake [00:20:39]: We pay prime plus a spread, and we can refinance any of the debt as rates go down. The terms are pretty good. The unfortunate thing is that Twitter has no nuance, so people say, “Venture debt bad.” But as with all things, there are specific tools and areas where you can be deliberate instead of using one tool as a hammer. Venture capital is not the hammer for everything. You have to explore and figure out what works.Swyx [00:21:12]: VC is usually the most expensive financing you can get.Jake [00:21:15]: Yeah. I also think people think about VC incorrectly from a capital-raising perspective. Most people think, “How do I raise as much money as possible from whoever is probably the best I can get at that time?” That's close to right, but what we've tried to do is figure out what unfair advantage we can buy with that equity.Jake [00:21:34]: It's the most expensive equity you're going to give away at that point in time, assuming the company keeps getting better. How do you use it to work with someone stellar who complements you? In the seed stage, I had never started a company. Ray Tonsing had good advice, and I could text him all the time. He was really fast. Awesome.Jake [00:22:01]: Then with John and Erica at Unusual, they said, “You roughly know what you're doing building a product. We'll mostly leave you alone and be available for advice.” Amazing. Then we got to Series A and the business was an operational tire fire because we didn't know how to scale a business. Work with Erica, and Jordan is over at Redpoint, so bonus.Jake [00:22:28]: Now we've raised from TQ and FPV as we're moving into enterprises. Every step of the way, we've asked: who can we partner with at this specific time to unlock the next section of the journey? I don't know enterprise sales. As an engineer, I can eyeball what features we might need, and we have wonderful people internally who can help. But you want boardroom dynamics where everyone is aligned and asking, “How do we win this?” instead of bickering about strategy.Data Centers in Space and the Physics of ComputeSwyx [00:23:31]: You had a tweet about data centers in space. Why no data centers in space?Jake [00:23:37]: It's not “no data centers in space.” My hot take is that I think it is solvable. I've just never seen anybody solve it.Swyx [00:23:49]: You said, “How are you going to dissipate that much heat in a vacuum?” You're making a physics claim.Jake [00:23:55]: I haven't seen anybody prove how you're going to dissipate that much heat in a vacuum. It doesn't mean it's not possible. It just means nobody has brought it up yet.Swyx [00:24:05]: Astrophage.Jake [00:24:06]: I don't know what that is.Swyx [00:24:07]: The Martian thing. Okay, you're very logical.Jake [00:24:09]: It could work. A lot of people are putting the cart before the horse. They say, “We're going to put data centers in space.” Okay, but how? “We have time to figure it out.” It's like in The Martian where they ask how they're going to intercept something and say, “We'll figure it out.”Swyx [00:24:36]: Making a bet on human invention is weird because you blind trust that it can be solved. But with physics, there are first-principles bounds you can put on it. Maybe not. Maybe you're asking to travel time or break a fundamental thermodynamic law.Jake [00:24:57]: I don't know how VCs do this either. How do you know what's not possible and a grift versus what's possible but sounds completely insane? “We're going to put data centers in space.” Coin flip as to which it is, and I guess you'll know in 10 years. That's one cycle.What Agents Need: Versioning, Observability, and 1,000x ScaleSwyx [00:25:23]: Moving back to agents. The branching, fast spin-up, and orchestration you do feels like pre-work that happened to be exactly what agents want. What do agents want differently than humans?Jake [00:25:37]: They want the ability to version things. It's not that different; it materializes slightly differently. Agents want a way to test changes incrementally. Engineers have feature flags. Is there a reason agents can't use feature flags? I don't think so.Jake [00:25:54]: They want version control. Can we use Git or not Git? That one is up in the air. I think something outside Git will emerge for how we version these things over time. They need observability. You need to query what happened, when it happened, which steps failed, traces, logs, metrics, and all the rest. They need network, compute, and storage. They need to write files, save files, iterate on files, and snapshot file systems.Jake [00:26:25]: A lot of what humans needed is in line with what agents need. Branching and forking are not different; we're just moving 1,000 times quicker. It can look like you need something massively different, but what you need is something massively better than what existed. You need orchestration massively better than Kubernetes. You need networking probably better than Envoy. It goes all the way down the stack.Jake [00:26:55]: If the workload profile doesn't change so much as it gets massively compressed because you need thousands of these things, what assumptions change? etcd is going to melt. You need to replace it with something. You can go all the way down the stack and say, “That part has to change, that part has to change, and that part has to change.”Jake [00:27:19]: The interesting thing about the super-exponential curve is that you have to build systems where you can rip out those parts at any time because a new bottleneck might emerge. You get good at parallel agents, and a different part of the system breaks. So it's similar to what humans needed, but at 1,000x scale.Jake [00:27:55]: How do you do code review in the age of agents?Swyx [00:28:00]: You throw more agents at it.Jake [00:28:01]: You don't. But then who reviews for CVEs and all these other things?Swyx [00:28:07]: More agents.Jake [00:28:08]: And that's how we hit the inference wall. You can continually throw agents at the problem, but I think there's a limit to the number of agents you can throw at a problem.CLI, Agent Handles, and Closing the LoopSwyx [00:28:24]: You already had a CLI before it was cool. How is the shape of what you're exposing changing, if at all?Jake [00:28:28]: CLIs have always been cool. The CLI changes because we think about how to give Claude, Codex, ChatGPT, or any model a handhold.Jake [00:28:50]: A CLI is a single command: deploy, get logs, and so on. Things that were prohibitively annoying to humans are not annoying to agents. They're nice. If I handed you a CLI with 40 arguments and 600 flags, you'd think, “I'm never going to use all of this.” But if you hand it to an agent, it says, “This is excellent. I have so many handles to work with.”Jake [00:29:24]: If you're going to expose things to agents that way, you want as many handles as possible where they can get information, query dynamic information, and close the loop quickly. Most problems right now are about how to close the loop as quickly as possible. Where does the agent get stuck, and how can you remove that?Jake [00:29:49]: Telemetry is important. If you can tell where the agent gets stuck from the CLI and say, “12% of people deviate from the happy path because of this, and now I add this argument and drive it down to 2%,” you massively increase the rate of loop closure.Jake [00:30:03]: That's how we think about not just the CLI, but every point in the dashboard. It's a user journey: I hear about Railway. I get something deployed. I get my first green build or aha moment. I see an endpoint, logs, whatever. Then I iterate. The iteration loop is indefinite. The user wants to deploy a new thing, a Postgres instance, change code, and keep iterating.Jake [00:30:36]: If you focus on the iteration loops and what's blocking them from closing quickly, one thing we say internally is: you never want to be waiting on compute anymore. You always want to be waiting on intelligence. If you're waiting on compute, there's a bottleneck that needs to be destroyed because eventually that bottleneck becomes so large that another workflow emerges to change it.Jake [00:31:04]: We've built a product where you push code, build it, and so on. But I fundamentally believe the push-pull loop is going away. We'll get to a point where you make a small change in production, that change is versioned across your infrastructure, you're working alongside copy-on-write versions of your database and infrastructure, and then you merge it in and it's instantaneously live. That's the holy grail of loops. The push-pull-rebuild thing is a point of friction that we're removing entirely.Canvas as Output: Dashboards, Context Anchors, and HyperstructuresSwyx [00:31:43]: It's incredibly fast. If anyone hasn't tried it, that fast feedback is great. My hot take is that Railway was famous for its canvas, which visualizes your infrastructure and lets you manipulate it visually. But that was for humans. For the next phase of growth, Railway CLI is more important than canvas.Jake [00:32:05]: The canvas is funny because it's a mechanism to show changes over time. You're right that previously we used it a lot as an input. Moving forward, its goal is more like an output. You would go to the canvas, make changes, see them, and watch your infrastructure evolve. Now agents have access to the CLI and can make those changes. So the canvas becomes an output: what information does the human need at this moment to make suitable decisions about control requests? Do I approve this or not?Jake [00:32:57]: It also has to be an anchor for your context, a port in the storm. Think of it like layers in a file system. You start with a project, then drill down into services, then into a function or code, because you want to represent the entire thing not just in your head, but in the canvas. Other people can share that representation, think on the same wavelength, and move quickly.Jake [00:33:33]: A lot of organizations get in trouble as they scale because all the context lives in someone's head. “How does this microservice work?” “I have no idea; go ask this person.” Then you have whole categories of products built around context discovery. A lot of that melts away if you have a solid hierarchy and can infinitely nest services, code, context, and everything else all the way down. That's what lets you build these structures over time.Jake [00:34:18]: It's also what lets us build what I've called hyperstructures: things that are way bigger. You look at the Golden Gate Bridge and ask, “How did we build that?” There's a meme that we lost the technology. To some extent, yes, because the coordination that built those things evolved and changed. We lost some of the art of building structure as we jammed everything into Slack.Swyx [00:34:52]: But you jam everything in Discord.Jake [00:34:53]: Same point. It doesn't matter. It's message passing and interrupts, message passing and interrupts.Swyx [00:35:00]: So you're arguing there should be something better and more structured than Slack?Jake [00:35:04]: Yeah. For sure. I think Slack is awful, and Discord is awful too.Central Station: Context Routing, Support, and Incident ClustersSwyx [00:35:09]: This is the equivalent of my mom test. What have you done that has your solution to this?Jake [00:35:15]: Internally, we've built a tool called Central Station that aggregates all the context from our users. Every piece of feedback, every customer support item, everything gets aggregated into clusters. If an incident is brewing, we can determine how many users are affected and break off a discussion based on that.Jake [00:35:40]: That is more helpful than long-running channels where you're trying to decide which channel to put something in. If you can dynamically aggregate information and dynamically route it to the right person based on context, it works better. We know internally that these four people are close to networking. If we see a networking thing, we can drill it down to those four people. If it's with this part, we can look at the commits. This is no longer a manual process internally.Jake [00:36:13]: If you go to station or help.railway.com, that's why we built it. We wanted to scale with a massive amount of leverage by aggregating feedback.Swyx [00:36:27]: This is built in-house?Jake [00:36:28]: Yep.Swyx [00:36:29]: I remember helping out on this one with Angelo in 2023. You scale a lot with a very small team.Jake [00:36:38]: Yeah. We're about 10 times bigger now.Swyx [00:36:40]: You have your full developer code here? Very cool.Jake [00:36:44]: If you go to railway.com/stats, we expose this as a pub-sub-able thing. It's all real-time metrics. There's a way to get it as JSON somewhere if you care.Jake [00:37:01]: We're big on trying to build everything in public and talk about what we're working on. We've had issues in the past, and we'll say, “Here's how we're fixing these things.” We've gotten compliments and flak for incident reports. We're always trying to make them better and talk with people.Incidents, Disclosure, and Progressive RolloutsSwyx [00:37:20]: You had a big one recently. I liked that it was scoped to 3,000. You presumably used Central Station. Talk through what happened and how you address it internally as a team.Jake [00:37:38]: Internally, this one really sucked. It had to do with an upstream provider that didn't do the behavior it said it documented, which is unfortunate given they wrote the RFC for how the behavior should work. We rolled those things out, and Central Station caught it initially when a couple users said caches weren't invalidating. We turned it off immediately.Jake [00:38:03]: When you roll out to a large user base of three million people, you get a lot of disparate behaviors. We tested in staging and had tests, but we hit an edge case. We've hardened those systems, and now we can make that better. But it was a tough one.Swyx [00:38:39]: I always wonder how private disclosure is supposed to work if people find an issue. Are they supposed to contact you first? When you run a platform, these things will happen. What channels should people pursue to quietly resolve it before it becomes a bigger incident?Jake [00:38:59]: There's responsible disclosure. We err on the side of over-disclosing and letting you know something is wrong versus having your provider gaslight you. We've erred on sharing those things more publicly, even if they impact a small subset of users. That's a decision we've made internally. We have four values. One is honor. The honorable thing is to notify people to the widest degree at which they may have been affected or there was an issue, and then confront it head-on: why did it happen, what can we do better?Swyx [00:39:45]: Not the whole user base. That's because of incremental rollouts and other things?Jake [00:39:50]: Yeah. Progressive rollouts.Swyx [00:39:54]: That should be the norm at all large platforms.Jake [00:39:58]: It should. A variety of companies do this. There's the quote that Meta runs 10,000 different versions of Meta. To our earlier point about agents, they need the same thing. They need shadow traffic and all these other things. We've built so much ceremony around production being sacred that we need to make it trivially easy to test different behaviors in a safe environment. Then you can make mistakes in a safe environment.Safe AI SRE: Customer Agents, Forked Environments, and Production ParityAlessio [00:40:30]: Do you see a world where these things get automatically caught, not necessarily by your agent, but by your customer's agent? The cache invalidation issue seems easy to check if you know to look for it.Jake [00:40:44]: It's hard because to determine it, we almost need to hook into your observability infrastructure. That's why we have the template loop on the platform: so you can roll things out progressively. You can roll out to Johnny Vibe Coder initially, or push a shard that someone consumes at their own leisure. Or you can roll it out over weeks: 0.1% of people, 1% of people, early adopters, then all the way up. That's the non-deterministic version control we talked about earlier.Jake [00:41:30]: I believe that's where most things should go, because most companies end up building staged rollout systems in-house. It's the same thing built again and again at every company. There's a massive opportunity to consolidate developer debt.Alessio [00:41:45]: You should have a free tier. Model providers give free tokens if you let them use the data. You could give free compute if someone is the number-one shard that goes out and lets you plug into their observability.Jake [00:41:55]: We do that. That's why we talked about the impact on 3,000 people. We start with lower-impact people. Larger companies on the platform are last to receive those rollouts so they have a version of the platform that's deeply stable.Alessio [00:42:16]: I have three services, so I'm sure I get the first rollout. You can nuke my thing at any time. There are all these SRE agent companies. Observability people also want agents that fix upstream problems. You have your own agent in the canvas now. How do you see that playing out?Jake [00:42:39]: It's the stacking entropy problem. If you don't have primitives to make iteration in production safe, it becomes difficult. If you're an observability provider saying, “Here's the fix to this error,” assume 80% are good and make sense. But in the last 20% long tail of complex issues, if you let somebody stamp it, you create an opportunity for an incident.Jake [00:43:08]: That's why forked environments are important. People have staging, but it always drifts from production. You need primitives, workflows, and experience built first-party on the platform so you can fork any service at any point in time.Jake [00:43:33]: I think of the canvas as a sheet of transparency paper. The agent is a little guy you push up into the canvas. It should say, “I need to copy that service and that service so I can test these two things.” It gets a read-only copy of production. Anything that's PII gets marked as a transform when we clone the database, create a copy-on-write version, or read from it. Then the agent makes changes and asks, “Does this actually work?” as close to production as possible.Jake [00:44:22]: That's how close you have to be, or you get massive drift. The system becomes unstable. You see this with massive systems built on Docker for local, Kubernetes for production, and a specific thing for something else. That complexity slows developers and becomes unstable at scale, making it hard to iterate. We want to compress that way down and say, “As close to prod as possible is where we want to be.”From AISRE Skeptic to Agent BelieverSwyx [00:45:00]: I was texting Erica for questions, and she says you were originally not a believer in AISRE. Have you come around on it?Jake [00:45:10]: I flipped, but I'm still not a believer in AISRE if you don't have the primitives to make it safe. If you unleash AISRE on production infrastructure without safe primitives for copying volumes and making sure things are fine, it's going to nuke your production database. It's not a matter of if, but when. I'm a big believer in making those loops safe.Jake [00:45:33]: I was a deep AI skeptic until 2023. In 2024, I thought, “Maybe I can roughly make this thing do it.” In 2025, I thought, “Now I can hold this.” Over winter break, everybody came back saying, “It's almost impossible to hold this.”Swyx [00:46:01]: Did you see this on the Claude docs? CloudBot? OpenCloud?Jake [00:46:06]: It's gotten to a point where it's harder to hold it wrong than to hold it right. There's a scene in Avengers where Vision picks up Thor's hammer and says it's terribly well-balanced. It self-balances and works well. I'm a deep believer at this point that this will be the dominant species: assembly, C, C++, JavaScript, words.Swyx [00:46:35]: It feels like a big jump.Jake [00:46:37]: It is. But it's not like you abandon CPU-based discrete logic and move straight to fuzzy logic. You need both. Your skills should call code or applications or some static structure. You can use skills to distill what the procedure should be or how the code should act.Jake [00:47:02]: I'm coming to a thesis: you need three points. You need a clear spec defining the system, the code, and the tests. When you say it out loud, if you've been in engineering long enough, you're like, “Of course. That's an RFC, tests, and code.” But they all matter. Having them together lets them reinforce each other: the spec and tests match, but the code doesn't, so reconcile it. Or the tests and code match but the spec doesn't, so reconcile that. That's the iteration loop.Jake [00:47:41]: That's why you're seeing people talk about software factories, docs, and reconciliation. Some of that is architectural astronomy if you don't implement it, but that loop is where most things will end up.Swyx [00:48:07]: For listeners, we've been talking about this on the pod for three years: the holy trinity of specs and tests. Itamar Friedman from Qodo is the reference if people want to look it up.Self-Modifying Infrastructure and the End of Push-Pull-RebuildSwyx [00:48:18]: One thing I want to mention on the OpenCloud idea is self-modification. I don't know how Railway would support it, but I have my OpenClaw, and I just tell it it has the Railway CLI and can do whatever. In theory, whatever capabilities or new infra it needs, it can call the Railway CLI, provision it, and add it to itself. The agent can modify its own infra.Jake [00:48:45]: It's nuts. I have a loop set up where you put the Railway CLI on top of something that runs on Railway. You're authenticated as whatever the current box is, and you can make any changes to it. Then you call Railway deploy, and it deploys itself.Jake [00:49:04]: It's like: “I need to spin up this instance of this environment. I already exist in this environment. Excellent, I have access to a Postgres instance now.” That's where we want to go with agentic, self-replicating infrastructure. That's your loop: iterate in production. You continue making changes. If it works, merge it upstream. If it doesn't, throw it away.Jake [00:49:37]: How do you make throwaway copies trivial to spin up and super cheap? The era of “I have an AWS instance with four vCPU and 16 gigs of RAM” is going to get destroyed. If you do that for agents, you need a thousand of those machines. It's prohibitively expensive compared with what we've spent a ton of time figuring out: the atomic unit of deploy, whether you call it isolates, sandboxes, or something else. Only pay for what you use, spin up instantaneously, and close the loop as quickly as possible.Jake [00:50:15]: If the system can self-replicate safely and say, “This is my environment, I'm making these changes,” it can come back with, “Does this look good? This is a new state of infrastructure given this prompt. I think I've solved it.” Then you go back and say, “Actually, it looks different.” It does the loop again. Then you say, “Cool. Apply.”Swyx [00:50:38]: That's retroactively obvious, which is the most useful kind. Any other comments on agent deployment on Railway?Jake [00:50:51]: It's getting better every day. I'm on X or Twitter. You can always yell at me about the parts not working as well as they should, because plenty of things should work way better.The New Serverless: Stateful, Long-Running, Pay-for-What-You-Use LinuxSwyx [00:51:04]: At this stage, when people want massively or embarrassingly parallel compute, they usually talk serverless. I feel like there's a new serverless compared to the previous five years of serverless. You're in that new bucket. Do you have comparisons or philosophical differences you want to call out?Jake [00:51:31]: It's somewhere in between. It's the ability to run stateful, long-running workflows or executions.Swyx [00:51:42]: Vercel has Fluid Compute, Cloudflare has some container thing, Google has App Runner and others.Jake [00:51:55]: That's where everything is roughly going, and it's why we've been working on this for six years. We believe users need access to a computer: a box that speaks Linux. They need to deploy what they want. Other systems change the surface area of what you can build. For us, users need a computer and need to deploy anything they truly want. That's why we've focused on the primitives: network, compute, storage. If we give you those and expose them so you can run things indefinitely, that's where we believe it's going.Jake [00:52:43]: Twitter has no nuance, so everyone says “servers” or “serverless.” It's always somewhere in the middle: I want to run it for a long time, but I don't want to provision the resource statically or pay for things I'm not using. That's been our thesis from day one: pay only for what you use, run it indefinitely, and it is full Linux.Swyx [00:53:12]: That's why I like the naming of Fluid. It's fluid. Flexible.Heroku, Focus, and Carrying the Torch Without Becoming the PastSwyx [00:53:18]: Another milestone is the Heroku official deprecation. You're one of the presumptive new Herokus. “New Heroku” has been a category for as long as I've been in developer tooling. It's finally happening. What was that like? Any behind-the-scenes of, “This is the moment”?Jake [00:53:42]: You have people where you're like, “You were running stuff on here? You, as this company?” It's crazy that names you would know are running on it and now coming to us saying, “We want to move a lot of this off.”Swyx [00:54:00]: Any behind-the-scenes on why Salesforce let Heroku stagnate?Jake [00:54:05]: I can only guess. It's hard when it's not your business. Salesforce's business is to build a great CRM. That's their focus. Then you acquire a compute business as an offshoot. A lot of early Meta people talk about focus. Boz has a write-up about how in the early days of Meta they had no money, so they were forced to focus. Then they turned on the money tree and had no reason not to split their focus.Jake [00:54:52]: But that dilutes your product. You get offshoots where you ask, “Is this the focus of the business?” If it's not core, it languishes. A lot of companies get in trouble when they split focus because they're fighting a multi-front war, not just externally but internally for alignment. Where are we going? What are we doing? What is our purpose?Jake [00:55:24]: If you're Salesforce-built and mission-driven, you want to work on Salesforce. Heroku is off to the side. It's not core to the business. Getting resources, budget, focus, and alignment internally becomes hard. It was a matter of time.Swyx [00:56:06]: Kudos for them to call it out instead of leaving it unknown.Jake [00:56:12]: Their release was a little odd. They called it out, but they didn't say they were shutting it down. Behind the scenes, I think they issued messages to people saying they should close accounts and that they were going to deprecate and remove things over time.Jake [00:56:30]: It's crazy because some of my first deployment experiences were on Heroku. You start with dragging things into an FTP server, then you try to get a deploy working, and then it's Heroku. It was the on-ramp for us. But the wheel turns. New things emerge. We're happy to carry the torch for a lot of that. But we don't want to be the new Heroku. We want to be the way people build and deploy software, and ultimately the way people monetize software over time.Swyx [00:57:19]: It's still a big crown to be the new Heroku. There are 50 companies that fought for that.Jake [00:57:23]: Everybody is holding some portion of it. We're happy to support people and companies. The platform works differently. The game loop is similar, but we've been dogmatic about where these things are going: primitives, agents, fan-out. Some things fit; some workflows need to change. We have an approximation of Heroku pipelines with the environment system. It's exciting. We've got a ton of people we can support, and it's growing a lot.Temporal, Workflow Engines, and State MachinesSwyx [00:58:12]: I have one more technical question about Temporal. I've sold my shares. You're a power user and one of our earliest customers. I met you through Temporal. You built on Temporal. You have complaints. This may be the most neutral and informed conversation anyone will hear about Temporal without someone working at the company.Jake [00:58:39]: That's fair. I've used Temporal for almost 10 years because of Cadence at Uber.Swyx [00:58:52]: Give people a sense of what Cadence was at Uber.Jake [00:58:57]: Cadence was the precursor to Temporal. It powers trip actions, rides, when you rent a Jump bike or scooter or car. You're running workflows for a period of time and saying, “This ride will run indefinitely until it finishes.” You attach information: you paused in this zone, so add this charge to the bill. When you end the trip, the workflow is done. That experience was powered by Cadence at the time.Swyx [00:59:34]: I used to say it's like programming the entire user journey top-down as one function.Jake [00:59:39]: It's a powerful idea and important. It's also important for the next phase of the agentic journey. You want an agent to do a specific task, be complete or incomplete on that task, and move on to the next thing. You need a way to manage workflows dynamically.Jake [00:59:59]: Temporal was always great in theory, and great when you got it working the way you wanted in production. But it required you to model the entire journey in your head. If you didn't, you could cause issues where replaying the state of the workflow causes non-determinism.Swyx [01:00:25]: Because it works on deterministic workflow history.Jake [01:00:28]: Exactly. I describe it as a jet engine. If you know how to operate it and run it, it's great. But you can't hand it to people trying to build complicated things if they don't have the whole state in their head.Jake [01:00:48]: We run our whole deployment pipeline on top of it. That's a reasonably complicated workflow: pre-commit hooks, signaling, queuing, and all the rest. We ran into the same thing at Uber. As you express a large workflow, it gets more complicated, with more states in the state machine that you have to map back to the workflow.Swyx [01:01:15]: It's a lot of ifs.Jake [01:01:16]: Exactly. At Uber, we built a system for doing the state machine and testing it. We've started to build some of those things here because it's grown heavily. It's not quite love-hate. When it works well, it works super well. But if someone who doesn't have full context puts something into the system that invalidates state or causes non-determinism, or spins off a ton of activities, you have to keep track of underlying SRE knobs like activity slots. Those should scale with memory, vCPU, and so on. It becomes a bear to scale.Swyx [01:02:10]: You need a capable sysadmin running things behind the scenes. If you moved off, what would you do?Jake [01:02:19]: We'd build our own workflow engine. We have a few internally that we've worked on.Swyx [01:02:27]: This is one of those classes of things you typically wouldn't vibe code, but I'm wondering if you can.Jake [01:02:33]: I still don't think you should vibe code it. You still want to run decent tests to make sure it works.Swyx [01:02:39]: Timo didn't invent that from scratch either. There are libraries you can run. On top of that, it's just a state machine that you have to map out. Ultimately, you define the instructions you want and run them through a state machine.Jake [01:03:00]: It's very doable. Workflow stuff is interesting. Restate is doing neat stuff here.Swyx [01:03:10]: You're tied into JavaScript. Are you a JavaScript maxi?Jake [01:03:13]: Internally, we have TypeScript, Rust, and Go. We don't add more languages. Actually, we have a little C because we write BPF code and hooks. But those are the languages.Swyx [01:03:28]: Is this for sidecars?Jake [01:03:32]: No. It's for the networking stack, volumes, and things like that. We use TypeScript a lot because it powers the dashboard, but we're moving a lot of workflow stuff off the dashboard stack and into the infrastructure stack.Railpack, Nixpacks, and Content-Addressable FilesystemsSwyx [01:04:00]: Cool. Any other technical infrastructure stuff? Railpacks?Jake [01:04:07]: We built an engine for determining dependencies based on source code. It's called Railpack. We built the first version, Nixpacks, on top of Nix, and then we moved.Swyx [01:04:17]: People have been trying to get me to adopt Nix and NixOS for four years. Is it ever going to be a thing?Jake [01:04:23]: I don't know. We're excited about it, but it has pain points. Think of it as a stack of versioned binaries at specific slices in time. If you want version X and version Y, you bloat the package space, which blows up image size and makes real-world workloads difficult.Swyx [01:04:53]: But you content-address it and cache it. In theory, there are optimizations.Jake [01:05:00]: In theory, yes. But with a large enough user base and disparate enough machines, you run into a problem Meta described in the XFAAS paper, their internal serverless system. It becomes difficult at scale unless you break out specific runtimes.Jake [01:05:24]: We didn't want to do that because we wanted to truly allow you to deploy anything. That was our initial thing with Nix. But we've moved toward interesting work around content-addressable file systems that can lazy-load anything from any point and page it into memory.Swyx [01:05:48]: Amazing.Jake [01:05:49]: The future is very bright. It's crazy, and it's going to be nuts.Coding Agent Spend, Roadmaps, and Token ROISwyx [01:05:54]: Founder journey stuff?Alessio [01:05:56]: Your cloud usage: you tweeted you're going to spend $300K this month?Jake [01:06:01]: I think we got to $200K.Alessio [01:06:02]: Coding agents?Jake [01:06:03]: Yeah.Swyx [01:06:04]: Across the company?Alessio [01:06:05]: You only have 35 people, so I'm sure they're not all spending $10K a month. What's the distribution?Jake [01:06:10]: I think I'm at about $25K. We have power users all the way down. We came back from winter break, and I basically said, “If you're writing code by hand, you're doing this wrong.” The tools are good enough now that you can move extremely quickly. There are issues and pain points, but you should be reviewing the code you are writing instead of writing it by hand.Jake [01:06:40]: Architectural patterns matter more now than ever, but you shouldn't spend your time generating code you would write. If you know how to write it, ask the agent to write it and reconcile it until it looks like you would have written it yourself.Jake [01:06:58]: People misconstrue my propensity to push people toward agents as connected to our growth and some reliability bumps. They're not necessarily related. The tools are good enough to move extremely quickly and build things way larger than you could before.Jake [01:07:19]: To the earlier point about cooling data centers in space: I don't know. But with software, you can ask, “How would I build block storage from scratch? How would I do these things?” I have ideas because I have history and have read papers. Let me work them out and build massive test benches with thousands of tests, because those are now free to author. If you're not using AI systems to speed-run your roadmap and reconcile your existing system onto the future, you're missing a large point of what's happening.Alessio [01:08:12]: What's the path to spending $3 million a month? Is it bound by ideas and things customers can absorb?Jake [01:08:19]: For most companies, it's bound by deployment at this point. That's why we've seen a massive boom in users and companies, from Fortune 50s down, asking how to get developers to move faster. You'll probably hit your CFO before any technical limits because they'll look at the eye-watering amount of money spent on tokens. Inference costs have to come down, but we're inference constrained now. There will be price discovery around what makes sense for an org to adopt.Jake [01:09:06]: I think you'll end up with the F1 driver concept. If someone is really adept at these things, it makes sense to put them in a $3 million car. If they're not, it probably doesn't make sense. You'll take a few people and say, “You can drive the F1 car. We need to go in this direction. Figure out if it works and prototype it.”Jake [01:09:33]: We've done some of that and vastly accelerated our roadmap. We thought we'd ship something in a few years; now we can probably ship it in a few months because we validated it and don't have to build it incrementally. We can skip steps and move toward our vision.Alessio [01:09:58]: A lot of people are realizing the roadmap doesn't always have a business impact, so they say tokens are too expensive. But if your roadmap were built to make more money by the time you built it, you'd have token pricing for it, the same way you do with sales. You'd spend a billion dollars on sales if you knew you would get $2 billion of revenue.Jake [01:10:19]: Exactly. A naive way to measure this is the percentage of tokens that end up in production. If you can measure impact because those tokens end up in production, that's awesome. But the burden of proof will rise. Internally, we have a growing number of pull requests that haven't merged. The question becomes: how do you get this into production? It's about how quickly you can build and deploy software, which is exciting because that's our whole thing.The SDLC Shift: Prompt Requests, Feature Flags, and Safe RolloutsSwyx [01:10:56]: The SDLC is changing. One thesis is that the pull request is dying. It's going to be the prompt request. Beyond that, code review is also kind of dying if you have all the other systems in place. What else is changing about the SDLC?Jake [01:11:19]: The AISRE and the tools to make it happen. AISRE is pie-in-the-sky aspirational. What does it take to get an AISRE? What tools do you need to build?Swyx [01:11:32]: You should expose your tooling to customers at some point. The Central Station command center.Jake [01:11:39]: We have it for template maintainers. Template maintainers can deploy and maintain templates, and they get feedback. We're going to expose those things incrementally.Swyx [01:11:51]: Clustering around incidents. Everyone has a version of that, but I don't think anyone has solved it.Jake [01:11:56]: I won't say we've solved it internally, but it's gotten so good that we can see incidents forming pretty quickly. At some point, those will be things either someone else builds or we build. We've always built things purpose-built for us. If it makes sense to make it useful for users, monetize it, or turn that loop into a profit center instead of a cost center, we want to do that.Jake [01:12:28]: Pull request is definitely dying.Swyx [01:12:29]: Do you do first-party feature flagging and incremental rollout stuff?Jake [01:12:34]: We have a feature-flagging engine we built internally and will eventually roll out.Swyx [01:12:38]: I don't see it as a user. How come you didn't give us what you have?Jake [01:12:43]: We have to beta test it. We care a lot about the quality of the things. There's plenty we've used internally that doesn't make it all the way through the journey because it fails. It works for one service but not multiple services. We'd have to build it for multiple services and know that if we released it, we'd rebuild it again and again. Some things are worth that, but many inform the roadmap.Jake [01:13:18]: We don't want to dilute the experience by saying, “This works, but only for this service,” unless it's a core initiative. Over the next few months, we'll roll out things that work for a single service, then multiple services, then multiple services across the environment. You have to be deliberate. Otherwise you create broken disparate experiences and support load because people ask how to use the feature.Jake [01:13:52]: It's the earlier expansion and compaction pattern. You expand the company to get features, then compact and smooth them out so the experience is stellar. You told me in the hallway, “It's gotten so much better.” Internally we're saying, “This part really sucks. We need to make it significantly better.”Swyx [01:14:11]: I can attest to that over the last three years watching you build Railway. For listeners, feature flagging is a huge part of Uber culture. So much so that they have too many feature flags and another thing to remove feature flags. Facebook has Gatekeeper. Agents are going to need this. It's fundamental to incremental rollouts. OpenAI acquired Statsig. GPT-5 is routing and flagging through different models.Jake [01:14:56]: It's super important. If the software development lifecycle is going to change because we're doing things 1,000 times faster and 1,000 times more concurrently, what becomes important at scale?Jake [01:15:16]: Before I started Railway, I built a feature-flagging product and tried to sell it. It was an easier version of LaunchDarkly. I ran into a problem: anyone small enough to adopt your technology doesn't care about feature flags, and anyone large enough to need feature flags needs so much scale that you have to build out all the infrastructure. I scrapped it.Jake [01:15:42]: But what is old is new again. Companies are trying to move quickly, but you can't YOLO a vibe-coded thing straight into production. You need to say, “Here's my blast radius, my impact, and I want to shadow it for these users.” Feature flags. You're going to need the tools larger companies built to maintain their structures. Everything gets compressed by 1,000x so everybody can build those structures quickly.Jake [01:16:07]: That's exactly where we are: compressing the software development lifecycle, then expanding it and adding more new things.Cattle, Pets, and Clonable InfrastructureSwyx [01:16:15]: Another term that comes to mind for newer developers is “cattle, not pets.” People treat production like a pet. It has a name. You baby it and keep it alive. With cattle, you can mass farm, roll out, portion parts out, and kill them.Jake [01:16:37]: I think that might change. You can move toward having pets as long as you have a cloning machine for your pets.Swyx [01:16:52]: Yeah.Jake [01:16:52]: If you can snapshot every single thing at every frame, it doesn't matter if something gets obliterated because you have a snapshot of it. The things we've built right now are designed to block changes from the hermetically sealed DevOps line. You have to write a Dockerfile because you nee

Italiano ON-Air

Italiano ON-Air

Play Episode Listen Later May 20, 2026 5:33 Transcription Available


Può uno scherzo di matrimonio finire in un duello a colpi di spada... e con una citazione letteraria? In questa puntata, Alessio e Katia partono da un bizzarro aneddoto familiare per portarvi alla scoperta del romanzo più famoso (e a volte temuto!) della letteratura italiana: I Promessi Sposi di Alessandro Manzoni. Se volete capire davvero la cultura e la lingua italiana, questo è l'episodio che fa per voi!

Strategia Digitale
Podcast a Scuola: gli Studenti Raccontano CantiPod - dal Cantico alla comunicazione della Gen Z

Strategia Digitale

Play Episode Listen Later May 20, 2026 18:23


Come si fa a far podcast a scuola? In che modo il podcast può essere uno strumento educativo a supporto della didattica? L'insegnamento del podcast nelle scuole elementari, medie e superiori può rappresentare un'opportunità di sperimentare con un media digitale dall'identità fortemente analogica e di scoprire un'alternativa più consapevole ai social media?Scopriamolo con gli studenti della classe 3B Tecnico Grafica e Comunicazione dell'Istituto Silvio D'Arzo di Sant'Ilario d'Enza (RE), in particolare con Filippo, Alessio e Alex, che guidati dal professor Biagio Tornatore, docente e formatore certificato ASSIPOD Education, hanno creato CantiPod - dal Cantico alla comunicazione della Gen Z ( https://pod.link/1885350144 ), vincitore del premio Podcast a Scuola 2026 nella categoria Scuola secondaria di secondo grado.

Just End The Suffering
562-Knicks Advance, NFL Schedule Release And Marvel Catchup

Just End The Suffering

Play Episode Listen Later May 16, 2026 106:40


It's time to tip off a jam-packed episode of the Just End The Suffering podcast! Host Mike Phillips (⁠⁠⁠@MPhillips331⁠⁠⁠) kicks off the show by breaking down the new 2026 NFL schedule (1:17) with Nick Fraietta (⁠@NickFry_9⁠), his co-host from The Sky Guys podcast. Mike is then joined by Tom Bocchino of the Sorry To Interrupt (⁠⁠⁠@SorrySports⁠⁠⁠) podcast to recap the Knicks' sweep of the Philadelphia 76ers (38:46) and preview their Eastern Conference Finals matchup. Mike then wraps the show by catching up on the latest from Marvel (1:11:46) with Nick D'Alessio.Subscribe to the Just End The Suffering podcast on ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Apple⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠, ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Amazon⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠, ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠TuneIn⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠,⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ and⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Spotify⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠!Subscribe to ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Mike Phillips's channel⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ on YouTube!Subscribe to the ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Sorry to Interrupt podcast⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠!

Fluent Fiction - Italian
Melodies of Love: A Gift from the Heart in Springtime Firenze

Fluent Fiction - Italian

Play Episode Listen Later May 7, 2026 15:54 Transcription Available


Fluent Fiction - Italian: Melodies of Love: A Gift from the Heart in Springtime Firenze Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-05-07-07-38-19-it Story Transcript:It: Il sole di primavera illuminava le strade di Firenze, mentre Alessio camminava nervosamente verso l'ospedale.En: The spring sun illuminated the streets of Firenze, as Alessio walked nervously towards the hospital.It: Quel giorno, la città era un'esplosione di colori: i giardini erano in fiore e l'aria profumava di freschezza.En: That day, the city was an explosion of colors: the gardens were in bloom, and the air was scented with freshness.It: Alessio aveva ricevuto la notizia che Chiara, sua sorella, era stata ricoverata.En: Alessio had received news that Chiara, his sister, had been hospitalized.It: La futura madre aveva bisogno di riposo.En: The expectant mother needed rest.It: Alessio era ansioso.En: Alessio was anxious.It: Voleva fare un regalo speciale per Chiara, qualcosa che dimostrasse quanto fosse entusiasta all'idea di diventare zio.En: He wanted to give a special gift to Chiara, something that demonstrated how excited he was about becoming an uncle.It: Tuttavia, non riusciva a decidere cosa prendere.En: However, he couldn't decide what to get.It: Preso dallo stress, decise di chiedere aiuto a Livia, la sua fidanzata, che era sempre piena di idee brillanti.En: Overwhelmed by the stress, he decided to ask for help from Livia, his girlfriend, who was always full of brilliant ideas.It: Livia lo incontrò davanti a un mercato artigianale della città.En: Livia met him in front of a city artisan market.It: Era affollato, con bancarelle che esponevano oggetti fatti a mano, carichi di storia e amore.En: It was crowded, with stalls displaying handmade items, rich with history and love.It: "Guarda," disse Livia, indicando una bancarella con splendide scatole di legno intarsiate.En: "Look," said Livia, pointing to a stall with beautiful inlaid wooden boxes.It: "Queste sono bellissime.En: "These are gorgeous.It: E ascolta questo!"En: And listen to this!"It: Livia aprì una scatola musicale, e una dolce melodia riempì l'aria.En: Livia opened a music box, and a sweet melody filled the air.It: Era un carillon che suonava una ninna nanna.En: It was a music box that played a lullaby.It: Alessio fu subito colpito.En: Alessio was immediately captivated.It: La musica sembrava una promessa di protezione e amore.En: The music seemed like a promise of protection and love.It: Avrebbe accompagnato il bambino di Chiara nei suoi sogni.En: It would accompany Chiara's baby in his or her dreams.It: Senza esitare, acquistò la scatola musicale.En: Without hesitation, he purchased the music box.It: Più tardi, in ospedale, Alessio aprì lentamente la porta della stanza di Chiara.En: Later, at the hospital, Alessio slowly opened the door to Chiara's room.It: Lei era distesa sul letto, stanca ma sorridente nel vedere il fratello.En: She was lying on the bed, tired but smiling at the sight of her brother.It: "Ho qualcosa per te," disse Alessio, allungando il pacchetto.En: "I have something for you," said Alessio, extending the package.It: Chiara lo aprì con cura, e quando la melodia iniziò a suonare, i suoi occhi si riempirono di lacrime.En: Chiara opened it carefully, and when the melody began to play, her eyes filled with tears.It: "È bellissimo, Alessio.En: "It's beautiful, Alessio.It: Grazie."En: Thank you."It: Il momento era carico di emozione e le parole non servivano più.En: The moment was full of emotion, and words were no longer needed.It: Abbracciandosi, superarono la distanza che la vita a volte crea tra le persone.En: Embracing each other, they overcame the distance that life sometimes creates between people.It: Per Alessio, quel giorno divenne indimenticabile.En: For Alessio, that day became unforgettable.It: Non solo aveva scelto un regalo significativo, ma aveva anche imparato qualcosa di importante: la forza delle emozioni condivise e il valore del sostegno reciproco.En: Not only had he chosen a meaningful gift, but he had also learned something important: the strength of shared emotions and the value of mutual support.It: Firenze, con i suoi fiori e le sue promesse primaverili, aveva donato loro un nuovo legame.En: Firenze, with its flowers and springtime promises, had given them a new bond. Vocabulary Words:the sun: il solethe hospital: l'ospedalethe city: la cittàthe garden: il giardinothe air: l'ariathe news: la notiziathe mother: la madrethe gift: il regalothe idea: l'ideathe girlfriend: la fidanzatathe market: il mercatothe stall: la bancarellathe item: l'oggettothe history: la storiathe box: la scatolathe music: la musicathe melody: la melodiathe lullaby: la ninna nannathe promise: la promessathe love: l'amorethe protection: la protezionethe dream: il sognothe room: la stanzathe brother: il fratellothe package: il pacchettothe tear: la lacrimathe emotion: l'emozionethe bond: il legamethe spring: la primaverathe support: il sostegno

Italiano ON-Air

Italiano ON-Air

Play Episode Listen Later May 6, 2026 6:17 Transcription Available


Ti sei mai chiesto perché i venti in Italia hanno nomi così particolari? In questa puntata di

Fluent Fiction - Italian
Discovering Inspiration and Connection in Cinque Terre

Fluent Fiction - Italian

Play Episode Listen Later May 2, 2026 19:06 Transcription Available


Fluent Fiction - Italian: Discovering Inspiration and Connection in Cinque Terre Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-05-02-22-34-01-it Story Transcript:It: Le onde del Mar Mediterraneo lambivano dolcemente le coste rocciose di Cinque Terre, mentre Alessio camminava contemplativo lungo il sentiero.En: The waves of the Mar Mediterraneo gently lapped against the rocky shores of Cinque Terre, as Alessio walked contemplatively along the trail.It: La primavera era arrivata, portando con sé profumi freschi e cieli tersi.En: Spring had arrived, bringing with it fresh scents and clear skies.It: Era alla ricerca di ispirazione per i suoi dipinti, ma anche di qualcosa di più: una connessione con qualcuno.En: He was in search of inspiration for his paintings, but also something more: a connection with someone.It: Francesca, un'amica che abitava a Vernazza, aveva organizzato un'escursione guidata per mostrare ai viaggiatori le bellezze naturali della zona.En: Francesca, a friend who lived in Vernazza, had organized a guided tour to show travelers the natural beauties of the area.It: Alessio aveva deciso di partecipare, anche se all'inizio era riluttante.En: Alessio had decided to participate, even though he was initially reluctant.It: Le parole di Francesca lo avevano convinto: "A volte, l'ispirazione si trova nelle persone che incontriamo."En: The words of Francesca had convinced him: "Sometimes, inspiration is found in the people we meet."It: All'inizio del sentiero, Alessio incontrò Bianca.En: At the start of the trail, Alessio met Bianca.It: Aveva lunghi capelli castani e occhi pieni di curiosità.En: She had long brown hair and eyes full of curiosity.It: Era arrivata a Cinque Terre in cerca di avventure e nuove esperienze.En: She had come to Cinque Terre in search of adventures and new experiences.It: Solitamente viaggiava con amici, ma questa volta aveva deciso di affrontare il percorso da sola, desiderosa di qualcosa di diverso.En: She usually traveled with friends, but this time she had decided to face the path alone, eager for something different.It: "Piacere, Alessio," disse lui con un sorriso timido.En: "Nice to meet you, Alessio," he said with a shy smile.It: "Sei pronta per l'escursione?"En: "Are you ready for the hike?"It: "Non vedo l'ora di scoprire questi paesaggi," rispose Bianca, facendo un gesto ampio con la mano.En: "I can't wait to discover these landscapes," replied Bianca, making a broad gesture with her hand.It: Aveva un'energia contagiosa.En: She had an infectious energy.It: Insieme al gruppo, iniziarono a camminare, Francesca in testa come guida entusiasta.En: Together with the group, they started walking, Francesca leading the way as an enthusiastic guide.It: Durante la salita, Alessio e Bianca si trovarono spesso a camminare fianco a fianco, scambiandosi racconti di viaggio e risate.En: During the climb, Alessio and Bianca often found themselves walking side by side, exchanging travel stories and laughs.It: Ma Alessio era ancora pensieroso, incerto su cosa cercare davvero.En: But Alessio was still thoughtful, unsure of what he was truly searching for.It: Bianca, d'altro canto, non voleva legarsi troppo.En: Bianca, on the other hand, didn't want to get too attached.It: Le ferite del passato la rendevano cauta.En: Past wounds made her cautious.It: Eppure, con Alessio, sentiva qualcosa di diverso.En: Yet, with Alessio, she felt something different.It: Mentre avanzavano lungo il sentiero costiero, gli splendidi panorami sembravano riflettere nuovi sentimenti nei loro cuori.En: As they advanced along the coastal trail, the splendid vistas seemed to reflect new feelings in their hearts.It: Finalmente, raggiunsero una sporgenza con vista mozzafiato sul mare.En: Finally, they reached a ledge with a breathtaking view of the sea.It: Il sole iniziava a calare, tingendo tutto di un arancione caldo.En: The sun began to set, tinting everything in a warm orange.It: Alessio si fermò, incantato dalla bellezza del momento.En: Alessio stopped, enchanted by the beauty of the moment.It: Accanto a lui, Bianca restava senza parole.En: Next to him, Bianca was left speechless.It: "In momenti come questi, tutto ha un senso," disse Alessio, rompendo il silenzio.En: "In moments like these, everything makes sense," said Alessio, breaking the silence.It: "Ho viaggiato tanto per trovare ispirazione, ma ora capisco che è qui, davanti a me."En: "I have traveled so much to find inspiration, but now I understand it's here, right in front of me."It: Bianca lo guardò, colpita dalla sincerità delle sue parole.En: Bianca looked at him, struck by the sincerity of his words.It: "Anche io ho sempre cercato lontano," ammise.En: "I've always searched far away too," she admitted.It: "Ma forse, a volte, ciò che cerchiamo è più vicino di quanto pensiamo."En: "But maybe, sometimes, what we are looking for is closer than we think."It: Parlarono a lungo, mentre il sole continuava a scendere.En: They spoke at length, while the sun continued to set.It: Condividevano pensieri, sogni e paure.En: They shared thoughts, dreams, and fears.It: In quel momento, le loro barriere si sciolsero.En: In that moment, their barriers melted away.It: Alessio trovò non solo l'ispirazione per i suoi dipinti, ma anche la compagnia che non sapeva di desiderare.En: Alessio found not only inspiration for his paintings but also the companionship he didn't know he desired.It: Bianca, invece, scoprì nella fiducia un nuovo senso di appartenenza.En: Bianca, on the other hand, discovered in trust a new sense of belonging.It: Mentre la notte avvolgeva Cinque Terre, un nuovo capitolo si apriva per entrambi, intrecciando le loro vite come sentieri tra le rocce e il mare, sotto un cielo stellato.En: As night enveloped Cinque Terre, a new chapter opened for both, intertwining their lives like trails between the rocks and the sea, under a starry sky. Vocabulary Words:the waves: le ondethe shore: le costethe trail: il sentierothe spring: la primaverathe inspiration: l'ispirazionethe painting: i dipintithe friend: l'amicathe traveler: il viaggiatorethe area: la zonathe gesture: il gestothe energy: l'energiathe climb: la salitathe story: il raccontothe laugh: le risatethe panorama: i panoramithe heart: il cuorethe ledge: la sporgenzathe beauty: la bellezzathe moment: il momentothe silence: il silenziothe word: le parolethe thought: il pensierothe dream: il sognothe fear: la paurathe barrier: la barrierathe sunset: il tramontothe trust: la fiduciathe belonging: l'appartenenzathe chapter: il capitolothe starry sky: il cielo stellato

Italiano ON-Air
L'arte di respirare in italiano

Italiano ON-Air

Play Episode Listen Later Apr 29, 2026 4:50 Transcription Available


Ti senti stressato o hai sempre il "fiato sul collo"? Allora fermati un momento e... prendi fiato con noi!

Fluent Fiction - Italian
Artistic Freedom: A Defiant Stand Against Compromise

Fluent Fiction - Italian

Play Episode Listen Later Apr 28, 2026 17:41 Transcription Available


Fluent Fiction - Italian: Artistic Freedom: A Defiant Stand Against Compromise Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-04-28-22-34-02-it Story Transcript:It: Nel cuore della Primavera, la città festeggiava la Giornata della Liberazione.En: In the heart of spring, the city celebrated Liberation Day.It: Il sole splendeva alto, risvegliando il vecchio magazzino abbandonato alla periferia della città.En: The sun shone high, awakening the old abandoned warehouse on the outskirts of the city.It: Un luogo dimenticato da molti, ma non da Giorgio, Martina e Alessio.En: A place forgotten by many, but not by Giorgio, Martina, and Alessio.It: Era il teatro perfetto per un'esposizione d'arte fuori dall'ordinario.En: It was the perfect setting for an extraordinary art exhibition.It: Giorgio, con i capelli spettinati e una mente piena di visioni artistiche, stava in silenzio di fronte a una tela.En: Giorgio, with messy hair and a mind full of artistic visions, stood silently in front of a canvas.It: Attorno a lui, Martina posizionava sculture e quadri, organizzando ogni cosa con cura.En: Around him, Martina arranged sculptures and paintings, organizing everything with care.It: L'atmosfera del magazzino, con le sue alte volte e finestre rotte, creava un ambiente unico.En: The atmosphere of the warehouse, with its high ceilings and broken windows, created a unique environment.It: I fasci di luce che cadevano sul pavimento trasmettevano una sensazione di magia e speranza.En: The beams of light falling on the floor conveyed a sense of magic and hope.It: "Coraggio, Giorgio," disse Martina sorridendo, mentre appoggiava una mano rassicurante sulla sua spalla.En: "Come on, Giorgio," said Martina with a smile, while placing a reassuring hand on his shoulder.It: "Sarà un successo!En: "It will be a success!It: Alessio sarà colpito dalle tue opere."En: Alessio will be impressed by your works."It: Giorgio annuì, ma dentro di sé la paura montava.En: Giorgio nodded, but inside, fear was mounting.It: Dubitava di sé stesso, aveva paura di non essere all'altezza.En: He doubted himself, afraid he wouldn't measure up.It: Alessio, il gallerista noto e sofisticato, era un uomo dalle grandi ambizioni, sempre alla ricerca di nuovi talenti.En: Alessio, the well-known and sophisticated gallery owner, was a man of great ambitions, always in search of new talents.It: Finalmente arrivò Alessio, con un passo deciso e l'aria di chi non aveva tempo da perdere.En: Finally, Alessio arrived, with a determined step and the air of someone who had no time to waste.It: Girava per il magazzino, osservando ogni opera d'arte con sguardo critico.En: He wandered through the warehouse, observing each artwork with a critical eye.It: Giorgio sudava freddo, il cuore che correva veloce.En: Giorgio was cold with sweat, his heart racing.It: Quando Alessio raggiunse Giorgio, sorrise.En: When Alessio reached Giorgio, he smiled.It: "Le tue opere sono interessanti," disse.En: "Your works are interesting," he said.It: "Ho un'offerta per te.En: "I have an offer for you.It: Un posto nella mia galleria, ma voglio che tu dipinga in uno stile diverso."En: A place in my gallery, but I want you to paint in a different style."It: Giorgio esitò.En: Giorgio hesitated.It: Questa era l'occasione che aspettava, ma c'era una condizione.En: This was the opportunity he had been waiting for, but there was a condition.It: Martina, che ascoltava, lo guardò incoraggiandolo con gli occhi.En: Martina, who was listening, looked at him encouragingly with her eyes.It: Finalmente Giorgio rispose: "No, grazie."En: Finally, Giorgio replied: "No, thank you."It: Alessio alzò un sopracciglio, sorpreso.En: Alessio raised an eyebrow, surprised.It: "Sei sicuro?"En: "Are you sure?"It: "Sì," rispose Giorgio con più coraggio di quanto si sentisse.En: "Yes," replied Giorgio with more courage than he felt.It: "Voglio restare fedele a me stesso e alla mia arte."En: "I want to stay true to myself and my art."It: Mentre Alessio se ne andava, Giorgio sentì una nuova forza dentro di sé.En: As Alessio left, Giorgio felt a new strength within himself.It: Forse non aveva ottenuto il contratto, ma aveva conquistato qualcosa di più prezioso: la fiducia in sé stesso.En: Perhaps he hadn't secured the contract, but he had gained something more precious: self-confidence.It: E con Martina al suo fianco, sapeva che, un giorno, il mondo avrebbe riconosciuto il suo talento.En: And with Martina by his side, he knew that one day the world would recognize his talent.It: L'esposizione divenne un successo tra gli amici e la comunità locale, simbolo di libertà artistica e autenticità.En: The exhibition became a success among friends and the local community, a symbol of artistic freedom and authenticity.It: Giorgio aveva finalmente trovato il suo posto, non in una galleria famosa, ma nel cuore della sua comunità e soprattutto, dentro sé stesso.En: Giorgio had finally found his place, not in a famous gallery, but in the heart of his community and above all, within himself. Vocabulary Words:the heart: il cuorethe spring: la primaverato celebrate: festeggiarethe outskirts: la periferiaabandoned: abbandonatoforgotten: dimenticatoextraordinary: fuori dall'ordinariocanvas: la telasculpture: la sculturaatmosphere: l'atmosferabeam: il fasciomagical: magicoreassuring: rassicurantefear: la paurasophisticated: sofisticatoambition: l'ambizionedetermined: decisocritical: criticoto hesitate: esitareopportunity: l'occasionecondition: la condizioneto encourage: incoraggiareeyebrow: il sopracciglioto feel: sentireto gain: conquistareself-confidence: la fiducia in sé stessofreedom: la libertàauthenticity: l'autenticitàto recognize: riconosceretalent: il talento

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0
Physical AI that Moves the World — Qasar Younis & Peter Ludwig, Applied Intuition

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

Play Episode Listen Later Apr 27, 2026 72:21


From building Applied Intuition from YC-era autonomy tooling into a $15B physical AI company, Qasar Younis and Peter Ludwig have spent the last decade living through the full arc of autonomy: from simulation and data infrastructure for robotaxi companies, to operating systems for safety-critical machines, to deploying AI onto cars, trucks, mining equipment, construction vehicles, agriculture, defense systems, and driverless L4 trucks running in Japan today. They join us to explain why “physical AI” is not just LLMs on wheels, why the real bottleneck is no longer model intelligence but deployment onto constrained hardware, and why the future of autonomy may look less like one-off demos and more like Android for every moving machine.We discuss:* Applied Intuition's mission: building physical AI for a safer, more prosperous world, powering cars, trucks, construction and mining equipment, agriculture, defense, and other moving machines* Why physical AI is different from screen-based AI: learned systems can make mistakes in chat or coding, but safety-critical machines like driverless trucks, autonomous vehicles, and robots need much higher reliability* The evolution from autonomy tooling to a broad physical AI platform: starting with simulation and data infrastructure for robotaxi companies, then expanding into 30+ products across simulation, operating systems, autonomy, and AI models* Why tooling companies came back into fashion: Qasar on why developer tooling looked unfashionable in 2016, why Applied Intuition still bet on it, and how the AI boom made workflows and tools central again* The three core buckets of Applied Intuition's technology: simulation and RL infrastructure, true operating systems for vehicles and machines, and fundamental AI models for autonomy and world understanding* Why vehicles need a real AI operating system: real-time control, sensor streaming, latency, memory management, fail-safes, reliable updates, and why “bricking a car” is much worse than bricking an iPad* Physical machines as “phones before Android and iOS”: Peter explains why today's vehicle and machine software stack is fragmented across many operating systems, and why Applied Intuition wants to consolidate the platform layer* Coding agents inside Applied Intuition: Cursor, Claude Code, internal adoption leaderboards, and how AI tools are changing engineering workflows even in embedded systems and safety-critical software* Verification and validation for physical AI: why evals get harder as models improve, how end-to-end autonomy changes simulation requirements, and why neural simulation has to be fast and cheap enough to make RL practical* From deterministic tests to statistical safety: why autonomy validation is shifting from binary pass/fail requirements toward “how many nines” of reliability and mean time between failures* Cruise, Waymo, and public trust: Qasar and Peter discuss why autonomy failures are not just technical issues, how companies interact with regulators, and why Waymo is setting a high bar for the industry* Simulation vs. reality: why no simulator perfectly represents the real world, how sim-to-real validation works, and why real-world testing will never disappear* World models for physical AI: hydroplaning, construction equipment, visual cues, cause-and-effect learning, and where world models help versus where they are not enough* Onboard vs. offboard AI: why data-center models can be huge and slow, but onboard vehicle models need millisecond-level latency, low power, small size, and distillation-like efficiency* Why physical AI is not constrained by model intelligence alone: the hard part is deploying models onto real hardware, under safety, latency, power, cost, and reliability constraints* Legacy autonomy vs. intelligent autonomy: RTK GPS in mining and agriculture, why hand-coded path-following worked for decades, and why modern systems need perception and dynamic intelligence* Planning for physical systems: how “plan mode” applies to robotaxis, mining, defense, and multi-step physical tasks where actions change the state of the world* Why robotics demos are not production: the brittle last 1%, humanoid reliability, DARPA Grand Challenge-style prize policy, and the advanced engineering gap between research and deployment* Applied Intuition's hard-earned lessons: after nearly a decade, Peter says they can look at a robotics demo and predict the next 20 problems the company will hit* Qasar's advice to founders: constrain the commercial problem, avoid copying mature-company strategies too early, and remember that compounding technology only matters if you survive long enough to see it compound* Why 2014 YC advice may not apply in 2026: capital markets, AI company dynamics, and the difference between building in stealth with a deep network versus building as a new founder today* What Applied is hiring for: operating systems, autonomy, dev tooling, model performance, evals, safety-critical systems, hardware/software boundaries, and engineers with deep curiosity about how things workApplied Intuition:* YouTube: https://www.youtube.com/@AppliedIntuitionInc* X: https://x.com/AppliedInt* LinkedIn: https://www.linkedin.com/company/applied-intuition-incQasar Younis:* X: https://x.com/qasar* LinkedIn: https://www.linkedin.com/in/qasar/Peter Ludwig:* LinkedIn: https://www.linkedin.com/in/peterwludwig/Timestamps00:00:00 Introduction: Applied Intuition, Physical AI, and 10 Years of Building00:01:37 Physical AI vs. Screen AI: Why Safety-Critical Changes Everything00:02:51 The Origin Story: Tooling, YC, and the Scale AI Comparison00:05:41 The Three Buckets: Simulation, Operating Systems, and Autonomy Models00:11:10 Hardware, Sensors, and the LiDAR Question00:14:26 The Operating System Layer: Why Vehicles Are Like Pre-Android Phones00:19:13 Customers, Licensing, and the Better-Together Stack00:21:19 AI Coding Adoption: Cursor, Claude Code, and the Bimodal Engineer00:26:41 Verifiable Rewards, Evals, and Neural Simulation00:31:04 Statistical Validation, Regulators, and the Cruise Lesson00:40:25 World Models, Hydroplaning, and Cause-Effect Learning00:43:34 Onboard vs. Offboard: Latency, Embedded ML, and Distillation00:50:57 Plan Mode for Physical Systems and Next-Token Prediction Universally00:53:04 Productionization: The 20 Problems Every Robotics Demo Will Hit00:58:00 Founder Advice: Constraints, Compounding Tech, and Mature-Company Mimicry01:05:41 Hiring Philosophy: Hardware/Software Boundary and Engineering Mindset01:08:50 General Motors Institute, Education, and the Curiosity MindsetTranscriptIntroduction: Applied Intuition, Physical AI, and 10 Years of BuildingAlessio [00:00:00]: Hey everyone, welcome to the Latent Space Podcast. This is Alessio, founder of Kernel Labs, and I'm joined by Swyx, editor of Latent Space.Swyx [00:00:10]: And today we're very honored to have the founders of Applied Intuition, Qasar and Peter. Welcome.Qasar [00:00:17]: You guys really know how to turn it on to podcast mode. That was, you guys are real pros at this.Qasar [00:00:23]: They were just joking around right before this, and then they flipped it pretty quick.Alessio [00:00:29]: Oh, yeah, it's good to have you guys. Maybe you just wanna introduce yourself so people know the voice on the mic and they'll know what they're hearing.Peter [00:00:33]: Oh, sure. Yeah, I'm Peter Ludwig. I'm the co-founder and CTO of Applied Intuition.Qasar [00:00:38]: And my name is Qasar Younis. I am the CEO and co-founder with Peter.Alessio [00:00:42]: Nice. Can you guys give the high-level overview of what Applied Intuition is? And I was reading through some of the Congress files, when you went out there, Peter, and eighteen of the top twenty global non-Chinese automakers, you two guys, you have customers in agriculture, defense, construction. I think most people have heard of Applied Intuition tied to YC when it was first started, and then you were kinda in stealth for a long time, so maybe just give people the high-level overview of what it is today, and then we'll dive into the different pieces.Peter [00:01:10]: Yeah. So at Applied Intuition, our mission is to build physical AI for a safer, more prosperous world. And so we work on physical AI for all different types of moving systems, everything from cars to trucks to construction and mining equipment, to defense technologies. And we're a true technology company, so we build and sell the technology, and we sell it to the companies that make the machines. We sell it to the government, really anyone that wants to buy a technology to make machines smart.Physical AI vs. Screen AI: Why Safety-Critical Changes EverythingQasar [00:01:38]: Yeah. And I think in the broader AI landscape, a lot of the focus, rightfully so in the last, three years has been on large language models, and so everything fits in a screen. Like, whether it's code complete products or things like that. And what's different about us is we're deploying intelligence onto a lot of things that don't have screens. they're physical machines. There are sometimes screens within the cabin or for example of a car or a truck or something like that, but most of the value we provide is putting intelligence that is in safety critical environments. So that those two words are really important because learn systems can make mistakes if you're asking for, like, some, so something like, “Tell me about these podcast hostsQasar [00:02:28]: that I'm about to go meet.” But you can't do that obviously when you run, like, as an example, we run driverless trucks in Japan right now, as we speak. We can't have errors. Those are L4 trucks. Yeah.Alessio [00:02:40]: Yeah. Was that always the mission? I remember initially, I think people put you and Scale AI very similarly for some things about being kinda like on the data infrastructure side of things. What was the evolution of the company?The Origin Story: Tooling, YC, and the Scale AI ComparisonPeter [00:02:51]: Well, from the very beginning, we always wanted to, really be a technology company that helped generally push forward the industrial sector. And so we started off working in autonomy. Our very first customers were robotaxi companies. And we started off doing a lot of work in simulation and data infrastructure. And then over the years, we've expanded our portfolios. Now we have, over thirty products, and it's a pretty broad technology play within the landscape of physical AI.Qasar [00:03:19]: Yeah, I think the Scale reason is because we're all YC Universe companies. But it was a very different company. Scale, was, is more of a services company, data labeling company fundamentally. We started and still are, do a lot of tooling. So like, you think developer tooling is now in vogue again, thanks to the AI boom. But honestly, ten years ago, it was out of vogue. It w Like, doing a tooling company in 2016, 2017 was not, like, the thing to do because, I don't know if you remember, the VCs generally, their views was that toolings are They're just workflows, and workflows ultimately are not really interesting. And we've gone and come, full circle with that. But when we started the company, our kind of it's kinda like in the periphery of what the company wants to be. It was like, from our earliest days, like, we wanna deploy software on physical machines, like on cars and on trucks and things like that. And obviously, we didn't know that the transformer boom was gonna happen. We didn't know that autonomy systems would become end-to-end. Those things we didn't know. And why that's important when autonomy systems become end-to-end, it is just now those models can be generalized to, multiple form factors. And so back nine, ten years ago, tooling was a great way, and still is a great way to, build the technology and sell technology to our end customers, a lot of them who wanna build this stuff themselves. And so we just offer like a spectrum of solutions from you can just use like one part of a development suite of tools all the way to buying the full thing. The way to think about the company, or at least the way we think about the company is, as Peter said, a technology provider. It's kinda like, what NVIDIA does or what an AMD, but we just don't do chips.Qasar [00:05:06]: We don't do silicon. But we're a technology provider fundamentally. And I think even, we used to joke when we started the company, like, we're not the guys to build, like, Instagram. Like that was just towards That's not our That's just not us in a most fundamental way. IAlessio [00:05:20]: You have thoughts.Qasar [00:05:21]: Yes.Qasar [00:05:22]: Well, it's, it's I mean, I think it's just like what And I mean, we worked on Maps and stuff, Google Maps. Consumer products are extremely difficult for a lot of different reasons. It just, I think doesn't scratch the itch. I think we're like Michigan guys who are kind of more of that traditional engineering kind of a realm, or lineage. we used to jokeThe Three Buckets: Simulation, Operating Systems, and Autonomy ModelsPeter [00:05:41]: I gotta say, though, what was clear ten years ago was that there was so much more that was possible with software and AI in vehiclesPeter [00:05:47]: and that was generally the space that we started in ten years ago.Peter [00:05:51]: And the precise path that we've taken over the years, I think we've been strategic, and we've adjusted to make sure that we're actually building stuff that's valuable to the market. And like, the technology has changed so much. Like our own technology stack has completely changed, I would say, roughly every two years. And so now we've probably done, let's say, four complete evolutions of our own technology stack. And I sort of see that cadence roughly keeping up.Peter [00:06:13]: And so the way even we think about engineering is almost on this two-year horizon, we're preparing ourselves that, hey, like, we wanna invest the appropriate amount, but then also be very dynamic as the research gets published and as our research team figures out new advancements and adapting to that.Qasar [00:06:27]: Yeah. One thing that has been consistent is the type of people we've, we've recruited. It's engineers who are fall into the sometimes very traditional, like, GoogleQasar [00:06:38]: -gen suite, but way different from, other companies. We are hiring folks who really know the intersection of hardware and software, who know really low-level systems. Obviously, traditional ML researchers and folks who've, actually, put ML systems into production. That's been pretty consistent. I think that, like, you look at the mix of our engineering, eighty-three percent of the company is engineering, so it's, like, a giant list.Qasar [00:07:05]: A lot of engineers.Alessio [00:07:06]: Which, by the way, a thousand engineersQasar [00:07:07]: Yeah. A thousand engineers.Alessio [00:07:08]: that's on your website, so I imagine it's up to date.Qasar [00:07:11]: It is, it is up to date, yes. Yes.Alessio [00:07:12]: okay. And then forty-plus founders.Qasar [00:07:15]: Yeah. We would tend to also, This was more luck than strategy. But we've recruited a lot of ex-founders. It's been a great place for founders, YC and non, ‘cause obviously I know a lot of the YC folks. It's kind of like we recruit a lot of Google people.Qasar [00:07:33]: For them to exercise both their technical and non-technical skills because, we're, we're, we're on the applied side. We have a research team that we do fundamental research, we publish, and we've, we've had great traction there. But fundamentally, the business wants to take this intelligence and deploy it into production and there's, like, a certain type of person that's more interested in that.Alessio [00:07:54]: Yeah. You mentioned the tech stack, Peter, so I just wanted to give you some rein to just go into it. I'm interested in where Wayve Nutrition, starts and ends in some sense, what won't you do? What, do you do that's common among all the verticals that you cover?Peter [00:08:10]: There's a few buckets of work that we do, and we've been at this for almost ten years now, so the technology's pretty broad. But we got startedQasar [00:08:17]: Yeah, with a thousand engineers, like, you could work on lots of things.Peter [00:08:19]: There's lots of stuff, yeah, espe-especially with AI tools to help.Peter [00:08:22]: So we got our start in simulation and simulation tooling and infrastructure. And so generally, if you're trying to build a very complex software system that involves moving machines, you need to test that, and the best way to test it is it's a combination of virtual developments, a simulation, and then also obviously real world testing.Peter [00:08:39]: And then there's a very careful process of that correlation between the simulation results and the real world results and ensuring that the simulator is in fact accurate to that. Simulation's a very deep topic.Peter [00:08:49]: We have a whole suite of products in that, and we could talk for many hours about that specifically. But that is one part of what we do as a company. Reinforcement learning as a subpart of that is also super critical. I think a lot of the a lot of the best advancements happening in a lot of these AI systems right now in some way relate to reinforcement learning, and with now we have lots of compute, and you can do tons of interesting things for reinforcement learning. The second bucket of work that we do is on operating systems technology. true operating systems. Like, think about, schedulers and memory management and middleware and message passing and highly reliable networking and data links. Like, the reality is, if you want to deploy AI onto vehicles, you need a really good operating system. And when we were getting deeper into that space, there wasn't really anything that we were happy with.Peter [00:09:39]: Like, things existed, absolutely, and we were using what was available in the market, and as an engineering organization, we roughly realized these things aren't great. We think we can do this better, and so let's, let's build something. And that was then the that was the moment of inspiration that started our operating systems business, which is now a very real business for us. And in order to write and run great AI, you need a great operating system, and so that-that's what got us into that. And then the third bucket that we work on, it's, it's true fundamental AI technology. Models, we do a lot of work in, as mentioned, the foundational research, but then the also the world models and the actual autonomy models that are running on these physical machines, and that's across cars, trucks, mining, construction, agriculture, and defense, and so that's both land, air, and sea.Qasar [00:10:31]: And also, a smaller subsector of that third bucket is the interaction of humans with those machines.Qasar [00:10:38]: So that's a multimodal, experience. Historically, if you're moving a dirt mover or any of these machines, there are, like, buttons you press, whether they're actual physical tactile buttons or something like a touch screen. That's just That fundamentally is changing to where you're just talking to the machine and the machine and you're teaming with the machine.Alessio [00:10:58]: Voice?Qasar [00:10:59]: Yeah, voice, absolutely, yeah.Alessio [00:11:00]: Oh.Qasar [00:11:00]: And also the machine just being aware of who is in the cabin, what their state is. you can think from a safety systems perspective, the most simple version of this is, like, the driver is tired, right? They're, they're if you get those alerts when you're driving your car and saysHardware, Sensors, and the LiDAR QuestionQasar [00:11:15]: -maybe take a coffee break, that take that times, a couple of order of magnitudes up. But this concept of teaming man and machine is important. When you think about running agents or just running, different instances of, Claude and doing work for you in the background, you can take that analogy out, almost copy and paste and put it into, like, a farm, where you have a farmer who's running a number of machines. So where they interact with the machine is where there's maybe a critical decision or a disengagement or something like that, but generally speaking, the agent on the physical machine is running and making decisions on the behalf of the farmer until there's something maybe critical. And that's also what we work on. So that's not pure autonomy. It's a little bit of a mix, but it falls under, autonomy. In the automotive sense, that's typically defined in SAE levels as an L2++ systemQasar [00:12:05]: -with a human in the loop. But just take that idea, to other verticals.Alessio [00:12:09]: Yeah. You've not mentioned hardware at all, like sensors or obviously we you mentioned you don't do chips. I think even in AV there's, like, a big, cameras versus lidars. Like, what are, like, in your space maybe some of those design decisions that you made, and are they driven by the OEM's ability to put things on the machinery? And like, how much influence do you guys have on co-designing those?Peter [00:12:32]: Yeah. So we don't make sensors. Like, we're, we're not a manufacturer. Obviously, we use a lot of sensors in our autonomy products. in terms of what actually goes on the vehicles, we have a preferred set of sensors that we, let's say fully support, and then our customers, they can sort of choose from those. And obviously if there's a very strong opinion on supporting something else, we'll add that to the platform as well. And the lidar question is at this point sort of the age-old,Peter [00:12:59]: topic in autonomy, and the state of the industry right now is lidar is hands down a useful sensor, specifically for data collection and the R&D phase of autonomy development. if you see, for example, a Tesla R&D vehicle, it actually has lidar on itPeter [00:13:17]: to this day, right? In the Bay Area we see these. you'll see, like, Model Ys or Cybercab that have lidars on them just driving around. So it's, it's useful because it gives you per pixel depth information. So if you can pair a lidar with a camerand you can say that, well, this camera's looking this direction, this lidar's looking this direction, and now for each pixel of the camera I can see how far away is that pixel. you can actually then use that as a part of your model training, and then the that depth information then becomes a learned, a learned state of the camera data. And then when you're doing the production system, you can now remove the lidarPeter [00:13:52]: and now you can actually get depth with just the camera. And so that difference between, like, a highly sensored R&D vehicle and then the down-costed production vehicle, we use that across our whole portfolio of products. And of course the end goal is you want super low cost and super reliable.Peter [00:14:08]: And then in certain use cases you have some more, bespoke things. Like in defense as an example, you do things at night oftentimes, and so you care about sensors like infrared, more so than And you don't, you don't wanna be putting energy out, so you don't wanna use lidar or radar.Peter [00:14:23]: but you still need to be able to see at nighttime. So yeah, we work the whole gamut.The Operating System Layer: Why Vehicles Are Like Pre-Android PhonesAlessio [00:14:27]: Cool. So that's kinda like on the hardware level. Then on the OS level, how does that look like? What is, like, unique? my drive- I drive a Tesla. Whenever I drive some other car that has a screen, it always sucks.Alessio [00:14:38]: It's on, like, cheap Android tablet. It's like, it's laggy and all of that. What does the OS of, like, the autonomy future look like?Peter [00:14:46]: When most people, it's really what you just described. When you think about operating system in a vehicle, you're thinking about the HMI, right? The human machine interface, and absolutely that's a an important part of it, but that's actually only one thin layer on top. So when we talk about operating systems for, like, AI in vehicles, there's many layers that go deep into the CPU critical realm and embedded systems, and you're talking about the real time control ofPeter [00:15:13]: let's say the electric motors or the engine and the actuators, and you have different redundancies for different, let's say, the steering actuation in the vehicle. And all of these things, need very core support in the in the operating system. And then of course for autonomy you have real time sensor data that's streaming in, and the latencies there are really important, right? If you try to Imagine you try to run Microsoft WindowsPeter [00:15:35]: like streaming your sensor data in or controlling the vehicle. Like, the latencies are gonna be absurd. Like, you can never do that. And so what's special about what we do is we really have this system level thinking, right? So we're looking at, we care about every performance characteristics of the entire system, and then we also, because we're doing a lot of the software or all of that software, we can fine-tune and control all of those things. So we can very carefully tune in the latencies for every aspect of the system. We can carefully tune in the memory management. We can have the right, fail-safes and fallbacks, for different things. ‘Cause you have to account for what if, what if there is a critical failure? What if there's a cosmic ray that flipsPeter [00:16:14]: a bit in the middle of the processor that causes some, malfunction? And you have to have a fail-safe to all of that, and so the core operating system is a part of that. And then the one last thing, which is a lot less exciting but is, actually a very big topic, is reliability of updates.Peter [00:16:30]: so the I have a Tesla and you get updates fairly frequently, right?Peter [00:16:36]: Once a month. Most companies that are making vehiclesPeter [00:16:40]: are basically never doing updates, and they're And even if they are doing updates, they're usually only updating maybe one module. Maybe they're updating the HMI module. But they're not able to update, let's say, the CPU critical parts of the system.Peter [00:16:51]: You have to go into the dealer for that. And so with our operating system now we can actually enable highly reliable updates of any system in the vehicle, and that's way easier said than done. Like, there's lots of technical, technically deep stuff, in the tech stack to do that in a way that you're not going to accidentally brick a vehicle.Peter [00:17:08]: And right? If, imagine yourAlessio [00:17:10]: That would be bad.Alessio [00:17:11]: Bad.Peter [00:17:11]: Bricking a car is a very expensivePeter [00:17:13]: and honestly, like across the industry maybe one of the most just pure impactful things that we've done is we've just, we're, we're now enabling the industry to actually do software updates.Alessio [00:17:22]: Just to clarify as well, who is the customer for this? Like, I assume a lot of hardware manufacturers have their own firmware, and I'm sure some of them would just have you write it for them because you're experts. And others would have their own. Like, who pays for this? Who invites you into the house? Is it, is it the end user, or is it, is it the manufacturer?Peter [00:17:41]: Yeah. So let me make an analogy firstly on the on the fragmentation of software. So physical machines today are more akin to the state of the phone market before Android and iOS existed, right? So I worked on Android at Google by the way many years ago, and part of the reason that Larry at Google decided to get into Android was they wanted to run Google products on a bunch of phones, and they bought all of these phones from the industry, and it turned out they had like 50 different operating systems on these phones. And it was virtually impossiblePeter [00:18:17]: for Google to make their app run on all 50 devices equally well. And so the solution was, well, actually what if, what if they created-A really great operating system and made it attractive to all of these phone makers, and that was sort of the genesis for what Android was and why Android existed. It was a way for Google to get their products onto really wide diversity of devices. The state of the physical, industry right now, it's a little bit like that. Like, there's yes, these companies have firmware, but they have so many different operating systems, it's so fragmented, and to actually get a modern AI application to run on these vehicles, you actually, you first have to consolidate the operating system, and so that's, that's why we've done that. And then, your specific question was who are our customers? It's, it's, generally it's the companies that are making these machines.Peter [00:19:06]: And we're, we're, we're selling our technology to them to really simplify the architecture and then enable these AI applications to run on them.Customers, Licensing, and the Better-Together StackSwyx [00:19:13]: How much is reusable across? Like, do you have, like, one OS that is just configured for everything, or is there some more customization that is needed?Peter [00:19:22]: Yeah, highly reusable. So the fundamental technology is quite universal, right? So things that we do have to think about though are, like, chipset support. And so if you're, if you're coding, let's say, an LLM and you have start with an assumption that, “Hey, oh, I'm gonna, I'm gonna use CUDA, and I'm gonna run this, on an NVIDIA chip,” then you don't really have to think about the hardware in that sense. Like, you're just, “Okay, I'm just I'm in the CUDA/NVIDIA ecosystem, and I'm, I'm going to use that.” But the hardware, especially in safety critical systems, it's a lot more diverse. There's not one or one or two players. There's a bunch of different chipsets that we have to support. And so our operating system doesn't just run on, like, the equivalent of X86. It has to, it has to run on a number of different architectures from chips from a bunch of different companies. But again, we've been working on this for a long time now, so we have, we have support for all of those chipsets. And then when you want to then run the AI applications, we can then do that reliably across now a variety of providers.Qasar [00:20:19]: And I think that is, like, heavily inspired by Android, right? Android has a huge suite of testing and it's a reliable operating system that runs on thousands of devices. And we think we can, we can do the same in all these physical moving machines, with the difference that we're really in a safety critical realm. Android isn't.Alessio [00:20:40]: So on Android, I don't need to use Gmail, I can use Superhuman. Like, what about your machinery? Like, can people bring somebody else's automation to it, or is it kinda like all-in-one?Qasar [00:20:50]: You have to use us. No. Yeah. we're If, Yeah. Yeah, it's totally open. Yeah.Peter [00:20:56]: Yeah. our philosophy is that we are a technology company, and so we license our technology to customers to use how they want. And so if a customer wants to If they wanna license our autonomy tech and our operating system, then great, we'll license those. If they just wanna license the operating system and then use different autonomy tech, that's fine also, and we have great documentation andSwyx [00:21:17]: Or if they wanna use developer tooling.Peter [00:21:18]: Yeah, exactly.AI Coding Adoption: Cursor, Claude Code, and the Bimodal EngineerSwyx [00:21:19]: It's, like, a better together if, obviously, if you, if they work together. Is it all C++ I assume is with different compile targets?Peter [00:21:27]: We use a lot of C++.Peter [00:21:28]: Rust is sort of a hot, the new hot kid on the blockPeter [00:21:32]: for a bunch of things as well. But yeah, the lower level you get, especially when you get to real-time constraints, you hit C++ at some point, and at some point maybe you work your way into assembly when needed.Swyx [00:21:44]: Oh, damn.Alessio [00:21:46]: I'm curious about the coding agent adoption, just, like, since you're mentioning more esoteric languages. Like, what's the adoption internally? What have you learned?Peter [00:21:55]: Yeah. We use everything. So Cursor was, I think the hottest tool in the company for a good while. Now Claude Code, I think has taken the reign on that. We have a internal leader, leaderboard that we use just to sort of encourage adoptionPeter [00:22:09]: with-within the company. And yeah, it's, they're phenomenally useful. it's, Honestly, we take inspiration from some of those tools also in how we're adapting some of that mindset of thinking to the physical realm. Like if it's so easy to build an app for this or that thing that lives just on a screen, we can We're taking now a lot of the same ideas and applying that to, “Okay, well, if you wanted a physical machine to do something, how easy can we make that, using our own tooling and platform as well?”Alessio [00:22:40]: Are you changing any of, like, the OS architecture, kinda like the way you expose services to, like, be more AI friendly or?Peter [00:22:48]: Yeah, absolutely. The in the early days of our tools infrastructure work, it was a lot about, You had engineers that were experts in certain topics, but the things that you're dealing with, they're oftentimes more mathematical or more abstract, where actually GUI tools are very useful for certain things. Like as an example, we have a product we call Sensor Studio, which is, it helps you design the sensor suite for your autonomous vehicle, whether, again, it could be a car, it could be a drone, could be a mining equipment, could be a robot. And you place sensors in different places. You There's different, There's a library. You can understand what are the trade-offs that you're making in the design of that system, and that was, like, a very, a very GUI intensive, thing ‘cause it's a little more like a CAD tool in that senseSwyx [00:23:37]: YepPeter [00:23:37]: if you've seen CAD tools. Nowadays, though, right, we expose all of the underlying APIs for that and now using, AI agents, you can actually configure a sensor suite with just text and likely reach a better result than you could've through the GUI in the past, and we're taking that thinking now through the whole product portfolio.Swyx [00:23:57]: Another thing I was thinking about is just in terms of, like, AI, adoption, does it change your hiring at least a little bit, or how do you, how do you sort of manage engineers, differently?Peter [00:24:08]: Yeah. absolutely, it does. we, I think like every company in the Valley right now, are evolving our hiring practicesPeter [00:24:16]: because the skills required to be effective are changing so fast, right? you used to really select for just rote implementation ability and now it is more the AI engineer skill set, right? Where it's like, yeah, how to implement, but actually-Just banging out code is no longer the core job, right? It's, it's actually knowing what questions to ask, knowing how to tie, how to tie together these different AI tools. And so the interviews that we give now I think are way harder than they've ever been.Peter [00:24:46]: But we also allow, right, selective use of AI tools to solve the problems. And I think in that you start to see more of a bimodal distribution of engineers, right? You start to see like wow, there's, there's this subset of people that they really get it. Like they're, they're all in and they've, they've clearly invested the hours needed to learn these tools and how to be effective.Peter [00:25:09]: And then there's sort of the group of people that haven't done that, and that the productivity gap is just enormous. And so we're, we're trying to obviously select for the people that are really into this.Qasar [00:25:20]: I first wrote the my AI engineer piece three years ago, and when I first wrote about it, I was like, “Actually, not everyone should be an AI engineer,” ‘cause I think there's a there's an extremist stance where well, every software is an engineer is an AI engineer. And my actual example of people who should not be adopting AI was embedded systems and operating systems, and database people. Are they adopting AI?Peter [00:25:41]: I think it's the classic bitter lesson, topic, which is the Six months ago I would've said the same thing, but it's, it's becoming super useful for every domain.Qasar [00:25:53]: I'm sure.Peter [00:25:54]: Right? Like,Peter [00:25:56]: there was, I think six months ago, or maybe a year ago, if you tried to use, let's say the latest Claude model for writing shaders, GPU shaders, the results were probably underwhelming. And if you use the latest model now to do that kind of task, you're a little bit blown away, like, “Wow, that actually worked. That's amazing.” And we see the same thing in the embedded realm. No question though, especially when you get into safety critical systems, the human validation isPeter [00:26:25]: is 100% key. Like I You're not gonna trust your life to a an AI written software that's, that's not been very carefully, checked by humans. And so I think now the really the challenge is about that appropriate level of human validation for these safety critical systems.Verifiable Rewards, Evals, and Neural SimulationAlessio [00:26:41]: How do you think about, yeah, touching on the simulation side, I think verifiable reward and reinforcement learning is, like, the hottest thing. What have you done internally to build around that? And like, what gives you What makes you sleep at night? Like, if somebody's like, just web coding something or likeAlessio [00:26:57]: wants to try something new, you have like a good enough system. Because I think the opposite is also true, is like if it's super easy to write anythingAlessio [00:27:04]: then it puts a lot of work on like the verifiableAlessio [00:27:07]: side of it. Like, what does that look like for people?Peter [00:27:10]: Yeah. So verifiability, a broader bucket of like evaluations, right? Like how do you evaluate the results that you're, you're getting? I think this is probably the hardest problem right now, because the As the models get better, it can be harder and harder to find the faults on the system.Peter [00:27:29]: And so like the problem of doing proper eval to find those faults, like that problem also keeps getting harder as the models get better. But it's no less important than it's ever been, right? You still there are still going to be edge cases that are not met and whatnot. And so it's, it's a big area of investment for us. On the reinforcement learning topic, the key thing is there's all these new requirements that come to be in the latest generation of these technologies. So for example, end-to-end is the big thing right now in autonomy and physical AI, which is you can now train these models that can effectively take sensor data in and then put control signals out, and get really good results out of that. But the way that you train and improve those models is really different from the previous generations. And so to do reinforcement learning on an end-to-end model, you now need to actually simulate all the sensor data, right? So then this becomes a we call our, work in this neural simulation, but it'sPeter [00:28:26]: think of it like a hybrid of Gaussian, splatting and diffusion methods, and where you really care about performance. Like performance is everything. If you can't do enough simulation fast enough and cheap enough, you actually can't get results that are worthwhile, in the end. It also gets to a lot of our work in embedded systems, which is like performance critical work, and that performance optimization, performance criticality, it carries over to a lot of the model training work. because, like, the only way to make it affordable is it has to be really fast.Qasar [00:28:58]: I think it's worth a few minutes talking about our own, evolving thoughts on verification and validation withinQasar [00:29:05]: kind of, traditional simulators, which are, you can think of like vehicle dynamics or something like that, which you're just taking textbooks and taking those formulasQasar [00:29:13]: and putting them into software, to like now this neural sim/world model universe. I think that's an interesting topic.Peter [00:29:20]: Yeah. So in more traditional development, right, you oftentimes would have, more black-and-white answers to questions.Peter [00:29:28]: And so the in Europe as an example, there's, a regulatory, system, it's called Euro NCAP. It's the European New Car Assessment Program, and as part of that, the vehicles have to pass a bunch of tests, and those tests actually, include, safety systems. So automatic emergency braking for a child that runs in front of a carPeter [00:29:51]: or let's say an occluded child that runs out and you hit it. And so you have You end up with sort of these binary answers of like, well, did the car under test pass this specific test? And there's a very well-known set of test casesPeter [00:30:05]: that the vehicle has to pass. And that was how the industry worked, let's say, until 10-ish years ago. But what's changed now is with these models, everything is statistics, right? Like you no longer have a black-and-white answer, but it's like, well, how many orders of magnitude or how many nines of reliability can I get in the system, and how can I, how can I prove that to be true? And the big unlock honestly for physical AI as an industry is that these models are just becoming much more reliable. Right? Things like things actually work a lot better. It's like the number of nines you can get out of these systems are now good enough that it actually becomes cost effective to really deploy these things. And so the big shift in, so verification and validation has been from a little bit more of a Again the past it was strictly requirements, and are you meeting or not? And now it's more of a statistical, verification and validation case where it's all about how many nines of reliability and meantime between failures, that sort of thing.Statistical Validation, Regulators, and the Cruise LessonSwyx [00:31:04]: And is the target audience regulators or even the customers are yeah, if you I imagine the customers are bought in, and it's mostly regulators that need to be satisfied.Peter [00:31:15]: We do work with the US government, we do work of course with the European governments and the government of Japan, and the government is not like an AI lab by any means.Peter [00:31:25]: So Swyx [00:31:26]: They just care about the outcome.Peter [00:31:27]: They care about the outcome.Peter [00:31:28]: And so we do education, in that regard, and like so sort of teaching about, “Hey, this is how we think validation should be done, and this is an approach that we think is reasonable,” and how to think about like when is a driverless system actually safe enough to go on the roads and that sort of thing. But I wouldn't say that the government is asking for it. It's like we're more teaching the government in that, in that sense. It's honestly, it's more so for our own, our own comfort, right? Like, we want to build very safe systems, and then of course our customers care deeply about that as well. But in that context we're also typically educating our customers.Qasar [00:32:01]: Yeah. Our first, our first core value is on round safety. So I think we can't underline enough that, us also verifying and validating that the systems that we're deploying are safe to us is probably as important as, like, some regulator or a customer saying,Swyx [00:32:19]: Of course. Okay. Yeah.Swyx [00:32:20]: You have to satisfy yourselves.Peter [00:32:22]: As I say, as a whole across the world, regulation oftentimes it's like a almost lowest common denominator. But like, you really have to substantially exceed what the regulators are expecting to make good products.Swyx [00:32:33]: Yeah. One thing I often talk about, I think and I try to make this relatable to the audience also, is Cruise, where they had an accident that basically ended the company. I wonder if people overreact to single incidents, because incidents are going to happen regardless, right? ‘Cause it's a statistical thing, but as long I don't know if regulators understand that, you cannot extrapolate from a single incident, but we do because that's all we have to go on. And your sample sizes are necessarily gonna be lower than, I don't knowSwyx [00:33:00]: consumer driving.Qasar [00:33:01]: Yeah. I think the Cruise example wasn't a technology failure. there was The real, compounding issue there was just how did the company talk to the regulators and what was their kind of behavior, and I think that became more of the issue. If you look,Peter [00:33:19]: It isn't It definitely was a technology failure, but it was made much worse by theSwyx [00:33:23]: Put the car back on the woman.Qasar [00:33:25]: Yeah. And let me put it another way. There is a version where Cruise still exists.Swyx [00:33:29]: right. Right.Qasar [00:33:30]: Right. It'sSwyx [00:33:30]: It was like the last strawQasar [00:33:31]: ItSwyx [00:33:31]: in like a long chain ofSwyx [00:33:33]: like issues.Qasar [00:33:33]: So do you feel like ATG had that horrific accident or someone actually dying, because, that was a homeless person crossing the street? So yeah, I think we can't understate enough that ultimately, like, statistical validation of something, that's one part of it, but it's not the only part of it. Like, consumer and let's say, mainstream adoption of these technologies is also gonna be part of that conversation. I think companies like Waymo are doing a lot of service positively to the industry in the sense of they're, they're setting a high benchmark and they're showing, kind of in a very responsible way how to, how to deal with these. There have been Waymo incidences as well. They've just not been as significant as the Cruise one that you mentioned. But yeah, so I think you'll just continue to see that. I think probably the long term question is really gonna be, again, around Like it is very clear humans are way worse drivers statistically.Qasar [00:34:29]: Like, there's no, there's no debate. And so at what point But we're emotional animals.Swyx [00:34:34]: Yeah. So my thing is, like, we have to get to a point as a society where we accept horrific accidents that would never happen by a human because statistically we understand that it is safer overall. In the same way that planes, they're safer, than I think they're the safest mode of transport that we have.Qasar [00:34:50]: Yeah. it's more dangerous to drive to the airport than it is to get on a flight.Qasar [00:34:53]: So if you're everQasar [00:34:54]: if you're ever getting nervous about getting on a plane, just think “I just gotta get to the airport.”Swyx [00:34:58]: Yes, we're flying.Qasar [00:34:59]: If I get to the airportQasar [00:35:00]: I'll be good.Swyx [00:35:00]: But then it's, planes also concentrate the tail risk if planesQasar [00:35:03]: Yeah. AndPeter [00:35:04]: And I was, I don't think we honestly have to worry about there ever being, accidents from these systems that are like much worse than what humans would cause, ‘cause humans do terrible things.Peter [00:35:14]: Like, people fall asleep at the wheel all the time.Swyx [00:35:16]: I have.Swyx [00:35:17]: Like, I'll call, I've been a drowsy driver.Peter [00:35:19]: Kinda drunk drivers, and that'sPeter [00:35:20]: that's the extreme end of the example. But these AI systems, you have redundancies, you have fallbacks. Like, there's many things have to go wrong for there to actually be a something catastrophic because there's, there's so many, fallbacks that these systems have.Alessio [00:35:36]: your simulation is like so vast because there's so many use cases. What are, like, maybe things that worked in a simulation and then you put it out and it's like, “F**k, this isAlessio [00:35:45]: this just did not work at all?”Peter [00:35:47]: Yes.Alessio [00:35:47]: IsPeter [00:35:47]: That's maybe a bit of a misconception, about simulation there. So let me go a little bit, more technical on this. So at first go, no simulation is going to represent the real world. There's always a process of this, sim to real matchingPeter [00:36:02]: where you actually, you need the real world feedback to basically feed into the parameters that are being used in the simulator, and you have to do that, it's like this validation flow, a number of times until you can get some confidence that, like I think the simulator is now accurately representingPeter [00:36:19]: what's gonna happen in the real world. Now, if you have a situation where you've done that full validation and you thought that it was accurate and then there's something different, those are much trickier cases, and that's, that absolutely can happen, but really I think the validation process is a really important part. You can never skip the simulation validation process, like where you're actually ensuring that, hey, the actual, my sim to real gap here is small enough that I can trust these simulation results. And there's, there's so many fun things that you can do when you get into it. Like, I'll, I'll give one fun example that came up recently is like in these humanoid robotics, systemsOverheating actuators is a real problem, right? So obviously phenomenal demos. IPeter [00:37:01]: The most amazingAlessio [00:37:02]: For 10 minutes.Peter [00:37:03]: The most amazing I can get. I love, I love watching robots do acrobatics like everybody but the these systems actually overheat, right? If, like, And one of the ways you can use simulation though is you can actually have that, the temperature of those actuators be one of the parameters that's representedPeter [00:37:18]: in the simulation. And if you're doing reinforcement learning over a certain task, then the robot can actually adjust its motions in the simulation to account for the fact that, oh, it knows that as it's moving, it's actually beginning to overheat this motor. But if you didn't have that parameter of, let's say, the heat of that motor represented in the simulation initially, then your RL policy might It will disregard that. And now you run that on the robot and the robot will overheat and fail.Alessio [00:37:43]: I guess the question is, like, how do you have all of these parameters taken care of while also understanding the deployment environment? Like, temperature is like a great example, right? WellAlessio [00:37:53]: why did you make my robot worse when it runs in like a freezer?Alessio [00:37:57]: So it actually shouldn't worry about that. it's like, yeah, how do you design these simulations?Peter [00:38:02]: This is honestly the This is what makes simulation so hard, right? it's because you Simulation is fundamentally about you're trying to optimize the development of a system, right? Like, how can I build this system faster and better and cheaper and what are all the levers that I have to actually accomplish that? And because simulation's just a software program, you can, you can change it a lot more easily than you can hardware systems. And then what's particularly awesome about the let's say, world models and using that as a part of simulation is now the simulation doesn't just scale with, let's say, adding new math equations inPeter [00:38:36]: but we can actually scale the simulation environment now with additional real world data and that also unlocks a whole new field of robotics.Qasar [00:38:46]: There is a meniscus line where you cross where still doing real world testing is better. there's, in this, sim-to-real gap, you can reproduce reality at exceedingly expensive costs and this So nothing is free. So really you have to you're finding that line where you're getting great performance, you're getting great feedback, whether it's on the training side or on the eval side, but it's way cheaper than doing it in the real world. At some point it, that doesn't make sense. And so even, from our earliest days in autonomy, our view was you're still gonna do real world testing. You There's, there's not, there's not this, magical land where you're not gonna do that. And maybe even like a more nuanced version of this in like traditional software development is, most of your testing for software in a vehicle, 95% of that can be like traditional CI/CD kind of, flows that you would have in traditional web development. But once you have Now you, let's say you have a truck. Well, you can do like 4% of those in like a rig which has all the components, the electrical and electronics of a truck, but doesn't have, it doesn't have the tires and it doesn't have the And then you have the 1%, which is actually the vehicle. There's something There's a similar analogy in terms of using simulation for intelligent systems. You can do a lot in a simulator, but in using world models, but ultimately it's, it's physical AI. So you're gonna deploy it on physical machines andQasar [00:40:17]: the freezer example comes to, comes to light.Alessio [00:40:20]: The world model thing has been to me the hardest thing toAlessio [00:40:22]: wrap my head around. Like we have Faith Eliyon on the podcast.World Models, Hydroplaning, and Cause-Effect LearningQasar [00:40:25]: We've been doing a small series with like another Intuition company, General Intuition as well.Qasar [00:40:31]: yeah, and I mean, lots of, lots of coverage on NeRFs and yes.Alessio [00:40:34]: Yeah. It feels like we talk with about, the heliocentric system, right? It's like in a world model, if you just feed visual data, the model might learn that the sun spins around the Earth. It makes sense, right? And it's like, well, not really. And I think what are like some of these other things that like hydroplaning is one thing I think about, is like can a world model understand hydroplaning and like what amount of water like causes it to happen? And it's like, yeah, to me it's like I don't understand how you guys do it. I guess it's like the real thing is like when you're doing both cars and the highway in Japan versus the excavator in a mine in,Qasar [00:41:13]: ArizonaAlessio [00:41:13]: wherever you're Arizona, wherever you're deploying them.Alessio [00:41:15]: How much of it are you relying on the world models to like generate the simulations for you and then try and close the gap after versus like giving the world models as a tool to your engineers to like curate the simulations if that makes sense?Peter [00:41:28]: Yeah, totally. So yeah, I can say at a pure engineering level, I think if you're hoping to do real world deploys and you're purely relying on a world model approach, you probably won't get to something that works, before you go bankrupt. So there is just a very practical mindset of like, world models are amazing and they're extremely useful for a lot of use cases, but there are a lot of other things that you need to do to actually get something started and something deployed and working. most fundamentally, world models are all about It's understanding the world, but also understanding what's going to happen. It's like the cause-effect relationship.Peter [00:42:01]: Right? And so like it, right, if you have a take some sort of construction tool, and that construction tool is gonna be doing some work on the Earth in some way, it's gonna be moving earth, the world model needs to understand that cause-effect relationship. Like, okay, when I, when I take this material from here and put it over there and now I have things that are over here and not over there anymore and that cause-effect, relationship. data obviously is a is a big problem. The hydroplaningPeter [00:42:26]: one is actually a really great example because it's actually quite non-obvious sometimes. Right? It's like, well, it's, it's raining and well this road, has, let's say the appropriate curvature to it so the water is running off the road and cars are driving faster here and then you approach a road that's very flat and water is now puddling on that road and all of a sudden cars are driving slower because when they were driving faster they were starting to lose control. And there are a lot of visual nuance, very nuanced visual cues in the scene and so I do think in the world model concept there's a good chance that the model actually would learn that you should just drive slower when these visual cues exist, and that's obviously the beautiful-The beauty of, these kinds of models where they just, they learn these non-obvious things.Swyx [00:43:14]: It doesn't need to know about hydroplaning to know that it needs to drive slower.Peter [00:43:17]: Yes.Swyx [00:43:17]: I guess it's Yeah. I wanna ask questions about, also deploying models. I presume, like, you use a lot of these world models for training data and simulation, but what about deploying it onto the systems in production? Presumably you have you have, like, GPUs on deviceOnboard vs. Offboard: Latency, Embedded ML, and DistillationSwyx [00:43:36]: but they're I keep saying on device. What's the what's the right term for that?Peter [00:43:40]: On machine.Swyx [00:43:41]: On machine.Peter [00:43:41]: Or embedded, yeah.Swyx [00:43:42]: Yeah. What is the embedded world like? because for people who are not used to that world, this is very alien.Peter [00:43:49]: Yeah. So it's actually We call it onboard and off board.Peter [00:43:52]: So like, onboard software and off board software.Peter [00:43:54]: And the great thing about off board software is you don't have to care about time, and you can run really large models, right? So you can, you can say, “Well, this model, I don't care if it takes one second for it to give me a result or 10 seconds for it to give me a result, because we have time.” And the models can be really big, and they can run, in a data center or on a on a huge GPU and you can obviously have distribute to compute, et cetera. But onboard you don't have any of those benefits. You're like, “Well, I need I have this many milliseconds where I need an answer from this model.” And so a lot more of the energy then is about, think of it more like distillation and it's like truly efficiency and like, literally every fraction of a millisecond counts. And you can't have a situation where the model takes too long because then the vehicle can't actually function.Peter [00:44:42]: And so you can, you can still use a lot of the same techniques, and the models themselves you can think of as like a derivative of larger models that you can run offline, and then you're, you're trying to just get a model that is still performs really well but it's, it's a it's smaller, small enough version that you can then run on this embedded system where you care about latency and power.Qasar [00:45:03]: Yeah. And I think like, the broader point I think which, maybe is not obvious but it's worth saying is in physical AI world, we're not really constrained right now by, like, the intelligence of the models. It's actually what Peter's talking about, it's actually deploying them inSwyx [00:45:19]: The hardware they give you.Qasar [00:45:21]: Yeah. On the hardware you give you.Qasar [00:45:22]: And so And there's just a reality is of safety critical systems. So those end up being the your limiting factorsQasar [00:45:29]: rather than, let's say, a limiting factor for, a foundation model companyQasar [00:45:34]: is gonna be just capital maybe or researchers.Qasar [00:45:38]: So we're, we're in that way dealing with, for us as people who kind of come in that realm with like a very interesting Those constraints force creativity.Swyx [00:45:47]: And I imagine, nobody was deploying or giving you the hardware for transformers back in 2018, whatever, but now they are. What's the evolution like? just peel back the curtains a little bit.Peter [00:45:59]: Yeah. Transformers first off, I think the paper was originally published in 2017.Swyx [00:46:02]: 2017.Swyx [00:46:02]: So there's no time.Peter [00:46:04]: And ISwyx [00:46:05]: But I'm just saying I guess I'm saying, like, embedded ML systems usually, like, a lot less parameters, a lot less compute, and now, like, orders of magnitude more.Peter [00:46:14]: Yeah. absolutely. what I was gonna say though was I think in the in the original paper in 2017, maybe it's in the last paragraph, somewhere in the paper they talk about, like, “Oh, by the way, this technique might be useful for, like, images and videos as well.”Peter [00:46:30]: These last subjects.Peter [00:46:31]: And it took a few years for that impact to really hit. But like, now, we're seeing transformers are everywhere.Swyx [00:46:39]: Yeah. Vision transformers.Peter [00:46:40]: And then then the compute just keeps getting better and better. But you do have this fundamental trade-off, right? It's like you have power, you have cost, and performance and like, getting the right, getting the right mix of those things in an embedded package that can also be, like, shaken and baked in all thePeter [00:47:00]: conditions that these things have to have to operate in. But yeah, I think that they're only going to keep getting better and so we also try to plan our strategy understanding that, we know the rate of improvements of these systems.Swyx [00:47:11]: Yeah. So like, Google just released the Gemma 2B modelSwyx [00:47:15]: that effective 2B model. Is that useful to you guys or is that too big?Peter [00:47:18]: You can run that model on an embedded system, definitely.Peter [00:47:21]: the So yes, it's, it's useful in that regard. The bigger question is, like, what do you use it for in an embedded system? Like, you actually need to customize it quite a bit to make it useful for something. But yeah, you could run a two billion parameter model, definitely.Swyx [00:47:35]: It also interesting, like, what percent is a custom ML model that only does that thing versus a generalist LLMSwyx [00:47:41]: which probably is not that useful actually for your context.Peter [00:47:46]: Like, you, like, you can imagine different use cases, right?Peter [00:47:48]: So theSwyx [00:47:49]: The voice stuff, yes.Peter [00:47:49]: Yeah, the voice test. Totally, yes.Peter [00:47:51]: So for the actual, autonomy elements, that's 100% in-house. We do every bit of that, the data simulation, the model, everything. But when you get into the more generic use cases like voice or voice assistant kind of thing, that's where these more generalist models like Gemma actually can be quite, can be quite useful.Swyx [00:48:09]: Yeah. And then there's also obviously a trade-off between, like, what percent must you do on machine, versus just call home.Peter [00:48:16]: Yeah. It's all about latency.Swyx [00:48:17]: Latency.Peter [00:48:17]: It's all about latency. Yeah.Swyx [00:48:18]: Yeah. Well, like, I think actually in a lot of contexts, especially in the US, you can just have a connection to the web.Qasar [00:48:26]: Yeah. I think though most of our universe is everything has to be fairly, embedded and local because just the nature of Even in the US there's a lot of likeSwyx [00:48:39]: PatchinessQasar [00:48:40]: don't haveQasar [00:48:41]: have coverage, right? And if you look at, like, the old world of autonomy within mining, which is, like, long before transformers and kind of, neural networks, in the like CNN and kind of a universe, they were really just hand-coded, systems. They were just like, this machine is gonna run to that place with thisPeter [00:49:03]: That was our GPS, like very accurate GPS.Qasar [00:49:05]: Yeah. And so that worked, and that worked for 20 years, so why would we actually need to use transformers or kind of more modern end-to-end systems? Mainly because you can only really run a path and run backwards. That provided a lot of value, but m-Not as much as you get when the machine is actually intelligent. It's, it's seeing, it's perceiving, it's acting in a dynamic world.Alessio [00:49:28]: I looked up RTK, real-time kinematic, one to two-centimeter accuracy.Qasar [00:49:32]: Yeah. Fantastic. But the and fantastic in faraway lands where there's not gonna be cell phone coverage.Peter [00:49:39]: Yeah, so it's widely used on the legacy mining and agricultural autonomy systems today. So like, for example, a combine that can be precise within one or two centimeters as it's driving down the field, they use RTK.Qasar [00:49:53]: Yes.Peter [00:49:53]: But it's, it's expensive.Qasar [00:49:54]: Yeah. And it's, it's, it's autonomy, but it's not intelligent in the way that I think all of usQasar [00:49:58]: if in twenty-six we'd be talking about intelligence.Alessio [00:50:00]: In one of your blog posts, you mentioned research on large scale transformers that are similar to those doing modern generative AI. What are, like, the big differences other than, “You're absolutely right. I should steer the car, so you probably wanna remove that?”Peter [00:50:14]: We have a diversified bet strategy internally, and the reason we've done that is because we operate in now a bunch of industries, a bunch of geographies, and each of the approaches has, obviously a different risk to them.Peter [00:50:27]: And so like, we're not going to put all of our eggs in a single basket for a single approach because that approach may no

The Last Standee
111: AAA (Ascension, Cozy Stickerville, Tin Mint Games)

The Last Standee

Play Episode Listen Later Apr 27, 2026 73:13


Sorry we are late- again! We checked in with the doctor, it was fine, here's our writ! Well, jokes aside, WELCOME to episode 111: Three one's for three A's! The renowned A-Team (Alexis, Audrey and Alessio) are doing the honours in this episode! Right after the Standee Catch-up, we have Alexis talking about Ascension (the Deckbuilder game) - in its Third Edition incarnation to be precise, and with a namely great Steam adaptation too (assuming we can find the link...!). After that, it's Audrey's turn with Cozy Stickerville - a legacy sticker-based campaign you can play exactly twice - why is that great and why stickers? Listen to the episode to find out! Closing the episode, Alessio exploring one niche within the niche: Tin Mint games with three notable titles: Gamma Guild, Tin Helm and Judgemint of the Realm Lords! That's it for today! Hope you recharged your batteries with this episode! (Got it? Because it's AAA!)

Italiano ON-Air
Polpette o lecca-lecca? Uno scherzo... di buon gusto!

Italiano ON-Air

Play Episode Listen Later Apr 22, 2026 6:55 Transcription Available


Il 1° aprile è passato, ma gli scherzi continuano a far discutere. In questo episodio, Katia e Alessio ci raccontano come una notizia nata per gioco tra Ikea e Chupa Chups stia diventando una realtà gastronomica bizzarra: il lecca-lecca al gusto di polpette svedesi!Dagli schermi dei computer "bloccati" in ufficio alle tradizioni storiche del Pesce d'Aprile, esploreremo insieme la cultura dello scherzo in Italia e scopriremo perché questo annuncio "visionario" verrà presentato proprio durante la prestigiosa Milano Design Week.

Fluent Fiction - Italian
Brushstrokes & Dreams: A Tale of Art and Ambition in Rome

Fluent Fiction - Italian

Play Episode Listen Later Apr 19, 2026 17:43 Transcription Available


Fluent Fiction - Italian: Brushstrokes & Dreams: A Tale of Art and Ambition in Rome Find the full episode transcript, vocabulary words, and more:fluentfiction.com/it/episode/2026-04-19-07-38-19-it Story Transcript:It: La luce del sole di primavera ornava Piazza Navona di una luminosità dorata.En: The spring sunlight adorned Piazza Navona with a golden brightness.It: Le fontane zampillavano allegramente, accompagnate dal rumore delle chiacchiere e delle risate dei turisti e dei romani.En: The fountains cheerfully splashed, accompanied by the sound of chatter and laughter from tourists and romans.It: Oggi era una giornata speciale, c'era la gara annuale degli artisti di strada.En: Today was a special day; there was the annual street artists' competition.It: Tra la folla c'erano Giulia e Alessio.En: Among the crowd were Giulia and Alessio.It: Giulia era emozionata.En: Giulia was excited.It: Era la sua prima partecipazione alla competizione.En: It was her first participation in the competition.It: Voleva vincere, voleva che il suo talento fosse riconosciuto.En: She wanted to win; she wanted her talent to be recognized.It: Sognava di diventare un'artista a tempo pieno.En: She dreamed of becoming a full-time artist.It: Con occhi sognanti, osservava la piazza, respirava profondamente e si preparava a dipingere.En: With dreamy eyes, she observed the square, breathed deeply, and prepared to paint.It: Alessio si trovava poco lontano.En: Alessio was a short distance away.It: Con la sua esperienza di molti anni, era un concorrente temuto.En: With his many years of experience, he was a feared competitor.It: I suoi paesaggi colpivano sempre per la precisione e i dettagli.En: His landscapes always impressed with their precision and details.It: Ma oggi osservava Giulia con interesse.En: But today, he watched Giulia with interest.It: Notava la passione nei suoi occhi.En: He noticed the passion in her eyes.It: Giulia iniziò a dipingere.En: Giulia began to paint.It: Il suo stile era diverso, moderno e unico.En: Her style was different, modern, and unique.It: Esitava.En: She hesitated.It: Doveva seguire uno stile tradizionale, come gli altri, o osare con la sua visione?En: Should she follow a traditional style, like the others, or dare with her vision?It: Decise di rischiare.En: She decided to take a risk.It: Con pennellate sicure, portò sulla tela colori vivaci e forme audaci.En: With confident brush strokes, she brought vibrant colors and bold shapes to the canvas.It: La piazza si animò attorno a lei.En: The square came to life around her.It: I passanti s'arrestarono, attratti dalla sua opera.En: Passersby stopped, attracted by her work.It: Ammiravano la freschezza del suo lavoro.En: They admired the freshness of her piece.It: Giulia sentì la loro energia e dipinse con maggiore fiducia.En: Giulia felt their energy and painted with greater confidence.It: Alessio osservava dalla sua postazione.En: Alessio watched from his spot.It: I suoi occhi seguivano il movimento del pennello di Giulia con curiosità e ammirazione.En: His eyes followed the movement of Giulia's brush with curiosity and admiration.It: Quando l'ultimo colore fu posato sulla tela, il sole stava calando.En: When the last color was laid on the canvas, the sun was setting.It: Arrivò il momento della premiazione.En: The moment for the awards arrived.It: Giulia non vinse il primo premio.En: Giulia did not win the first prize.It: Era delusa, ma non abbattuta.En: She was disappointed, but not discouraged.It: Si avvicinò alle sue opere con un sorriso.En: She approached her works with a smile.It: Era il primo passo del suo sogno.En: It was the first step of her dream.It: Alessio si avvicinò a lei.En: Alessio approached her.It: "Hai talento", disse con sincerità.En: "You have talent," he said sincerely.It: "Posso aiutarti a crescere come artista, se vuoi."En: "I can help you grow as an artist, if you want."It: Giulia lo guardò sorpresa e grata.En: Giulia looked at him surprised and grateful.It: "Davvero?"En: "Really?"It: chiese emozionata.En: she asked excitedly.It: Era l'inizio di qualcosa di nuovo.En: It was the beginning of something new.It: La gara si concluse, e i colori della sera avvolsero Piazza Navona.En: The competition concluded, and the colors of the evening enveloped Piazza Navona.It: Giulia tornò a casa con un sorriso.En: Giulia returned home with a smile.It: Non aveva vinto il premio, ma aveva guadagnato qualcosa di più prezioso: la fiducia in se stessa e l'opportunità di imparare da un maestro esperto.En: She hadn't won the prize, but she had gained something more precious: confidence in herself and the opportunity to learn from an experienced master.It: Alessio e Giulia camminarono per la piazza, parlando d'arte e di sogni.En: Alessio and Giulia walked through the square, talking about art and dreams.It: Primavera a Roma era magica, e per Giulia era l'inizio di un'avventura indimenticabile.En: Spring in Rome was magical, and for Giulia, it was the beginning of an unforgettable adventure. Vocabulary Words:the sunlight: la luce del soleadorned: ornavathe fountains: le fontanesplashed: zampillavanothe chatter: le chiacchierespecial: specialethe participation: la partecipazionerecognized: riconosciutothe dream: il sognothe vision: la visionethe risk: il rischioconfident: sicurovibrant: vivacibold: audacithe passersby: i passantitheir energy: la loro energiagreater: maggiorethe curiosity: la curiositàadmiration: l'ammirazionedisappointed: delusathe awards: la premiazionetalent: il talentothe opportunity: l'opportunitàthe master: il maestroexcitedly: emozionataunforgettable: indimenticabilethe adventure: l'avventurathe square: la piazzasurprised: sorpresafeared: temuto

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0
Notion's Token Town: 5 Rebuilds, 100+ Tools, MCP vs CLIs and the Software Factory Future — Simon Last & Sarah Sachs of Notion

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

Play Episode Listen Later Apr 15, 2026 77:17


For all those who missed out on London, see you in Miami next week!Notion, the knowledge work decacorn, has been building AI tooling since before ChatGPT, with many hits from Q&A in 2023 and unified AI in 2024 and Meeting Notes in 2025. At the end of their last Make user conference, Ryan Nystrom teased Notion 3.0's Custom Agents - and they are finally embracing the Agent Lab playbook!Sarah Sachs and Simon Last of Notion join us for a deep dive into how Notion built Custom Agents, why it took years and multiple rebuilds to get right, and what it means to turn a productivity tool into an agent-native system of record for enterprise work.We go inside the product, engineering, evals, pricing, and org design decisions behind one of the most ambitious AI product efforts in software today — from early failed tool-calling experiments in 2022 to agent harnesses, progressive tool disclosure, meeting notes as data capture, and the long-term vision for software factories and agentic work.We discuss:* Sarah and Simon's path to launching Notion Custom Agents, and why the feature was rebuilt four or five times before it was ready for production* Why early agent attempts failed: no tool-calling standard, short context windows, unreliable models, and too much complexity exposed to the model* The “Agent Lab” thesis: not just wrapping a model, but understanding how people collaborate and building the right product system around frontier capabilities* How Notion thinks about roadmap timing: not swimming upstream against model limitations, but also building early enough that the product is ready when the models are* Why coding agents feel like the kernel of AGI, and how Notion is thinking about “software factories” made up of agents that spec, code, test, debug, review, and maintain codebases together* How Sarah runs AI engineering at Notion (“notes from Token Town”): objective-setting over idea ownership, low-ego teams comfortable deleting their own work, and a culture designed to swarm around fast-changing opportunities* The “Simon Vortex,” company hackathons, and why security gets pulled in early rather than late* How Notion organizes AI: core AI capabilities and infrastructure, product packaging teams, and a broader company mandate that every product surface must increasingly work for both humans and agents* Why prototypes have become much easier to build internally, and how “demos over memos” changes product development inside a tool the whole company already uses every day* Notion's eval philosophy: regression tests, launch-quality evals, and “frontier/headroom” evals that intentionally only pass ~30% of the time so the company can see where model capabilities are going* What a “Model Behavior Engineer” is, and why Notion treats eval writing, failure analysis, and model understanding as a distinct function rather than just software engineering* The changing role of software engineers in the age of coding agents, and why the new job looks less like typing code and more like supervising a rigorous outer system of agents, PRs, and verification loops* How the “software factory” should work: specs, self-verification, bug flows, subagents, and minimizing human intervention while preserving the invariants that matter* A live walkthrough of a Notion Custom Agent handling coworking space tenant applications by triaging email, enriching applicants with web search, and writing structured data into a Notion database* How agents compose inside Notion: shared databases as primitives, agents invoking other agents, “manager agents” supervising dozens of specialized agents, and memory implemented simply as pages and databases* Notion's take on MCP vs CLI: why Simon is bullish on CLI's self-debugging nature, where MCP still makes sense, and how Sarah thinks about capability, determinism, permissioning, and pricing alignment* The evolution of Notion's internal agent harness: from early JavaScript coding agents, to custom XML, to Markdown and SQL-like abstractions, to tool definitions, progressive disclosure, and a much shorter system prompt* Why Notion cares about teaching “the top of the class,” building for sophisticated operators rather than abstracting away too much capability for everyone* How agent setup works today: agents that can configure themselves, inspect their own failures, and edit their own instructions — with guardrails around permissions* How Notion prices Custom Agents: credits as an abstraction over tokens, model type, serving tier, web search, and future sandbox costs; why usage-based pricing was necessary; and how “auto” tries to match the right model to the right task* Why Notion is not eager to train a foundation model, where they do fine-tune and optimize today, and why retrieval/ranking is one of the most important investment areas as more searches come from agents rather than humans* Why Meeting Notes became one of Notion's strongest growth loops: not just as transcription, but as high-signal data capture that powers search, custom agents, follow-up workflows, and the broader system of record for company collaboration* Why Notion is more interested in being the place where collaboration data lives than in building hardware themselves — and how wearables or other capture devices may eventually feed into that systemSarah SachsLinkedIn: https://www.linkedin.com/in/sarahmsachsX: https://x.com/sarahmsachsSimon LastLinkedIn: https://www.linkedin.com/in/simon-last-41404140X: https://x.com/simonlastFull Video EpisodeTimestamps* 00:00:00 Introduction and launching Notion Custom Agents* 00:01:17 Why Notion rebuilt agents four or five times* 00:03:35 Building for where models are going, not just where they are* 00:05:32 The Agent Lab thesis, wrappers, and product intuition* 00:08:07 User journeys, leadership, and low-ego AI teams* 00:13:16 The Simon Vortex, hackathons, and bringing security in early* 00:16:39 Team structure, demos over memos, and building for agents* 00:20:25 Evals, Notion's Last Exam, and the Model Behavior Engineer role* 00:27:37 Evals as an agent harness and the changing role of software engineers* 00:30:42 The software factory: specs, verification, and agent workflows* 00:32:18 Live demo: a custom agent for coworking space applications* 00:35:08 Composing agents, manager agents, and memory as pages* 00:38:15 Notion Mail, Gmail, native integrations, and tools* 00:39:43 MCP vs CLI and the cost of capability* 00:44:13 When Notion uses MCP vs building its own integrations* 00:47:43 The history of Notion's agent harness rebuilds* 00:55:35 Power users, public tools, and the setup agent* 00:58:01 Self-fixing agents, permissions, and “flippy”* 01:01:13 Pricing, credits, and choosing the right model automatically* 01:09:01 Why Notion isn't training its own frontier model* 01:14:07 Retrieval, ranking, and search built for agents* 01:17:27 Meeting Notes as data capture and workflow automation* 01:21:18 Wearables, hardware, and Notion as the system of record* 01:23:45 OutroTranscript[00:00:00] Alessio: Hey everyone. Welcome to the Latent Space podcast. This is Alessio founder of Kernel Labs and I'm joined by swyx, editor of the Latent Space.[00:00:11] swyx: Hello. Hello. We're back in the beautiful studio that, uh, Alessio has set up for us with Simon and Sarah from Notion. Welcome.[00:00:18] Sarah Sachs: Thanks for having us.[00:00:19] Alessio: Thanks for having us. Yeah.[00:00:20] swyx: Congrats on the launch recently the custom agents, finally it's here. How's it feel?[00:00:26] Sarah Sachs: We ship things slowly. So it had been in Alpha for a little bit and at the point at which is it's an alpha, um, there's a group of people that are making sure it's ready for prod, and then there's a group of people working on the next thing.So sometimes some of these launches are a bit delayed satisfaction, so it's quite nice to remind yourself all the work you did because we do have a habit of like. Being two or three milestones ahead. Uh, just ‘cause you have to be, you know, you can't get complacent. Um, but it's been great that people understood how this is helpful.And I think that's just easier in general building AI tools today than it was two, three years ago. People kind of get it and so that user education, um, there's just, it was our most successful launch in terms of free trials and converting people and things like that. It was really successful, so yeah.But there's a lot to build.[00:01:12] swyx: Making it free for three months helps.[00:01:16] Sarah Sachs: Yep.[00:01:17] Simon Last: It was definitely super exciting for me because it's probably the fourth or fifth time that we rebuilt that.[00:01:22] swyx: Yes.[00:01:23] Simon Last: And I mean,[00:01:24] swyx: you've been building this since like 20, 22.[00:01:26] Simon Last: Yeah, I mean, like, it was even right when we got access to like GPT four in late 20 22, 1 of the first ideas we had is like, oh, okay, let's make an agent that I, we used the word assistant at the time, there wasn't really the word, the word agent yet, but, oh, we'll give an access to all the tools the notion can do, and then it, we run in the background like, like do work for us.And then we just tried that many times and it just. Was too early. Um,[00:01:48] swyx: I need to force you to like double click on that. What is too early? What didn't work?[00:01:52] Sarah Sachs: We were fine to, like, before function calling came out. We were trying to fine tune with the Frontier Labs and with fireworks, like a function calling model on notion functions.This is right when I joined. I joined because, um, we needed a manager as Simon was needed to be able to go on vacation. So, uh, that's, that's around when I joined, so you can speak much more to it.[00:02:11] Simon Last: Yeah, we did partnerships with both philanthropic and open AI at different times, uh, to try to, at the time the, I mean, when we first tried, there wasn't even a constant of like tools yet.We, we sort of designed our own like, like tool calling framework and then we tried to fine tune the models to, uh, to use it over multiple turns. Um, and because it, it didn't work well out the box, I think. Yeah. The models are just too dumb and the context thing was also way too short.[00:02:37] Alsesio: Yeah.[00:02:37] Simon Last: Um, and yeah, we just kind of banged our head against it for a long time.Uh, unfortunately it was always like, there was always like sort of. Glimmers that it was working, but um, it never felt quite robust enough to be like a useful, delightful thing. Um, until I would say, uh, the big unlock was probably like Sonic 3.6 or seven, uh, early last year. And that's when we started working on our agent, which we shipped last year.Um, and then, and then uh, uh, custom agents, kinda a similar capability and that, that one just took longer because we, we just wanted to get the reliability up a lot higher. ‘cause it's actually running in the background.[00:03:14] Sarah Sachs: And the product interface of like permissions and understanding, you know, this custom agent is shared in a Slack channel with X group of people and has access to documents that are surfaced to Y group of people.And the intersect experts, Y might not be whole. And so how do you build the product around making sure administrators understand that permissioning took multiple swings.[00:03:35] Alsesio: Everything is hard back at the end of the day. Yeah. I'm curious, like when the models are not working, how do you inform the product roadmap of like, okay, we should probably build, expecting the models to be better at some reasonable pace, but at the same time we need to, you know, you had a lot of customers in 2022.It's not like you were a new company or like no user base.[00:03:54] Simon Last: Yeah, I mean I think there's always the balance of, you know, like you want to be a GI pilled and thinking ahead and building for where things are going. Uh, but also you wanna be like shipping useful things. And so we always try to like, like keep a balance there.You know, we. We try to take clear, like a portfolio approach. You know, we're always working on multiple projects and, and we're always trying to work on, you know, maintaining things where that have already shipped, like, like shipping new things that are like eminently working well and make them really good.And, and then we wanna always have a few projects that are a little bit crazy. Um,[00:04:23] Alsesio: and what are the a GI peel projects that you have today? I'm curious about, uh, you don't have to share exactly what you're working on, but I'm curious what are things today that maybe in 18 months people will be like, oh, obviously this was gonna work[00:04:35] Sarah Sachs: 18 months.[00:04:37] Alsesio: Yeah, 18 months is, you know,[00:04:37] Sarah Sachs: it's a long time and Yeah. Yeah.[00:04:39] Simon Last: I mean, there's a number of things happening. I think one thing that's becoming more clear is I think like, like, uh, coding agents are the kernel of EGI, sort of, everything is a coding agent. Mm-hmm. I think that's, that's sort of one, one direction.Um, and then, yeah, the exciting thing about that is sort of your agent can sort of bootstrap its own software and capabilities and actually debug and maintain them. And so yeah, we're, we're, we're thinking a lot about that. And then, yeah, like, like another category of things that I'm, I'm really excited about is like, uh, we call the software factory also.People are using this, uh, this, this sort of word. Um, basically it just means can you create sort of like a, as automated as possible, a workflow for developing debugging. Mm-hmm. Merging, reviewing, and maintaining a code base and a service where there's a bunch of agents working together inside, and like, like how does that work?[00:05:28] Sarah Sachs: If you think back to your initial question, like, why did this take so long? I think something,[00:05:32] swyx: I didn't say that, but Yes. Okay. Go ahead.[00:05:34] Sarah Sachs: Why, what, what changed over the three and half years of trying[00:05:37] swyx: it? Exactly. Right. Because most people always say like, it didn't work yet. Then reasoning models came, then it worked.I was like, okay, let's go a little[00:05:43] Sarah Sachs: bit. That's, I mean, that's part of it, but I think the other part of it that I actually think is really what will set notion apart for every new capability is we have like. Two skills that are crucial when it comes to frontier capabilities. One is not letting yourself swim upstream.So like quickly realizing if you're just pressing against model capabilities versus not exposing the model to the right information, not having the right infrastructure set up. That and of itself is the skill of intuition. And the second is to see, okay, you're not swimming upstream. Which direction is the river flowing and what is like, how do we think ahead about the product and start building it even if it's not great yet, so that when it is there, we're ready for it.Right? And like those can sometimes feel like counterintuitive things. Like we can be trying to fine tune a tool calling model when they don't exist yet. And that the trick is to not do that for too long, but realize that there was something there. And we've had a lot of things which like, um, we're just like not swimming in the right direction with the streams.I think we had multiple versions of transcription before we got meeting notes, right? Oh, I gotta talk[00:06:39] swyx: about that. Yeah.[00:06:40] Sarah Sachs: Yeah. Um, and so. I, I, I think that like we, we really closely partner with the Frontier Labs on capabilities and we also have to have strong conviction on, as those capabilities move.Notion is about being the best place for you to collaborate and do your work. And how does that narrative change if the way that we work changes?Yeah.[00:06:58] swyx: Yeah. You told me you were a fan of the Agent Lab thesis, and this is, this is kind of it, right?[00:07:02] Sarah Sachs: Right. I show that thesis to so many candidates. Like I have it as like micro chrome autofill.Um, at this point, like it's one of my most visitations[00:07:10] swyx: because like, is this the, here's why you should work in notion and not open, open eye. I, it's like,[00:07:14] Sarah Sachs: here's, here's what's different about it.[00:07:16] swyx: Yeah.[00:07:16] Sarah Sachs: And here's why. It's not just a rapper. I actually think more and more people understand it's not just a wrapper.[00:07:21] swyx: Yeah.[00:07:22] Sarah Sachs: Um, and by the way, like in the beginning, parts of what we build are wrappers on functionality. That works well, of course, but that's not really the most, um. I would say that's not the product that, that drives revenue. And that's not necessarily always what users need.[00:07:35] swyx: I mean, you know, notion is the AWS wrapper, but like the, the wrapper is very beautiful and like very, very well polished.So[00:07:40] Sarah Sachs: like the analogy,[00:07:41] swyx: like[00:07:42] Sarah Sachs: the analogy that I've been coming back to his Datadog in AWS[00:07:45] swyx: Yeah.[00:07:46] Sarah Sachs: So, uh, Datadog could not exist with, without cloud storage. Right. That it's kind of fundamental that that works. Um, and AWS has like a CloudWatch product, but Datadog is an expert on understanding how people want observability on the products they launch.And we're experts in understanding how people wanna collaborate, and that's really where our expertise lies.[00:08:04] swyx: Totally.[00:08:04] Sarah Sachs: Um, regardless of the tools that we use,[00:08:07] Alsesio: I'm kind of curious how you think about implicit versus explicit expertise. I feel like Datadog is half and half implicit and explicit. It's like they understand across markets and industries what engineering teams usually look for.With notion, it's almost like more of the expertise is at the edge because you as a platform, you're like so horizontal that the end user is not really the same. Mm-hmm. Like with Datadog, the end user is always like, yeah, an engineering lead, a kinda like SRE related person with notion. It can be anything.So I'm curious how you put that expertise into a product versus, you know, obviously it, WS cannot build notion. It's, that doesn't quite work in this case, but[00:08:44] Simon Last: it's, it's a little bit differently shaped. I think, you know, a classic vertical SaaS, like the data is kind of like that. They understand their individual customer very deeply.It's kinda a narrow slice, um, notion has always been super horizontal. And our, our task has always been to sort of balance these two somewhat opposing forces of like, we're listening to our customers and what they want us to build. It's a broad slice. And then also we're thinking about like, okay, how do we decompose what they want into, uh, nice primitives that are, that are really nice to use and we'll, we'll get us like as much bang for the buck as possible.And then, you know. Maintain the whole system, make it all like, like super clean and nice to use.[00:09:22] Sarah Sachs: We still have user journeys. I mean, we still focus on like core. I actually think the failure of our team is when we focus too much on what are cools that are, what are tools that are[00:09:31] Simon Last: mm-hmm.[00:09:31] Sarah Sachs: Cool tools. I actually think that's when we make have the least velocity because you still need some sort of focus on a user journey.So like for instance, we'll all sit down every Friday and look at the P 99 of like the most token exhaustive custom agent transcript and just look at why it didn't do well and cut a bunch of tasks. Like we still focus on like, this has, like this should work. Email triaging should work. Mm-hmm. Right. And similarly, like when we're talking about before building, um, chatting, um, before we started filming about, okay, how can I do PDF export?Well that's functionality that then merits. Maybe we should build a tool that has access to a computer sandbox in a file system and the ability to write code. Right? Right. Um, but it's because we're thinking about the fact that our users to do their, to do their daily work, need to export PDFs, not because we're like, Hmm, I think a computer tool could be cool.Like, let's just see what happens. Mm-hmm. Like we, we have to focus on some user journeys, otherwise we just don't have like, enough strategy to, to prioritize.[00:10:29] swyx: I think there's a lot of like really strong opinions that you've had. Do you have like sort of like a towel of Sarah Sachs? Like, you know, like what, how do you run your team?Like I feel like you just have accumulated all these strong opinions. Obviously part, part of this is your, your token town thing.[00:10:43] Sarah Sachs: I think the TAs working with Service X is, um, you'd have to, it depends who you ask. Um, I think it depends if you're on my team or a partner Right. Or a vendor.[00:10:54] swyx: Yeah. There other people want to run their teams the way that you're Yeah.You're like bringing these things. And then also similarly, uh, Simon, when you did the custom agents demo, you had like, well, we've been using custom agents and here's the super long list of everything that we do. No humans ever read it. Right? That's what you said. I was like,[00:11:07] Sarah Sachs: yeah. So I think for, for me, um, something that I learned very quickly and became very comfortable with was that my job was not to be the ideas per person or the technical expert.My job was to make it so that everybody understood the objective, had a resource to help prioritize what they should work on, and had an avenue to prioritize what they thought was important. And I think that's true with all, all leadership, but I think especially on the AI team. Almost all of our best ideas come from prototypes, from people that have a cool idea because they saw a user problem, and it's a huge disservice if all of those ideas have to pass, like the sniff test of what me and a product partner or Simon and Ivan decided were the direction, right?Because a lot of what we're doing is leaning into capabilities, so. I think that's the first thing is like, I don't really view like the role of engineering leadership as like, uh, hierarchical, nor has it ever been, but especially now, like very willing to change direction based on, um, like proof is in the pudding.Yeah. And like, and I think we have rebuilt our harness three or four times. And when you do that, then the second rule of engineering leadership is like you need to build a team that's comfortable deleting their own code and is very low ego and is driven by what's best for the company. And, um, doesn't write design docs because they think it's their promotion packet.Right. And that's a culture that notion had long before I joined, but like our willingness to just swarm on different problems and um, redo things that we've built before because something has changed. Like, there's a lot of friction that can happen at companies when you do that. And it doesn't happen at Notion.And because it doesn't happen when new people join. Like they don't wanna be the ones that are saying, we shouldn't do this. I wrote that code. So then it's, you know, you, you create a culture that everyone thoughts and that culture comes directly, I think from Simon and Ivan though, um, because they're very open-minded.[00:12:50] swyx: Anything that you,[00:12:50] Simon Last: you'd add? I'm not a manager, like, like, like Sarah is. Um, a lot of my role is really to try to think a little bit ahead, make sure that we're, we're building on the right capabilities and then like the prototyping stuff. And yeah, it's really, really critical to always just be starting again.It's like, okay, this is new thing. What does this mean? What if we just rethought everything or wrote everything? And so I, I'm, I'm basically just doing that in a loop every six months.[00:13:16] swyx: Yeah. Do you believe in internal hackathons for this stuff?[00:13:19] Sarah Sachs: I think there's like two different versions. So one is like, we just have a, a, a solid bench of senior engineers that come and go on what we call the Simon Vortex and Productionizing what we built, right?Because when you're in the Simon Vortex, the velocity is super high. The direction changes daily, and it's meant to be like the equivalent of a SC Works lab. We don't need to do hackathons for that. We need to have senior engineers that we trust to come in and out of those projects. For instance, like management boundaries are really loose.Like you report to him, but you work for her right now. Yeah. That's something that when we hire managers, it's important they don't care about because we tend to form more structures. Yeah. Don't be too[00:13:54] swyx: territorial.[00:13:55] Sarah Sachs: We form more. It's after we ship things, not not before, just historically. Um, the second thing is we do have companywide hackathons.Actually we just had our demos day for the hackathon we had last week this morning. That's more for people that aren't directly working on the project, feeling like they have the time to pause and learn how to make themselves more productive or how they would use notion custom agents to build something.Or part of the hackathon was actually encouraging everyone across the company to build their own agentic tool loop, calling from scratch. Follow like an every blog post on how to do what I think because we want[00:14:26] swyx: just with the compound engineering one. Yeah.[00:14:28] Sarah Sachs: We want everyone to use cloud code in the company or whatever the coding agent they please and understand that fundamental.So we set aside a day and a half. We're all leadership, encourage everyone on their teams across the company to do it. So we have hackathons like that. I would say like kind of facetiously, like everything we build is a little bit like a hackathon until it graduates and puts on big boy pants and as a product ops rollout leader and has a assigned data scientists and stuff like that,[00:14:54] swyx: security review enterprise stuff,[00:14:56] Sarah Sachs: actually security reviews one of the things that we bring in first because it just slows us down way more and, um, causes a lot of tension and they build better product if they're involved early.So, um, that is probably the first person to get involved in something that's the[00:15:09] swyx: right PR approved answer.[00:15:10] Sarah Sachs: No, but it's not just PR approved. It like, um, um, it's[00:15:13] swyx: actually real. It's actually real. It's like, um, I'm just saying scar[00:15:15] Sarah Sachs: tissue.[00:15:15] swyx: Yeah,[00:15:16] Sarah Sachs: because like, you know, my background's also, I worked at Robinhood for a number of years.Yes. So like, uh, compliance and things like that, um, are a little bit more, you learn the hard way when it doesn't come naturally.[00:15:26] Simon Last: Yeah. I think the. The hackathon is really important for uplifting the general population, but like, if that's the only way you can build new things, you're kind of toast. I mean, it, it has to be like the daily processes, like, you know, building these new things.Um, and it has to be about, I think like, I think in the AI era a lot more leverage accumulates to the most curious and excited people. And so it's like we're all about just like activating that energy. You know, like if someone's protesting something on the weekend that they're excited about and it's important, that should be the main thing that we're doing.Yeah. Um, it's not a hackathon that we schedule once a quarter, it's just like, yeah. Daily process. Part of the culture.[00:16:02] Sarah Sachs: I mean, that's how we shift image generation and notion now. It was always this thing that would be kind of nice to have, but it wasn't really clear where that was necessarily aligned in product priorities.It'd be a lot of work. And we had someone on the database collections team, Jimmy, who was like. I really wanna do image generation for cover photos and inside notion. And we're like, if you wanna build it, like it's, do it please. Like we encourage you. We gave ‘em all the resources of working directly with Gemini and being able to like track the token usage and it working through endpoints.We gave them eval, support, everything, and then became a, a full project.[00:16:34] Alsesio: Yeah.[00:16:35] Sarah Sachs: That's why you can't have like ego as a, a leader. Like that's, that's how we work.[00:16:39] Alsesio: What's the size of the team today, both engineering and overall?[00:16:43] Sarah Sachs: I manage, uh, the team. That's what we'll call it. Core AI capabilities and infrastructure.That's about 50 people. But then we have per i partner teams that do packaging. So how it shows up in the corner chat versus custom agents versus meeting notes, that's another 30, 40 people. And, and then every team that has a product service at Notion that a user can interface with owns the tool that the agent interfaces with the editor team.The team that did CRDT for offline mode is the same team that handles how two agents, um, edit competing blocks. Mm-hmm. Right? It's the same problem. The team that built the underlying SQL engine is the same team that owns how the agent asks it to run a SQL query, and it does it performantly. And so from that regard, anyone working on product engineering is tasked with making them work for customers that are humans and agents because over time the majority of our traffic will be coming from agencies using in our interface, not humans.And so. Our objective is to make it so that the whole product org is building for agents.[00:17:40] Alsesio: Yeah. How has it changed internally? The activation bar is kind of lowered a lot. Like anybody can kind of create a prototype very, somewhat easily, especially if you're like an existing code base. Have you raised the bar on like what type of prototype people need to bring forward to gonna be taken?Not like seriously, but like, you know what I[00:17:58] Simon Last: mean? Yeah. I think the bar is lowered in many ways. Be like, one thing our, uh, our team built that is really cool is our, uh, our, our design team made a whole separate GitHub repo, uh, called the, the design Playground. And it's basically just to create a bunch of like, like helper components and you, uh, for, for quickly a throwing together UIs.And it's become like actually quite sophisticated. Like it has like an agent in there and like, uh, that's pretty fun. So like, we pretty much, like, they don't do mocks, they just make like, like full, full prototypes.[00:18:27] swyx: Here it is. It works.[00:18:28] Simon Last: They give you like a u rl. They're like, okay, all right. So we have to make the, like the real production version of that.Um, and then for engineers. A prototype looks like just making it a feature flag that actually works. Like that's sort of the bar.[00:18:39] Sarah Sachs: Something to understand that's really unique about notion. One of the reasons I joined we're super lucky is no one uses Notion in their job as much as people that work at Notion.[00:18:46] Simon Last: Of course.[00:18:47] Sarah Sachs: So I think there's very few companies, maybe if you worked on Chrome I guess, but like everything that we ship, we ship internally first and get a lot of really quick feedback. And also sometimes our dev instance is totally borked and you have to change a bunch of flags to get things done. And that's kind of like, but everyone, so people that do it ticketing, people that do supply chain procurement, recruiting, everyone is using the same instance of notion with like a lot of flags on for these prototypes people build.Um, and so we have this, Brian Levin, one of the designers on our team, I think evangelize this concept of demos over memos.[00:19:18] swyx: Ooh, too[00:19:20] Sarah Sachs: good. Um, which has been, uh, very good for building demos, and I think it's put a big pressure point on us to have really strong product conviction, because if anything can be demoed, you really need a strong filter of making sure that if you know, you're doing X amount of work, you're making the, you're, you're focusing on one tower, you're not just building a really flat hill.Right. That's actually where I think there has to be more conviction from our PMs, um, and our designers and, and well, the company really to have conviction of what journey we're going on.[00:19:52] Simon Last: But overall, I feel like it works pretty well. Like people, almost all the engineers have good enough taste to realize that like, this prototype doesn't actually make sense in the product, or, or it does.So it's not that common that I would see a prototype. It's like, oh, this makes no sense. Mm-hmm. It's like, you know, people are doing reasonable things and, and, and then it's just a matter of. Which things we build first and then often just, just figuring out how to turn it on and off. There's our, in the, in our like experimental chat ui, there's this, there's probably like, like a hundred check boxes in there.[00:20:22] Sarah Sachs: Kills me[00:20:23] Simon Last: the things you could turn on and off.[00:20:25] Sarah Sachs: Uh, but I think that, okay, so that is kind of true, Simon, but like being the person that manages the evals team, like there is a level of intensity that it adds to the platform team. So, you know, if we're gonna do image generation and notion, all of a sudden the way that we do attachments and the way that we, um, our LLM completion like cortex talks and expects tokens back and now it's getting images back.Like there's a lot of platform work that we do need to, like solidify a little bit. So sometimes it'll be in dev for a couple weeks before it makes it to prod just because we still have to like, make it robust, make it HIPAA compliant, ZDR compliant, figure out the right contracting with the vendor, whatever it is.And we need to eval it because we want the team. To still maintain what they build. That's the one thing is like if we have a bunch of prototypes, it can't just be like a small group of people that then maintain whatever end prototypes. So we have invested a lot of people in an eval and model behavior understanding teams that, we call it agent dev velocity.So your dev velocity building agents can be faster if we invest in that platform. And so we have a whole org dedicated to Asian, um, platform velocity so that you can build your own eval and then maintain it once you ship it. So if a new model release comes out and we, every[00:21:38] swyx: team maintains their own eval,[00:21:40] Sarah Sachs: we maintain the eval framework.Every team owns their own evals and a lot of them we've integrated to Optin, to ci, or we run them nightly and we have a team, uh, a custom agent that triggers to a team to look at the major failures. That's really critical because if we have like all these different surfaces now, a lot of it's on the same agent harness, so it's easier to maintain.It's just packaging of different agent harnesses, but new functionality of the agent. Let's say that like we wanna update like. Uh, you know, they deprecated, sonnet, um, four or whatever it is and we need to auto update. Are[00:22:11] swyx: they already? That's so, okay. Yeah. Actually wasn't that long ago.[00:22:14] Alsesio: Theywere[00:22:14] Alsesio: just 3.5.[00:22:15] Sarah Sachs: 3.537. Just got deprecated.[00:22:18] swyx: 3 7, 5 0.2 or, yeah. No,[00:22:20] Sarah Sachs: it's not. 5.2 is five point. Five point no. Yeah, five four is 40% more expensive than five two. So if they deprecated five two, you would hear they can, you would hear from me about that one. Um, but, uh, another conversation to have.[00:22:35] swyx: I have a cheeky evals question for you.Have you noticed any secret degradation from any of the major model providers?[00:22:40] Sarah Sachs: Secret degradation,[00:22:42] swyx: like. During the War Bay, when it's high traffic, it suddenly gets dumber.[00:22:47] Sarah Sachs: Yeah. I mean, not just between the, I mean, we definitely notice flakiness, we've definitely noticed, particularly for some providers, that things are slower during working hours and[00:22:57] swyx: there's a latency argument.Yes. Not a quality argument.[00:22:59] Sarah Sachs: No. I think the quality difference that's interesting is, um, even though companies that say they're selling the same, a, it's really into like quanti quantization, but like companies that say they're selling the same model through different vendors, whether it be through first party or Bedrock, Azure, et cetera.We do see different qualities sometimes, and that's not necessarily what's advertised.[00:23:21] swyx: Yeah. Kidney went to the point of like, if we, they shipped like this, like eval across all the providers and it was like very obvious we were secret equalizing and it was very,[00:23:28] Sarah Sachs: yeah. But[00:23:29] swyx: that's very embarrassing.[00:23:30] Sarah Sachs: You know, um, we hire Subprocess to figure that out for us.So we just wanna understand where it's regressing or where it's optimized. And sometimes we're okay with regressions that optimize latency if they're the appropriate regressions. Our job is to make sure we have the evals to understand the changes that are important to us. And even like when we're partnering with labs on pre-releasees of models, they'll send us multiple snapshots.And this is less about quantization, but more just regressions. Like they have shipped models that were not the snapshots that we wanted, and they have changed the snapshots that they shipped based on the feedback that we give. Because our feedback tends to be more enterprise work focused and not coding agent focused.And definitely those can be bummers, like, you know, uh, we know that this wasn't the version you wanted, but we'll help you make it work. I mean, we always make it work, but that definitely happens.[00:24:16] Alsesio: Yeah. Do you have, um, failing evals that you're just hoping, oh, that will have success eventually when a good model comes out?[00:24:23] Sarah Sachs: Uh, I mean, yeah. So I think. I mean, I could talk about this for 60 minutes, so I will limit myself. I think it's a real issue when people say evals and it's just like, that's quality, that's like unit, I mean, it's like saying testing. It's not just unit tests, right? So. We have the equivalent of unit test.Regression test. Those live in ci, those have to pass a certain percent, you know, within some stochastic error rate. Then we have, as you're building a product, evals of these aren't passing right now, and this is launch quality. So we have a report card and we need to, on these categories, you know, be it 80 or 90% of all of these user journeys to launch, and then what we have what we call frontier or headroom evals, where we actively wanna be at 30% pass rate.And that's actually been a effort that we took in partnership with philanthropic and OpenAI in the past maybe two or three months, because we actually hit a point where our evals were saturated and we weren't able to really give insightful feedback other than it wasn't worse. And not only is that not helpful for our partners, it's not helpful for us to understand where the stream is going.You know, going back to that analogy. And so we spent a lot of time thinking about. What notions last exam looks like, right? Mm-hmm. Not just humanities, last exam. Ooh, notions last exam. Mm-hmm. And, um, there's a lot of, you know, dreams about what that would look like. I know we've talked a lot about benchmarking, um, swix, but, uh, yeah.Notions last exam is a big thing inside the company and we have people, full-time staff to it exclusively. Mm. We have a data scientist, a model behavior engineer, and an full-time, um, evals engineer just dedicated to the evals that we pass 30% of the time.[00:25:56] swyx: What you're hiring for[00:25:57] Sarah Sachs: MBEs? I am hiring[00:25:58] swyx: What is an MBEA[00:25:59] Sarah Sachs: model?Behavior Engineer Model. Behavior engineers started with a title data specialist before I joined when they were working with Simon on like, uh, Google Sheets and like Simon just needed someone to look through Google Sheets and say, yes, no, this looks bad. This looks good. Right? And so we hired people with kind of diverse linguistics background.We had like a linguistics PhD dropout. Mm-hmm. And a Stanford ate new grad. And they're amazing. And they formed a new function basically. And over time we've built a whole team, um, with a manager who's now kind of reinventing what that role is with coding agents. So they used to be kind of manually inspecting code.Now they're primarily building agents that can write evals for themselves or LLM judges. There's a really funny day I can send you the picture where Simon, about a year and a half ago, was teaching them how to use GitHub. Um, and they're on the whiteboard and it was like, okay, I think it would be so much faster if our data specialists learned how to use GitHub and like learned how to commit these things in Dakota.And, and that was then and now I think, you know, coding has been a lot more accessible. Um, but moving forward it's this mix of like data scientist PM and prompt engineer because there's craft in understanding like even like what models can and can't do things. How do we define like that headroom? How do we define like what a good journey is?Um, is this model better or not? Why is this failing? There's some qualitative work, but then there's also like a lot of instinct and taste to it, and that's not necessarily software engineering. And so we have like very firm conviction and we have had for a number of years now that that is its own career path and we have always welcomed the misfits, so to speak.So we really firmly believe that you don't need an engineering background to be the best at this job. And that's what's quite unique about this particular role.[00:27:37] Simon Last: Yeah, this is something that I've been pretty excited about recently is we made an effort basically to treat the eval system as like an agent harness.So if you think about it, like, you know, you should be able to have an agent end-to-end, download a dataset, run an eval, iterate on a failure, debug, and, and then implement a fix. And ultimately you should be able to, you know, drive the full time process with a human sort of observing the, you know, the outer uh, system.So yeah, we went, went pretty hard on that. And that's, that's worked extremely well so far. It's like basically just to turn it into a coding agent, uh, uh, problem.[00:28:11] swyx: Your coding agent or just whatever[00:28:13] Simon Last: harness No coding agent. Yeah, code, cloud code. It should be totally general. Yeah. I think if it would be a mistake to like, like fix it on any, any particular coding agent.At the end of the day, it's just like CLI tools.[00:28:21] Sarah Sachs: It's like the same way that you would've a coding agent write the unit test. You should have a coding agent write the eval.[00:28:26] swyx: Yeah.[00:28:26] Sarah Sachs: But there's a lot of supervision in that still. We just don't believe that supervision has to come from software engineers because a lot of it is like, um, kind of you XREE and whatever, and these are the people that also triage failures and tell us where we should be investing next.[00:28:40] swyx: Yeah. I'm gonna go ahead and ask a spicy question. Is there a data, there are no software engineers at Notion.[00:28:46] Simon Last: Um,[00:28:46] Sarah Sachs: what does it mean to be a software engineer?[00:28:47] swyx: Exactly.[00:28:48] Simon Last: I mean, I think the way things are going is like we're on some continuum where. If, if you look back three years ago, humans were typing all the code and then we had auto complete, you're typing list of the code.Then we had sort of like filling agents, filling lines, and now we're getting into like agents doing longer range tasks where you can debug and implement a fix and then verify it works and you know, get your, get your PR even like, like Merion deployed. I think we're sort of just moving up the abstraction ladder and then the human role becomes more about observing and maintaining the outer system.There's a string of agents flowing through, like me prs what's going off the rails. Like what do I need to approve? Is there like a learning or memory mechanism that that works? So it's kind of a hard engineering problem. There's a, you know, there's, there's a lot to do there. I think we're just sort of moving up stack[00:29:34] Sarah Sachs: the same transition machine learning engineers have made, right?Like I haven't looked at a PR curve in a while.[00:29:39] swyx: Yeah. You used to do this stuff and now, um, auto research can do it,[00:29:42] Sarah Sachs: right? Like I think it depends on what you define as a software engineer.[00:29:46] swyx: Yes. It's, that's changing for sure.[00:29:49] Sarah Sachs: I think every software engineer in notion this summer went through like this, um, sheer, um, one of our engineering leads of the company called it, like every software engineer is going through the, the, uh, identity crisis that every manager goes through, where all of a sudden they realize their ability to write code is less important than their ability to delegate in context switch.And I think that is a transition out of being a software engineer. But[00:30:12] Simon Last: yeah. Yeah, there's a critical difference to being a manager, which is that like, it is actually very deeply technical. The problem, you know, humans are very like, like, like fuzzy and you can't like treat a team of humans like a, like a rigorous system where like, you know, prs like, like flow through and can be in like a block status and then what happens when they're blocked, right.With a set of agents, you actually can do that. And, and, and I think it's actually, there's a lot of interesting technical rigor that that goes into that it's like it's a technical design problem. Ultimately.[00:30:42] Alsesio: What is the design of the software factory that you're building?[00:30:46] Simon Last: Yeah, I mean, I think we're. Trying a lot of different things.I mean, ultimately you want to design a system that requires as little human intervention as possible, but like still maintaining the in variance that, that you care about. So yeah, we're exploring a lot different ideas there. I mean, I think I could talk about a few things I think are important there.Like, one thing I think is really important is, um, having some kind of like specification layer you can just commit marked on files. Mm-hmm. That works pretty well, but[00:31:15] swyx: it's nice to be notion man. I'm just saying like the spec, like Yeah. The natural home for specs is notion.[00:31:21] Simon Last: Yeah. Right. It can be a database of pages.Yeah. I mean, it needs to be something that is, you know, human readable and I viewable and I think that's pretty key. Another really key component is like the, the self verification loop. Yes. You need really, really good testing layers, basically. And that's a really deep, uh, uh, problem. But by getting that right, you know, and then, and then it's kinda like the workflow of like.What happens when there's a bug? How does it flow into the system? Like, is it like a subagent working on it? How does it make a PR and how does that get reviewed? And me, and then, you know, so there's like the, the flow or process.[00:31:56] swyx: Yeah. Cool. Uh, you know, one thing we did work out before you guys came in was this demo or this[00:32:01] Simon Last: agents[00:32:02] swyx: agent demo.Uh,[00:32:03] Simon Last: so every,[00:32:04] Alsesio: every time we do an episode, we try the product. Right. I don't think there's ever been an episode that I haven't tried. Yeah. Um,[00:32:11] swyx: and we, we try, try is a, a big word. Like since day one lane space has been on Notion, but this is the, this is the net new thing. Yes.[00:32:18] Alsesio: So this is for Nel Labs, which is the space we're in.So next week we're opening applications for tenants. So there's a web form, let me, we got this form done here. Uh, so, uh, before. Uh, the workflow would be I get an email, then I look at the person. It was like, should I spend time talking to this person? Then I respond, they respond back. So I build this. So the name it came up for on its own.Can you maybe h how do, how does it come up with its own name?[00:32:43] Simon Last: Yeah, that's a pretty app name. It's, it, it is just a random, it's a random, a name generator.[00:32:47] Alsesio: Oh, that's funny. It just came,[00:32:49] Simon Last: the fact that it picked that is, is kind of hilarious. I'm pretty sure it's just determined,[00:32:54] Sarah Sachs: resilient collector. I, I think I've never looked at the code for that.I've never second guessed it. I think it's kind of like a madlib situation.[00:33:00] Simon Last: Yeah, I think you're right. Yeah. It's, it's totally a, a deterministic. Oh, I thought it was great. Yes. Although, although when the, if you use the AI to set itself up, it can update its own name, so. Okay. Um,[00:33:11] Sarah Sachs: how did you create it? It, did you just do[00:33:12] Alsesio: classroom?I,[00:33:13] Sarah Sachs: okay.[00:33:13] Alsesio: I did, yeah. I'll say just check my inbox for applications for a coworking space. Keep a people, so it created the database for me. Which I have here. And I guess database is like an notion table because everything is notion. Um, and then whenever um, an email comes in, like here, it just creates a new role for the person.Mm-hmm. And then it uses web search to enrich the mm-hmm. The profile. So it kind of like searches the web and it's like, this is who this person is, this is when they say they wanna move in and kind of updates everything else. This is, I mean, it's not a GI, but to me, I don't wanna do this work. So it feels like, I mean, it took me maybe like 15 minutes to set up the whole thing.Um, and I really like that most of the information should live here. You know, it is not like some other tool asking me[00:34:01] Sarah Sachs: Yeah.[00:34:01] Alsesio: To like, bring my stuff there. It's like I would've probably already created an ocean thing.[00:34:06] Sarah Sachs: Mm-hmm.[00:34:06] Alsesio: So[00:34:07] Sarah Sachs: most of our biggest use cases and gains are from. That extra layer of human involvement in the process to make it so right.And so like one of our biggest use cases is bug triaging. So if someone posts something in Slack, can you just have a custom agent that lives there that has its own routing constitution of what team this belongs to, creates a task in your task database and then posts in that Slack channel, right? Like that's like one of the first things that we built internally, I think.And it's completely changed the way that notion functions as a company. Nothing falls through, well, most things don't fall through the crack. We don't know what we don't know. But it's not replacing people, it's replacing processes.[00:34:44] Alsesio: Yeah.[00:34:44] Sarah Sachs: Right.[00:34:45] Alsesio: And I'm curious how you think about composability of these things.So the other one I was working on is like a. These filler. So whenever somebody signs up as a tenant, kind of he'll sell the lease for them. There should probably some agent that is like office manager agent mm-hmm. That can handle the request, make the lease, and then, uh, give them a ADA access to the office and all of that.How do you think about that feature?[00:35:08] Simon Last: Yeah, so I mean, there's, there's two ways you can compose. One way is by using like the data primitives. So you can, you know, you, you could give, you have one agent, uh, be writing to the database and there's another agent that's walked in the database. So that's, that's one way that they, they can coordinate that's like a little bit more decoupled and mm-hmm.Works really well. Or you, you can couple them. So I, I think it's actually not released yet. Releasing it like next week is, uh, in the settings for an agent, you can give access to invoke any other agent.[00:35:34] swyx: Hmm.[00:35:34] Simon Last: So you can have them just. Just, uh, uh, talk directly. So[00:35:37] swyx: you, was there a limit on like, number of recursions or just,[00:35:40] Simon Last: um, probably,[00:35:42] swyx: you know what I mean?Like, you can just get an infinite loop that way there's[00:35:45] Simon Last: some kind of Yeah,[00:35:46] Sarah Sachs: I think it's, there is actually a number somewhere.[00:35:49] swyx: I believe I'm just, you know, like, you're, you're, someone's gonna screw up. You[00:35:51] Simon Last: should you try to see[00:35:53] swyx: Yeah. I mean, everything's gonna be paperclips.[00:35:55] Simon Last: Oh, yeah. Yeah. But, uh, but, but that's really useful.Yeah. So we, you know, like I just, I, I helped, uh, someone internally the other day, they had, they had built like over 30 custom agents for, uh, for our go to market team doing all kinds of different things. You know, for example, like researching, you know, like, like filling information about, about a customer or like, like triaging customer feedback or like, uh, something like that.Literally over 30 of them. And, and then he, and then he even made like a database of all the agents and then he is like, okay, and, and now I'm getting 70, over 70 notifications per day with just the agents are blocked on various things. Uh, and then I was like, oh, okay, cool. You know, the obvious thing to do there is to make a manager agent,[00:36:32] Sarah Sachs: right?[00:36:33] Simon Last: That's gonna sort of blocks be another abstraction layer in between your, your, uh, uh, 30 agents. Uh, so yeah, we, we send out with like a manager agent and then has access to invoke all the other agents and it's sort of like, like watching and observing them and then it sort of, it just creates a layer of abstraction.So instead of 70 notifications per day, it's like, like five. And then, and then the manager agent can help like, uh, debug and fix any problems with the,[00:36:54] swyx: does this is a concept of like an inbox or something like piece, you're basically saying that they can message each other?[00:37:00] Simon Last: Yeah.[00:37:01] Sarah Sachs: Well[00:37:01] swyx: they use the system of record, which, which is[00:37:02] Sarah Sachs: notion, so we[00:37:03] Simon Last: actually, yeah, we didn't make any special concepts at all.[00:37:06] swyx: They're interested to the motion notifications that I would've got,[00:37:09] Sarah Sachs: they can just like write a task to a database that the other agent's task to listening to, or they can actually call a web book to the agent, like they can just add the agent. Okay.[00:37:17] Simon Last: Yeah, I mean, this is something that, that we're still working on.I, I think we, you know, like, like generally, generally the way we do these things is, you know, you first make it possible, maybe like a sort of janky way. So I, I, I think the way I set ‘em up is like, you know, we created like a new database that was sort of like issues mm-hmm. That the custom agents were, were experiencing, and then gave them all access to file an issue and then the manager has access to, to read the issues.Um, and that works pretty well, essentially like, like give it its own like internal issue tracker just for the agents. And then, you know, if that becomes a, a concept that seems useful, generally maybe we will think of how to package it in. But I mean, generally we try to just keep it to composing the primitive if we can.You know, another example of this is we have no built-in memory concept. Memory is, is just pages and databases. And so if you wanna give a memory, just give it a page and give it. Edit access to that page and the[00:38:03] swyx: human can edit it. Agent can edit[00:38:04] Simon Last: it. Yeah. And so that works, that pattern works extremely well on it.And you know, depending this case, you can have it be just a page or it could be an entire database with, you know, or, you know, I can have sub pages is is pretty on what you can do with that.[00:38:15] Alsesio: So when I was setting this up, uh, I connected my inbox and it was like, do you wanna use Gmail or Notion Mail? And I'm like, I don't wanna use Eater, I just want you to do it.I'm curious how you think about, you know, notion, mail, notion, calendar, all of these kind of ui ux interfaces, full stack[00:38:29] Simon Last: notion.[00:38:30] Alsesio: Yeah. When like at the same time you have the agents abstracting them away from you in a way, you know, how do you spend like the product calories so to speak?[00:38:37] Simon Last: Yeah, I mean, I think it's pretty important that you don't have to use, not your mail to connect to the mail capability.So we can just connect to Gmail or, or whatever you want, uh, to use. And we're thinking of the mail service as being really great to the extent that it's really agent built, right? So maybe the mail app is just sort of a prepackaged agent that helps you automate your, your inbox.[00:39:00] Alsesio: Yeah, the auto labeling is great.Think[00:39:03] Sarah Sachs: the, when we, um, integrate with Gmail for instance, we have a series of tools available that are available via MCP or API to Gmail. When we integrate with Notion Mail, we have the Notion Mail engineering team to build us the, um, exact right tools that optimize latency, optimize performance and quality.They own that quality. Um, there's product leads there. They're directly thinking about the user problems that happen in mail. So it tends to be when we build integrations and connections, we build natively first. Um, and then think about, um, extending them generally just because it's also easier. Mm-hmm. Um, um, to build natively first.Um, so that tends to be how we phase things out.[00:39:43] swyx: Talking about integrations, you prompted me, so I gotta ask. M-C-P-C-L-I. What's going on? What's the[00:39:48] Simon Last: Yeah. Opinion. I think, I mean, I'm, I'm definitely bullish and excited about cli. I think there's a few really cool things about cli. So one really cool thing is like, um, is that it's in the terminal environment, so it gets a bunch of extra power.So it, you know, for example, it can like, like paginating and cursor through like long outputs. Um, and it has a progressive disclosure inherently. Uh, so, you know, you don't see all the tools at once. It's just, you see the CLI wrapper and you can like use the, the help commands and, and, and read files. And then I think the most important thing that's, that's super cool is that there, it's also inherently a, a bootstrapped.So if there's an issue, uh, the agent can debug and fix itself within the same environment that it uses the tool.[00:40:30] swyx: Mm.[00:40:30] Simon Last: Right. Like, you know, I think I saw a tweet this morning. Someone said, you know, my agent didn't have a browser, so I asked it to make all a browser tool and within a hundred lines of code, it gave itself a little browser, like, like wrapping the, the, the chromium API, um.That's pretty incredible. And then if there was a bug, it would just immediately try to fix it. Mm-hmm. Right. On the other hand, if you use an, you know, if you use like of, of the Chrome dev tools, MCP, I've had this issue where like, like sometimes the transport gets like messed up. If it gets messed up, the agent has no way to fix itself.It, it no longer has a browser, it's, it's not broken. Right. I think that's, that's pretty fundamental, but I would say like a lot of the, the bad things about it can be fixed. Uh, so I think like, as a progressive disclosure, that can be fixed with, with right harness. Like, it, it obviously doesn't make sense to show it all the tools all the time.That's not really inherent to the MCP protocol. It's just like how you wrap it and use it.[00:41:16] swyx: There's many poorly built MCPs because we didn't know.[00:41:19] Simon Last: Yeah, yeah. I mean it was just early, like, like the obvious thing is, uh, you know, to start with is, is to just show it all the tools and it's like, okay, now we have a hundred tools.Yeah. And like the tool calling actually works. So let's of[00:41:28] swyx: your success[00:41:29] Simon Last: give it a way to like, like filter to source the tools. So yeah, I would say like broadly speaking, I'm really bullish on cli. I'm still bullish on CPS and in a certain environment. I think in, in particular, CP is really great for when you want sort of like a narrow, lightweight agent.I think there's, there's definitely a lot of use cases where, where you don't want like a full coding agent with a compute run time. And also you want it to be like more tightly permissioned. MCP inherently has a really strong permission model, like all you can do is call the tools. A CLI is a little bit murkier.It's like, can I access the, if PI token are you, like, properly sort of like re-encrypt the token so it can't like exfiltrate it, it introduce a lot of like, like new issues, which are. Real and hard to solve. And MCP is just like the dumb simple thing that works and it that it's pretty good.[00:42:12] Sarah Sachs: I'll add two more perspectives, not from it working well for Notion, but how notion like commits to both platforms.Notion is dedicated to being the best system of record for where people do their enterprise work. So we will always support our MCP and so far as other people are using cps, right? So regardless of our perspective, we've put a lot of effort into our MCP and we have a fantastic team that we're building, um, to do more there.And the second thing I'll say, I think, um, we all think a lot, but lately I've been thinking a lot about making sure there's a value alignment and pricing, um, with capability.[00:42:43] swyx: Literally our next question[00:42:44] Sarah Sachs: and. Needing language to execute deterministic tasks feels wasteful and requiring on a language model to interface with third party providers seems wasteful for tasks that don't require it.And particularly because our custom agents are using usage-based pricing. We think of pricing as like the barrier of entry for use of our product, and we're quite committed to making sure that it's not wasteful. Um, not just because it's a bad deal for our customers, but it's also bad business. We wanna have as many buyers, like there's a, there's an elasticity of demand and so if we can have our agents properly execute code that calls on CLI deterministically, it's a one-time cost, right?Versus constantly having a language model integrate with an MCP over and over and over and paying those like repeated token fees and it's happening outside the cash window, then you're paying for it over and over and over and it's just kind of unnecessary and less deterministic when it doesn't have to be.[00:43:36] Alessio: Yeah, the open-endedness I think is like, the main thing is like, well, if I go write code to just call an API, I would never use an MCP. But then you need an NCP sometimes when you know what to call, but you don't want it to restart versus like, I think the it built a browser from scratch is like, it's great when you're doing it on your own, but like if your customers were having your AI write a browser from scratch every time and you had to pay the token cost of that, yeah.You'd be like, no, no. The Chrome dev tools CP is actually pretty great. Just use that. I'm curious, how do you make that decision? Like should it be. Just straight API call very narrow. Should it be an MCP? Should it be super open-ended?[00:44:10] Sarah Sachs: Do you mean for when we ship notion capabilities or when we add capabilities to[00:44:13] Alessio: notion[00:44:14] Sarah Sachs: AI or,[00:44:14] Alessio: I mean, you might have a capability that the only way to do is an open-ended agent, like an agent with a coding sandbox.[00:44:21] Sarah Sachs: Yeah. In Notion ai they're not explicit, not We also ship an MCP.[00:44:24] Alsesio: Yeah. Yeah. In B,[00:44:25] Sarah Sachs: yeah.[00:44:26] Alsesio: Internally. Okay. Like is there ever a discussion of like, we're not gonna ship it because we're not able to tie it down? Or are you happy to just like,[00:44:33] Sarah Sachs: um, no. I mean, there are a lot of things where we choose not to use MCP because we wanna add more high touch to quality.I think search an agent to find is like the largest instance of that, where we have. Um, slack and linear and Jira search and notion that is not using necessarily the search MCP functionality that is provided by those companies. And that's because it's quite critical we think, to how our agent trajectories work is for us to have a little bit more control on the functionality of the search journey.And so it usually comes from quality and there's a long tail of things and that's why we built an MCP client or an MCP server, excuse me, so that people can connect whatever they want. There's that long tail, right. But we, for search particularly, I would say that's like the primary entry point, but there are other connections as well that it's a little bit of secret sauce a

Midrats
Episode 754: European Navies' Lessons, with Alessio Patalano

Midrats

Play Episode Listen Later Apr 13, 2026 64:05 Transcription Available


The last four years' conflicts from the Strait of Hormuz through the Red Sea to the Black Sea have presented a raft of lessons to the navies of Europe. How are they positioned to address the lessons, and what moves are already taking place?Returning to the Midrats Podcast to discuss this and related topics is Alessio Patalano.Alessio is a Professor of War and Strategy in East Asia and senior fellow at the Center for Statecraft and National Security at King's College London, where he specializes in maritime strategic issues.SummaryIn this episode, Alessio Politano, Mark, and Sal engage in a deep discussion on the evolving landscape of naval security, strategic innovation, and the importance of historical and contemporary insights in shaping maritime defense policies. Main topics include:The significance of maritime history and its influence on current naval strategiesChallenges facing the UK Royal Navy and European navies amid funding and technological gapsModern threats in the Red Sea, Persian Gulf, and beyond, including missile and drone warfareInteroperability and technological advancements in NATO naval forcesThe strategic importance of autonomous systems and undersea infrastructure resilienceTimestamps:00:00 - Introduction and overview of current naval strategic challenges02:11 - Major recent regional conflicts and their global implications03:09 - Mritime strategy and how history informs modern security04:48 - The importance of understanding maritime history in policy making05:45 - Lessons from past empires and their relevance today07:36 - Strategic literacy among policymakers and military leaders08:49 - The impact of natural disasters and supply chain disruptions (e.g., Japan 2011)10:28 - Europe's response to emerging naval threats and fleet modernization efforts11:51 - The role of anti-access and area denial (A2/AD) systems in modern warfare13:23 - Challenges faced by European navies in resource allocation and modernization14:48 - The Red Sea operations: European and NATO approaches to maritime security17:01 - Lessons learned from Ukraine and how they influence fleet development18:24 - The state of the Royal Navy's readiness and funding issues19:48 - Upgrades and challenges regarding naval guns and missile defense systems20:45 - British Navy's current strategic considerations and historical perspective22:23 - Political and financial factors impacting UK naval capabilities23:13 - The importance of strategic investments and capability development26:33 - The role of autonomous systems and unmanned vessels in future naval missions33:24 - Regional missile threats, focusing on Iran and Chinese developments37:18 - Europe's plans for missile defense and cooperation with the U.S.44:36 - The significance of interoperability and joint exercises50:07 - Building resilience through technology, autonomy, and international collaboration55:09 - Critical infrastructure protection in the Baltic and North Sea62:57 - Future trajectories for European and Asian navies63:13 - Alessio's upcoming projects and publicationsResources & Links:Books by Alessio PatalanoThe Sun Also Rises — by Ernest HemingwayFleet Tactics and Naval Operations, Third Edition — by Wayne Hughes:Centre for Statecraft and National Security at King's College LondonBooks by Sam J. TangrediProject BeehiveRussia probing of the UK seabed resourcesNATO's Baltic Sentry

Lytes Out Podcast
John Alessio - Fighting in the 90s | MMA History Podcast

Lytes Out Podcast

Play Episode Listen Later Apr 6, 2026 131:37


Join hosts Mike Davis and Joey Venti on the MMA History Podcast for Episode #323 with John "The Natural" Alessio. The former UFC Welterweight title challenger, 50-fight veteran, and Canadian MMA pioneer discusses his remarkable career from outlaw shows in British Columbia to the bright lights of the UFC, PRIDE, King of the Cage, Pancrase, and Super Brawl.Alessio turned pro as a teenager after spotting a Gracie Jiu-Jitsu bumper sticker at a gas station. He shares wild early stories, including no-rules biker-backed events in a Sikh temple (with mid-fight police interventions and headbutts allowed), mismatched weight bouts in Japan, altitude struggles in tournaments, training at the Shark Tank with Ken Shamrock and legends like Pete Williams, and his UFC title shot against Pat Miletich at UFC 26 at age 20, making him the youngest fighter ever to compete for a UFC championship.Now a veteran with the Las Vegas Metropolitan Police Department on the FBI fugitive apprehension task force, Alessio also talks police jiu-jitsu training, life after fighting, and the raw, unregulated era of the sport he helped build. Packed with OG tales of Jason Height, Eddie Millis, John Peretti, Stefan Patry, Shawn Tompkins, and more.If you love early MMA history, Canadian pioneers, and unfiltered fighter stories, this is essential listening.Subscribe for more interviews and deep dives preserving mixed martial arts history.0:00 MMA history podcast intro 0:32 Joey Venti's guest introduction 0:54 interview start 0:57 current life/ career 3:49 50 fight club members from Canada 4:33 training accident with Javier Vasquez 6:45 Jason Heit9:11 beginnings in MMA10:34 Bikers supporting early MMA in Canada15:06 Ultimate Battle 4 man tournament 16:38 Adam Ryan 17:29 John Alessio vs David Harris19:33 Extreme Challenge 4 man tournament 23:39 dealings with Bill Mahood24:21 Kalib Starnes vs Nate Quarry 26:29 connection to Pancrase27:06 meeting Phyllis Lee 27:27 meeting Eddy Millis in Japan 28:04 John Alessio vs Kosei Kubota30:51 quitting Job to be a full time fighter 32:37 early support from friends 33:29 John Alessio vs Jason Meaders34:23 never turning down any fights 37:52 John Alessio vs Egan Inoue41:42 John Alessio vs Jay R Palmer43:14 John Alessio vs John Chrisostomo43:39 partying with the Lions Den 46:42 Lions Den initiation 49:03 John Alessio vs Pat Miletich52:40 UFC not acknowledging John Perretti 55:32 parents trusting fight career 58:14 arm barred by Pat Miletich 59:35 John Alessio vs Joe Doerkson1:00:34 getting cut from the UFC after title fight 1:03:29 Andy Anderson story after UFC event 1:08:34 declining involvement with the Bikers1:12:52 cancelled bout with Yves Edwards1:15:24 John Alessio vs Thomas Deny1:16:44 starting bad seed incorporated 1:17:59 dealings with Bobby Hoffman 1:19:07 Aaron Brink story 1:20:56 Bobby Hoffman story 1:24:40 fighting in the early MMA era 1:27:46 training camps with Billy Rush 1:28:53 Jeremy Horn not self promoting 1:31:08 John Alessio vs Sean Pierson1:31:37 interactions with Stephan Patry1:33:01 John Alessio vs Jason Black1:37:00 head kick KO to Sean Pierson 1:38:59 becoming a well rounded fighter 1:40:21 developing jiu jitsu with Javier Vasquez 1:41:17 interactions with Rafiel Torre1:42:58 John Alessio vs Chris Brennan 1:46:52 John Alessio vs Eiji Mitsuoka1:49:16 John Alessio vs Jason Black1:50:40 John Alessio vs Jason Black1:51:56 training with Karo parisyan1:52:57 John Alessio vs Ronald Jhun1:55:25 John Alessio vs Jonathon Goulet1:57:46 coached by Shawn Tompkins2:01:10 partying with Bas Rutten 2:01:40 John Alessio vs Shannon Ritch2:02:50 training with Dan Henderson and Matt Lindland 2:04:12 interview wrap up 2:05:23 cancelled bout with Joey Villasenior2:08:02 cancelled bout with Nick Diaz 2:10:56 outro/ closing thoughts#MMAHistory #JohnAlessio #UFC #EarlyMMA

The Phone Hacks
Dude, Who's Been In My Car? with Alessio Carducci

The Phone Hacks

Play Episode Listen Later Apr 5, 2026 45:24


First timer Alessio Carducci joins us to tell us about the pokies, his piece of shit car and his gala show he's doing at the comedy festival. Recorded at Up The Duff studios at the Albion Hotel in Collingwood(Good beers and even better parmas) Oooh yeah. Join the PATREON HERE - Just $7 (AUD) for bonus eps and content - get tons of behind the scenes hacks and pranks and help keep this podcast going! Go watch Capper's special Hold Me Closer Tiny Cancer HERE Follow CAPPER and ROHAN and PHONE HACKS on Instagram Subscribe where you're listening and leave a review to get the word out thereSee omnystudio.com/listener for privacy information.

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0
Marc Andreessen introspects on The Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"

Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

Play Episode Listen Later Apr 3, 2026 76:20


Fresh off raising a monster $15B, Marc Andreessen has lived through multiple computing platform shifts firsthand, from Mosaic and Netscape to cofounding A16z. In this episode, Marc joins swyx and Alessio in a16z's legendary Sand Hill Road office to argue that AI is not just another hype cycle, but the payoff of an “80-year overnight success”: from neural nets and expert systems to transformers, reasoning models, coding, agents, and recursive self-improvement. He lays out why he thinks this moment is different, why AI is finally escaping the old boom-bust pattern, and why the real bottleneck may be less about models than about the messy institutions, incentives, and social systems that struggle to absorb technological change.This episode was a dream come true for us, and many thanks to Erik Torenberg for the assist in setting this up. Full episode on YouTube!We discuss:* Marc's long view on AI: from the 1980s AI boom and expert systems to AlexNet, transformers, and why he sees today's moment as the culmination of decades of compounding technical progress* Why “this time is different”: the jump from LLMs to reasoning, coding, agents, and recursive self-improvement, and why Marc thinks these breakthroughs make AI real in a way prior cycles were not* AI winters vs. “80-year overnight success”: why the field repeatedly swings between utopianism and doom, and why Marc thinks the underlying researchers were mostly right even when the timelines were wrong* Scaling laws, Moore's Law, and what to build: why he believes AI scaling laws will continue, why the outside world is messier than lab purists assume, and how startups can still create durable value on top of rapidly improving models* The dot-com crash and AI infrastructure risk: Marc's comparison between today's AI capex boom and the fiber/data-center overbuild of 2000, plus why he thinks this cycle is different because the buyers are huge cash-rich incumbents and demand is already here* Why old NVIDIA chips may be getting more valuable: the pace of software progress, chronic capacity shortages, and the idea that even current models are “sandbagged” by supply constraints* Open source, edge inference, and the chip bottleneck: why Marc thinks local models, Apple Silicon, privacy, trust, and economics all point toward a major role for edge AI* American vs. Chinese open source AI: DeepSeek as a “gift to the world,” why open models matter not just because they're free but because they teach the world how things work, and how open source strategies may shift as the market consolidates* Why Pi and OpenClaw matter so much: Marc's claim that the combination of LLM + shell + filesystem + markdown + cron loop is one of the biggest software architecture breakthroughs in decades* Agents as the new “Unix”: how agent state living in files allows portability across models and runtimes, and why self-modifying agents that can extend themselves may redefine what software even is* The future of coding and programming languages: why Marc thinks software becomes abundant, why bots may translate freely across languages, and why “programming language” itself may stop being a salient concept* Browsers, protocols, and human readability: lessons from Mosaic and the web, why text protocols and “view source” mattered, and how similar principles may shape AI-native systems* Real-world OpenClaw use: health dashboards, sleep monitoring, smart homes, rewriting firmware on robot dogs, and why the most aggressive users are discovering both the power and danger of agents first* Proof of human vs. proof of bot: why Marc thinks the internet's bot problem is now unsolvable via detection alone, and why biometric + cryptographic proof of human becomes necessaryTimestamps* 00:00 Marc on AI's “80-Year Overnight Success”* 00:01 A Quick Message From swyx* 01:44 Inside a16z With Marc Andreessen* 02:13 The Truth About a16z's AI Pivot* 03:29 Why This AI Boom Is Not Like 2016* 06:33 Marc on AI Winters, Hype Cycles, and What's Different Now* 10:09 Reasoning, Coding, Agents, and the New AI Breakthroughs* 12:13 What Founders Should Build as Models Keep Improving* 16:33 AI Capex, GPU Shortages, and the Dot-Com Crash Analogy* 24:54 Open Source AI, Edge Inference, and Why It Matters* 33:03 Why OpenClaw and PI Could Change Software Forever* 41:37 Agents, the End of Interfaces, and Software for Bots* 46:47 Do Programming Languages Even Have a Future?* 54:19 AI Agents Need Money: Payments, Crypto, and Stablecoins* 56:59 Proof of Human, Internet Bots, and the Drone Problem* 01:06:12 AI, Management, and the Return of Founder-Led Companies* 01:12:23 Why the Real Economy May Resist AI Longer Than Expected* 01:15:53 Closing ThoughtsTranscriptMarc: Something about AI that causes the people in the field, I would say, to become both excessively utopian and excessively apocalyptic. Having said that, I think what's actually happened is an enormous amount of technical progress that built up over time. And like for, for example, we now know that neural network is the correct architecture.And I, I will tell you like there was a 60 year run where that was like a, you know, or even 70 years where that was controversial. And so, so the way I think about what's happening is basically, I think, I think about basically the, the, the period we're in right now is it's, I call it 80 year overnight success, right?Which is like, it's an overnight success ‘cause it's like bam, you know, chat GPT hits and then, and then oh one hits, and then, you know, open claw hits and like, you know, these are open, these are, these are like overnight, like radical, overnight transformative successes, but they're drawing on an 80 year sort of wellspring backlog, you know, of, of, of, of ideas and thinking it's not just that it's all brand new, it's that it's an unlock of all of these decades of like very serious, hardcore research.If I were 18, like this is a hundred, this is what I would be spending all of my time on. This is like such an incredible conceptual breakthrough.swyx: Before we get into today's episode, I just have a small message for listeners. Thank you. We will not be able to bring you the ai, engineering, science, and entertainment contents that you so clearly want if you didn't choose to also click in and tune into our content.We've been approached by sponsors on an almost daily basis, but fortunately enough of you actually subscribed to us to keep all this sustainable without ads, and we wanna keep it that way. But I just have one favor to ask all of you. The single, most powerful, completely free thing you can do is to click that subscribe button.It's the only thing I'll ever ask of you, and it means absolutely everything to me and my team that works so hard to bring the in space to you each and every week. If you do it, I promise you will never stop working to make the show even better. Now, let's get into it.Alessio: Hey everyone, welcome to the Lidian Space Pockets. This is CIO, founder Kernel Labs, and I'm joined by s Swix, editor of Lidian Space.swyx: Hello. And we're in a 16 Z with a, uh, mark G and welcome.Marc: Yes, yes. A and what, half of 16? Something like that. A one. Exactly,swyx: exactly. Uh, apparently this is the, the final few days in your, your current office.You're moving across the road.Marc: Uh, we're, yeah. We have a, we have some, we have some projects underway, but yeah, this is actually, oh, this is the original. We're in actually the original office. We're in the, we're in the, we're, we're in the whole thing.swyx: It's beautiful. Yeah. Great.Marc: Thank you.swyx: So I have to come out, uh, this is a, you know, I wanted to pick a spicy start in October, 2022.I just made friends with Roone and, uh, I wanted to give him something to sort of be spicy about. And I said, uh. Uh, it'll never not be funny. The A 16 Z was constantly going. The future is where the smart people choose to spend their time and then going deep into crypto and not in ai. And that was in October 22nd, 2022.And Ruen says there was an internal meeting in a 16 Z to reorient around Gen ai. Obviously you have, but was there a meeting? What, what was that?Marc: I mean, I don't, look, I've been doing AI since the late eighties.swyx: Yeah.Marc: So I, I don't know, like all that, as far as I'm concerned, this stuff is all Johnny cum lately.Yeah. You, I mean, look, we've been doing ar entire existence. I mean, we've been doing AI machine learning deep, you know, deeply. We've been doing this stuff way from the beginning. Obviously a AI is just core to computer science. I, I, I actually view them as like quite, uh, quite continuous. Um, you know, Ben and I both have computer science degrees.Um, you know, we, we both, Ben, Ben and I actually both are world enough to remember the actual AI boom in the 1980s. Yeah. There was like a, there was a big AI boom at the time. Um, and there was a, was names like expert systems. Um, and they of like lisp and lisp machines. Uh, I, I coded in lisp. I was coding a lisp in 1989.When that was the, the language of the AI future. Um, yeah. So this is something that we're like completely, you completely comfortable with. I've been doing the whole time and are very enthusiastic aboutswyx: is there a strong, like this time is different because, uh, my closest analog was 20 16 17. It was an AI boom.Mm-hmm. And it petered out very, very quickly. Um, we, it just, it just in terms of investingMarc: sort of, sort of,swyx: yeah. Investment, investment excitement.Marc: Although that's really when the, the, the Nvidia phenomenon really, it was, I would say it was in that period when it was very clear that at, at the time it, the vocabulary was more machine learning, but it, it was very clear at that time that machine learning was hitting some sort of takeoff point.Alessio: Yeah.Marc: Well, and as you guys, you guys have talked about this at length on, on your thing, but, you know, if you really track what happened, I think the real story is, it was, it was the Alex net, uh, basically breakthrough in like 2013. That was the, that was the real knee in the curve. Um, and then it was obviously the transformer breakthrough in 17.Alessio: Yeah.Marc: Um, and then everything that followed. But, but, you know, look, machine learning, you know, there were, you know, look, uh, I mean look, I've been working, you know, I've been working with, uh, one of my, you know, kind of projects working with Facebook since 2004. Um, and on the board since 2007, and of course, you know, they, they started using machine learning very early, um, and, you know, have used it basically, you know, for like 20 years for, you know, content, you know, feed optimization and advertising optimization.And obviously many, you know, financial services. You know, many, many, many companies, many different sectors have been doing this. And so it's like one of these things, it's like, it's not a, it's not a single thing. Like it's, it's like, it's like layers, right? Yeah. Um, and, and the layers arrive at different paces and, but they kind of build up.swyx: Yeah.Marc: Uh, they kind of build up over time and then, and then, yeah. And then look, in retrospect, it was 2017 was kind of the, you know, the key, the key point with the trans transformer and then. And then as you guys know, there was this really weird like four year period where it's like the, the transformer existed and then it was just like,swyx: let's go.Yeah.Marc: Well, but, but it was just, but, but between 2020, but between 2017 and 2021, I mean, that was the era of which like companies like Google had internal chat Botts, but they weren't letting anybody use them.swyx: Yeah.Marc: Right. And then, you know, and then OpenAI developed Chat GT or GPT two, and then they told everybody, this is way too dangerous to deploy.Right. Yeah. You know, we can't possibly let normal people, normal people use this thing. And then you, you guys, I'm sure remember AI Dungeon, um mm-hmm. So the o for, there was like a year where like the only way for a normal person to use GP T three was in, in AI dungeon.Alessio: Yeah.Marc: And so you, you, we would do this, you'd go in there and you'd pretend to play Dungeons and Dragons.In reality, you're just trying to talk to talk to GPT. And so there was this, you know, there was this long, you know, and I, you know, the big, big companies, you know, big companies are cautious and, you know, the big companies were cautious. It, it, by the way, it took open ai. You know, they, they, they talk about this, it took open AI time to actually adjust, you know, kind of re redirect their researchswyx: path.I, I think, uh, let say Rosewood, right? Uh, the, the dinner that founded OpenAI was right there.Marc: Right, right. But that, that dinner would've taken place in 20swyx: 18Marc: 19. The formation of OpenAI Uhhuh as late as 2018.swyx: Uh, uh, sorry. Uh, no, I'm, I'm, I'm, I'm wrong. Probably It should be 20. Yeah. They just celebrated a 10 year anniversary, so it it is 2025.Yeah, so, so 2015?Marc: Yeah. 2015. Yeah. 2015. But then, uh, um, Alec Radford did G PT one in what, probablyswyx: mm-hmm. 17, 18,Marc: yeah. 17, 18. So it, yeah. For, and then, and then they didn't really, and then GPT three was what? 2020? 2020.swyx: 2020.Marc: Because that became copilot immediately. Even open ai, which has been, you know, the leader of, of this thing in the last decade, you know, e even they had to adapt and, and, and lean into the new thing.And so. Um, yeah, I, I think it's just this process of basically sort of wave after wave layer after layer, you know, building on itself. And then you kind of get these catalytic moments where, where the whole thing pops and, and obviously that's what's happening now.swyx: Is it useful to think about will there be any ai, winter?‘cause there's always these patterns. Like, is this, in the summer is something I constantly think about because do I get, do I just like. Just get endlessly hyped and just trust that I will only be early and never wrong or right. Well, are we, will there be a winter?Marc: So there's something about, say the following.There's something about AI that has led to this repeated pattern. Um, and, and, and you guys know this,swyx: it's summer, winter, summer,Marc: winter, summer, winter, summer, winter. And it goes back 80 years. Yeah. 80 years. Uh, so the original neural network paper was 1943. Right. Which is, which is amazing. Uh, that it was, it was far back that long.And then there was you, if you guys have ever talked about this on your show, but there was this, uh, there was a big, uh, there was an a GI conference at Dartmouth University in 1950. 55. 55, yeah. And they got a NSF grant to, uh, for the, all the AI experts at the time to spend the summer together. And they figured if they had 10 weeks together, they could get a GI, uh, at the other end.And they got their, by the way, they got the grant, they got the 10 weeks and then, you know, 1955, you know. No, no. A GI. And like I said, I, I lived through the eighties version of this where there was a big, a big boom and a crash. And so, so there is this thing, and there, there is something about AI that causes the people in the field, I would say, to become both excessively utopian and excessively apocalyptic.Um, and, and it's probably on both sides of like the, the, the boom bus cycle. You, you kind of see that play out. Having said that, I think what's actually happened is like just, and you know, and we now know in retrospect like an enormous amount of technical progress that built up over time. And like for, for example, we now know that neural network is the correct architecture.And I, I will tell you like there was a 60 year run where that was like a, you know, or even 70 years or that was controversial. And, and we now know that that's the case. And so we, we now, you know, everything we're building on today just sort of derives from the original idea in 1943. And so, so in retrospect, we, we now know that like, these, these guys are right.They, they, you know, they would get the timing wrong and they thought, you know, capabilities would arrive faster, or they were, it could be turned into businesses sooner or whatever, but like, they were fundamentally, the, the scientists who worked on this over the course of decades were fundamentally correct about what they were doing.And, and the, and the payoff from, from, from all their work is happening now. And so, so the way I think about what's happening is basically, I think, I think about basically the, the, the period we're in right now is it's, I call it 80 year overnight success, right? Which is like, it's an overnight success.‘cause it's like bam, you know, chat, GPT hits and then, and then oh one hits, and then, you know, open claw hits and like, you know, these are open, these are, these are like overnight, like radical, overnight transformative successes, but they're drawing on an 80 year sort of wellspring backlog, you know, of, of, of, of ideas and thinking it's not just that it's all brand new, it's that it's an unlock of all of these decades of like very serious, hardcore research.Um, and thinking, and look, there were AI researchers who spent their entire lives. They got their PhD. They, they worked for, they've researched for 40 years. They retired in a lot of cases, they passed away and they never actually saw it work.swyx: Yeah. It's all sad.Marc: It is. It is sad. It's sad. Knewswyx: Jeff Hinton was like the last guy.Marc: Yeah. Yeah. Well, there were the guys, uh, was a guy, Alan Newell. I mean, there's tons of John McCarthy. You know, John McCarthy was like one of the inventors in the field. He's one of the guys who organized the Dartmouth Conference and you know, he taught at Stanford for 40 years. Wow. And passed, you know, passed away, I don't know, whatever, 10, 10 years ago or something.Never, never actually go. Got to see it happen. But like, it is amazing in retrospect, like, these guys were incredibly smart and they worked really hard and they were correct. So anyway, so then it's like, okay, you know, say history doesn't repeat, but it rhymes. It's like, okay, does that mean that there's gonna be another, like, you know, basically boom buzz cycle.And I, I will tell you, like, let, like in a sense, like yes, everything goes through cycles and, you know, people get overly enthusiastic and overly depressed and there's, there's a time, there's a timelessness to that. Having said that, there's just no question. Um, so the form, the foremost dangerous words in investing this time are, this time is different.Do you know the 12 most dangerous words investing? No. The four most d foremost dangerous words in investing are this time is different. Yeah. Um, the 12 most dangerous words. And so like, I'll tell you what's different. Like now it's working like, like there's just no, I mean, look, there's just no question.And by the way, I, I'll just give you guys my take. Like L LLMs, like from, from basically the Chad G PT moment through to spring of 25. I think you could still, I think well intention, well, and of. Form skeptics could still say, oh, this is just pattern completion. And oh, these things don't really understand what they're doing.And you know, the hall hallucination rates are way too high. And, you know, this is gonna be great for creative writing and creating, you know, Shakespeare and so sonnets and, you know, as, as rap lyrics or whatever, like, it's gonna be great and all that stuff, but we're not gonna be able to harness this to make this relevant in, you know, coding or in medicine or in law or in, you know, you know, kind of feels that, you know, kind of really, really matter.And I think basically it was the reasoning breakthrough. It, it was oh one and then R one that basically answered that question basically said, oh no, we're gonna be able to actually turn this into something that's gonna work in the real world. And, and then obviously the coding breakthrough over the, over basically the coding breakthrough that kind of catalyzed over the holiday break was kind of the third step in that.Mm-hmm. Where you're just like, alright, if, if, you know, if Linus Tova is saying that the AI coding is no better than he is like. Like, that's, that's never happened before. That's theswyx: benchmark.Marc: Yeah. That's never happened before. And so now we know that it's, it's gonna sweep through coding and, and then, and then we, we know, you know, we know that if it's gonna work in coding, it's gonna work in everything else.Right. It's just then, because that's, that's like, that's like, that's like the hardest in many ways. That's the hardest example. And how everything else is gonna be a, a derivative of that. And then on top of that, we just got the agent breakthrough, you know, with Open Claw, which is fantastic. Which is amazing and incredibly powerful.And then we just got the, the, um, the auto research, uh, you know, the, the self-improvement. You know, we're now into the self-improvement breakthrough. And so the, so the way I think about it is we've had four fundamental breakthroughs in functionality, l OMS reasoning, uh, agents, um, and then, uh, and, and then now RSI, um, and, and they're all actually working.Um, and so I'm, I'm just, as you like, you can tell I'm jumping outta my shoes. Like, like this is, like this is it like this, this is the culmination of 80 years worth of worth of work, and this is the time it's becoming real.Alessio: Yeah.Marc: I, I'm completely convinced.Alessio: I think the anxiety that people feel is like during the transistor era, yet Mors law, and it's like, all right, we understand why these things are getting better.We understand the physics of it. Yeah. With ai, it's. It's so jagged in like the jumps where like, like you said, it's like in three months you have like this huge jump like, and people are like, well this can keep happening. Right? But then it keeps happening,Marc: it'll keep happening.Alessio: And so like how do you think about also timelines of like what's we're building?I think we always have this question with guests, which is like, you know, should you spend time building harness for a model versus like the next model just gonna do it one shot in the lead space. Right. And how does that inform, like how you think about the shape of the technology? You know, you talk about how it's a new computing platform.If you have a computing platform, then like every six months it like drastically changes in what it looks like. It's hard to build companies on top of it.Marc: Yeah. So, so a couple things. So one is like, look, the, the Moore's law was what we now call a scaling law. Like Moore's Law was a scaling law and for your younger viewers, more Moore's Law was every chip chip chips either get twice as powerful or twice as cheap every, every 18 months.And that, and that and that, you know, that it's gotten more complicated in the last few years. But like that, that was like the 50 year trajectory of, of, of the computer industry. And then, and then by the way, and that's what took the mainframe computer from a $25 million current dollar thing into, you know, the phone in your pocket being, you know, a million times more powerful than that.Like that, you know, for, for 500 bucks. And so that, that was a scaling law. And then, and then, and then key to any scaling law, including Moore's Law and the AI scaling laws is, you know, they're not really laws, right? They're, they're, they're, they're predictions, but when they work, they become self-fulfilling predictions because they, they, they, they, they set a benchmark and, and then the entire industry, right?All the smart people in the industry kind of work to make sure that, that, that actually happens. And so they, they kind of motivate the breakthroughs that are required to, to keep that going. And, and in and in chips, that was a 50 year, that was a 50 year run. Right. And it, it was amazing. And it's still happening in, in some areas of, of chips.I think the same thing is happening with the, the core scaling laws. The core scaling laws. In, in, in ai, you know, they're, they're not really laws, but like they, they are basically. There are predictions and then they're motivating catalysts for the research work that is required to be. And, and, and, and by the way, also the investment, uh, dollars, um, uh, you know, required to basically keep, you know, keep the curves going and, and look, it, it is, it's gonna be complicated and it's gonna be variable and they're, you know, there're gonna be walls that are gonna look like they're fast approaching, and then they're gonna be, you know, engineers are gonna get to work and they're gonna figure out a way to punch through the walls.And obviously that's, you know, that's been happening a lot, you know, and then look, there's gonna be times when it looks like the walls have, you know, the, the, the laws have petered out and then they're gonna, they're gonna pick up again and surge and then, and then, and then it, it appears what's happening to the eyes is there's not multiple, you know, multiple scaling laws.Um, there's multiple areas of improvement. And, and I think, you know, I don't know how many more there are already yet to be discovered, but there are probably some more that we don't know about yet. You know, they, like, for example, there's probably some scaling law around, um, world models and robotics that we don't fully understand, you know, kind of acquisition of data at scale in the real world that we don't fully understand yet.So that, that, that one will probably kick in at some point here. There's a bunch of really smart people working on that. Um, and so, yeah, I, I think the expectation is that, that, you know, the, the scaling laws generally are gonna continue. Yeah. The, the pace of improvement will continue to move really fast.Um. To your question on like what to build. So, uh, I'm a complete believer the scaling laws are gonna continue. I'm a complete believer the capabilities are gonna keep getting amazing, um, you know, leaps and bounds. Uh, the part where I kind of part ways a little bit with how, what I would describe as the AI purists, um, you know, which is, which I would characterize as like the people who are.In many ways, the smartest people in the field, but also the people who spend their entire life, like at a lab, um, and have, have, I would say, have very little experience in the outside world. Um, the, the, the nuance I would offer is the outside world of 8 billion people and institutions and governments and companies and economic systems and social systems is really complicated.Um, and, um, and doesn't, you know, it it 8 billion people making collective decisions on planet Earth is not a simple process of like, just like you see this happening now. It's like a bunch of AI CEOs have this thing, which is just like, well, there's just this, they just all have this kind of thing when they talk in public where they're just like, well, there's these, these obvious set of things that so society to do.Alessio: Mm-hmm.Marc: And then they're like, society's not doing any of those things. Right. And it's like, how can society not, you know, what, whatever their theory is, how can society not see x, y, Z? Mm-hmm. And the answer is, well, society is number one. There's no single society, it's like 8 billion people. And they like all have a voice, and they all have a vote, like at the end of the day of how they, they react to change.And then, you know, it just like, it's just human reality is just really complicated and messy. Um, and, and, and so the specific answer to your question is like, as usual, it depends. Um, you know, it, it depends. Look, pe there's no question people are gonna, like, there's no question they're gonna be companies.It's already happening. There are companies that think that they're building value on top of the models and then they're just gonna get blissed by the, by the next model. There's no question that's happening. But I think there's no question also that just the process of adaptation of any technology into the real and into the real messy world of humanity is, is just going to be messy and complicated.It's, it's not going to be simple and straightforward. It's gonna be messy and complicated. And there are gonna be a lot of companies and a lot of products, um, uh, and in, in fact entire industries that are gonna get built to, to, to basically actually help all of this technology actually reach real people.Alessio: The amount of capital going into these companies, I mean, Dario talked about it on the Door Cash podcast and Door Cash was like, why don't you just buy 10 x more GPUs? And he is like, because I'm gonna go bankrupt if the model doesn't exactly hit the, the performance level. How do you think about that?Also as a risk on, you know, you guys are investors, open AI and thinking machines and world apps. It seems like we're leveraging the scaling loss at a pretty high rate, right? Like how comfortable, I guess, do you feel with the downside scenario, like, and say like things Peter out, you think you can kind of like restructure uh, these build outs and uh, you know, capital investments.Marc: Yeah. So should start by saying, so I live through the.com crash, um, and I can tell you stories for hours about the.com crash and it was horrible. No, it was awful. It was, it was, it was apocalyptic by the way. The, a lot of the.com crash was actually at the time, it was actually a telecom crash. It was a bandwidth crash.Like the, the thing that actually crashed, that wiped out all the money with the tele, the telecom companies.swyx: GlobalMarc: crossing. Global, global, yeah.swyx: I'm from Singapore and they, they laid so much cable o over over our oceans.Marc: Actually there was a scaling law in the.com. Era. And it was literally the, the US Commerce Department put out a report in 1996 and they said internet traffic was doubling every quarter.Um, and, and actually in 1995 and 1996, internet traffic actually did double every quarter. And so that became the scaling law. And so what all these telecom entrepreneurs did was they went out and they raised money to build fiber, anticipating that the demand for bandwidth is gonna keep doubling every quarter.Doubling every quarter though is like, you know, grains of chess and the chessboard, like at some point the numbers become extremely large. Right. And, and, and it really, and really what happened was the internet. The internet by the way, continuously kept growing basically since inception. And it's, you know, it's, it's continuously grown.It's never shrunk. And it's grown really fast compared to anything else. Mm-hmm. You know, in, in, in human history. But it wasn't doubling every quarter as of 19 98, 19 99. And so there was this gap in the expectation of what they thought was a scaling law versus reality. And that's actually what caused the.com crash, which was the, it they, they way over companies like global crossing way overbuilt fiber, which is sort of the, and by the way, fiber, telecom equipment, you know, so all the, all the networking gear, you know, and then, and then by the way, the actual physical data centers, like that was the beginning of the, of the, of the data center build and then, and the data center overbuild.And so you had that, but it was, it was literally, I think it was like $2 trillion got wiped out, right? It was like Jesus, it was like a big, it was. And by the way, the other, the other subtlety in it was the internet companies themselves never really had any debt. ‘cause tech, tech companies generally don't run on debt, but the telecom companies run on debt.Physical infrastructure companies run on debt. And so the companies like Global Crossing not just raise a lot of equity, they also raise a lot of debt. So they're highly levered. And so then you just do the thing. It's just like, okay, you have a highly levered thing where you're, you're just over, you're overbuilding capacity.Demand is growing, but not as fast as you hoped. And then boom, bankrupt. Right. And, and then it, and then it's like they say about the hotel industry, which is, it's always the third owner of a hotel that makes money. It has to go bankrupt twice, right? You have to wash out all of the over optimistic exuberance before it gets to actually a stable state.And then it makes money. So by the way, all of those data centers and all of those, all the fiber that they're in use, it's all in use today. Yeah. But 25 years later. But it, it, it took, and actually the elapsed time was, it took 15 years. It took 15 years from 2000 to 2015 to actually fill, fill up all that capacity.The cautionary warning is the, the overbuild can happen. Um, and, and, and, and, you know, you, you get into this thing where basically everybody, everybody who basically has any sort of institutional capital, it's like, wow. It's just, I, I don't know how to invest in these crazy software things. For sure I can put build data centers and for sure I can buy GPUs that I can deploy, you know, compute grids and, and all these things.Um, and so, you know, if you're a pessimist, you could look at this and you could say, wow, this is like really set up to be able to basically replicate, you know, what we went through, what we went through in 2000. Obviously that would be bad. The counter argument, which is the one I I agree with, which is the counter on, on the other side is a couple things.One is the companies that are investing all the, the companies that are investing the money are like the bluest chip of companies. And so back, back, back in the, in the do, like Global Crossing was like a, it was like an entrepreneur. It was like a, a new venture, but like the money that's being deployed now at scale is Microsoft, and, you know, and Amazon and Google, Facebook and Facebook and Nvidia and, you know, these, these, these, and, and now you know, by the way, open ai philanthropic, which are now at like, you know, really serious size, um, you know, as companies with, you know, very serious revenue.These are very large scale companies with like, lots, lots of cash, lots of debt capacity that they've, they've never used. And so th this is institutional in a way that, that really wasn't at the time. And then the other is, at least for now, every dollar that's being put into anything that results in a running GPU is being turned into revenue right away.Like so, and you guys know this, like everybody's starved for capacity, everybody's starved for compute capacity and then, you know, all the associated things, memory and, and, and interconnected and everything else. Um, data center space. And so e every dollar right now that's being put into the ground is turning into revenue.And, and it, and in fact, I actually think there's an interesting thing happening, which is because everybody starve for capacity, the models that we actually have that we can use today are inferior versions of what we would have if not for the supply constraints. That's true. Um, if Right pose a hypothetical universe in which GPUs were 10 times cheaper and 10 times more plentiful mm-hmm.The models would be much better. ‘cause you would just allocate a lot more money to training and you'd just build better models and they would be better. Um, and so we're, we're actually getting the sandbag version of the technology.swyx: Yeah. No. Everything we use is quantized because the, the labs have to keep the, the full versions,Marc: right?swyx: LikeMarc: we're not even getting the good stuff.swyx: Yeah.Marc: But, but getting the good stuff, it's, it's just, even if technical progress stops. Once there's like a much bigger build of like GPU manufacturing capacity and memory, you know, all, all the things that have to happen in the course of the next five or 10 years.Once it happens, even the current technology is gonna get, gonna get much better. And then as you know, like there's just like a million ways to use this stuff. Like there's just like a million use cases for this. Mm-hmm. Like, it, it, you know, this isn't just sending packets across a, a thing, whatever, and hoping that people find something to do with it.This is just like, oh, we apply intelligence into every domain of human activity. And then it works like incredibly well. Yeah. Um. Here's what I know, here's what I know. Um, in the next three or four year, it's like somewhere between three or four years out, basically everything is selling out. So like the, the entire supply chain is, is, is, is sold out or, or, or selling out.And so there, there's no, like, we're just gonna have like chronic supply shortage for, you know, for years to come. Um, there's going to be a response from the market that's gonna result in an enormous, you know, it's happening now. An enormous flood of investment in a new fab capacity and ev you know, every, everything else to be able to do that, at some point the supply chain constraints will unlock, you know, at least to some degree that will be another accelerant to industry growth when that happens.‘cause the products will get better and everything will get cheaper. Um, and so, so I know that's gonna happen. I know that, you know, the deployments, you know, the, the actual use cases are like really compelling. And then, like I said, you know, with reasoning and agents and so forth, like, I know they're just gonna get like much, much better from here.And so I, I, I know the capabilities are like really real and serious. I also know that the technical progress is not going to stop. It. It, it is excel. It is, is accelerating. Like the, the breakthroughs are are tremendous. I mean, even just month over month, the breakthroughs are really dramatic. And so, you know, I think if you were a cynic and there, there are cynics, you can look at 2000, you can find echoes.But I can't even imagine betting it that this is gonna like somehow disappoint and, you know, at least for years to come, I think it would be essentially suicidal to make that bet. Yeah. Um, it was that Michael Burry, uh, uh, that'sswyx: anMarc: interesting guy, huh? We'll pick on a guy. We'll pick, let's pick on one guy.We'll pick. Well ‘cause he did, he he came out with, it was, it was the, heswyx: doesn't mind.Marc: It was the Nvidia short. Right. He came with the Nvidia short. And then if you guys probably talked about this, which is the, the analysis now that like the current models are getting better faster at such a rate that if you are running an Nvidia, if you're running an Nvidia inference chip today, that's three years old, you're making more money on it today than you did three years ago because the pace of improvement of the software is, is faster than the, the, the depreciation cycle, the chip.And then my understanding is Google is running. I don't if they've, I don't know exactly what, uh, these are rumors that I've heard or maybe it's public, but, um, I think Google's running very old TPUs, very profitably. Ference. Yeah. And very profit and very profitably. Yeah. Um, and so, so it actually turns out, as far as I can tell, it's actually the opposite of the Beery thesis is actually.He was actually 180 degrees wrong. It's actually the, the, the, the old Nvidia chips are getting more valuable, which is something that's like literally never happened before. Like it's never been the case that you have an older model chip that becomes more valuable, not less valuable. And that, and again, that's an expression of the just ferocious pace of software progress.Ferocious pace of capability payoff. Yeah. Uh, that you're getting on the other side of this. And so I just, the idea of betting against that, like.swyx: Yeah. Yeah. Well, one ofMarc: my, it seems like an invitation to get your face ripped up.swyx: One of my early hits was like modeling the lifespan of the H 100 and h two hundreds and, and going like, you know, usually they advise like four to seven years and it was, you know, maybe you sort of realistically haircut cut it down to two to three.Yeah. But actually it's going up and not down. Yeah. And, and uh, that's, I mean that's, I think that's the dream. Uh, we are finding utilization and I think utilization solves all problems. Like, you can, you can find use, use cases for even like the poor, like even memory, we're having a shortage. Right. And, and even like the, the shittier versions of, of memory that we do have, we are finding use cases for it.So like That's great.Marc: Yeah.Alessio: How, how important is open source AI and kinda like edge inference in a world in which you have three years of supply crunch. Like, do you think in the, like, you know, if you fast forward like five years, like how do you think about inference, uh, in the data center versus at the edge?Marc: Well, so just to start, yeah. So I think, I think open source is very important for a bunch of reasons. I think edge, edge inference is very important for a bunch of reasons. I, I think just practically speaking, if we're just gonna have fundamental construc, supply crunches for the next, I mean, you, you guys know if you just project forward demand over the next three years, right?Yeah. Relative to supply, one of the, its main predictions you can do is what's gonna, what, what's gonna happen to the cost of, of inference in the core, uh, over the next three years? And like, it may rise dramatically, right? Like, so, so what is, and then is, is, you know, like the, the, the big model competition are subsidizing heavily right now.Right? Right. And so, so what's the, what will be the average person's, you know, per day, per month token cost, you know, three years from now to do all the things that they want to do. And I, I don't know, it's gonna. I mean, I have, you guys probably have friends, I have friends today who are paying a thousand dollars a day for open claw, for claw tokens to run open claw.Right? And so, okay. $30,000 a month. Right? And, and by the way, those, those friends have like a thousand more ideas of the things that they want their claw to do, right? Yeah. And so you, you could imagine there, there's like latent demand of up to, I don't know, five or $10,000 a day of, of, of tokens for a fully deployed, you know, per personal agent.Uh, and obviously consumers can't pay that, right? And so, so, but it gives you a sense of the fu of the fu of the future scope of demand, right? And so, so even, even if there's a 10 x improvement in price performance, that still, you know, goes to a hundred dollars a day, which is still way beyond what people can pay.Mm-hmm. So there's just gonna be like. Ferocious to me, by the way. The agent thing, the other interesting thing is I think the agent thing, so up until now, a lot of the constraints of GGPU constraints, I think the agent thing now also translates into CPU constraints. Mm-hmm. Right?swyx: CPU memory.Marc: Yes. CPU memory, right?And so, like the entire chip ecosystem is just gonna get wait,swyx: wait for network constraints, that that will be the killer.Marc: It's all bottleneck potentially for years. And so, so I, I think that Brad, and, and I think it's actually possible, I mean, generally inference costs are gonna keep coming down, but I think the, let's put it this way, the rate of decline, I think may level out here for a bit because of these supply constraints.And then at some point, maybe the lab stops subsidizing so much and that, that, that again, will be, be an issue. And so there's just gonna be so much more demand for inference than, than can be satisfied. Um, you know, kind of with the centralized model. And then, and then, you know, you guys know this, but like all the, just the dramatic, I mean just the dramatic innovations that have happened in the Apple silicon to be able to do, uh, inferences, it's quite amazing the level of effort being put.Like the open source guys are putting incredible effort into getting, you know, this recurring pattern where the big model will never run on a pc, and then six months later mm-hmm. Oh, it runs in a pc, right? It's like amazing. And there's very smart people working on that. So there's all that. And then look, there's also, you know.There's also like other, there's other motivators. There's other motivators which is just like, okay, how much trust are the big centralized model providers? You know, how much trust are they building in the market versus, you know, how much are, you know, at least for, in certain cases with some people, for certain use cases, people being like, well, I'm not willing to just like, turn everything over.So there, there, there's all the trust issues. Um, by the way, there's also just like straight up price optimization. There's many uses of AI where you don't need Einstein in the cloud. You just need like a, a a, a smart local model. There's also performance issues where you want, you know, you want, you know, you're gonna want your doorknob to have an AI model in it.Right. You know, to be able to, you know, do, um, you know, to be able to do access control. Um, obviously like everything with a chip is gonna have an AI model in it. Mm-hmm. And it, a lot of those are gonna be local. Um, and so, yeah. No, like I think, I think you're gonna have ti and then you're gonna, by the way, also wearable devices, you know, you don't wanna do a complete round trip.You want, you know, you, whatever your smart devices are, you want it to be like super low latency. Yeah.swyx: The question, do we care who makes it? Yeah. One of the biggest news this week was the collapse of AI two, the Allen Institute. Mm-hmm. One of the actual American open source model labs. Yeah. Um, and, uh, I'm not that optimistic on, on American open source.Yeah. Like you, you guys invested in MIS trial and MIS trial's doing extremely well outside of China. That's about it.Marc: Yeah. We'll see. We'll see. I look, I, number one, I do think we care. Uh, I do think we, I do think we care who makes it. Um, I would say this, the, the, the, the previous presidential administration wanted to kill it in the us Oh yeah.They wanted to drown in the bathtub. Um, and so they wanted to kill it. So at least we have a government now that actually like, actually wants it wants it to happen. And youswyx: earned to councilMarc: and Yeah. And the new and the P pcast. Yeah. So the, the, you know, this admin for whatever other political issues people have, which are many, you know, this administration has, I think a very enlightened view and in particular an enlightened view on AI and in particular on open source ai.Uh, and so they're very supportive. Um, my read is the Chi. The Chinese have a very, the various Chinese companies have a very specific reason to do open source, which is, they, they, they don't fundamentally, they don't think they can sell commercial, uh, AI outside of China right now. And or at least specifically not, not in the US for a combination of reasons.And so they, they kind of view, I think, open source AI as a bit of a loss leader against basically domestic, uh, you know, paid, paid services. And then kind of an, you know, kind of an ancillary products. You know, they're, they're very excited about it, by the way. I think it's great. I think it's great that they're doing it.Um, you know, I think Deeps seek was like a gift to the world. Um, I think. The great thing about open source, open source, the, the, the impact of open source is felt two ways. One is you, you get the software for free, but the other is you get to learn how it works, right? And so like the paper, the paper, the paper and, and the code, right?And the code. And so, like, for example, I thought this was amazing. So open comes out with L one and it's an amazing technical breakthrough, and it's just like, absolutely fantastic. But of course they don't explain how it works in detail. And then of course they hide the, they hide the reasoning traces, right?And, and then, and then, and then everybody's like, okay, this is great, but like, who's gonna be able to replicate this? Are other people gonna be able to do this? You know, is their secret sauce in there? And then our one comes out and it's just like, there's the code and there's the paper, and now the whole world knows how to do it.And then, you know, three months later, every other AI model is, is adding reasoning. And so, so you get this kind of double, like even if the Chinese models themselves are not the models that get used, the education that's taken place to the rest of the world, the information diffusion, you know, is incredibly powerful.So that happens and then, I don't know. We'll, we'll see. You know, there are a bunch of American, you know, open source, you know, ai, uh, model companies. I mean, look, there's gonna be tremendous, you know, there already is. There's, you know, there's gonna be tre there's tremendous competition, uh, among the primary model companies.You know, there's, depending on how you count, there's like four or five, you know, big co model companies now that are, you know, kind of neck and neck, uh, in different ways. Um, uh, you know, and, and, and, um, you know, and then obviously Bo Bo both X and then MetAware involved are, you know, both have huge, you know, huge attempts to, you know, kind of, to kind of leapfrog underway.And then you've got, you know, a whole fleet of startups, new companies, including a whole bunch that we're backing, that are, you know, trying to come out with different approaches. And then you've got whatever it is. I don't know how, how many, how many, like main line foundation model companies are there in China at this point?It's probably six. It'sswyx: five Tigers is what they call it. Yeah. Uh, Quinn is in questionable because there's change in leadership,Marc: right?swyx: Yeah.Marc: But that, does that include, that includes like Moonshot,swyx: yes. Can deep seek, uh, uh, ZI, um, Quinn oh one is in there.Marc: Right. And then, um, and by dance and, and then you see,swyx: ance would be like the next tier ance.They weren't as prominent. They weren't, didn't haveMarc: a leading. Yeah. But they, you at least, you know, ance is very inspiring and presumably they have more stuff coming and Tencent probably has more stuff coming and, and so forth. And so, so, so like, look, here, here would be a thing you can anticipate, which is there are not these markets, there are not going to be between the US and China right now, there's like a dozen primary foundation model companies that are like at scale, at, at some level of a critical mass.It's not gonna be a dozen in three years, right? Like, it just because these industries don't bear a dozen, it's, it's gonna be three or you know, there's gonna be three or four big winners or maybe one or two big winners. And so there's gonna be like a whole bunch of those guys that are gonna have to figure out alternate strategies.Um, and I think like open source is one of those strategies. And so I, I think you could see like a whole, i, I, I think the questions like, who's gonna do open source? I think that could change really fast. I, I think that, that, that's a very dynamic thing. I think it's very hard to predict what happens. And, and I think it's very important.swyx: NVIDIA's doing a lot.Marc: Well, I was gonna say. Well, exactly. And then you're got Nvidia and then, and then, you know, just to, again, indu, there's an old thing in business strategy, which is called, uh, commoditize Compliments. Commoditize the compliment. That's right. And so if your Jensen is just kind of obvious, of course, you wanna commoditize the software.Yeah. And he's, and to his enormous credit, he's putting enormous resources behind that. And so maybe it, maybe it's literally Nvidia and I think that would be great.Alessio: Yeah. Uh, narrative violation to European projects, uh, in the, uh, damn.swyx: I'm hosting my, uh, Europe, uh, conference soon. And I got both of them.Alessio: They got us.They got us. MarkMarc: finished. They got us, us. Well, wait a minute. Where was Peter? So where was Steinberger when he did? In AustriaAlessio: was, yeah, yeah, yeah.Marc: He was in what? He was in Vienna. Oh, he was in Vienna. And then where is he now?swyx: Uh, he's moving to sf.Marc: Okay. Okay. Alright. Okay, there we go. And then, yeah, the PI guy, right?The PI guys are European.swyx: Yeah, they're also, they're buddies inAlessio: Australia. Mario's also there. Yeah.Marc: Right. And are they, yeah, they haven't announced yet. Any sort of change changed or have theyAlessio: No, they're, they have a company there.Marc: Okay. Got, okay. Good.Alessio: Good, good,good.Alessio: Um,Marc: yeah, good.swyx: Anyways, I think pie and open cloud very important software things and, and I just wanted you to just go off on what you think.Marc: Yeah. So I think in co the, the combination of the two of them I think is one of the 10 most important softwares. Openswyx: Claw got all the attention, but Right. Talk about pie,Marc: pi pie's, kind of the Yeah. PI's, PI's kind of the architectural breakthrough for those of us who are older. There was this whole thing that was very important in the world of software basically from like 1970 to, I don't know, it still is very important, but like 19, from 1973 to like basically the creation of Linux, which is basically this, this thing used to call like the Unix mindset.Like so, so, ‘cause there were all these different, you know, theories. There are all these different operating systems and mainframes and, and then you know, all these windows and Mac and all these things. And then there was this, but kind of behind it all was this idea of kind of the Unix mindset. And the Unix mindset was this thing where basically you don't have these, like, like in the old days, like, like the operating system that like made the computer industry really work, like in the 1960s mm-hmm.Was this thing called o os 360, which was this big operating system that IBM developed that was supposed to basically run everything. And it was this like giant monolithic architecture in the sky. It was like a, you know, it was like a giant castle. Um, of software. And, and by the way, it worked really well and they were very successful with it.But like, it was this huge castle in the sky, but it was this thing, it was almost unapproachable, which is like, you had to be kind of inside IBM or very close to IBM. And you had to really understand every aspect, how the system worked. And then the, the Unix sky is originally out of at and t and then out out of Berkeley, um, you know, came out and they said, no, let's have a completely different architecture.And the way architecture's gonna work is we're gonna have, we're gonna have a, a prompt and, and a, and a shell. And then, and then we're gonna, all, all the functionality is gonna be in the form of these discreet modules, and then you're gonna be able to chain the modules together. Mm-hmm. Yeah. And so like the, the, the op, it's almost like the operating, operating system itself is gonna be a programming language.Um, and then that led led to the, the, the sort of centrality of the shell. Um, and then that led to sort of, uh, you know, basically chaining together Unix tools. And then that led to the emergence of these, these scripting languages like Pearl, where you, you could basically kind of very easily do this, and then the shells got more sophisticated and then, and then, and then look like, you know, that, that, that number one, that worked and that, that was the world I grew up in.Like I was, I was a Unix guy. You know, sort of from, call it 1988 to, you know, kind of all, all the way through my work and it worked really well. It, it's in the background, um, you know, nor normal people don't need to, didn't need to necessarily know about it, but like, if you were doing like system architecture, application development, you, you, you knew all about it.Um, and then, you know, it's been in the background ever since. And, you know, look, your Mac still has a Unix shell, you know, kind of in there, and your iPhone still has a Unix shell kind of buried in there somewhere. So they're kind of in there. And then, you know, the Windows shell is kind of a, you know, sort of a weird derivative of that.But, um, you know, but look, the inter, the internet runs on Unix, um, and that smartphones, actually, both iOS and Android are Unix derivatives. And so, you know, kind of Unix did end up winning. But, but anyway, and then we just started taking that for granted. And then, and then so, so basically the, the way I think about what happened with Pie and then with Open Claw is basically what those guys figured out is, I always say the, the great breakthroughs are obvious in retrospect, right?Which is the best kind, the best kind. They weren't obvious at the time or somebody else would've done them already. Um, and so there is a, like a real conceptual leap, but then you look at it sort of the backwards looking and you're just like, oh, of course. Mm-hmm. Like the, the, to me those are always the best breakthroughs.Well, actually language models themselves are like that. It's just like, oh, next token completion. Oh, of course.swyx: Yeah. What other objective mattered?Marc: Yeah, exactly. But, but like it, right. But she's even saying it wasn't obvious until somebody actually did it. Right. And so the conceptual breakthrough is real and deep and powerful and, and very important.And so the way I think about pie and olaw is it's basically marrying the, the language model mindset to the un to the Unix, basically shell prompt mindset. And so it's, it's basically this idea that what, what, so what is an agent, right? And as, as, and as you know, like many smart people who have been trying to figure out what an agent is for, for, for decades, and they've had many architectures to build agents and the whole thing.And it turns out what is an agent. So it turns out what we now know is an agent is the following. It's, so it's a language model. And then above that, it's a ba, it's a bash shell. Um, so it's a, it's a Unix shell, and then it's, and then the agent has access, uh, has access to, to the shell. And, you know, hopeful, hopefully in a sandbox, maybe in, maybe in a sandbox.So it's, it's the model. Um, it's the shell. Um, and then it's a fi, it's a file system. Um, and then the state is stored in files. And then, you know, there's the markdown format for the, you know, for, for the files themselves. And then, and then there's basically what in Unix is called Aron job. There's a loop and then there's a heartbeat for the, there's heartbeat and, and the thing basically Wake Wakes up.Wakes up. So it's basically LLM plus shell, plus file system, plus markdown, plus kron. And it turns out that's an agent. And, and, and every part of that, other than the model is something that we already completely know and understand. And in fact, it turns out that like the latent power of the Unix shell is like extraordinary because basically like all, like, there's just like an, there's just enormous latent power in the shell.There's enormous numbers of Unix commands, there's enormous number of command line interfaces into all kinds of things already in the, you know, your entire, I mean your entire, just to start with, your computer runs on a shell. If you're running a Mac or a, or, or a phone, your computer, your computer's running on a shell, uh, already.And so like the full power of your computer is available at the command line level. Um, and then it turns out it's really easy to expose other functions as a command line interface. And so like this whole idea where we need like MCP and these like product mm-hmm. Fancy protocols, whatever, it's like, no, we don't, we just need like a command, command line thing.So that's the architecture. And then it turns out what is your agent? Your agent has a bunch of files starting a file system. And then there's the thing that just like completely blew my mind when I write my head around it as a result of this, which is like, okay. This means your agent is now actually independent of the model that it's running on.Because you can actually swap out a different LLM underneath your agent and your, your agent will change personality somewhat. ‘cause the model is different, but all of the state stored in the files will be retained.swyx: Yeah. Different instruction set, but you just compiledit.Marc: Right, exactly. And it's all right.It's like right. Swapping out a ship and recompiling, but it's, it's still, it's still your agent with all of its memories. Um, and with all of its capabilities. And then by the way, you can also swap out the shell, uh, so you can move it to a different execution environment that is also, is also a b shell, by the way, you can also switch out the file system, right.Uh, and you can, and you can, and you can swap out the, the, the heartbeat for the, the crown framework, the, the loop that the agent framework itself. And so your agent basically is ba basically at the end of the day, it's just. It's just, its files. Um, and then, and then there's of course it a openswyx: call.Marc: Yeah, it's, it's basically, it's, it's just the files.Um, and then by the way, as a consequence of that, the agent and then the agent itself, it turns out a couple important things. So one is it, it's, it, it can migrate itself, right? And so you're, you can instruct your agent, migrate yourself to a different, uh, runtime environment, migrate yourself to a different file system, migrate yourself to a different, you know, swap out the language model.Your agent will do all that stuff for you. And then there's the final thing, which is just amazing, which is the agent is the agent actually has full introspection. It actually, it actually knows about its own files and it could rewrite its own files. Right. Which by the way, is basically no widely deployed software system in history where the, the, the thing that you're using actually has full introspective knowledge of how it itself works and is able to modify itself.Like that, that, I mean, there have been toy systems that have had that, but there, there's never been a widely deployed system that has that capability and then that leads you to the capability. That just like completely blew my mind when I wrap my head around it, which is you can tell the agent to add new functions and features to itself and it can do that.Extend yourself. Yeah. Right? Extend, extend yourself. Like extend yourself. Give yourself a new capability. Right? And so, and so literally it's just like you run into somebody at a party and they're like, oh, I have my open claw, do whatever, connect to my eat, sleep bed, and it gives me better advice and sleep.And you go home at night and you tell your claw, or if they're at the party, by the way, you tell your claw, oh, add this capability to yourself. And your claw will say, oh, okay, no problem. And it'll go out on the internet and it'll figure out whatever it needs and then it'll go out to claw code or whatever.It'll write whatever it needs. And then the next thing you know, it has this new capability. And so you don't even have to, like, you can have it upgrade itself without even having to, without having to do anything other than tell it that you want it to do that. And so anyway, so the, the combination of all this is just, I mean, this is just like a massive, incredible, I mean, it's just incredible.Like if I, if I were, if I were 18, like this is a hundred, this is what I would be spending all of my time on. This is like such an incredible conceptual breakthrough. Yeah. And again, pe people are gonna look at it and they already get this response. People are gonna look at it and they're gonna say, oh, well, where's the breakthrough?‘cause these, the, all of these components were already known before. Mm-hmm. But, but this is the key, the key to the breakthrough was by using all these components that were known before, you get all of the underlying capability of that's buried in there. And so all, and so for example, computer use all of a sudden just kind of falls, trivi, trivial.Of course it's gonna be able to use your computer. It has full access to the shell. Right. And then, and then you just, you, you give it access to a browser, and then you've got the computer and the browser and, and often away it goes. And, and then you've got all the abilities of the browser also. Um, yeah.And so, and so the capability unlock here is profound. My friends who are, you know, deepest into this, are having their claw do like a, like, literally like a thousand things in their lives. They have new ideas every day. They're just like constantly throwing new challenges at the thing. And by the way, it's early and, you know, these are, you know, these are prototypes and there are, you know, as you guys know, there's security issues.Yeah. And, and so, you know, there's a bunch of stuff to be ironed out, but the, the unlock of capability is just incredible.swyx: Yeah.Marc: And I, I have absolutely no doubt that everybody in the world is gonna, is gonna have at least, you know, an agent like this, if not an entire family of agents. And w

Aquí Telenovelas
_Qué pasa en la familia Nodal__ Jorge D_Alessio y Marichelo_ desde el amor

Aquí Telenovelas

Play Episode Listen Later Apr 2, 2026 61:29 Transcription Available


Just End The Suffering
554-2026 March Madness Second Weekend Recap With Troy Mauriello

Just End The Suffering

Play Episode Listen Later Mar 30, 2026 89:42


It's time to break down the latest from March Madness on the latest episode of the Just End The Suffering podcast! Host Mike Phillips (⁠⁠⁠⁠⁠⁠⁠@MPhillips331⁠⁠⁠⁠⁠⁠⁠) kicks off the show by reacting to the opening weekend of the MLB season (1:42) for both the Mets and Yankees. Mike is then joined by Troy Mauriello (⁠⁠⁠@TroyMauriello⁠⁠⁠) to recap the second weekend of March Madness (8:58) and preview the Final Four. Mike then wraps up the show by recapping the Season 2 premiere of Daredevil: Born Again (57:28) with Nick D'Alessio.Subscribe to the Just End The Suffering podcast on ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Apple⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠, ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Amazon⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠, ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠TuneIn⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠,⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ and⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Spotify⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠!Subscribe to ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Mike Phillips's channel⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ on YouTube!