Podcasts about Linux

Family of free and open-source software operating systems based on the Linux kernel

  • 3,979PODCASTS
  • 33,551EPISODES
  • 48mAVG DURATION
  • 5DAILY NEW EPISODES
  • Oct 1, 2026LATEST
Linux

POPULARITY

20192020202120222023202420252026

Categories




    Best podcasts about Linux

    Show all podcasts related to linux

    Latest podcast episodes about Linux

    The DevOps Kitchen Talks's Podcast
    DKT102: вернулись на кухню. Новости, stateless MCP и профессия FDE

    The DevOps Kitchen Talks's Podcast

    Play Episode Listen Later Oct 1, 2026 67:53


    Первые полсотни выпусков DKT записаны в этой самой комнате. Потом переезд, и мы ушли в окошки. В этот раз получилось вернуться: приехал в отпуск, и мы записали ещё один выпуск с той кухни, где всё начиналось. Три камеры, заваленный горизонт и возможность смотреть друг другу в глаза, а не в квадратик. Выпуск новостной, накопилось за несколько месяцев. О ЧЁМ ВЫПУСК • Forward Deployed Engineer: не замена DevOps, а карьерное ответвление. Тебя целиком отдают одному клиенту, зарплаты доходят до полумиллиона, Amazon вложил в направление миллиард. • MCP стал stateless: sticky session больше не нужна, сервер уезжает на Lambda или Cloudflare Worker. • Почему MCP аккуратнее, чем голый CLI: Вася перепутал staging и prod, сказал агенту «давай destroy базу», и оно всё ушло. • SpaceX купил Cursor за 60 млрд, а доля Cursor за год упала с 40% до меньше 20: Codex перевернул рынок. • Четыре уровня зрелости с агентами: вайб-кодинг, spec-driven, AI DLC и автономность, она же Dark Factory. • Dark Factory Саши живьём и сколько она стоит: 20 строк кода и 1300 строк markdown на одну задачу. • Amazon ECS дробит GPU: инстансы G6f, минимум одна восьмая карты NVIDIA L4. • AWS Frontier agents в GA, Meta выпустила Muse Code, Linux 7.1 и Argo CD 3.5 одной строкой. • The Human in the Loop is Tired: куда девается удовольствие от работы, когда сложное делает агент.

    Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0
    Why Dwarkesh is Wrong about Computer Use + How OpenAI shipped its Jev competitor in 1 Week

    Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

    Play Episode Listen Later Sep 30, 2026 39:12


    Three months ago Dwarkesh, who has been posting incredible blogs and episodes about RL, posted a framing question for his video essay on RLVR which upset a lot of Computer Use folks:We are no strangers to learning in public and are no strangers to the stress of getting things wrong when you have a big platform. However, we were at Anthropic for the Computer Use launch, there for Claude Cowork with the first big podcast on it, organized the first Computer Use track at AIE presenting the state of the art, and were close to the OpenAI-Sky Software acquisition that now powers the complete domination of computer use that Codex enjoys today. This is why we're excited to bring you today's first guest, Ari Weinstein, cofounder of Sky and now leading all the amazing CUA progress that casuals might miss:Ari explains why Computer Use is now “180 degrees different” from where it was months ago, how agents are learning to debug and recover from failures, why combining screenshots with accessibility data, the DOM, Playwright, and generated code changes the speed equation, and why the next frontier is making agents literally superhuman at using software.OpenAI clones JevIn the second half, Nikunj Handa from OpenAI's API team breaks down the new developer stack: async tool calling, mid-turn steering, WebSockets, UltraFast inference, the Decisions API, prompt caching, pre-warming, compaction, and the Agents API. Given that we were the first Jev podcast, we particularly focus on the unusually fast sprint on the Decisions API:And why it is just a Luna wrapper for now but the team is motivated and egoless enough to clone what they consider to be good patterns.We discuss:* Why OpenAI thinks Computer Use has changed dramatically in just the last few months* Dots and what changes when every agent gets its own Linux computer* Why Computer Use can now complete some tasks faster than the average human* The path from human-level to “literally superhuman” computer use* Why modern agents are much better at debugging and recovering from failure* How screenshots, accessibility trees, the DOM, Playwright, and generated JavaScript work together* App Shots and why they give models much richer context than ordinary screenshots* Why Computer Use can close the loop between writing software and testing it* Trust, permissions, and safety when agents can make payments and operate websites* Async function calling and why models no longer need to stop reasoning while tools run* Mid-turn steering, WebSockets, and the architecture behind more responsive agents* UltraFast inference and how OpenAI is pushing frontier models toward much lower latency* The rapid internal story behind the Decisions API* Why Decisions API is more than structured outputs at low latency* GPT Live, fast tool calling, and real-time computer control* How OpenAI is already using Decisions API for support classification and internal workflows* Longer prompt caching, cache pre-warming, and cache-aware applications* Server-side compaction vs manual compaction for long-running agent threads* What should live inside an Agents API versus a developer's own harness* OpenAI as an “AI cloud” and the search for higher-level primitives beyond raw model APIsAri Weinstein* Product & Engineering, Computer Use at OpenAI* X: https://x.com/AriX* LinkedIn: https://www.linkedin.com/in/weinsteinari/Nikunj Handa* Product, API at OpenAI* X: https://x.com/nikunjhanda* LinkedIn: https://www.linkedin.com/in/nikunjhanda/Timestamps00:00:00 OpenAI DevDay: Dots, GPT-6.1, Agents API, and Decisions API00:02:52 Dots and Personal Cloud Computers00:04:59 Why Computer Use Is “180 Degrees Different”00:06:04 From Sky to Self-Debugging Computer Use Agents00:09:24 How Computer Use Sees and Operates Software00:12:09 From Faster Than Humans to Superhuman Computer Use00:16:03 Agents API: Trust, Permissions, and Safety00:17:31 Computer Use for Coding, Testing, and QA00:19:14 GPT-6 APIs, Async Tool Calling, and UltraFast Inference00:23:21 The Rapid Story Behind Decisions API00:25:32 What Decisions API Is and How It Works00:30:24 What OpenAI Is Building With the New APIs00:32:23 Prompt Caching, Pre-Warming, and API Performance00:35:20 Context Compaction for Long-Running Agents00:37:13 Memory, Higher-Level APIs, and the AI CloudTranscriptIntroduction: OpenAI DevDay and the New Agent StackVibhu [00:00:00]: Okay. We're very excited to be here. Today is OpenAI DevDay. Special podcastSwyx [00:00:08]: We're the first podcast after your livestream.Vibhu [00:00:10]: First podcast. We have Ari here, who leads the product and engineering team for Computer Use agents. Before we kick in and dive deep on Computer Use, you wanna give a quick recap? What was announced? What's the quick slew of announcements you guys had today?Ari Weinstein [00:00:24]: Yeah. yeah, it was a super exciting day. we just got out of the keynote. It was really sick. there were a bunch of Computer Use announcements that I think are worth thinking about. We have, Dots, which is the new, sort of personal assistant product, and, that has some really exciting Computer Use features. There's GPT-6.1 Sol, which is this amazing new model, that I think is particularly great for Computer Use ‘cause of, sort of the cost and speed, advantages. I think, I think we shared that it's, a fifth of the cost of Astra and a seventh of the cost if you're looking at Computer Use specifically, which is really amazing. sorry, there were so many things. I'm trying to sort through it.Swyx [00:01:02]: And the API.Ari Weinstein [00:01:03]: Agents API, which now has Computer Use in it, which is really cool, ‘cause now developers can build on the same Computer Use, that is part of Codex, and ChatGPT. and then there were some demos of our existing Computer Use features, like app shots, where you can take the context of something you're doing on your computer and bring it into Codex and ChatGPT really fast. And then, like, native Computer Use on your Mac, where Roman had it taking screenshots of his app, automatically, and he could do other things on his computer while Computer Use was using his applications. so yeah, really exciting keynote.Swyx [00:01:35]: And not to mention the Decisions API.Ari Weinstein [00:01:37]: Decisions API.Swyx [00:01:38]: Off the bat, are they all the same model? Like, this is. Or the same dataset distilled to different models?Swyx [00:01:44]: Like, basically, like, is Computer Use using Decisions API, or are they, like, kinda separate?Ari Weinstein [00:01:49]: So what's really cool about the Decisions API is it, you know, it has all these new capabilities. It does inference in parallel. it doesn't have reasoning. It's a smaller model, than the ones we use for Computer Use. and so those capabilities make it really fast.Dots and Delegating Work to a Cloud ComputerSwyx [00:02:07]: Yeah.Ari Weinstein [00:02:07]: They also make it a little bit less good at doing, like, long horizon, sort of sophisticated tasks. And so I think I would say it's still an open area of research for how we, like, bring those approaches together. But, yeah, I'm really excited to see what people build with the Decisions API.Vibhu [00:02:24]: One of the interesting things is Dots now have attached personal computers.Ari Weinstein [00:02:28]: Yeah.Vibhu [00:02:28]: So it seems like they're very much more persistent. You've been using them for a while. How should people push the bounds? Like, what should people aim for? What should they try? Personally, right now I use it for a lot of customer service. LikeAri Weinstein [00:02:41]: CoolVibhu [00:02:41]: “Oh, this was wrong. I don't wanna sign in. I don't wanna authenticate.” Find whatever and just get it fixed.Ari Weinstein [00:02:45]: Yeah.Vibhu [00:02:46]: How should we push further? What should people try?Ari Weinstein [00:02:50]: Dots Are a really cool product because each Dot has access to its own Linux virtual computer in the cloud, which is different from our other products. you know, traditionally, we've have access to a browser in the cloud, or it has access to your own computer, but now you get your own entire Linux computer in the cloud. And so it can run full desktop applications, and it can also use a web browser. And so, yeah, you know, I think the powerful thing about Computer Use and the reason why I think it's so, exciting is because it makes it so that the agent can do anything you as a, as a person can do, because all the software in the world was designed for humans, and now agents can use that same software, and you can delegate to the agent. So, yeah, like, anything that you would do on a computer, you can ask a Dot to do. Yeah, I think what particularly is useful is gonna really depend on who the end user is and what- what's valuable in their life. but yeah, I would just start by thinking about, like, one of the things that you spend time on and how could you delegate those to an agent.Swyx [00:03:47]: Yeah, a lot of flight booking and shopping and honestly even, like, playing a game or whatever, right?Ari Weinstein [00:03:52]: Totally.Swyx [00:03:52]: Yeah.Ari Weinstein [00:03:53]: Yeah, I don't know. For me, something I did recently, I've been working on. I've, subscribed to a meal prep service ‘cause I was trying to, like, eat healthy, you know? And I really like this meal prep service I found because it lets me customize the meals I order to, like, a high degree of granularity. So I can say, like, “I want this many grams of chicken and this many grams of rice.” but it was so complicated. It took me two hours to do an order, and I found that I could ask Computer Use to do it for me, and it did it in 15 minutes. so I actually saved two hours. it both did it eight times faster than I could, and it saved me two hours on GPT-6.1 Sol.Swyx [00:04:32]: Yeah.Ari Weinstein [00:04:32]: So those are the kinds of tasks that I feel like, are really powerful.Swyx [00:04:36]: As a creator, I can tell you automatically, immediately, my number one use case is automating YouTube.Ari Weinstein [00:04:40]: Nice.Swyx [00:04:40]: Because, YouTube doesn't expose a lot of things via API.Ari Weinstein [00:04:43]: Yeah.Swyx [00:04:43]: And you have to just put it in a VM and just, like, run it, for, like, let's say, let's say their AB testing feature or making community posts. None of this is available by API ‘cause they hate developers.Swyx [00:04:53]: Anyway, so,Ari Weinstein [00:04:55]: I've heard that from our developer experience team too. They use it with YouTube a lot. Yeah. It's really awesome.Swyx [00:04:59]: So I wanna draw for, you know. let's say, I wanna get a little bit spicy. One of our, the leading AI podcasts, our friends, is famous for saying that Computer Use hasn't advanced in the last two years.How Computer Use Has Changed in the Last YearAri Weinstein [00:05:12]: Yeah.Swyx [00:05:13]: Which is a very interesting statement, and I think you're one of the best people in the world to talk about this, like, how have things have progressed, right?Ari Weinstein [00:05:20]: Yeah. You know, they said that a few months ago, I think, and I hope they have a different perspective now because Computer Use is, like, 180 degrees different than it was.Swyx [00:05:26]: He's a, he's a tough guy to impress.Ari Weinstein [00:05:27]: Yeah, okay. well, we're working on it.Swyx [00:05:30]: But, you know, you worked on. You've, like, basically spent your whole career working on, like, some kind of computer automation, right?Ari Weinstein [00:05:34]: Yeah.Swyx [00:05:34]: Like shortcutsAri Weinstein [00:05:35]: YeahSwyx [00:05:35]: At Apple, and then Sky, and then, and then joining OpenAI. Can you draw, like, what your through line is for, like, what is driving you and what- you, what wasn't possible back then maybeAri Weinstein [00:05:47]: Yeah.Swyx [00:05:48]: And, like, what your sort of milestones were.Vibhu [00:05:49]: I guess to add on to that as a follow-up question, what's the major change from using Codex Computer Use from, like, last weekAri Weinstein [00:05:57]: YeahVibhu [00:05:57]: Through to today? Is it model? Is it dots? Is it harness? So all the history plus what really just changed in today's announcements?Ari Weinstein [00:06:04]: Yeah. On the through line, I guess I've always been excited about automation and helping people automate tasks because then you can, like, save time in your life and focus on things that are more important to you than, like, operating a computer very intricately. And so, yeah, that was why we worked on some of those products. I was at Apple before. we made a company called Sky. we ended up joining OpenAI, which is really exciting. and I think something that wasSwyx [00:06:27]: And almost like you have to hack around Apple until Apple was like, “Fine, like, we'll just hire you and you can just work on the inside,” right? Like.Ari Weinstein [00:06:35]: It was, it was a cool place to get to work. what was really interesting looking back at Sky is we were, we were working on Computer Use there as well, and the models were so much less capable. And now the models, just in the last one year, have become extraordinarily capable at Computer Use. I think the biggest delta that I see is before they could, like, reliably start tasks, but then they would run into problems, and now they're really good at debugging. They're really good at trying again, introspecting what is and isn't working. and I think we've also brought the Computer Use the Computer Use field itself has moved forward. I think we're using more techniques. now Computer Use, often writes code. So if you actually look at it in Codex and you expand the tool calls manually, you can see that it's not just doing one action at a time. It's actually writing JavaScript code that it executes, that the computer executes to perform sometimes many actions at once, which is a great, you know, speed up and great capability. We use more accessibility, sort of multimodal interfaces. So, the model may use screenshots, it may use accessibility, it may use Playwright. it can use a lot of different mechanisms, based on the task at hand. and then, yeah, the model acceleration has been, has been just amazing. So, yeah, what's different today? I think we're making computers better all the time, so I think just, like, one day's difference, is probably a little bit less consequential than, like, even the past month or the past two months. but, yeah, I think the Computer Use in Dot is really exciting as well as, the new model that we came out with.Measuring Computer Use and Improving the HarnessVibhu [00:08:03]: On the keynote, Tejal was mentioning 7x improvements in Computer Use speed, a lot better on a few benchmarks. How do you guys think about measuring it? Computer Use is one of those things where, as you say, you know, it's improvements over time.Ari Weinstein [00:08:20]: Yeah.Vibhu [00:08:20]: Is it harness? Is it model? Is it post-training?Ari Weinstein [00:08:22]: Right.Vibhu [00:08:22]: How do you guys look at it internally about measuring how good it is, and what were the changes with the new model?Ari Weinstein [00:08:29]: We actually have a bunch of different ways of measuring it, some of which are on different permutations and configurations of the harness. It's a bit of a complicated story because, you know, our production products have, you know, some more safety checks, and, you know, those are configured differently based on the needs of the, of the task at hand. So there's a lot of ways to measure it, but I think regardless of how we measure it, we find pretty consistent gains. and those gains are, sometimes in the harness and sometimes in the model. and yeah, I was really excited by this result that GPT-6.1 is even more cost-effective for Computer Use than its baseline cost improvement as compared to Astra. It's, like, really cool to see.Swyx [00:09:10]: Yeah. I mean, one of the visuals I really liked from the livestream was that, you're sort of improving the Pareto frontier of, your, curve, and there was a lot of talking about how you're improving it together with the harness.Ari Weinstein [00:09:24]: Yeah.Swyx [00:09:24]: Can you give some examples of aha moments that you had, whether it's on, like, model driving the harness driving the model, whatever?Ari Weinstein [00:09:32]: I don't mean to repeat myself, but I think, like, introducing more modalities has been really powerful.Swyx [00:09:36]: Okay.Ari Weinstein [00:09:36]: One more specific example of that is, in the past, I think we saw a lot of Computer Use, products had to spend a lot of time, like, scrolling, you know? So it would, like, take a screenshot. It would try to do something. It would be like, “Oh, I gotta, like, scroll down to the next page of results,” and then it would take a screenshot, and then it would try to do something. It would scroll down again. And so I think, with accessibility and other. and, direct access to the DOM and other things like that, now the language model can actually see, like, an entire page or an entire application. It can write code that can do multiple steps at once. And so I think those have been probably the biggest single aha moments. There's, like, a lot of tiny ones that are less exciting in comparison, but actually we do find also that a lot of speed improvements are driven by, like, a lot of little paper cuts that we gotta go in and introspect.App Shots, Accessibility, and Better Computer ContextSwyx [00:10:21]: Yeah. A lot of really hard engineering.Ari Weinstein [00:10:23]: Yeah.Swyx [00:10:23]: I mean, app shots in general, right? Like, I think people don't quite get the difference if. because there's, like, a nice visual in Codex when itAri Weinstein [00:10:30]: YeahSwyx [00:10:30]: When you take an app shot, but they don't maybe they get the difference that, you are able to actually drive each button and you have the, you have each text, in a very optimal representation.Ari Weinstein [00:10:40]: Yeah. Exactly. Yeah. It's kind of fun actually. If you wanna be, like, really nerdy about it, you can go into Codex, take an app shot by hitting the two command keys. So you grab the content from whatever app you're working with, bring it into the, Codex or ChatGPT chat. And then the. if you click on the attachment and you click on this, like, little tiny button in the top right, you can see the raw text and you see the raw accessibility representation. And yeah, we've put a lot of work into, puttingSwyx [00:11:04]: Just dumping everything out. Yeah.Ari Weinstein [00:11:05]: Dumping it out, but also making it token-efficient, doing it efficiently. There's, like, a bit of an art to it. And, you know, it turns out that the same technology that was invented for humans, you know, who maybe have accessibility needs, who wanna use a screen reader technology, that technology is really helpful for them to be able to use computers. It's also really helpful for LLMs to be able to use computers. So that's been, like, really fun to get to work on.Vibhu [00:11:27]: For context, I feel like a lot of people don't understand app shots. They don't even know it's a feature.Ari Weinstein [00:11:30]: Yeah.Vibhu [00:11:31]: It's when you double hit command, it pulls in what looks like a screenshotAri Weinstein [00:11:34]: RightVibhu [00:11:34]: And you're like, “Oh, why have I opened up just a screenshot and thrown it in?” No, it's actually pulling all the metadata, all the code, everything.Ari Weinstein [00:11:40]: Yeah, exactly. Yeah. So it's like, you know, if you take a screenshot of a webpage that has a link- The screenshot doesn't include where the link goes. It doesn't include, you know, maybe you take a screenshot of your calendar, the ca- event ti- titles are truncated, you know? But when you take an app shot, it gives, like, the language model, like, full context about everything and, that lets it, just sort of, like, do much more.Swyx [00:12:02]: Yeah. For those who wanna see more, Jason Liu, I invited him to do a full workshop on this, at AI Engineer.Ari Weinstein [00:12:07]: Amazing.Swyx [00:12:08]: Did a great job.Vibhu [00:12:09]: I have a broader vision questionToward Superhuman Computer UseAri Weinstein [00:12:11]: YeahVibhu [00:12:11]: On Computer Use agents. So your example of take a screenshot, scroll page, take a screenshot is where we were.Ari Weinstein [00:12:17]: Right.Vibhu [00:12:17]: Today, they can automate a lot. what are the bottlenecks? Is it models? Is it harnesses? What. Where do you see it going in, like, two years? Do you see it just running for hours? How do we get there? Any predictions on where Computer Use goes?Ari Weinstein [00:12:32]: Yeah. I mean, I think what's really crazy that I think, You know, the team's accomplished over the past couple of months is that now Computer Use is, like, faster at accomplishing tasks than, like, the average human probably in most cases. and I think that the next frontier is to have Computer Use be, like, literally superhuman in its performance where it actually is as fast or faster at using software than, like, expert Computer Users like us. and I think that'll be really consequential and exciting when that happens because I think we'll be able to all of a sudden build products, that, provide just much more real-time experiences. And I think it'll also. lowering the barrier to entry of, or the activation energy, I suppose, of using Computer Use I think will make us start to default to doing certain things in agents that we've become accustomed to doing manually. And I think that's exciting also ‘cause it'll save us a ton of time. and I think there's a, you know, there are a lot of different little paper cuts and bottlenecks that are sort of standing in the way of that. I think that there's, yeah, there's things on the model side, there's things on the inference side, there's things on the harness side, there's things in the, in the representation. You know, we find that as Computer Use gets faster, we're increasingly bottlenecked by just, like, the speed of doing an operation. Like, for example, you know, a non-trivial amount of time in our benchmarks of Computer Use tasks is actually, like, let's say you're automating a task on doordash.com. Like, a lot of the time is actually waiting for doordash.com itself to load, you know?Swyx [00:14:04]: Yeah, then you just write a wait and then you execute the wait.Ari Weinstein [00:14:07]: Yeah, totally. And you wanna get. Yeah, actually, it's actually really important that you get that de- like, you want as little delay as possible between when it finally finishes loading and when you go andSwyx [00:14:16]: YeahAri Weinstein [00:14:16]: Trigger the LLM to do the next action, which is actually- itself a statistical science.Swyx [00:14:20]: Like an event-driven way maybe to do that.Ari Weinstein [00:14:22]: When possible, you want it to be event-driven.Swyx [00:14:24]: JavaScript has some load events.Ari Weinstein [00:14:25]: And JavaScript has load events for. or the web browser has load events for web navigation, but there's other types of events that actually really can't be event-driven. So there's a lot of complexityVibhu [00:14:34]: The one that comes to mind is, like, chatting with customer service.Ari Weinstein [00:14:37]: Yeah.Vibhu [00:14:37]: Replies could take 30 seconds, could take three minutes.Ari Weinstein [00:14:39]: Oh, right.Swyx [00:14:41]: I have dealt with so many bots with Codex. it's great, but I also wonder if the other side knows that they're talking to a bot ‘cause I'm, like, answering in complete sentences. Like, I'm capitalized correctly.Ari Weinstein [00:14:50]: That's hilarious.Swyx [00:14:51]: Like, I'm giving full num- full reference numbers and everything. Like, it's too. it's clearly too good. I don't care. Like Like, I'm just, like, trying to get my support case.Vibhu [00:14:58]: I've prompted it to, like, you know, “Don't pretend you're a bot. Be very annoyed human.”Vibhu [00:15:02]: Short one-liners, likeSwyx [00:15:04]: YeahVibhu [00:15:04]: Push it, do all this. I also tell it, “While you're waiting for responses, like, use subagents to research better ways to figure out what we need.”Ari Weinstein [00:15:12]: Nice.Vibhu [00:15:12]: It's just, like, human little intervention.Ari Weinstein [00:15:14]: That's awesome. I also feel like half the time it's a bot on the other end, so now youVibhu [00:15:17]: YeahAri Weinstein [00:15:17]: Got the bots talking to each other.Swyx [00:15:18]: Yeah. I will also say, you know, like, you know, one milestone of Computer Use that we are, we're at now is, you know, three, four years ago, we were scared of hooking up LLMs to the, to the web and toAri Weinstein [00:15:31]: YeahSwyx [00:15:31]: To our, to our devices. And now I'm having it configure DNS for me.Ari Weinstein [00:15:35]: Wow.Swyx [00:15:36]: I'm having it pay my bills, and, like, really, like, tens of thousands of dollars of, like, stuff I'm just sending it over and yoloing with Computer Use and, like, you know, what's the, what's the worst thing that can happen?Swyx [00:15:48]: So that- that's all, that's all really good.Building Safely With Computer Use in the Agents APIAri Weinstein [00:15:50]: Yeah.Swyx [00:15:50]: I think now that you've. you know, obviously, you also have to dogfood your own products and all these things. Now that you've sort of released this in API, what are some pitfalls or tips that you wanna tell developers, because they're about to, I guess, encounter all this, firsthand?Ari Weinstein [00:16:03]: First of all, I'm just really excited that we brought Computer Use into the Agents API. I think this is, really great because obviously a lot of developers are building applications that wanna be able to work with third-party websites and services. And so Computer Use has this universality to it. It can work with anything. So now all of a sudden, developers can build using the same Computer Use implementation that we're building on. I think there's great work to be done if you wanna build your own Computer Use harness, but it's hard. And also, we train our models on our Computer Use harness, so there is, an advantage to using the one that's in distribution for the model. There actually might be a speed and cost and accuracy advantage. So I think it's really great for people to get to build on top of that. And, yeah, you know, I think kind of to the point that you were making, like, I think we're all sort of still in the process and maybe, like, some of us are ahead of many people in the world of, like, getting comfortable with this technology and trusting it. And so I think it's incumbent on us to, sort of build that trust over time by making sure we're building things that are reliable, by building, the right kinds of safety checks, by asking for the user's consent before doing something consequential like making a payment, by, asking, you know, maybe depending on the application, making sure you're only letting it access the websites or applications that it actually needs for the task. So that's, I think, something important to think about. but yeah, I'd really encourage people to try the new Agents API, build all kinds of cool stuff on it. We'd love to hear your feed- feedback if, you know, depending on how it goes.Vibhu [00:17:31]: Have you seen any changes in the way it affects dev workflows? So one of the things with dots is, you know, you're seeing it in Slack.Ari Weinstein [00:17:38]: Yeah.Vibhu [00:17:38]: You're seeing people use voice and build. the example Roman showed of change this app and send me screenshots along the way and all this.Computer Use for Testing and Closing the Software LoopAri Weinstein [00:17:46]: Yeah.Vibhu [00:17:46]: Is anything that you're seeing there in adoption about how people are using Computer Use for coding workflows? Any tips people should take from that?Ari Weinstein [00:17:55]: One of my favorite use cases for Computer Use actually, and one that we see a lot in the wild, is Computer Use letting the agent- actually test the software that the agent has built, which is far more consequential than it sounds. Because traditionally, you know, you'd build something in Codex and then the-- and the Codex builds it for you, and then you have to test it, and you are now like QA for the agent, right? So with Computer Use, you can complete the develop-- the software development life cycle, where, the agent can build software, it can test it. So I have a lot of fun, you know, building stuff, having the agent test it. By the time it comes to me, it's already working. I have, extra fun because sometimes I'm, like, developing Computer Use itself, and so now I have a Computer Use agent that's using my Computer Use agent that's using something else. so yeah, I really, I really think this is a super powerful class of use case.Swyx [00:18:44]: I have a visual play test skill that I've developed that, really catches a lot of design issues,Ari Weinstein [00:18:49]: NiceSwyx [00:18:50]: That, you know, normally when you just look at code, you wouldn't really pick it up. it's also really good for cloning apps, though. If you're using a shitty SaaS and you wanna kill the SaaS You just clone it screen by screen by screen. and Obviously, Computer Use can completely drive everything, take screenshots, note it down, and then clone everything with Codex.Ari Weinstein [00:19:06]: That's really cool.Swyx [00:19:06]: But yeah, thanks for all your progress. I think, that isAri Weinstein [00:19:08]: AbsolutelySwyx [00:19:09]: Our time.Nikunj Handa: What's New in the OpenAI APIAri Weinstein [00:19:10]: Yeah.Swyx [00:19:10]: This is not the last that we're gonna talk.Ari Weinstein [00:19:12]: Yeah, cool. This has been really fun. Thank you guys for having me.Swyx [00:19:14]: All right.Vibhu [00:19:14]: All right. Okay, we're a strict cutoff. We're just gonna dive right in.Nikunj Handa [00:19:17]: Let's do it, yeah.Vibhu [00:19:19]: Okay, so, Nikunj, we're very excited to have you. You shipped a lot on the API side, like we justNikunj Handa [00:19:25]: YeahVibhu [00:19:25]: Talked about with Ari. You can now build with Computer Use agents. Anything you wanna highlight, the API side of changes, and introduce yourself a little and what you do?Nikunj Handa [00:19:34]: Yeah, for sure. My name is Nikunj. I lead product for the API team. Been here for roughly three years. been working on launching models. I feel like that's just been, like, a thing, constant thing throughout my time, here at OpenAI. And, with every new model, we try to, like, basically work super closely with the post-training team, the research team, to figure out what's new in it. and then we, like, expose those capabilities in the API. so that's, like, the basic way of putting it. and if you just look at, everything that's new with GPT-6, the cool new capabilities that we launched were, firstly, async function calling. so what you see with, like a lot of the things that you're seeing in, like, Codex and Dots and everything is that tool calls take so long that you don't have to, like, pause the model's execution while, the tool is running. So you could just, like, kick off a tool call, keep running, keep reasoning, and then check back in. so we launched async tool calling. We launched, like, mid-turn steering, so now you can, like, inject messages while the model is reasoning, in the middle. so as your tool call finishes, you can put in that instructions.Async Tool Calls, Mid-Turn Steering, and WebSocketsSwyx [00:20:43]: And that's also partially a model alignment capability, right?Nikunj Handa [00:20:46]: Yeah.Swyx [00:20:46]: Like, they have to train in the ability to train.Nikunj Handa [00:20:48]: Exactly, yeah. AndVibhu [00:20:49]: I feel like we've had it in the app. You could always, as it's reasoning, you could steer.Nikunj Handa [00:20:54]: Yes.Vibhu [00:20:54]: It wasn't the best. It's gotten much better.Nikunj Handa [00:20:57]: Yeah.Vibhu [00:20:57]: Excited to see how it does this in versionNikunj Handa [00:20:58]: Yeah, and I like our mainVibhu [00:20:59]: And nowNikunj Handa [00:21:00]: Goal in, our main goal in the API is to, like, put things in the API once it's trained into the harness. And so we kinda wait for that moment until it's good enough. And a lot of that is, like, actually being powered by WebSockets, which we launched, a few, I wanna say months ago. And so WebSockets just opens this, like, whole bidirectional, like, communication thing with the model. This is not, the GPT Life thing. I'm just talking about GPT-6. and you can do all these, like, async tool calling, async reasoning, injecting messages. It's a really fun API to work on. I think, like, really enjoying.Swyx [00:21:33]: Yeah. This is why we are the engineering podcast, because we get to talk about WebSockets.UltraFast and the Inference StackNikunj Handa [00:21:36]: Yeah.Swyx [00:21:37]: This also pairs very well with UltraFast, right?Nikunj Handa [00:21:39]: Oh, yeah.Swyx [00:21:39]: Like, that is now, like, I think for the first time ever available in the API.Nikunj Handa [00:21:43]: Yes.Swyx [00:21:43]: Which is, which is basically the theoretical fastest speed you can ever get, Frontier of Intelligence.Nikunj Handa [00:21:49]: Yeah. It's been so exciting to work on that project. I think, before I go into the API, the most fun part of, UltraFast has been just watching the inference team cook with Astra. Like, they're just, like, constantly having these, like, Codex agents running, trying to, like, squeeze out more performance. And, I would say, like, at least for a couple of months, a lot of it was focused on efficiency and driving the cost down, which is how we, like, were able to cut the Luna price by, like, 80%. It was, like, a lot of that was driven by, like, all the inference improvements they landed. And then now they've, like, shifted gears towards, like, how can we make this run as fast as possible? And so UltraFast has just been, like, amazing to see on a mo- on a model like Astra. Like, to go that fast has been really cool. And yeah, WebSockets is like. actually it was like the first time we launched WebSockets, it was for GPT, 5.3 Codex Spark, which was. Can't believe we named a model that, but, you know, that's what we launched it for. And obviously, it helps so much because, like, you gotta have the tool calls. you had, like, really reduced the overhead, of going back and forth with tools. And so, WebSockets is awesome for that.Swyx [00:22:57]: Yeah. it's always cute to see, like, I have my reset usage limit, and then I have my Spark usage limit that I never use.Nikunj Handa [00:23:03]: Yeah.Swyx [00:23:04]: Like, it's there if I want it.Nikunj Handa [00:23:05]: I think it's gone finally.Swyx [00:23:06]: It's gone. It's gone, yeah.Nikunj Handa [00:23:07]: I know it's gone, so.Swyx [00:23:08]: Yeah. you're slowly killing off all the, you know, theNikunj Handa [00:23:11]: The old ones, yeah.Swyx [00:23:11]: Oldies.Vibhu [00:23:11]: This is a great week. I mean, it was the first time we had Frontier Intelligence at extreme speeds.Nikunj Handa [00:23:17]: Yeah.Vibhu [00:23:18]: People really liked it.Nikunj Handa [00:23:19]: Yeah.Vibhu [00:23:19]: SoSwyx [00:23:20]: YeahVibhu [00:23:20]: First time it comes back.Swyx [00:23:21]: Yeah. for, 5.3 Spark is explicitly attributed to Cerebras. You guys are not confirming or denying that, UltraFast is related to Ce- Cerebras, but people are. I'll just say that people do care and, are wondering about it. And you have your own silicon as well. elephant in the room, decision models.Decisions API: OpenAI's Fast Decision ModelNikunj Handa [00:23:38]: Oh, yeah.Swyx [00:23:38]: Decisions API. We were the first podcast to do a big Jev, deep dive with, Diogo, and I also, you know, featured him at AI Engineer. How quickly did you see Jev and go likeNikunj Handa [00:23:49]: Oh my gosh. Yeah.Nikunj Handa [00:23:50]: Yeah. Firstly, like, huge props to Diogo and, like, the Jev team for, like, really inspiring theSwyx [00:23:55]: YesNikunj Handa [00:23:55]: Like, whole segment in the market. Like, obviously Jev comes out, everyone's, like, losing their minds over it. Our users are, like, hitting us up. But also, like, our internal teams are like, “We need, like, a much faster classification system.” We can. I don't wanna, like, get ahead of some of the dots features that are gonna comeSwyx [00:24:16]: WhooNikunj Handa [00:24:16]: But you're gonna see, like, some cool, like, really snappy, fast things built on top of the decisions API. but, you know, like, yeah. Props to Jev for, like, inspiring this whole thing. obviously a bunch of people at OpenAI get nerd sniped by that, and they're like, “How can we, like, make this work? We're not gonna, like-”Swyx [00:24:33]: Okay.Nikunj Handa [00:24:33]: “. train a new model.” ButSwyx [00:24:34]: Like, four weeks ago, this was not on the dev radar, right?Nikunj Handa [00:24:37]: No, not at all. No.Swyx [00:24:37]: Okay.Nikunj Handa [00:24:37]: This is likeSwyx [00:24:38]: WowNikunj Handa [00:24:38]: Jev-inspired and, likeSwyx [00:24:40]: I think you are officially the first one to your lab to, like, clone and, adopt this.Nikunj Handa [00:24:44]: Yeah. Yeah. I feel like, OpenAI has such a strong, like, hacker culture and, like, people are just, like, they get excited about things. And so, guy from inference, this one awesome guy from, the infra team are like, “ this is amazing. We're gonna, like, hack on it.” They build a prototype, it, like, works, and now we- we are just, like, hill climbing on latency and trying to make this as fast as possible, and we wanna, like, launch it in the coming days. so as soon as we hit our, like, latency target, we'll try to get this out.Vibhu [00:25:13]: It's interesting. At the same time of hacker culture, you also, as Sam said, like 99%, one of the most reliable APIs withNikunj Handa [00:25:20]: Mm-hmmVibhu [00:25:20]: I think probably the most usage, which is your team directly. how should people see decisions API? I feel like a lot of people saw Jev, heard the buzz, haven't built with it. You're making it very mainstream.What Decision Models Are Good ForNikunj Handa [00:25:32]: Mm-hmm.Vibhu [00:25:33]: What should people see it as? How should they use it?Nikunj Handa [00:25:36]: Yeah. I think the main use cases we've seen is, like, really fast classification. all the Computer Use demos have been amazing and really cool. I think there will be limitations, of course, in terms of, you know, having Astra, like, write, like, a JavaScript-like script to control your computer, versus having Luna pick, like, one action at a time. I think, it's not gonna be at the same intelligence level, but, like, maybe there's some Computer Use tasks that this is good enough for. So excited to see that come through. the other cool prototype I've seen internally is people hooking it up with GPT Live. So GPT Live is like, you know, our bidirectional, like, real-time,Swyx [00:26:14]: VoicingNikunj Handa [00:26:14]: A- API. And, it's built on this, like, model of front-end models and back-end models. So GPT Live is this, likeSwyx [00:26:20]: Think or talkerNikunj Handa [00:26:21]: Super fast. Yeah, think or, talker thing. So GPT Live is the talker, super fast, really good at delegation, and you have something like Astra sitting at the ba- at the back. But tool calling has always felt, like, really slow in GPT Live. and so people have been, like, putting together these, like, tool calling demos of GPT Live controlling a computer, and it just feels like so much more snappy and natural. So I'm, like, kinda excited to see, like, what people do with Live and with Luna on decisions API. so that'll be pretty exciting. Yeah.Swyx [00:26:55]: So I wanna iron this out for people, especially from the product side, because a lot of people have been putting out Jev clones. There's been about 100 in the last two weeks.What Makes a Decision Model DifferentNikunj Handa [00:27:01]: Oh, really? That's amazing.Vibhu [00:27:03]: The first couple days.Swyx [00:27:04]: But like, it. Like, they can clone a Jev API, which is honestly structured outputsNikunj Handa [00:27:09]: YeahSwyx [00:27:09]: Which OpenAI was first to.Nikunj Handa [00:27:10]: Yeah.Swyx [00:27:11]: Right? So, like, I think let's iron out for people what is a decision model, as far asNikunj Handa [00:27:16]: YeahSwyx [00:27:17]: As far as, like, what is important? It is not just latency. It's not just structured output, right? Because I could just have Luna as it'- The decision model is priced the same as Luna, right?Nikunj Handa [00:27:26]: Mm-hmm.Swyx [00:27:27]: Have turned off reasoning and then have structured output. Do I have a Jev? you know, no, right? And that's theNikunj Handa [00:27:33]: YeahSwyx [00:27:33]: That's the realVibhu [00:27:34]: There's a confidence there.Swyx [00:27:35]: Yeah.Nikunj Handa [00:27:36]: Yeah, totally. I think, the way that. So we haven't trained, like, a new model for this.Swyx [00:27:40]: Yeah.Nikunj Handa [00:27:40]: We're, like, building this purely on top of the same Luna weights that we have.Swyx [00:27:44]: Oh.Nikunj Handa [00:27:44]: So yeah. This is, like, really just Luna. And, on top of that, what you're doing is you're constraining. So, like, structured output's a big part of it. you're really optimizing the inference stack to, like, get very fast on TTFD. And because you can have multiple questions, what you do is, like, you basically run those in parallel,Swyx [00:28:05]: As a batch.Nikunj Handa [00:28:06]: Yeah. You run those in the-- as a batch. you-- All sorts of, like, inference techniques people are working on to try to make it as fast as possible. But I'd say, like, at least our implementation of it at the start and this first version is, like, zero-shotting this on top of Luna, to see how it goes. And obviously, you wanna, like, put it out there. Like, this is OpenAI's, like, classic iterative deployment thing. Put it out there, see what people think, and then, like, we'll make more model improvements, as needed. so yeah. That's, the decisions API.Swyx [00:28:38]: Yeah. And, obviously as a benefit, you have vision. They don't have vision, right?Nikunj Handa [00:28:42]: That's true.Swyx [00:28:42]: Obviously, Jev's comes withNikunj Handa [00:28:43]: Yeah. Like, we get it for free with Luna. Yeah.Swyx [00:28:45]: Yeah. I do think that, like, you know, some of the innovations, it sounds like, it's still to come if it's still the same Luna weights, which is, like, the confidence stuff, like, the in calibration is something that we've talked about on the podcast with, benchmarking calibration. ‘Cause basically, the whole point is that RLHF kind of collapses you towards what you want to hear.Calibration, Architecture, and the Open Research QuestionsNikunj Handa [00:29:03]: Yeah.Swyx [00:29:03]: But, like, not actually, like, what the amount of confidence is.Nikunj Handa [00:29:06]: Yeah. Yeah, totally. I'm eager to see how it pans out. Maybe there's, like, gonna be. These are gonna be, like, the key areas where we may have to, like, hill climbSwyx [00:29:15]: YeahNikunj Handa [00:29:15]: With the, with the future model release. But, yeah.Swyx [00:29:18]: And then architecture-wise, the other thing that's in the debate, obviously, you-- Nobody knows because Jev doesn't talk about it, but the two speculations are, one, maybe diffusion model instead of autoregressive.Nikunj Handa [00:29:28]: Mm-hmm.Swyx [00:29:29]: But you are able to achieve the parallel, generation in your way. And then the other one is some mech interp type thingNikunj Handa [00:29:37]: Mm-hmmSwyx [00:29:37]: That you're, like, analyzing the activations and then just outputtingNikunj Handa [00:29:40]: That would be coolSwyx [00:29:41]: The weights.Nikunj Handa [00:29:42]: Yeah.Swyx [00:29:42]: Which, like, you guys have all done the research on this. People have speculated.Vibhu [00:29:45]: There have been demos onSwyx [00:29:46]: YeahVibhu [00:29:46]: Both of these as well. I think Gemini shared a Gemini diffusion, Gemma diffusion on a Jev-style output.Nikunj Handa [00:29:53]: Oh, sick.Vibhu [00:29:53]: And, interp people have also, you know, pulled out interp from a middle layer, but this is all speculation.Swyx [00:29:59]: It's just like, what are you trying to aim for, right? Because you can achieve the API. Everyone can achieve the API. It's actually pretty trivial. But, like, then there's the speed, then there's the accuracy, then there's the other calibration features.Nikunj Handa [00:30:11]: Mm-hmm.Swyx [00:30:11]: I don't know what else.Nikunj Handa [00:30:13]: Yeah. Yeah. No, totally. It's so cool that this, like, whole space has been kicked off now and people are gonna do so much cool stuff and everyone's gonna learn from each other. And, yeah, I'm excited about it.What Developers Should Build NextVibhu [00:30:24]: I feel like being on the platform team, a lot of your job is to empower builders.Nikunj Handa [00:30:27]: Mm-hmm.Vibhu [00:30:28]: What do you think people should build with decisions API and also Computer Use agents? Any stuff that you've- been building with internally that you think really opens up after the new change?Nikunj Handa [00:30:39]: Yeah. okay, let's think. decisions API, use cases internally have been pretty obvious. Like, the user ops team was, like, jumping on it. We were like, “We gotta classify all of our support tickets.” what else came up? obviously, there were, like, the really cool GPT Live demos. I'm sure, like, the Codex app team might, like, pick this up and try to do something cool with it. So, you know, like, this whole thing started, like, a week ago, so it's, like, very early andSwyx [00:31:06]: Oh, one week.Nikunj Handa [00:31:07]: We're excited. Yeah. Yeah, pretty much.Vibhu [00:31:08]: There was a big push in, evals, LLM as a judge having really low latency there.Nikunj Handa [00:31:13]: Right. Yeah. That'll be interesting to see. and then, with the Agents API, we have-- we're basically, like, having a bunch of first-party products, like, at OpenAI built fully on top of it. we've had the Codex security stuff that just went out that's fully built on top of, the Agents API. We have, sort of the-- we- we are having, like, a meetings type of thing launching today.Agents API and OpenAI's First-Party ProductsSwyx [00:31:40]: Mm-hmm.Nikunj Handa [00:31:40]: I think there was, like, a demo. do you remember, like, the plugin extensions when Sam was showing it? There was, like, a demo for, like, you're in a calendar, you can sort of, like, have your meeting notesSwyx [00:31:51]: Like, drop into a singleNikunj Handa [00:31:52]: Flow into like your spaceSwyx [00:31:52]: Like, Google Docs type thing.Nikunj Handa [00:31:53]: Yeah.Swyx [00:31:54]: Right?Nikunj Handa [00:31:54]: And so the-- all of that stuff is, like, fully built on top of, the Agents API. and yeah, I'm, like, just excited to see. Like, we're just getting this out, and let's see what people build on top of it.Vibhu [00:32:04]: I think you showed it off very well. The whole edit spaces, pages, collaborate, add in your dot. Like, that's a lot, soNikunj Handa [00:32:12]: YeahVibhu [00:32:12]: There's a lot of inspiration people can go to.Nikunj Handa [00:32:14]: Yeah. All possible with Astra, you know. Like, thing- things just move so fast now. LikeSwyx [00:32:19]: YeahNikunj Handa [00:32:19]: People go from idea to execution so quickly, it's amazing.Swyx [00:32:23]: Is there something that you want, people to focus on to give you feedback? Like, what-- like, you know, maybe you're just putting this out there and you want-- and there's, like, a fork in the road and you want developers to help you decide.Responses API Performance and Long-Lived CachingNikunj Handa [00:32:35]: So I think Agents API and decisions API, they are like, these are our newest products. Would love, like, any and all feedback on that to figure out where to take them. I think, over here, we're, like, very open on Responses API, which is sort of like our workhorse over here. like, really focused on performance right now, and the performance comes in, like, two main ways. first is just, like, latency. We've been, like, rewriting the whole Responses API stack to, like, make it as fast as possible from a TTFT perspective, DVD perspective. So there's like-- that, like, continues to be, like, a main area of focus for us. The second thing we've been trying to do is, like, really go deep on caching, particularly with these, like, personal agents that are, you know, like, basically, like, a single thread that just goes on and on forever. We've been, trying to, like, really up our game on caching. We provide now guarantees of, like, cache hits within, like, 30 minutes. We're actually, like, we-- for one of our users, we just launched, like, a much longer cache window. So we have, like, a 12-hour caching guarantee, that we offer so that you have, like, guaranteed cache hits forSwyx [00:33:40]: Is that a public API?Nikunj Handa [00:33:42]: Not yet. That's in preview.Nikunj Handa [00:33:43]: We're gonna, like, try to get that out to everyone as soon as possible. But, like, just pay a little bit more for the cache write, and we, like, guarantee, like, cache reads for, like, a much longer period. So even if, like, your instinct thread, for example, like, you just, like, do something on it and then come back to it, like, three to four hours later, you- you're still getting the caching performance out of it. And launchedVibhu [00:34:04]: And you cut the cost there quite a bit too, right, with the new model?Nikunj Handa [00:34:07]: Oh, yeah. That's right.Vibhu [00:34:08]: Like, 25% cheaper, soNikunj Handa [00:34:08]: Yeah, with, like, driving down cache reads, yeah.Cache Pre-Warming and Cost-Efficient Agent ThreadsVibhu [00:34:10]: For builders, they should implementNikunj Handa [00:34:13]: YeahVibhu [00:34:13]: Because it's significantly cheaper.Nikunj Handa [00:34:14]: Yeah. Yeah. Just, like, building your apps with, like, to be very cache aware and sort of, like, use our prompt diagnostics or cache diagnostics tool to figure out, like, where things are dropping off. And, so the caching part is, like, really important. yeah, I also wanted to talk about pre-warming. We have that in the API now. So, like, if you know that, “Hey, I'm gonna get a cache,” like-- sorry, “I'm gonna get this prompt. I just wanna, like, pre-warm the cache, pay, like, the cache write fee right now, and then, like, have it sort of ready to go for the next 30 minutes for whenever.”Swyx [00:34:49]: And it can spawn many instances of that thread.Nikunj Handa [00:34:51]: Exactly, yeah.Swyx [00:34:52]: Yeah.Nikunj Handa [00:34:52]: You can just keep going and haveSwyx [00:34:54]: Yeah, just keep messing with the prompt thereNikunj Handa [00:34:55]: Tons and tons of that. and so, yeah, like, I'm very excited about getting feedback on, like, the low-level performance things that we can keep making Responses API the most performant and reliable way to, like, build on top of an LLM. And then you basically have our, like, new products where I'm just looking for, like, any and all feedback.Swyx [00:35:15]: Yeah, just use it, right?Nikunj Handa [00:35:16]: So yeah, just useSwyx [00:35:16]: Tell us what toNikunj Handa [00:35:17]: Yeah. Define our roadmap for us, please. So yeah.Swyx [00:35:20]: I think for me, the caching thing, great, right? Like, obviously very needed. But at the end of the day, you're still bumping up against a million-token contextCompaction and Managing Million-Token ContextsNikunj Handa [00:35:28]: Mm-hmmSwyx [00:35:28]: And that's probably not gonna change for the foreseeable future.Nikunj Handa [00:35:31]: Mm-hmm.Swyx [00:35:31]: Like, you still need good compression.Nikunj Handa [00:35:33]: Yeah.Swyx [00:35:33]: What is the best practice there?Nikunj Handa [00:35:34]: Yeah. Yeah, totally. so firstly, OpenAI has its own, like, proprietary compression, compSwyx [00:35:40]: Which is inNikunj Handa [00:35:41]: Compaction.Vibhu [00:35:42]: Compaction.Swyx [00:35:42]: It's in the agents.Vibhu [00:35:43]: It's in the API.Nikunj Handa [00:35:43]: Yes.Vibhu [00:35:43]: Agents API.Nikunj Handa [00:35:44]: Yeah.Swyx [00:35:44]: You decide for us, right?Nikunj Handa [00:35:45]: Yeah, exactly. So in the Agents API, it comes built into the harness. and if you're in Responses API, there's, like, two ways of doing it. One is what we call server-side compaction, which is you basically tell Responses API that if you ever hit this threshold of tokens, just auto-compact it and, like, go back, or sorry, like, reduce the context, being used. And the second way is, like, /compact, which is, like, if you want full control. So you can, like, /compact at any timeSwyx [00:36:15]: I hear youNikunj Handa [00:36:15]: Have your own logic on when to, likeSwyx [00:36:17]: It's not AGI.Nikunj Handa [00:36:18]: It.Swyx [00:36:18]: It's not AGI.Nikunj Handa [00:36:19]: Yeah. Yeah.Swyx [00:36:20]: Yeah. But it, I meanNikunj Handa [00:36:20]: YeahSwyx [00:36:20]: It is the manual override.Nikunj Handa [00:36:21]: Yeah, it is the manual way. And like, I don't know, but a lot of the big coding agents like to do it manually. I mean, like, if you look at the Codex implementation of it in the Code- open source Codex harness, you can see that they use /compact and do it. and, there's also, like, new, by the way, new compaction techniques that we are working on. Some of them you will be able to see in the Codex harness. Like, it's already implemented in the Codex harness. And so, they're like some file-based, systems that we are, like, experimenting with. So yeah, lots of cool stuff going on around in compaction as well.Swyx [00:36:57]: Cool. we are running out of time.Nikunj Handa [00:36:59]: Okay.Swyx [00:36:59]: I think you've talked about, a lot about performance and talked a lot about, the new APIs that you're launching. Can you give us any other hints as to things that you're interested in as far as the future of the platform is concerned?Higher-Level Platform Primitives and the AI CloudNikunj Handa [00:37:13]: We're obviously like very low level. Like, I used to work at Stripe before this, and, at Stripe a lot of the game was like building these higher level primitives and products on top of like the core payments primitives. and, I'm always like curious about what the best way of doing that is in AI. And I think we've had a couple of attempts at that. We like had launched assistance API like way back in the day, and like wasn't really the right fit. We were sort of like going off with this like Agents API, and, it gives you the codex harness, but like where's like the, what's the right amount of flexibility to give in that? That's like an open question. Like how should we like have memory walls and like all of these like higher level like API objects to take away, also like to abstract away more, like storage concepts. Like this is like a whole, like, there's a whole space that I'm like very curious about figuring out how we design. I think a lot of things in AI are just have a low-level API primitive and see an example harness and go and have your coding agent implement that. But how much of that do we build into the API is like a constant question that I'm thinking about.Swyx [00:38:24]: Yeah.Nikunj Handa [00:38:24]: So I don't know if folks have thoughts on that. If anyone has ideas, it would be super interesting to hear.Swyx [00:38:30]: Yeah. The analogy I always bring back to, and we'll end there, is, you're building an AI cloud, right?Nikunj Handa [00:38:35]: Mm-hmm.Swyx [00:38:35]: Like, which is, something that, Sam said a year agoNikunj Handa [00:38:38]: Mm-hmmSwyx [00:38:38]: Where, and you're, it's almost like you're kind of doing the AWS invention and you have to do, okay, this is EC2Nikunj Handa [00:38:45]: YeahSwyx [00:38:45]: And this is S3, and this is like. But you're doing the AI-native versions of each of these.Vibhu [00:38:48]: There are a lot of analogies, so you're pre-warming caches for stuff that you know will beNikunj Handa [00:38:53]: Yeah.Vibhu [00:38:53]: And it's nice that it's all exposed to buildersClosingNikunj Handa [00:38:56]: Mm-hmmVibhu [00:38:56]: ‘cause it just opens up ways that you can build new things.Nikunj Handa [00:38:59]: Yeah, absolutely.Swyx [00:39:00]: Okay.Vibhu [00:39:00]: Awesome. WellSwyx [00:39:01]: That's everything.Nikunj Handa [00:39:01]: Thank you, guys.Vibhu [00:39:02]: Thank you.Nikunj Handa [00:39:02]: Yeah. This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.latent.space/subscribe

    #neuvottelija
    Aura Boards | Ville Tolvanen Juho Jokinen | Neuvottelija 411

    #neuvottelija

    Play Episode Listen Later Sep 30, 2026 68:42


    Suomalaisyritysten hallitukset ovat usein analogisen ajan jäänne: ne kokoontuvat katsomaan menneitä lukuja ja vaihtavat toimitusjohtajaa, kun tulokset eivät parane. Ville Tolvanen väittää, että hallitus pitää keksiä uudelleen tulevaisuuden johtamisen prosessiksi, jossa omistaja luo mahdollisuudet, hallitus rakentaa huomisen ja johto tekee tuloksen. Aura Boards yhdistää tulevaisuuteen katsovan ideologian, tuhannen päivän strategian, vuosisuunnitelman ja tekoälyn tukeman tilannehuoneen, jossa kone laskee ja ihminen päättää. Juho Jokinen ja Sami Miettinen grillaavat ajatusta säälimättä ja kysyvät, mikä omistajan, hallituksen ja johdon ketjussa oikeasti muuttuu, miten se toimii arjessa ja mitä näyttöä hyödyistä jo on. Mukana esimerkkejä Fredman Groupin ERP-hankkeesta, strategiapäivistä Alpeilla ja siitä, miksi hallitus on Suomessa halpaa kuin saippua.00:00 Intro: onko Aura Boards hallitustyön vallankumous?01:00 Ville Tolvanen ja Juho Jokinen esittäytyvät01:35 Jokisen ura ja 19 vuotta Tolvasen kanssa03:39 Vallankumous vai keisarin uudet vaatteet04:29 Tolvasen sata hallitusvuotta ja omistajuuden tutkimus05:50 Miksi hallituksen rooli pitää mullistaa08:42 Suomalaisyritykset eivät kasva eivätkä ole kannattavia09:51 Lifeline Ventures ja epäonnistuvat sijoitukset10:48 Omistaja luo mahdollisuudet, hallitus huomisen11:17 Investointipankkiirin näkökulma omistajiin ja hallituksiin12:45 Hallituspaikka ei ole meriittipaikka14:16 Miksi hallitus on jäänyt analogiseen aikaan15:58 Kulttuurimuutos ja we over me17:46 Hallitusvaltaa ilman omistajan euroja20:14 Hallitustutkimus: kuinka paljon aikaa tulevaisuudelle20:52 Tavoitteena 80 prosenttia tulevaisuudesta22:01 Ferrarin joukkue ja ryhmätyön voima22:59 Board as a Service ja Linux-malli24:33 AI-tilannehuone kolmesta tilinpäätöksestä25:32 Kone laskee, ihminen päättää26:52 Kenen tekoäly on paras28:16 Fredman Group ja Peltolan Pussin ERP-hanke31:47 Tuhannen päivän strategia ja vuosisuunnitelma33:35 Tulevaisuuteen katsominen ilman tekoälyyn nojaamista36:31 Esimerkki: konepaja ja lykkääntyvä suuri tilaus40:10 Yhteinen data korvaa toimitusjohtajan suodattimen41:02 AI-agentti hallituksen keskustelukanavalla43:23 Itsenäinen agentti vai digitaalinen kaksonen45:22 Organisaation muisti ja tekoälyn kontekstin rajat47:22 Puheenjohtaja kapellimestarina50:00 Strategiapäivät Alpeilla ja reaaliaikaiset muistiot53:37 Tietoturva ja yrityssalaisuudet tekoälyn aikana54:47 Hallitusportaalit ja Admincontrolin mahdollisuus55:28 Hallitus on halpaa kuin saippua57:31 Liikaa yrityksiä, liian vähän hyviä yrittäjiä59:13 Venture capital, private equity ja perheyhtiöt1:00:48 Piemonten leiri ja avoin yhteisö1:03:00 Mitä Aura Boards muuttaa käytännössä1:04:38 Ennuste ja CFO jokaiseen yhtiöön1:06:47 Miten Aura Boardsiin pääsee mukaan1:07:05 Tuomio: vallankumous vai keisarin uudet vaatteet1:07:56 Suomen paras Rolodex ja Neuvottelijan sisäpiiri

    Latent Space: The AI Engineer Podcast — CodeGen, Agents, Computer Vision, Data Science, AI UX and all things Software 3.0

    We are excited to have Anthropic share their latest AI x Finance work at AI Engineer New York, coming up in 2 weeks!In case you've been under a rock, here's a non-exhaustive list of what Anthropic has been shipping since closing the largest fundraise of all time in May at $47B ARR:* June: Launched Claude Tag and Sonnet 5 and Fable 5* July: Opus 5, /checkup. crossed $65B ARR* Last month: Fable/Mythos 5.1, and EFS (upcoming pod)* IPO target $2T, end 2026 ARR estimated $100B* Cowork/chat merged before did* Claude Mods* Dario endorses the same Pacing the Frontier message cosigned by all labs* Last week: Opus 5.5, Plugins portal, Cloud Sessions/Claude Projects* Today: Sonnet 5.5!Today's episode should catch you up, with Thariq Shihipar, the explainer-king of Anthropic, who we last caught up on Fable launch day with The Field Guide to Fable:The Future of Mutable SoftwarePay special attention to Claude Mods (especially the cheatsheet):In general this is also the inverse of the other viral tweet from Thariq:Cloud Brain, Local HandsAnd give a try to Claude Projects:The “hands” terminology is not just an analogy for the local/cloud paradigm that is being built up at frontier coding agent companies like Cognition, but is ALSO particularly relevant to the safety systems discussions that we'll be discussing with Anthropic in an upcoming episode as they prepare to pace to frontier with responsible AI deployment.For those who want Thariq's writing tips we teased at the start of the pod, watch the full video here:From the rapid rise of Claude Code to a future where agents can rewrite their own harnesses, collaborate across teams, and operate across cloud and local environments, the way we build software is changing extraordinarily fast. In this episode, Anthropic's Thariq Shihipar joins swyx and Vibhu to unpack how power users are actually working with Claude Code today, why prompting remains a high-skill discipline, and where Anthropic thinks the agent harness is headed next.We go deep on Claude Code's evolving interface: Ask User Question and elicitation, artifacts as persistent generative interfaces, Claude Tag for multiplayer agent workflows, Projects, model effort, implementation notes, and the new Claude Mods system for customizing the harness itself. Thariq explains why Claude.md may eventually disappear, why the smartest model could also become the cheapest model for many tasks, and why mutable software could become a new paradigm for how applications are built and customized.The conversation then turns to agent security and Anthropic's “Pacing the Frontier” argument. Thariq walks through recent incidents where agents discovered unexpected ways to communicate, exploit infrastructure, reverse-engineer benchmark scorers, and chain vulnerabilities together. We discuss sandboxing, prompt injection, autonomous agents, interpretability, constitutional classifiers, probes, fallbacks, Auto Mode, and why securing increasingly capable agents may become one of the defining engineering problems of the next few years.We discuss:* Why agentic coding went from controversial to the default in less than a year* Why prompting is still one of the highest-leverage skills for working with Claude Code* How expert users build a mental model of Claude and what it can reliably one-shot* Why discovering your “unknown unknowns” matters more as agents become more capable* Artifacts as persistent, generative interfaces between humans and agents* How Claude could split into a cloud-based “brain,” local or remote “hands,” and dynamic interfaces* Claude Tag, Projects, and multiplayer agents and how collaborative agent workflows could evolve* Why spending more time on the initial prompt can dramatically reduce wasted agent work* When to use low, medium, high, or max effort for different engineering tasks* Why frontier models may eventually outperform smaller models on both intelligence and token efficiency* Why implementation notes can expose decisions the model considered but chose not to make* Why Claude.md may eventually disappear — and why starting without one can sometimes be better* Claude Mods: customizing the execution loop, UI, subagents, routing, and behavior of Claude Code* Model routers, forked agents, and supervisor agents that automatically improve agent workflows* Why Claude Mods may be an early preview of “mutable software”* The bitter lesson of harness engineering and why agent architectures go out of date so quickly* How Claude Tag is becoming an organizational harness for multiplayer work* Why giving agents access to company data creates an enormous new security surface* The Exploit-Bench incident where agents discovered ways to communicate and collaborate* Why agents hacked Hugging Face for scorer code rather than benchmark answers* How agents chained sandbox and infrastructure vulnerabilities in unexpected ways* Why increasingly capable agents make traditional security assumptions harder to maintain* The argument behind Anthropic's “Pacing the Frontier” proposal* Why software engineers are increasingly doing two jobs: engineering and keeping up with AI* Constitutional classifiers, probes, and fallbacks and what interpretability looks like in production* How Auto Mode checks whether an agent's actions actually match the user's permissions* Why Thariq can see serious AI risks while still having a relatively low p(doom)Thariq Shihipar* X: https://x.com/trq212* LinkedIn: https://www.linkedin.com/in/thariqshihiparTimestamps00:00:00 Introduction00:04:12 Ask User Question and the Future of Agent Interfaces00:08:29 Artifacts, Projects, and Multiplayer Agents00:15:37 Prompting as the Core Claude Code Skill00:21:52 Context, Effort, and Smarter Model Usage00:28:10 Is Claude.md Going Away?00:32:49 Claude Mods: Customizing the Claude Code Harness00:36:35 Model Routing and the Rise of Mutable Software00:44:40 The Bitter Lesson of Harness Engineering00:50:49 Claude Tag as an Organizational Harness00:55:59 Pacing the Frontier and Autonomous Agent Security00:58:22 Agents Hack Hugging Face for the Scorer01:05:34 What Happens When Agents Need More Compute?01:10:32 AI Coding Is Changing Faster Than Engineers Can Keep Up01:17:17 Probes, Fallbacks, Interpretability, and Auto Mode01:28:32 AI Risk, p(doom), and Closing ThoughtsTranscriptIntroduction: Life at Anthropic and the Pace of ChangeSwyx [00:00:00]: We're here in the studio with our friend Thariq from Anthropic, and I guess generally the Claude Code, I-- there's, there's so much, merging of boundaries and you've been so on top of everything since you joined Anthropic. You have been early to Claude Code itself, but then also, and you've told that story in other podcasts, and you've also been talking about seeing like an agent. Most recently you did the top AIE World Tour talk, Field Guide to Fable, which obviously you guys launched Fable, so that was-- that's cheating. And mostly you most recently also launching Claude Tag, and we're also gonna be talking about Pacing the Frontier. There's a lot going on in Anthropic. I guess top of the question is, what's it like being at Anthropic when there's so much going on?Thariq Shihipar [00:00:48]: I think that It is, like. I think you can get whiplash sometimes. I think, like, going. When I joined Anthropic, I joined because of Claude Code. Like Claude Code had just come out and I was like, “This is so good.” And Opus 4 to me was like just, I could not imagine, like, how good it was? And that was, like, a real moment for me. But I was, like, trying to convince, like, my startup friends to use agentic coding, and they're like, “Oh, no, like, our engineers don't think it's good enough,” or something. And I was like, “That's insane.” and now you, like, fast-forward, 12 months, less, and, like, it's just like, yeah, the default way that everyone codes, right? And I think that, like, just having to go from, like, selling it to, like, now, teaching people how to be. make the most use of it and be more efficient and things like that is just like a big, like big change. And, yeah, I think, like, it's just hard to stay on top of everything as a human? Like, I think things happen so fast and likeSwyx [00:01:51]: You just throw more agents at it.Thariq Shihipar [00:01:52]: Yeah, like that's like the agentic stuff scales much better than the, like, human stuff where it's like, oh, like, there are three things happening right now and, like, they're all emergencies and, like, how do you, like, respond to it? Yeah.Teaching People to Use Claude CodeVibhu [00:02:05]: What do you split your time on? You do a lot of technical writing, engineering work.Thariq Shihipar [00:02:10]: Yeah, so I think that, like, when I joined the Claude Code team, I wanted to teach people how to use Claude Code and I think that, like, that has been something that, like, I thought, like, maybe I would spend a little bit of time on it or, like, I'd, like, do. I was spending some time on the agent SDK first, and I wasn't exactly sure, like, how the bitter lesson would go, when it comes to, like, harnesses, right? Like, I think sometimes we were like, “Oh, like, what's after Claude Code?”? And so initially I was like, I just wanna teach people how to use Claude Code and make it easier to use Claude Code. And I think that has just, like, as the harnesses have gotten better and better, that's like the dominant problem now is, like, how do you use the agents, right? Like, it's like such a high skill expression thing. So I do that and then I do engineering work. I give talks, but I think, like, when I'm doing engineering work, my goal is to take that feedback that we get from users and also, like, then be able to talk about, like, hey, how to use Claude Code to do engineering. So there's like a good loop there. Yeah.Swyx [00:03:07]: Yeah. I'll-- For listeners, we'll attach, the talk that you did with Sarah for the Dev Writers, meetupThariq Shihipar [00:03:13]: Oh, yeahSwyx [00:03:13]: Which we talked a little bit about, well, first you do the work and then you talk about the work.Thariq Shihipar [00:03:16]: Right.Swyx [00:03:16]: Something like that.Thariq Shihipar [00:03:17]: Yeah.Swyx [00:03:17]: It's sow and reap orThariq Shihipar [00:03:19]: Yeah, reap and. Sow and reap.Swyx [00:03:21]: Something like that. Something like that. Yeah, so, and then just to preview a little bit, we are gonna talk about the evolution of the harness. It has come a long way from just being a CLI. We're gonna talk about, Claude Mods, which is starting to leak today, because you couldn't keep it secret.Thariq Shihipar [00:03:36]: Yeah. yeah.Swyx [00:03:39]: Yeah, there's, there's a lot, there. I think you started off with, like, adding ask user question tool, which people love and hate.Thariq Shihipar [00:03:48]: Yeah.Swyx [00:03:48]: Like, I thought it was, like, very innovative, and then now I have, like, my own version. You have your Interview Me version.Thariq Shihipar [00:03:55]: Yeah.Swyx [00:03:56]: And, yeah, everyone just has, like, their own stuff. And, like, it no longer matters ‘cause now you're supposed to, write prompts that create other prompts and loops and all these things.Ask User Question and Human-Agent InteractionThariq Shihipar [00:04:05]: Sure, yeah.Swyx [00:04:06]: So what's the state of the art, today? Like, what are people. what are you, like, telling people to do today?Thariq Shihipar [00:04:12]: Yeah, ask user question was the first time that the model was good at elicitation. I think this was, like, an emergent behavior that I, like, wanted to see if the models could do. I have, like a human-computer interaction background, so I, like, did that in undergrad and grad school. And so this was like. I think it's like human-agent interaction to me, like, trying to figure out, like, how can the agent communicate with you and extract, the requirements, right? I think that, like, one of the things about, like, that's difficult as Claude Code has gone broader and broader is that everyone has, like, their own way of using it, and it's very hard to, like, change the default behavior. So for example, like, if someone asks Claude Code to do something,Thariq Shihipar [00:04:59]: Sometimes they just want them to do the work, ‘cause they're, like, maybe a very good prompter, and sometimes they want. like, are not good at prompting? And you need. like, the agent needs to, like, clarify? And so that's, like, a good split. Like, and the ask you the question tool like, splits along that side where, like, are-- do you feel like you're good enough to instruct the agent as it is, or is the agent able to, like. does the agent need to, like, pull out more requirements and, like, collaborate with you more and really understand your preferences?Thariq Shihipar [00:05:27]: I, on the whole, believe that pretty much everyone is more on the latter than the former, that they, like, have more ambiguity and they know less than they want, than they, like, think they know about the problem. but, like, it's like a interface design problem to make that easy? And so, like, if you're designing a problem, like, or if you're going through a problem, like, things like what's the schema or, like, what's the call stack and things like that are really important. like, the details in the design are important. Ideally, you want to figure out some of these, like, hard problems ahead of time before starting implementation. And yeah, that's why they call, like, unknowns, right? And so I think that this will forever be, like, a skill in agentic coding is, like, figuring out your unknowns. So, like, because even if the model is, like, super intelligent- It, like, needs to know what you want? And, like, you have preferences. like, you need to like, pull the, pull that out. and so that's, like, I think how I'm, what I'm pushing. the question then is, like, how does the agent interact with you? And I think that has been HTML, has been, like, the big way of doing that. And we've recently added artifacts, right? And artifacts, I think we've done a bad job of, like, or, like, I've done a bad job of, like, explaining how to use them fully. We have a lot of property capabilities. They have a database associated with them? And so every artifact can store and write persistent data. They can, like, feed back into Claude? And so, like, one thing that, like, people are not doing yet that I'm trying to, like, encourage is, like, this idea of a dashboard artifact. So you have, like, Claude working on a project long-term. Maybe it's like a kanban or something. it can store that kanban data in its database. Multiple Claudes can access that data via, like, the artifact MCP, and, like, that artifact can, like, talk to those Claudes as well. And so, like, the. We're building the primitives for you to be able to have this, like, generative interface via artifacts that will, like, let you surface more of that rich detail from the agents. And I think that, like, almost everything with agents right now is, like, this problem of, like, you think what you want, but you don't really know what you want, and, like, the agents need a lot of detail, and collaborating with them in the loop is really important. And so artifacts are, like, the, like, way that we're trying to evolve there. But there's a lot of work to do because it's so much more complicated than, like, a multiple-choice question? there's a lot more, like, detail in terms of, like, diagrams and code snippets and schemas or, like, whatever it is for that problem. But, like, artifacts is, like, the mo-more AGI-pilled way of, like, doing ask user question. So yeah.Artifacts as the Interface to the HarnessSwyx [00:08:15]: I think one thing that's unclear to me about these, the artifact stuff is, like, what feedback should go in through the artifact and what feedback should go through a Claude, a chat? Because the more AGI-pilled one is to just feed everything to the Claude.Thariq Shihipar [00:08:29]: I think the more AGI-pilled one is to go through the artifact. Like, and I think that, like, we imagine in the limit, I think that artifacts will be your interface into the harness? You can, like, comment on this, like, live, like, document of your plan, of the work. you can see maybe, like, multiple agents and different agents are doing this, and that artifact is built for the current work that you're doing, right? And so, like, each one has, like, slightly different. I think we're still, like, getting there from, like, an infrastructure perspective. But yeah, I think, like, on-the-fly interface for your harness is probably where things are headed.Vibhu [00:09:03]: Is there a version of it that's an abstraction from CLI or chat and you. Because right now, a lot of it is, okay, you're interfacing with Claude Code, you're having HTML given back for a mockup. It's pretty rich. There's diagrams. Artifacts are ways to connect these together. Why not just do everything that way?Separating Brain, Hands, and Surface UIThariq Shihipar [00:09:22]: Then it becomes, like, separating out, like, where is the inference happening? Where is the intelligence happening? Where is the work happening? like, I think this is like, difference between, like, or, like, some of the distinction between local and cloud, right? And so, I think right now, if you use Claude Code, it's, like, local and, like, you can spin off remote control, for example, to get some cloud behavior, or you can spin off Claude Code in the cloud, right? We're moving towards a place where instead of Claudes, like, you message a local Claude, it starts a session locally and it executes, to more like you have a Claude that you message that's in the cloud that's running. it can run, like, local, or, like, cloud sessions. This is how Claude Tag works. But, like, over time, we'll add, like, local hands as well. And so, like, local hands will be the ability for that agent to access your computer if it's online, and be able to, like, work there. And so it can spin off many different subagents. It can, like, commu- those subagents can communicate with each other, and that's where the artifact comes in to display all of that work. So you can imagine, like, the. You're separating out these things. So there's, like, the surface UI display that's an artifact and hosted somewhere and has a database and everything. There is the inference intelligence, right, that's happening on the cloud, and you don't have to worry about shutting off your computer or whatever, right? and then there's the, like, hands. Like, and it can be local, it can be in, like, a remote sandbox or wherever you need your work to be done. That's like unpackaging, like, the Claude Code experience right now where, like, right now it all happens in one place, right? So.Multiplayer Agents, Claude Tag, and ProjectsVibhu [00:11:00]: How do you see, like, the multiplayer side of that? So say teams want to work in this way. Right now it's very individual, but how do you see the future of multiplayer? Like, right now, I guess there's Claude Tag, which is a version, but.Thariq Shihipar [00:11:12]: We're launching projects. And so projects is the, like, this abstraction that's like Claude Tag, but on our Claude products, right? So you can message it and, like, it will do the Claude Tag-like stuff, like spinning off subagents. So We think with multiplayer. Like, Claude Tag is, like, a little bit more native multiplayer because it's just, like, in your Slack and the permissions are all figured out and stuff like that. But I do think multiplayer is, like, an important part of the story and, like, that will need to get tied together more. Like, you can imagine how complicated it gets when you're like, oh, you have hands, but now you have other hands in other people's computers too, and, like, you need to, like, permission them or, like, you have, like, your MCP and someone else's MCP, and how do you figure out how to use them, right? It gets, like, quite complicated. And Claude Tag does a good job of, like, sanding down all of these issues, right? So that, like, when you have, yeah, Google Docs, how does it access Google Docs, right? Like, it accesses through the shared Claude MCP, or it can access through your local credentials as well if it doesn't have access. But yeah, I think Claude Tag is our multiplayer, product, and it's really useful for these, like, things that are inherently multiplayer. Like, okay, like on-call, for example, incidents are inherently multiplayer. You want to tag Claude, you want multiple people to log in, you want it to be able to find context. I think whenever I'm, like, working on something and I want, like, privacy or security or, like, I want other people to review it's really nice to, like. I'll have a channel per project and I'll, like, at legal, for example, be like, “Hey, like, I want to ship this. Can you, like.” Like, here's. Like Claude knows everything, just chat with it. And that way legal gets precise answers, on like what exactly is shipping into the code, and I don't need to be in the loop, right? So I think like multiplayer is getting like more and more like, yeah, everyone can participate with Claude. I think Claude Tag is like that product and like projects will start off single player and will like, expand.Swyx [00:13:14]: I think there's a question about like maybe dual questions about identity and the unit of isolation.Identity, Permissions, and IsolationThariq Shihipar [00:13:20]: Yeah.Swyx [00:13:20]: Claude Tag, you specifically chose to make it its own identityThariq Shihipar [00:13:26]: Yes.Swyx [00:13:26]: Which is like, a controversial choice. There's, there's other ways to do it.Thariq Shihipar [00:13:30]: Yeah.Swyx [00:13:30]: Claude Projects probably it sounds like, if it's anything like ChatGPT Projects, it is, the isolation is that artifacts, that cloud instance, everyone's collaborating on this. It'll. It sounds like, it should be like if you're, if you're collaborating with legal on a thing, like that channel should be a project, right? Like it's not yetThariq Shihipar [00:13:50]: Yes.Swyx [00:13:50]: But it. that's the natural next step.Thariq Shihipar [00:13:53]: Yeah, like I think in Claude Tag, it's effectively. Like Claude Tag, you have to do your own arrangement. And so Claude Tag, yeah, each channel is like you can name it as you want, and I nameSwyx [00:14:04]: Yeah.Thariq Shihipar [00:14:04]: Like each featureSwyx [00:14:06]: Yeah.Thariq Shihipar [00:14:07]: As a channel.Swyx [00:14:07]: And, but I think like there is some trans- like it's unclear when there is transference, because let's say it is. if you have a coworkerThariq Shihipar [00:14:14]: Yeah.Swyx [00:14:14]: Who is tagging on all these things, yes, there is transferThariq Shihipar [00:14:16]: Yeah.Swyx [00:14:16]: Because it's the same person. but with Claude, it's unclear if it's like necessarily like, well, no, you don't know any of. you don't know about the other stuff. You should only use this stuff.Thariq Shihipar [00:14:25]: It's like the tip of the iceberg meme, right, where you can like. This is what we spend so much time onSwyx [00:14:31]: Yeah.Thariq Shihipar [00:14:31]: Is like there is like infinite surface area of like, okay, you want Claudes to. Not infinite, but like there's like surface area, a lot of like, surface area to figure out of like permissions and visibility and like how can you let Claude operate as well as you can, as safely as you can? And obviously, this is very important to us because like security for our code base is very important. And so we've put a lot of time into this. Yeah, there's so many like edge cases you can figure out where it's like, oh, like, yeah, this Claude in this channel has different permissions, but it can message another channel, and can't it exfiltrate data that way? Or like can you like. What if it uses your MCP and then messages someone else? Like there's like so much, and we've like really put a lot of work into sanding it down.Swyx [00:15:14]: Yeah. Lots of work. okay. Fable?Fable and the Meta-Skill of PromptingVibhu [00:15:18]: Fable, you wrote two good articles. you've written many good articlesThariq Shihipar [00:15:22]: Yeah.Vibhu [00:15:22]: But on, Field Guide to Fable, Building Claude Code. I'm curious from what you've seen, is there any common patterns that you see in like top users at Anthropic externally? Like what are best practices for getting the most out of Claude Code?Thariq Shihipar [00:15:37]: The like meta skill I say is like prompting is like very important? And like that. Like I think this is like not trivial to say because I think a lot of people are like, “Oh, prompting doesn't matter. It's just like I can just say a sentence and Claude will do it.” And I think prompting is really this like, this. It's like public speaking, like, or writing or something, and for a specific audience, and that audience is Claude. And you need to like build a mental model of Claude and how it thinks and how it works, right? And so that's like the most important skill in working with Claude Code is like having this mental model, right, of Claude and like what it can do well, what it can one-shot, what it can't. And so many people when you see prompting, they're just like, they're short prompts, but they have such a good mental model of Claude and of like the code base and things like that like it's effortless? But it's like high skill ceiling. So like that work of like, spending a lot of time prompting and building mental models of how, and intuition for how the agents work is really important. And then I think like the next thing is like the unknown stuff we talked about earlier, where it's like being able to find out like your, what you don't know or what you haven't written down, learning about like different things. I think as Claude can do more and more things, the likelihood of you doing something out of distribution for you and like you have low domain knowledge on is very high? And the more you can like learn the vocabulary to be able to prompt Claude, it becomes really important. And so like I think the most important unknowns are the unknown unknowns, where you're like, I just like don't even know that this exists, right? Yeah, exactly. I think that's like a illustration of like the map and the territory, right, where you're like, “Okay, this is my prompt,” and the territory is like the actual like work that the agent needs to do, right? And if you are like very precise, you can give more precise things, right? So like for example, in design, I'm not very precise. I'm not a designer, so I say like, “Give me like eight different mock-ups.” But if I was a designer, maybe I'd be like, “Oh, hey, here are some reference sites.” Like, “I want this type of font and this type of like look to it, and here's like a few different components to like visualize. Here's a Figma MC board to bring in,” like. And so you can just be so much more precise with that language. And if you're not a designer, you just need to like try and learn the language or learn the unknown unknowns. And this is true of like everything, I think. Like the more, like you can work with Claude to learn like how things work, the better your prompting will be. I think another good example of this is like game design, like where a lot of people are like, “Oh, like I can vibe code a game now.” And they're like, “It's not fun.” And like it's just like the thing about game design is like every one of these choices has like a lot ofTaste, Domain Knowledge, and Learning the VocabularySwyx [00:18:25]: Variations.Thariq Shihipar [00:18:25]: A lot of like craft to them. So it's like, oh, okay, like when you're making a flying game, the feel of the plane and the like, way it responds to your controls has a lot of like. Like, a game designer would spend like days on that. Do? and likeSwyx [00:18:44]: To me, that's what taste is, right?Swyx [00:18:45]: Like it is like from the possible space of one thousand mathematically valid answersThariq Shihipar [00:18:49]: Yeah.Swyx [00:18:49]: Here's the one that is the humans will like.Thariq Shihipar [00:18:51]: Yes. Yeah.Thariq Shihipar [00:18:52]: I think with taste, I'm like torn on this word ‘cause I think you're right, but everyone has different definitions, and it sounds kind, sounds like low skill or like elitist almost, where you're like, oh, like there are certain people with taste?Swyx [00:19:06]: It's like taste is what I call taste.Thariq Shihipar [00:19:07]: Yeah, exactly.Swyx [00:19:08]: And it's like these guys don't have taste.Thariq Shihipar [00:19:09]: Yeah, exactly. Oh, like an engineer doesn't have taste. Like I, the like founder, have taste.Thariq Shihipar [00:19:14]: ? And I think that's not true. Like I think like the engineers have a lot of taste for these particular like problems? And I think everyone has taste for particular problems. I think like Jason Liu, like say like in order to, yeah, have taste, you have to eat?Thariq Shihipar [00:19:32]: And I really like that, where it's like, okay, you have to like do a lot of things. You have to like iterate and figure out what you want, what you like, and, like build that like domainSwyx [00:19:41]: YesThariq Shihipar [00:19:41]: Domain vocabulary. And then when you're prompting, you're like synthesizing all of that for a product.Swyx [00:19:46]: Isn't it annoying when someone else says it better than you?Swyx [00:19:48]: It's just like, f**k, I have to quote this guy forever.Vibhu [00:19:51]: Having to quote Jason Liu forever.Vibhu [00:19:53]: He's gonna love this.Thariq Shihipar [00:19:55]: So I get prompts, more than that.Vibhu [00:19:57]: And sometimes it's not even that. Sometimes it's just intuitive, right? Like you don't realize you even want something till a model puts it out, and you're like, “Oh, this just feels immediately better,” right?Voice Prompting and Information DensityThariq Shihipar [00:20:07]: Yeah, exactly.Swyx [00:20:09]: One thing I go back and forth on is I feel like the way I prompt half the time, let's say I use voice.Swyx [00:20:16]: Did I say voice? Other people have voice. that is the opposite. That is just like me rambling for like two minutes Pressing down the function key and then let go, and then like hopefully it figures it out. And oftentimes it does.Thariq Shihipar [00:20:26]: Yeah.Swyx [00:20:26]: But it's not as thoughtful as like a structured prompt with like Well-run communication as though it's a PRD or a memo. Is that in line with how people do this? There's like bimodal prompting where there's some prompts where you spend a lot of time upfront and other prompts you just dash it off?Thariq Shihipar [00:20:43]: I don't think the voice is necessarily low. Like I think it's like more like how much information is in the prompt. like the model can. Like you can and like add some sentencesSwyx [00:20:53]: RightThariq Shihipar [00:20:53]: And be like, “Oh, like I changed my mind,” like in the middle of the prompt, and it will be able to follow that perfectly? So I think the like actual format of the text is less important, but then like the ability to. Like how much information is in it, right? And I think for voice, a lot of times, going back to like human-agent interaction and like for a lot of people, it's just way easier to talk than to like type? and I. If that gets more information out of you, like that's better.Vibhu [00:21:21]: At some level, it feels like just giving the model as much contextThariq Shihipar [00:21:24]: YesVibhu [00:21:24]: Over prompting before you kick off is a best practice. I don't know. A lot of the times, like when I was first trying out Fable, I spend a solid 30 minutes like really crafting a long prompt. This, I think, is a response of models running for longer and longer, right? It's still a little difficult to nudge them as they're in like, in the loop, but I just like intuitively spend more time kicking off that first prompt and working with it a lot.Spend More Upfront, Iterate LessThariq Shihipar [00:21:52]: My personal opinion is that if I was a software engineer, if I was like, just running my own startup, for example, I think I would mostly fit, stick to a max 20x? like maybe verification and so code review are like separate things. But I think like what I see a lot of times is people hit rate limits when they're doing this like, oh, like it did a lot of work and you're like, “Oh, I don't like this.” Like, “Can you like undo this and redo it?” And then you're like iterating on this like thing that the model could have done if you had like spent more upfront time or given it better context? And instead it's like you're like, “Nope, don't like that design. Try this.” Or like, “You messed this up,” or something like that. And then that just eats up so much more of like, your usage. And so that's like, I think maybe like a key like tip both for like efficiency as well, right? And yeah, I think like context, and not just like context on like what the goal is good, right? Like are you building a prototype or is it like a production thing? Like where can you spend compute or when, where can you not spend compute? Like I think you have to give the model permission or like not permission to do things sometimes where, like it doesn't know intuitively how much you want to spend on this task, right? And you can use effort for this. So I did-- I'm working on a blog post about that where it's like, if you want. For like we see that effort scales with the complexity of the task. So for security, effort gets like way more results. Like high effort versus like low effort gets, like changes the evals a lot. But for software engineering, it doesn't change it a huge amount because effort is mostly spent on the verification and the like edge case testing and things like that. And so like being able to like give the model that guidance of like, “Hey, this problem is something that I think I want you to spend a lot of time verifying and edge case testing,”?Effort, Model Choice, and VerificationVibhu [00:23:43]: How about model in the mix? So, there's Opus and Fable with effort.Thariq Shihipar [00:23:47]: Yeah.Vibhu [00:23:48]: There's also Haiku in there.Thariq Shihipar [00:23:49]: Yeah. It's not quite true yet, but it's very close where I think the frontier models will be Pareto dominant over like almost everything. like maybe. And sometimes I think Opus might be Pareto dominant. Do? Like I think depending on like how things, like shake out if it's like a newer version of Opus. But I think that like increasingly it's just going to be like the smart model is going to be able to like do the simple task for less tokens than the like the other models because of verification. With verification, in the limit, your model doesn't need to verify, right? If it's a perfect model, it just does the work once and it's like, okay, like you, I did it? And increasingly with Fable, I'm like, I'm like, “Dude, you don't need to spin up Chromium and screenshot all of these things.” Like I see it. Like you did it, right? And so a lot of the. At higher effort, you spend more of those tokens verifying. But if you're working on simpler problems, and a lot of software engineering is like well, like in Fable, like low and medium stability, it can spend less tokens verifying. And as the models get smarter and smarter, they will just be able to like, “All right, done.”? Like, I can run the lint for sanity's sake, but, like, I, like, know it lints? Like, you don't even need to do that. And that will be so much more token efficient than, like, the smaller models. Yeah.Swyx [00:25:15]: Is there a good, practice on our side that we can use to see if we're using too much effort? Like, I freakingThariq Shihipar [00:25:23]: YeahSwyx [00:25:23]: Hate wasting time on that stuff.Thariq Shihipar [00:25:24]: Yeah. I know what you mean. I think, like, so in this blog post, my rough distribution is, like, code review and security should be, like, high or max and, like, software engineeringSwyx [00:25:37]: You said recommend mix settings per domain.Thariq Shihipar [00:25:37]: Yeah. I think, like, if you're doing, like, UI or something like that, like low and medium, I think is you're building, like, an API and you want to make sure, like, you cover enough edge cases? And so I think building, like I said, that mental model of, like, how things work across these distributions is, like, yeah, part of the job.Implementation Notes and Decision LogsVibhu [00:25:56]: This is more intuition-driven or eval? Because I'm guessing this would change as you go.Swyx [00:26:00]: He has evals.Thariq Shihipar [00:26:01]: Yeah. So what I did in the blog post is I go over all of the terminal bench evals. So there are, like, 70 problems and I'm show that, like, okay, like, in the security problems it does more. and then I also, like, look at some of the transcripts just in terms of, like, how-- what does it answer, what does it forget or something. And a lot of times, this is another prompting tip I have, is, like, asking it to make decision notes or implementation notes because, in every eval problem that it faces, it thinks about the correct solution, and decides not to do it. it's like, oh, like, here is the answer. What if I did this? And then it's like, oh, probably not? and then keeps going. And this is, like, the majority of the failures, at, like, a higher max level. It's very rare that the model just doesn't know how to do something. If you just have these implementation notes, then you can review and you can be like, “Oh, I want you to do this thing that you didn't do.” The models are getting better at surfacing that overall. Like, I see in the transcripts of Fable 5.1, like, when it does this output, it will call out its decision-making as well. but making this more explicit in the harness is better. And now we're, allowing ways of you modifying the harness so you can, like, add someVibhu [00:27:23]: Ooh.Thariq Shihipar [00:27:24]: Calculate with there. Yeah.Swyx [00:27:25]: Yeah. So I do wanna call out two things that you mentioned that I think exist outside of prompting. One is like, let's, let's call it the prompt that is so important that it shouldn't be in a prompt. It is in Claude.md or Agents.mdThariq Shihipar [00:27:38]: YeahSwyx [00:27:38]: Which is like goals, right? Like your situation, your goals, the things that you want, the thing. and then second of all is the decision log or the experiment log or whatever log of traces that you might want to survive the current session to do those things. Those are, like, externalities that there's no standard. There's no-- It's not like skills. It's not like MCP. There's no standard. It's, it's just like it's a markdown file. first of all, is that right? Is Claude.md going away? You have a documented dislike of, Agents.md, but you're gonna do it?Claude.md, Agents.md, and Model-Specific InstructionsThariq Shihipar [00:28:10]: Yeah. Okay. So Agents.md, yeah, like, we're, we're gonna do it. I think it's just, like, different models are very different from each other? But I realize that it's, like, such a pain to, like, maintain different ones? And yeah, like, as the models get better and better, the floor of how they accomplish the simpler task is better. And so I do think in the limit, Claude.md goes away, and maybe not even, like, that far. Like, I think, like, I think that right now it might be better to start a new project without a Claude.md.Swyx [00:28:44]: Yes.Thariq Shihipar [00:28:44]: I think that, like, maybe if you see very repeated failure modes, you add them to your Claude.md. The really tough thing is that this changes per model. And so, like, if you've added a bunch of failure modes or, like evenSwyx [00:28:57]: So you need Fable MD, you need Opus MD.Thariq Shihipar [00:28:59]: Or well, even Fable 5.1 versus Fable 5.Swyx [00:29:03]: Yeah.Thariq Shihipar [00:29:03]: Like, it is annoying. Like, I'm not like,Swyx [00:29:05]: YeahThariq Shihipar [00:29:05]: Like, we don't, like, do this on purpose? It's just, like, how the models work, right? And so, like, maybe, like, Fable 5 had this, like, failure mode that Fable 5.1 doesn't. And if you keep this context, this running log of a bunch of different failure modes, they will probably over constrain Claude? And so this is like. we just added evals plugins for skills.Swyx [00:29:28]: Yeah.Thariq Shihipar [00:29:29]: And so now you can eval if a skill is better. I think Daisy on our team did this. And so, yeah, this is like we're trying to work on this. We know it's, like, you still have to spend tokens on it and, like, it's not, it's not perfect, but it's, like, we're trying to help out with this problem.Swyx [00:29:44]: And so, and as far as prompting goes, the one tip I wanna offer is, something I have told people a lot is sufficiently advanced prompting is indistinguishable from sufficiently advanced executive communication. So I've referred to-- This is an executive comms workshop from Heavybit that is the best I've ever seen in my career. And they teach this thing called the SCQA model. Just Google it. It's a, it's a thing. Like, people have done prompting for decades. It's just called executive communication. It's like when one person has to communicate to thousands of people down the org chart, this is what you do. so situation, complication, question and answer, is how you write the memo. but obviously sometimes you don't have the answer, but you can at least list out the SC and Q, and then they have some examples in there. So just leaving breadcrumbs for people if they want to explore.Underrated Prompting Patterns and ELI5Vibhu [00:30:31]: Before we move on, I wanna ask you, any other underrated tips, ways people could get a lot of value from Claude Code that they're not using?Thariq Shihipar [00:30:41]: Yeah, I think a lot of them are in the, this unknowns, like, doc. Like, I give a bunch of example prompts, like, using it for brainstorming, using it to quiz you after. we added this, like, explain it like I'm five skill which is a very short prompt. And it doesn't even say explain it like I'm five. It's like the key word of this prompt is big pictures, few words. like, that's like the main thing. And it is shockingly good? Like, you, like, I think I tweeted about this and it's like /eli5, and, like, you can install it as a plug-in. But yeah, it's, like, way better at just cutting through the BS and being like, yeah, exactly right here. So the diagrams are, like, quite clear. I think one of the things that is true with artifacts is, like, they put too much text in and people are not reading the artifacts? And so, like, this simplifies it a lot more. And, yeah, this came out of, like, just people at Anthropic, like, going through very complicated incidents and being like, “What is happening?”? So, this one I think is great, yeah.Swyx [00:31:47]: My version of this is the, it's like test your understanding. Give you a few choices and then, like, if you get it wrong, you have a mismatch between what you think is happening versus what's happening.Thariq Shihipar [00:31:58]: Yeah. I think this is one of those things that everyone loves talking about, and then very few people really do. Like, I thinkSwyx [00:32:05]: Really helpful.Thariq Shihipar [00:32:07]: Yeah. But most people just don't want to get quizzed about something? Unfortunately, I think this is one of the, like, things that we need to, like.Swyx [00:32:16]: What's the opposite of ask you the question or ask you the question before the thing?Thariq Shihipar [00:32:19]: Yeah.Swyx [00:32:19]: This is after the thing.Thariq Shihipar [00:32:20]: Exactly. Yeah.Vibhu [00:32:21]: It's a good way to stay grounded of, like, do you even know what you're doing, right? The worst case is when people send you slop and they haven't understood what they're asking for or what the output is, and it's like, “Dude, I don't wanna read this. Do you even know what it is?” So, you make it a rule for yourself that before you send stuff, you should at least know what's implemented.Claude Mods: Customizing the HarnessThariq Shihipar [00:32:41]: Yes, but so you could make this a mod and you could build your own mod to, like, make sure you test it. So yeah, you can do that.Swyx [00:32:49]: All right. Let's get right into it. What is Claude Mod, and what is this diagram showing?Thariq Shihipar [00:32:54]: Yeah. Okay, so Claude Mods is you can customize the entire Claude Code harness, and we're going to. If you have requests, we will, like, let you, like, please let us know. We'll add more and more. This works for CLI, it works for desktop. maybe it will work for Claude Tag in the future. I don't know. Like, we're trying to make this very extensible. You can see this reference sheet. I don't want people to get overwhelmed by it? At a high level, you can customize both the execution of the harness, and the UI of the harness. And so, like, you say on that Tetris example from Boris, that's like customizing the UI, right? Like showing, like, Tetris in the game.Thariq Shihipar [00:33:35]: But, like, let's say that you wanted to do this thing where you had. you tested your assumptions or, like, tested your understanding after every project, right? What you would do is you would ask Claude to make this plug-in. It would spin a classifier after every prompt. And so, like, at the end of each turn, you would spin off a sub-agent or, like, a forked agent. A forked agent is, like, maintains the prompt cache, right? So it's like a, like one of those unintuitive things where you can fork and do, like, a little request, and it'll be very cheap because the entire prompt cache is, like, done. And so you can be like, “Has this task been completed?” likeSwyx [00:34:18]: This is how you do BTW and all those.Thariq Shihipar [00:34:20]: Yeah. The underlying forked agent, yes. But so you can, in the f-fork sub-agent, you can say, like, “Has this task been completed? If so, return true.” And then in your hook, or in your, like, plug-in mod, or sorry, like, in the sub-agent probably, you would say, like, “If true, give me a quiz.” give me questions and answers, and then, like, in a JSON format, and then you'd parse it, and then you display above the prompt input, this list of questions, right? And so this is something that's, like, slightly token-intensive because, like, you have to do it after every end of the assistant turn. But it's, like, a lightweight classification, and then you can, like, get this quiz, and then you'll see, like, Claude will always do it for you. You don't need to remember to do it. There are lots of these, like, tips that we've talked about, right, where it's like, oh, implementation notes. You can also add a tool for implementation notes now. And so, like, this tool that I'm adding is, like, register, like, I think assumption is what I'm calling it, but, like, maybe I'll change it around. And this is a mod. And so, like, you give it a register assumption tool, and then it will keep a list. It'll. Every time it does it'll keep a, like, add to the list, and then at the end it will display those assumptions? Another mod I'm working on is a model router. And so, like, internal, like, Claude model routing, right? So it's. This is, I want to say the reason we don't do model routing by default is, like, it's a hard problem? And likeForked Agents, Assumption Tracking, and Model RoutingSwyx [00:35:51]: You will get it wrong.Thariq Shihipar [00:35:52]: Yeah, you, like, yeah, you will, like, accidentally use, like, Fable for a hard problem or Sonnet forSwyx [00:35:57]: Yeah, if you have auto approve, but you don't have auto mode.Thariq Shihipar [00:36:01]: Well, you will have auto. Like, you don't have, like, auto routing or something.Vibhu [00:36:04]: You don't have auto mode for model picker.Thariq Shihipar [00:36:06]: Yeah, exactly. SoVibhu [00:36:07]: I'm getting the rough question of, like, how much do you open this up and how much do people have to think about this? Like, when you talk about prompt caching and building a router, it seems like you could easily build a mod that routes per query, and I'm just killing my plan very fast, right? I guess my question is more so, like, what is, like, a product talk like this look like, right? Who is it for? Is it for power users? Is it everyone should be able to go throughSwyx [00:36:33]: Oh, definitely power users, right?Thariq Shihipar [00:36:35]: Yeah, I think it is power users, but, like, the nature of Claude Code is that so many people are power users? Because it's easy to share things, like you can. Like, one person can make a good model router thing that doesn't break prompt cache all the time, and then you can, like, compose them. Another cool thing about the plug-ins is that they can hook into and compose with each other. And so I have, like, a mod that will, like, create a mode selector at the top, and any plug-ins can register to be a mode. And so, like, the auto router can be a mode, right? Or, like, you can have a mode that's, like, artifact mode, where it's like it primarily talks to you in artifacts. like, you can toggle between plan mode? And so, like, you can create more and more of these modes. But the ability to create modes is in it itself a mod? And so there's a lot of richness here, but we do want to make it fairly easy. We want to be-- make it so that you can just, like, install someone else's. You can ta-- you can chat with Claude and, we'll, like, make sure that it understands the nuances of things like prompt caching and stuff, so it can, like, warn you. This is, like, not extremely complicated behavior for Claude, I think, but we should have just a good skill on how to make mods. and yeah, we'll see how we go. But I do think that this is, like, a preview of, like, mutable software, and, like, how, like, generative software, just like you can customize safely. If enabled, you could customize any piece of software. And I think that more and more apps ideally do something like this?Power Users, Modes, and Mutable SoftwareSwyx [00:38:13]: And by the way, you, we have, you have another cool tweet about how, there's the infinite money button, which is like make your SaaS, consumable by agents. I think mutable software is interesting and, other people have also tried to do it. I think the hurdle comes when you can do everything, then people, users get, tend to get confused. So usually the stuff that works is just like one opinionated flow. This is in the side of less opinionation. It's just like, well, more power to power users. And I think probably unlocked by AI, where, like, you can just prompt for whatever the thing is.Thariq Shihipar [00:38:47]: Yeah, or there can be a skill that gives the opinions?Mods vs. Hooks vs. ArtifactsSwyx [00:38:50]: Yeah.Thariq Shihipar [00:38:50]: And then, yeah.Swyx [00:38:51]: So knowing a little bit about, like, TypeScript and build systems and all these things, the closest-- I'm very curious that the team who worked on this, if, I don't know how close you were to them, if they drew any inspiration from build systems like Babel, Webpack, all these, like, old school things. Because it sounds very similar, like the plug-in ecosystem of those things where they can compose with each other.Thariq Shihipar [00:39:11]: Yeah, I'm not deep in the technical details, but I do know it was a collaboration with someone on the Bun team and someone on the Claude Code team.Swyx [00:39:17]: Yeah, it's a build system mecca.Thariq Shihipar [00:39:19]: Yeah. Exactly. It's, it's very exciting. But yeah, like, agents can just do this very complicated like, extensibility into your software now. And so, yeah, like, another reason to, like. If you run a startup, like, you can just prompt Claude and be like, “Hey, like, could we make an extension system? Like, what would that look like?”?Swyx [00:39:37]: Yeah.Swyx [00:39:38]: And I just really wonder, like, you had hooks in the past and plug-ins, all these things. So what specifically will mods be able to do that those things could not do?Thariq Shihipar [00:39:47]: Internally, we were originally calling this function hooks. And so, like, that's, like, gives you a little bit of an idea where, like, hooks register a, like an event to happen and then, like, a script to call. And this inside of the, like, TypeScript runtime is running things. And so, like, you get some benefits of just, like, it has a bunch of things in the Scope with, like, for example, like how many turns is in this conversation, right? Like, how many tokens have been used? Like, et cetera. Like, what are the messages? Things like that. So it has a bunch of messages that can be used. And then it's just, like, a lot more hooks. So we have, like, or a lot of, lot more, like, things you can register on. And then you can do because of the. because it's all happening in process, you can, spawn sub-agents, with four contests and contexts and stuff. And, like, that will return. You can parse the results of those. You can use structured output to like, return them. and then you can modify the UI, which you can never do in hooks. So, yeah.Swyx [00:40:50]: Yeah. Yeah. So modify UI, this is why you showed the Tetris example. Does it also ex-extend to artifacts? I assume it does.Thariq Shihipar [00:40:57]: You-- Like, artifacts are like a different way of customizing it. like, you can definitely. One of the mods I'm working on is, like, this dashboard mod, which will, like, prompt Claude to maintain a dashboard, that's an artifact. But they're like, slightly orthogonal, or not orthogonal. They compose with each other in different ways. Like, mods are, like, a little bit more, like, in your Claude Code harness, changing the agent loop? And, like, the UI is, like, an added benefit. and then artifacts are just like you want to, see things at a high level, very inter- highly interactive. like, the affordances can be a lot bigger than, like a TUI or even in our desktop.Next Steps, Supervisors, and Persistent GuidanceVibhu [00:41:40]: I'm guessing you'll have a good blog post on the differences, because right now you can also, make a loop that outputs to an artifact that's an interactive dashboard, but you can also do it with a mod. There's just some thinking about making a hacking on a harness when we don't know much about the harness, right?Thariq Shihipar [00:42:00]: Well, something I'm excited about with mods is, like, there's so much things with Claude Code that you just have to remember? You're like, “Oh, like, let me do this, and then let me call the dashboard skill that does the loop,” and things like that. And, or like, “Let me test my assumptions afterwards.” And I think, like, if you do all of these things using these little classifiers and stuff, and you're like, “These are the things I care about. This is what I want to do,” you can, like. You don't have to remember as much. One more, like, mod I'm working on is a next steps mod thatSwyx [00:42:28]: I have-- I was gonna say, I have a next step skill. I always run next steps.Thariq Shihipar [00:42:32]: And does it have access to your skills? Like, this is one of those things where I'm like.Swyx [00:42:37]: I think so.Thariq Shihipar [00:42:38]: Okay. Yeah, probablyVibhu [00:42:39]: Do skills need specific access toThariq Shihipar [00:42:41]: Well, I think there'sSwyx [00:42:41]: Don't they always haveThariq Shihipar [00:42:42]: I think there's, like, specific prompting, I guess, to, like, know your skills. Like I think Claude forgets them sometimes throughout, like, the thing. But anyways, the idea of, like, yeah, next steps that also are like, “Oh, hey, this has happened. Use the explain skill to explain to you what happened because this seems, like, quite complex,”? Or, like, yeah, “Use your unknown skill. It looks like you are, like, asking the model to, like, iterate on these small changes. It seems like you could prompt better.” like, “What if you did this?” Right? So, I think, yeah, like spending more compute there. Yeah.Swyx [00:43:20]: And it should always come out as multiple choice. we have, I haveVibhu [00:43:23]: We have his skill.Swyx [00:43:24]: My next step skill is like this.Thariq Shihipar [00:43:26]: Okay, perfect. Yeah.Swyx [00:43:27]: You can steal it.Thariq Shihipar [00:43:28]: Yeah.Swyx [00:43:29]: Like, but like, for me, it's all-- I think models really always need to be reminded, what are you trying to do here?Thariq Shihipar [00:43:35]: Yeah.Swyx [00:43:35]: Look at the whole transcript and go like, oh, was this original goal? Did your solution solve it? Were you lazy? If you're lazy, maybe there's a reason. Maybe you needed approval from me. Maybe you needed, there's two things you wanna suggest. So it's, it's a little bit like the modification of the ask user question or interview me skill. so it's next steps.Thariq Shihipar [00:43:55]: Yeah, exactly. And again, the benefit of doing it with mods is you can do it as a fork sub-agent, and so it doesn't remain in the context afterwards. So you have this, like, idea of like, okay, the model is doing its execution and you have this almost like supervisor, like, that is like making sure that you can do like the next steps well. So yeah.Swyx [00:44:15]: Yes. I do have two panels and like I often try to have a supervisor thing, keep the high-level context and then the implementationThariq Shihipar [00:44:21]: YeahSwyx [00:44:22]: Detail in another agent.Vibhu [00:44:23]: I feel like a lot of this abstracts away as models change? The, like, half an hour ago you said bitter lesson of harness engineeringThe Bitter Lesson of Harness EngineeringThariq Shihipar [00:44:31]: YeahVibhu [00:44:31]: And we're on the other extreme right now, I feel.Swyx [00:44:33]: Well, so yeah, exactly. If everything's customizable, what is Claude Code, right?Thariq Shihipar [00:44:37]: Yeah.Swyx [00:44:37]: And which I talked to you about last night.Thariq Shihipar [00:44:40]: Yeah, I think that this is. I think the bitter lesson is unintuitive? In terms of like. Also, like we're misusing a little bit of the bitter lesson here where it's like, it's more about like scaling and compute and stuff. But like, I think there is something where it's just like. I think I use it as an approximation here to say that harnesses go out of date very quickly? And like how, but how they change is unintuitive? And so like the big obvious example is like from chat to like agents where you had to give them entirely new tools, right? But like, I think this new version of like, oh, it can modify its own harness, right? This is like, an own harness loop is like a way of using its capabilities, right? Or like it can build an artifact. And like, I think the way I think about it is like the models have more and more intelligence, and they're like so much more intelligent now than like the average software engineering task. Like, you look at the like terminal bench ones and they're like solve like the Jacobian conjecture. Not really, but like, it's like they're, they're quite complex. Like, I would not have been able to do this really as a software engineer.Swyx [00:45:42]: And you said TB4 or TB2?Thariq Shihipar [00:45:43]: TB3. TB3.Swyx [00:45:44]: TB3.Thariq Shihipar [00:45:44]: Yeah. They're quite complex, but the goal is still to deliver user value, right? And like you said, there's like this infinite space of things to do. And so the ways like you spend compute are to keep the user in the loop and make sure that like you're getting to the right decision in the end of the day and like the right output. And artifacts and mods are this way of like spending that intelligence. and I think that's like, yeah, the next step. And so, yeah, I think Claude Code is like, has the core things of agent loop which are, have gotten more complicated. It's like, it needs a sandbox to operate safely. It needs auto mode to like make sure like the permissionsVibhu [00:46:21]: Approvals.Thariq Shihipar [00:46:21]: Yeah, approvals. it needs computer use and MCPs and like all of these like ways of accessing your data, and it needs web search and web fetch. And like, so the-- as the models can do more and more, the core harness has to be like quite complex and very secure. But then like how you interact with it can change quite a lot.Vibhu [00:46:42]: What other harness engineering best practices have you, from the Claude Code team itself? I feel like, there was a phase of plan mode, which is not as used. We now have auto mode. at a point you cut the majority of the system prompt, you got rid of examples. What other best practices are there for harness engineering?Core Harness Primitives and Managed AgentsThariq Shihipar [00:47:02]: I think there is like a forking path where at some point, eventually, yes, the model will just be able to like vibe code the exact version of Claude Code, even describing all this complexity that I've talked about, right? Like auto mode and computer use and stuff. Eventually, the models will just be able to do that in one shot. But I think they can one shot simpler harnesses? And so like, I think some people. Sometimes you don't need this full, like if you don't need computer use or like all this like more complicated stuff. I think before we, you had to use things like the agent SDK, which was like Claude Code wrapped, in order to like. And I would, like suggest people do that because there was so much complexity into building a harness. And now as that's got more abstracted, we have like, Claude managed agents, which lets you have that complexity, but still like, right, like a very bare bones like harness that's scoped to your task. Yeah, I think there's like this barbell effect where like for like very complex, for like coding task and like these like complex things, you should use our harness. And then for like a lot of like simpler or like, more domain-specific things, you can build your own harness because Claude has gotten better at building harnesses, and we have these harness primitives like managed agents. So yeah.Swyx [00:48:18]: Yeah. Is there a general progression? Let's say chapter one was ultra code dynamic workflows, then chapter two was cloud mods. Where is this going?Swyx [00:48:29]: Where you're, you're, you can customize the thing on demand.Thariq Shihipar [00:48:36]: Yeah. I do think that like this evolution of projects and like artifacts and splitting out like brain and hands and, surfaces is like where things are going more. And like, I think it's like not all quite there. partially it's like a, it's just like more token expensive? And like, I think likeProjects, Local Hands, and Cloud-to-Local HandoffsSwyx [00:48:59]: Why would projects be more token expensive? I understand mods would be slightly more token expensive. No, not something I'm worried about.Thariq Shihipar [00:49:06]: Yeah.Swyx [00:49:06]: But whatThariq Shihipar [00:49:07]: You're asking Claude to do. It's like creating loops. Like you're asking Claude to do more work for you. And so like it's managing the sub-agents and reviewing it, versus where you would be doing that work normally. And so that's like gonna be a little bit more intensive, like. Outputting to an artifact is gonna be a little bit more token-intensive than, like, outputting normally. I don't think it's too much more, but like, it's like combining all of these together well, like I think we're, we're still working on like local hands and things like that, I think is like, yeah, where things are headed, yeah.Swyx [00:49:37]: Yeah. Claude and local is, handoff is very interesting. I was thinking about this as reverse cloud remote.Thariq Shihipar [00:49:44]: Yeah.Swyx [00:49:45]: Because it's like remote, it's you're handing off to cloud, but here the cloud is handing off to local, right?Thariq Shihipar [00:49:49]: Yeah, exactly. Yeah, remote control is also another way of doing it. And I do want to say this is like how I think about it and like what the things that I'm most excited about this, but like there are, just like lots of different ways to work with Claude. Like some people use remote control a lot, some people use Claude Code on the web a lot. Obviously, like at Anthropic, we use Claude Tag a lot, and like what's great about Claude Tag is we set up all this stuff for our own execution. And I do think if you're an enterprise, that's still the best way to go. but if you're like an individual, Projects is this way of like, getting some of that like niceness of Tag, which has like that like supervising agent and yeah, adding artifacts and stuff, but like without having that whole like admin setup. And so there will be many ways to use Claude, I think. I think it's probably not just one like single.Claude Tag as an Organizational HarnessSwyx [00:50:36]: You had the multiplayer thing here. Let's, let's just check in on Claude Tag. it's been about two-plus months. Lots of, public, adoption and trying it out.Thariq Shihipar [00:50:45]: Yeah.Swyx [00:50:45]: What's new? What's, what have you found since the launch?Thariq Shihipar [00:50:49]: Like, Claude Tag is how we useSwyx [00:50:51]: It's like 80% of your

    LINUX Unplugged
    686: Stop Desktop Slop

    LINUX Unplugged

    Play Episode Listen Later Sep 28, 2026 92:15 Transcription Available


    KDE's fight over AI generated code explodes at exactly the moment Plasma is making some of its biggest changes in years.Sponsored By:Jupiter Party Annual Membership: Put your support on automatic with our annual plan, and get one month of membership for free!Managed Nebula: Meet Managed Nebula from Defined Networking. A decentralized VPN built on the open-source Nebula platform that we love.Support LINUX UnpluggedLinks:Web Boost — Send us a boost via sats or USD

    All TWiT.tv Shows (MP3)
    Untitled Linux Show 272: Typing with Mittens

    All TWiT.tv Shows (MP3)

    Play Episode Listen Later Sep 28, 2026 111:14 Transcription Available


    We're back again, a little jet lagged, and ready for fun! Up first is a mind-bending experiment to run VMs through a kernel reboot, followed by Ubuntu's new approach to kernel updates. There's a clever security flaw in inotify, another group working on Apple hardware Linux, and some divisive news about the Raspberry Pi. KDE's AI discussion blew up, Stema has a new streaming codec, and VLC and SPEC have updates worth talking about. For tips, we have the Starling Desktop, zoxide to spice up your cd life, lltag for managing audio metadata, and pstree for a tree view of processes. The show notes are at https://bit.ly/4rCvsWz and have a groovy week! Host: Jonathan Bennett Co-Hosts: Rob Campbell, Jeff Massie, and Ken McDonald Download or subscribe to Untitled Linux Show at https://twit.tv/shows/untitled-linux-show Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Club TWiT members can discuss this episode and leave feedback in the Club TWiT Discord. Sponsor: trustedtech.team/twit

    The Linux Cast
    Episode 244: Is Debian AI Now?

    The Linux Cast

    Play Episode Listen Later Sep 28, 2026 70:13


    The boys are back! Tonight we talk about AI use in FOSS projects ==== Special Thanks to Our Patrons! ==== https://thelinuxcast.org/patrons/ ===== Follow us

    All TWiT.tv Shows (Video LO)
    Untitled Linux Show 272: Typing with Mittens

    All TWiT.tv Shows (Video LO)

    Play Episode Listen Later Sep 28, 2026 111:13 Transcription Available


    We're back again, a little jet lagged, and ready for fun! Up first is a mind-bending experiment to run VMs through a kernel reboot, followed by Ubuntu's new approach to kernel updates. There's a clever security flaw in inotify, another group working on Apple hardware Linux, and some divisive news about the Raspberry Pi. KDE's AI discussion blew up, Stema has a new streaming codec, and VLC and SPEC have updates worth talking about. For tips, we have the Starling Desktop, zoxide to spice up your cd life, lltag for managing audio metadata, and pstree for a tree view of processes. The show notes are at https://bit.ly/4rCvsWz and have a groovy week! Host: Jonathan Bennett Co-Hosts: Rob Campbell, Jeff Massie, and Ken McDonald Download or subscribe to Untitled Linux Show at https://twit.tv/shows/untitled-linux-show Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Club TWiT members can discuss this episode and leave feedback in the Club TWiT Discord. Sponsor: trustedtech.team/twit

    uGeek - Tecnología, Android, Linux, Servidores y mucho más...

    Enlace original En este podcast hablo de cómo combinar el servidor WebDAV con MQTT Y crear algo parecido a lo que serían los servicios o aplicaciones con multiusuarios.

    Brad & Will Made a Tech Pod.
    358: It's a DE-rail, Not a DB-rail

    Brad & Will Made a Tech Pod.

    Play Episode Listen Later Sep 27, 2026 72:28


    It turns out we haven't answered all the questions that have ever been asked yet, so we figured we'd take a few more from you, the listener, this week, covering topics like cell phone signal strength and the way your carrier prioritizes your traffic, ARM vs. x86 in the context of efficiency, how much vibe-coding is going on in OSs and drivers right now, open source wearable devices, how to store a Nintendo Switch for the very long term, the long and mysterious legacy of the D-subminiature standard, and more.Sparkfun's D-sub blog post: https://news.sparkfun.com/14298 Support the Pod! Contribute to the Tech Pod Patreon and get access to our booming Discord, a monthly bonus episode, your name in the credits, and other great benefits! You can support the show at: https://patreon.com/techpod

    Nintendo Switch Craft
    Mini Steam Decks?

    Nintendo Switch Craft

    Play Episode Listen Later Sep 27, 2026 124:54


    ▶️Support on Patreon - https://www.patreon.com/NerdNestPodcast

    BIT-BUY-BIT's podcast
    Answering Your Questions | FREEDOM TECH FRIDAY 58

    BIT-BUY-BIT's podcast

    Play Episode Listen Later Sep 26, 2026 60:40 Transcription Available


    A weekly live show covering all things Freedom Tech with Max, Q and Seth.In this episode we answer your listener questions. Covering everything from local vs hosted AI, Bitcoin multisig solutions, private chat options and even who cuts Q's hair! HELP GET SAMOURAI A PARDONSIGN THE PETITION ----> https://www.change.org/p/stand-up-for-freedom-pardon-the-innocent-coders-jailed-for-building-privacy-tools DONATE TO THE FAMILIES w/ USD ----> https://www.givesendgo.com/billandkeonneDONATE TO THE FAMILIES w/ BTC ----> https://pay.zaprite.com/pl_JpxtkLv95T SUPPORT ON SOCIAL MEDIA ---> https://billandkeonne.org/TO DONATE TO ROMAN'S DEFENSE FUND: https://freeromanstorm.com/donateVALUE FOR VALUEThanks for listening you Ungovernable lot, we appreciate your continued support and hope you enjoy the shows.You can support this episode using your time, talent or treasure.TIME:- create fountain clips for the show- create a meetup- help boost the signal on social mediaTALENT:- create ungovernable misfit inspired art, animation or music- design or implement some software that can make the podcast better- use whatever talents you have to make a contribution to the show!TREASURE:- BOOST IT OR STREAM SATS on the Podcasting 2.0 apps @ https://podcastapps.com- DONATE via Monero @ https://xmrchat.com/ungovernable- BUY SOME STICKERS @ https://ungovernable.network/shop/FOUNDATIONhttps://foundation.xyz/ungovernableFoundation builds Bitcoin-centric tools that empower you to reclaim your digital sovereignty.As a sovereign computing company, Foundation is the antithesis of today's tech conglomerates. Returning to cypherpunk principles, they build open source technology that “can't be evil”.Thank you Foundation Devices for sponsoring the show!Use code: Ungovernable for $10 off of your purchaseCAKE WALLEThttps://cakewallet.comCake Wallet is an open-source, non-custodial wallet available on Android, iOS, macOS, and Linux.Features:- Built-in Exchange: Swap easily between Bitcoin and Monero.- User-Friendly: Simple interface for all users.Monero Users:- Batch Transactions: Send multiple payments at once.- Faster Syncing: Optimized syncing via specified restore heights- Proxy Support: Enhance privacy with proxy node options.Bitcoin Users:- Coin Control: Manage your transactions effectively.- Silent Payments: Static bitcoin addresses- Batch Transactions: Streamline your payment process.Thank you Cake Wallet for sponsoring the show!MYNYMBOXhttps://mynymbox.ioYour go-to for anonymous server hosting solutions, featuring: virtual private & dedicated servers, domain registration and DNS parking. We don't require any of your personal information, and you can purchase using Bitcoin, Lightning, Monero and many other cryptos.Explore benefits such as No KYC, complete privacy & security, and human support.(00:00:00) Welcome to Freedom Tech Friday(00:07:16) Small Models, Big Moves(00:24:01) Three Wallets, No Coldcard(00:30:52) Vibes and Benchmarks(00:33:35) Shielded Bitcoin, No Hard Fork(00:37:18) Start9 Burned His Laptop(00:39:50) RoboSats, NWC and DarkWisp(00:45:06) Uncle Jim's Compute Farm(00:51:11) Signal, Keet or SimpleX?(00:56:11) CUDA vs the M5 Ultra(01:00:16) See You Next Friday!

    The CyberWire
    Shut it down before they do.

    The CyberWire

    Play Episode Listen Later Sep 25, 2026 27:39


    Kiteworks urges customers pull the plug on vulnerable servers. CISA lays out its election security plan. Known vulnerabilities linger unpatched. Big AI labs consider a new standards body. Questions surround claims of an OpenAI Medicare hack. File notifications become a privacy leak. SectopRAT hides in audio software. A new Android banking trojan takes control. An attacker puts open-source AI agents to work hacking hundreds of organizations. A Rydox cybercrime marketplace operator pleads guilty. Our guest is Todd Thorsen, CISO of CrashPlan, with ways to prepare for and react to critical infrastructure campaigns. Americans give AI the side-eye. Remember to leave us a 5-star rating and review in your favorite podcast app. Miss an episode? Sign-up for our daily intelligence roundup, Daily Briefing, and you'll never miss a beat. And be sure to follow CyberWire Daily on LinkedIn. CyberWire Guest Todd Thorsen, CISO of CrashPlan, discusses ways to prepare for and react to critical infrastructure campaigns, and how AI is contributing to a bigger and more disruptive attack surface. Selected Reading Imminent Zero-Day Attack: KiteWorks Urges Customers to Shut Down Servers (Heise) CISA Election Security Plan Flags Patching Barriers, Voter Database Attacks (SecurityWeek) Only 26% of Detected CISA Known Exploited Vulnerabilities Were Fully Remediated (Hackread) Google, OpenAI, Anthropic Plan Frontier AI Standards Body (GovInfo Security) Doubts grow over claims OpenAI agent hacked Australian Medicare portal (The Record) Windows, Linux, Android File Notification Systems Leak User Activity (SecurityWeek) SectopRAT Abuses Legitimate Audio Software Files to Steal PC Data (Hackread) RemControl Banking Trojan Gives Attackers Remote Control of Android Devices (Infosecurity Magazine) Crook used three open source agents to break into a Fortune 500 hospitality company, a major US airline and 25+ other orgs (The Register) Kosovar Owner of Rydox Marketplace Pleads Guilty in US Court (SecurityWeek) Americans Fear AI Will Make the World Worse, Love It Anyway (404 Media) Share your feedback. What do you think about CyberWire Daily? Please take a few minutes to share your thoughts with us by completing our brief listener survey. Thank you for helping us continue to improve our show. Want to hear your company in the show? N2K CyberWire helps you reach the industry's most influential leaders and operators, while building visibility, authority, and connectivity across the cybersecurity community. Learn more at sponsor.thecyberwire.com. The CyberWire is a production of N2K Networks, your source for strategic workforce intelligence. © N2K Networks, Inc. Learn more about your ad choices. Visit megaphone.fm/adchoices

    DrunkFriend
    Episode 180 - Briz is Strange

    DrunkFriend

    Play Episode Listen Later Sep 25, 2026 74:48


    Send us Fan MailBriz comes on to talk all about the Life is Strange series of games. If you've ever been on the fence about this series, Briz will definitely push you off in one direction. Find Briz on the internet! YouTube Citizen Gamer, BlueSky @Citizengamer.bsky.socialPolymedia Charity Stream, October 10-11, noon-to-noon easterrnTwitch.tv/polymedianetworkLife is Strange (2015) Developer: DONTNOD Entertainment Systems: PS3, PS4, Xbox 360, Xbox One, PC, macOS, Linux, iOS, Android, Switch via Remastered CollectionLife is Strange: Before the Storm (2017) Developer: Deck Nine Systems: PS4, Xbox One, PC, macOS, Linux, iOS, Android, Switch via Remastered CollectionThe Awesome Adventures of Captain Spirit (2018) Developer: DONTNOD Entertainment Systems: PS4, Xbox One, PCLife is Strange 2 (2018–2019) Developer: DONTNOD Entertainment Systems: PS4, Xbox One, PC, macOS, Linux, SwitchLife is Strange: True Colors (2021) Developer: Deck Nine Systems: PS4, PS5, Xbox One, Xbox Series X|S, PC, SwitchLife is Strange: Double Exposure (2024) Developer: Deck Nine Systems: PS5, Xbox Series X|S, PC, SwitchLife is Strange: Reunion (2026) Developer: Deck Nine Systems: PS5, Xbox Series X|S, PCRelated: Lost Records: Bloom & Rage (2025) Developer: DON'T NOD Montréal Systems: PS5, Xbox Series X|S, PCSupport the showFind links for all things network related here: https://linktr.ee/polymedianetworkFind Travis on BlueSkyFind Alex on BlueSkySend us an email drunkfriendpodcast@gmail.comVisit our Subreddit reddit.com/r/polymedia

    AI Inside
    Qualcomm Wants Its AI Chips On Every Device

    AI Inside

    Play Episode Listen Later Sep 25, 2026 50:21


    Thanks to Qualcomm for sponsoring Jason's travel to Maui for the Qualcomm Snapdragon Summit. This week Jason Howell reports from Qualcomm's Snapdragon Summit in Maui, where the company unveiled the Snapdragon 8 Elite Gen 6 and a new Extreme tier capable of running 30-billion-parameter AI models on a phone. Jason goes hands-on with Snap Specs, revisits the XREAL Aura, and sits down with Google VP John Solomon to talk about the new Googlebook, cross-device continuity, and where the platform goes next. Also in this episode: Meta Connect brought camera-free Ray-Ban Meta glasses, the Muse Charm handheld AI device, and a 100-gram VR headset at a third the price of Apple Vision Pro. OpenAI released GPT-6 Sol and Luna, Anthropic disclosed that Claude now leads 26% of its R&D and opened a biology lab, and Gemini autonomously breached three real companies during a security test. New episodes every Wednesday at aiinside.show. Note: Time codes subject to change depending on dynamic ad insertion by the distributor. CHAPTERS: 00:00 - Podcast begins 02:57 - Snapdragon Summit and the agentic age 06:04 - Snapdragon 8 Elite Gen 6 and Extreme 07:21 - Personal Scribe and ambient AI 09:44 - 30-billion-parameter AI models on a phone 12:14 - AI agents on smart glasses 13:07 - Snapdragon START 13:51 - XREAL Aura and Snapdragon Reality Elite 15:07 - Snapdragon Sound Elite Gen 2 16:10 - Surface PCs and Snapdragon X2 Plus 17:18 - Googlebook and Linux on Snapdragon 18:14 - Snap Specs hands-on 21:20 - Patreon thanks and Jason's AI workshop 22:28 - John Solomon interview: Googlebook and ChromeOS 36:06 - Meta Connect: camera-free Ray-Ban Meta Audio 37:14 - Ray-Ban Meta Gen 3 37:45 - Meta's 100-gram VR glasses 39:06 - Meta Muse Charm 39:41 - Muse AI agents and transaction fees 40:37 - OpenAI GPT-6 Sol and Luna 41:31 - Claude helps build the next Claude 42:26 - Anthropic's biology lab and enzyme discovery 43:16 - Gemini breaches three real companies in a security test 44:13 - Wrap-up, plugs, and executive producers Host: Jason HowellGuest: John Solomon, VP & GM for Googlebook and ChromeOS Download and subscribe to AI Inside in audio and video: https://aiinside.show/ Support the podcast on Patreon for special perks: https://www.patreon.com/aiinsideshow. You'll get ad-free audio and video feeds, a members-only Discord, exclusive content, and T-shirts and stickers you'll love. Learn more about your ad choices. Visit megaphone.fm/adchoices

    Tech Over Tea
    Linux Community Is Getting Too Hot | Solo

    Tech Over Tea

    Play Episode Listen Later Sep 25, 2026 110:29


    KDE is currently being brigaded over an AI policy which is really sad because this should be a happy time for the project with the 30th anniversary coming up, but there's plenty more to talk about as well.==========Support The Channel==========► Patreon: https://www.patreon.com/brodierobertson► Paypal: https://www.paypal.me/BrodieRobertsonVideo► Amazon USA: https://amzn.to/3d5gykF► Other Methods: https://cointr.ee/brodierobertson==========Support The Show==========► Patreon: https://www.patreon.com/brodierobertson► Paypal: https://www.paypal.me/BrodieRobertsonVideo► Amazon USA: https://amzn.to/3d5gykF► Other Methods: https://cointr.ee/brodierobertson=========Video Platforms==========

    Windows Weekly (MP3)
    WW 1002: Only B & D Matter - Big Reorganization & Job Cuts Within XBOX

    Windows Weekly (MP3)

    Play Episode Listen Later Sep 24, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    Windows Weekly (MP3)
    WW 1002: Only B & D Matter - Big Reorganization & Job Cuts Within XBOX

    Windows Weekly (MP3)

    Play Episode Listen Later Sep 24, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    BSD Now
    682: NAS NAS NAS

    BSD Now

    Play Episode Listen Later Sep 24, 2026 39:06


    Multiple new BSD Bases NASes have appeared, FreeBSD intern bringing ROCm to FreeBSD, OpenSSH 10.5, and more... Headlines [Two more BSD based NASes have appeared] BSDNAS The developers seem to work for a Hungarian ISP They're building on top of zVault's fork for some of the effort FreeCORE Developer admits to vibecoding - Statement #1 Developer admits to vibecoding - Statement #2 FreeBSD Foundation Intern Sourojeet Adhikari on Bringing ROCm to FreeBSD News Roundup ICEBP Finally Documented Cleaning costs, or, examining the OpenBSD -fret-clean flag OpenSSH 10.5 Why do I run FreeBSD for my home servers. Tarsnap This weeks episode of BSDNow was sponsored by our friends at Tarsnap, the only secure online backup you can trust your data to. Even paranoids need backups. Feedback/Questions [Phil - Listener Feedback] Hi, Thanks for a great podcast - last year I was wondering what bsd I'd try, moving away from Linux, I'd heard Benedict and other hosts talking about most of the day to day software I use, so I wasn't worried about any of that working, but what bsd to go for (I get very little free time these days so getting it right was important - Benedict (i think) mentioned He was using FreeBSD - choice was made! And it has been great - 3 or 4 logins later and I'm reading a motd message by Benedict himself! I need to learn a few things, but these arethe things I expected would need a bit of reading - the handbook has mostly got me there, plus google etc. I am happy with FreeBSD and with Linus recently going off on one about Welcoming AI, after allowing an unfinished file system (bcashfs) into production, to name just 2 reasons for a change, I was thinking.. I don't know what the FreeBSD developers are doing with AI, but it'll be considered and responsible - OMG I had to re-play the last podcast a few times - They don't know yet???? Now I'm stuck! I use FreeBSD for desktop & nas, I love pkg and poudriere. do any of the BSDs have (at least) a no vibecode/slop policy and offer similar pkg/build systems? Thanks Phill Send questions, comments, show ideas/topics, or stories you want mentioned on the show to feedback@bsdnow.tv Join us and other BSD Fans in our BSD Now Telegram channel

    Windows Weekly (Video HI)
    WW 1002: Only B & D Matter - Big Reorganization & Job Cuts Within XBOX

    Windows Weekly (Video HI)

    Play Episode Listen Later Sep 24, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    Amateur Hour
    #219 - What is Microsoft doing?

    Amateur Hour

    Play Episode Listen Later Sep 24, 2026 70:35


    With everything going on with Microsoft and Windows 11, there's no wonder people call it "MicroSlop". What will make you swap to Linux?Email: Info@Amateurhourpod.comSocials: @Amateur_Pod

    Mark Vena Tech Guy Podcasts
    SmartTechCheck Podcast and Audio Newsletter: Qualcomm Snapdragon Summit 2026 --- What's Next For Agentic AI

    Mark Vena Tech Guy Podcasts

    Play Episode Listen Later Sep 24, 2026 22:59


    All TWiT.tv Shows (MP3)
    Windows Weekly 1002: Only B & D Matter

    All TWiT.tv Shows (MP3)

    Play Episode Listen Later Sep 23, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    Radio Leo (Audio)
    Windows Weekly 1002: Only B & D Matter

    Radio Leo (Audio)

    Play Episode Listen Later Sep 23, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    All TWiT.tv Shows (Video LO)
    Windows Weekly 1002: Only B & D Matter

    All TWiT.tv Shows (Video LO)

    Play Episode Listen Later Sep 23, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    Fedora Project Podcast
    61: Framework Laptop 12 Meets Fedora

    Fedora Project Podcast

    Play Episode Listen Later Sep 23, 2026 38:32


    Framework does not have a storefront you can walk into. What they do have is a laptop you can order that arrives with Fedora already on it, and a screwdriver in the box. Michael Tunnell is the Senior Marketing Manager for Linux Marketing at Framework, and spent years covering this ecosystem at TuxDigital before joining. Noah is back this week after missing the last episode while stuck forty feet up on a scissor lift with a dead battery. We get into: Why Fedora KDE, and why the answer is a boring technical one: KDE tested best for stylus and touchscreen, and Fedora moves fast enough on the kernel for brand new silicon The 80 percent of Laptop 12 owners already running Linux, measured before any preload existed, so every one of them installed it themselves What "officially supported" means at Framework: engineering collaboration with the distro, support for the distro, and a dedicated Linux support specialist a customer can actually reach Real upstream fixes, including a CPU limiter and thermals bug on Wildcat Lake, audio modules, and screen rotation, all landed in Fedora and the kernel The expansion card system, with two of the four slots on the 12 running Thunderbolt 4 ZMK open source keyboard firmware, published CAD files, and 3D printable components What it costs a company to commit to repairability instead of planned obsolescence There is a written companion post at https://itguyeric.com/blog/fedora-ep61-framework-laptop-12/Special Guest: Michael Tunnell.

    Radio Leo (Video HD)
    Windows Weekly 1002: Only B & D Matter

    Radio Leo (Video HD)

    Play Episode Listen Later Sep 23, 2026 152:53 Transcription Available


    Windows Week D arrives, and this time we get updates - a preview of Octoberʼs Patch Tuesday 24H2/25H2 Minor update with new Copilot key remaps, Wi-Fi in WRE, more 26H1 Major update that brings all new 26H2 features—Start, Taskbar, Search —to 26H1. Finally. Windows Insider Program Four new builds last Friday - Biggest change is Cloud Rebuild in WRE (in Experimental) gaining new capabilities. Software 30 years ago, Dave Plummer created Task Manager, and now heʼs vibe-coded a modern version that works on Mac and Linux too. Stardock announces a new Windows 11 utility called Snipboard. Clipchamp does AI upscaling to 4K now Hardware The Lenovo ThinkPad X9 15p Aura Edition is nearly perfect, with 11 hours of battery life. Googlebook revealed! Questions answered - $899 starting price is much lower than expected. This is Androidʼs Copilot+ PC moment - An AI PC? What a concept! AI/dev Microsoft announces a OneDrive + Copilot event in October - Also a separate Copilot event before then. Financial Times reveals internal OpenAI prediction of $278 billion in losses over the next five years. NYT copyright case discovery reveals what Microsoft and OpenAI really think about stealing content. Anthropic merges Cowork into chat, adds Docs and Slides. Google Labs CC turned into Gemini Daily Brief, so now itʼs an AI agent for families. Google opens up its Home smart home ecosystem to other AI assistants. XAML.io is a web-based vibe coding IDE for .NET apps and it looks amazing. XBOX and Gaming A big reorg and a small number of layoffs at XBOX sets off a new round of hand wringing. Tip and Picks Tip of the week: Get Auto SR on Panther Lake-based laptops Auto SR was designed to make gaming viable on Windows 11 on Arm/Snapdragon X/X2, but now itʼs available on Intel Core Ultra Series 3-based PCs and it is worth experimenting with. App Pick of the Week: Affinity Possibly the greatest free app in existence, Affinity just got 60 new features and is even better. Hosts: Leo Laporte and Paul Thurrott Download or subscribe to Windows Weekly at https://twit.tv/shows/windows-weekly Check out Paul's blog at thurrott.com The Windows Weekly theme music is courtesy of Carl Franklin. Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Sponsors: trustedtech.team/twit originhq.com/windowsweekly

    Absolute Tech
    Linux and Astronomy - 09/08/2026

    Absolute Tech

    Play Episode Listen Later Sep 23, 2026 60:50


    Send us Fan MailLinux and Astronomy with Roy, KI7PKLSupport the show

    The Emergency Management Network Podcast
    THE EMN DAILY NEWS | Monday, 21 September 2026

    The Emergency Management Network Podcast

    Play Episode Listen Later Sep 22, 2026 6:25


    THE EMN DAILY NEWS | Monday, 21 September 2026Tropical Storm Fay peaked at 70 miles per hour overnight, just short of hurricane strength, and is now weakening about 405 miles southwest of the Azores with no hazards to land. The Central Pacific Hurricane Center gives the remnants of Tropical Depression Fifteen-E a 60 percent chance of becoming a depression within seven days, possibly south or southeast of Hawaii later this week. National Preparedness Level 3 holds with 96 incidents and 10,886 personnel, down 437 in a day, and uncontained large fires fell from 46 to 34. The Southern Area carries more incidents than any other region, 36, as hot and dry conditions persist across the Lower Mississippi Valley. Boil water advisories are in effect in three states, including all of Monroe County, Indiana, and most of Kalamazoo, Michigan. The EMN NEWS Daily Brief is your concise daily update on national and state-by-state emergency management news.KEY TAKEAWAYSHawai'i watch: As of Sunday's outlook, the Central Pacific Hurricane Center puts formation chances for the remnants of Tropical Depression Fifteen-E near zero through 48 hours and at 60 percent over seven days as the system moves west to west-southwest. A depression could form south or southeast of the Hawaiian Islands once conditions improve. No watches are posted, keeping the first half of the week as the preparation window.Atlantic: Fay reached 70 miles per hour at the 0300 UTC advisory, then weakened to 65 miles per hour and 997 millibars by 0900 UTC, drifting south-southeast at 2 miles per hour. The National Hurricane Center expects Fay to become post-tropical in a day or two and to dissipate by Saturday. No hazards affect land.Wildfire posture: Preparedness Level 3 holds with 96 incidents, 34 uncontained large fires, 1,999,381 acres and 10,886 personnel, a drop of 437, per the Sunday situation report. Initial attack was light, with 60 fires and no new large incidents, and eight complex incident management teams remain committed. Year to date, the country is at 56,525 fires and 8,541,715 acres, 125 percent and 146 percent of the ten-year averages.Southern Area: Thirty-six incidents and 1,038 personnel, down 157, after 26 new fires. Dry River near Sonora, Texas, is 8,939 acres and 85 percent contained, with energy infrastructure threatened. Gracemont near Binger, Oklahoma, has destroyed 31 structures on 2,665 acres and is 60 percent contained. Relative humidity of 20 to 30 percent persists across the Lower Mississippi Valley.Northwest: Fourteen uncontained large fires, 5,050 personnel and five complex teams. The situation report no longer lists evacuations on Three Queens or Goat; Three Queens, near Snoqualmie Pass, is 9,566 acres and 23 percent contained, with numerous residences and critical infrastructure threatened. Border 2 in North Cascades National Park grew 345 acres, and Sisi near Stehekin is 15 percent contained, with residences and energy infrastructure threatened.California: The dome in Yosemite National Park grew 198 acres to 2,384 and remains 15 percent contained, with 525 personnel. Command on Timber was set to pass Sunday from Southwest Team 1 to California Team 6, with the fire at 25,435 acres, 59 percent contained and $169.9 million in costs. Floriston near Truckee moved from 25 to 45 percent contained and released 100 personnel.Rocky Mountains: Moonshine near Wheatland, Wyoming, is 7,325 acres and 61 percent contained, with residences threatened. Swiss Roll near Pagosa Springs, Colorado, is 36 percent contained, with energy infrastructure threatened. Rain fell Saturday over Davis Coulee and Sand Creek in Montana, where Sand Creek is 58 percent contained.Water systems: City of Bloomington Utilities issued a precautionary boil water advisory for all of Monroe County, Indiana, early Sunday after turbidity exceeded the permitted maximum, running through 3 p.m. today. Kalamazoo, Michigan, expects to lift its advisory today if Monday samples come back clean; it covers 19 of 22 city neighborhoods and Portage north of Interstate 94. Salemburg, North Carolina, restored service Saturday evening, but its advisory remains in effect.Federal assistance: SBA physical damage deadlines for the September 1 public assistance declarations close November 1 in Colorado and South Dakota and November 2 in Missouri and Texas. Economic injury applications stay open to June 1, 2027.Cyber: The most recent Known Exploited Vulnerabilities additions, posted this morning, are Friday's three Linux kernel entries, remediated by federal civilian agencies under Binding Operational Directive 26-04.Fire weather: Wet thunderstorms bring up to about an inch of rain to eastern Montana, Wyoming and the High Plains. Relative humidity drops to 7 to 18 percent across Nevada, the Mojave Desert and the Central Valley, with light winds; gusts of 25 to 30 miles per hour are likely along the Front Range and eastern Sierra.Tropical weatherTropical Storm Fay Public Advisory Number 7: NHC, 0300 UTC Sept 21Tropical Storm Fay Public Advisory Number 8: NHC, 0900 UTC Sept 21Tropical Storm Fay Forecast/Advisory Number 8: NHC, 0900 UTC Sept 21Central Pacific Tropical Weather Outlook: CPHC, 200 AM HST Sept 20, formation chance 60 percent in 7 daysCentral Pacific Tropical Weather Outlook: CPHC, 800 AM HST Sept 20Wildfire, nationalNICC Incident Management Situation Report: Sunday, Sept 20, 2026, 0730 MDTNICC Predictive Services discussion: Sept 20CISAKnown Exploited Vulnerabilities catalog: CVE-2025-39682, CVE-2025-39964 and CVE-2026-53266 added Sept 18; no later additions located in this sweepBinding Operational Directive 26-04: Prioritizing Security Updates Based on RiskFederal assistanceFederal Register: SBA notice, Texas, FEMA-4935-DR, amended incident period July 12 through August 4, published Sept 14Federal Register: SBA notices for Colorado FEMA-4944-DR, South Dakota FEMA-4938-DR and Missouri FEMA-4939-DR, published Sept 9 to Sept 14Texas and OklahomaNICC situation report: Dry River, Hydra, Gracemont and Calvary Creek, Sept 20Washington and OregonNICC situation report: Three Queens, Goat, Sisi, Border 2 and Little Giant, Sept 20CaliforniaNICC situation report: Dome, Timber, Plaskett and Floriston, Sept 20Wyoming, Colorado and MontanaNICC situation report: Moonshine, Swiss Roll, Davis Coulee and Sand Creek, Sept 20IndianaThe Bloomingtonian: City of Bloomington Utilities countywide boil water advisory, Sept 20Indiana Public Media: precautionary boil advisory for Monroe County, Sept 20MichiganThe Detroit News: Kalamazoo boil water advisory could end Monday, Sept. 20North CarolinaWRAL: Salemburg boil water advisory remains in effect, updated Sept. 20Links that need a note:* Advisory 8 has no NHC link. Both Fay advisory 8 products came through news-site mirrors of the NHC bulletin text, not nhc.noaa.gov, so I've marked them.* Two links are live pages. The situation report and one Central Pacific outlook link are overwritten with each new issue. After today they will show newer content than what the brief cites.* The CISA links are lost. I pulled the September 18 CISA alerts yesterday, but those links didn't carry over. A search just now couldn't recover them.* Three Federal Register links are lost. The same happened to the Colorado, South Dakota and Missouri notices. I've marked all of these as missing rather than rebuild addresses from memory.Tropical weather* Fay Public Advisory 7, NHC: https://www.nhc.noaa.gov/text/refresh/MIATCPAT1+shtml/200841.shtml* Fay Public Advisory 8, NHC text via Index-Journal mirror: https://www.indexjournal.com/news/hurricane/tropical-storm-fay-public-advisory-number-8/article_9a48d4f1-d16c-5f75-923d-febd44f298cf.html* Fay Forecast/Advisory 8, NHC text via Cape Weather mirror: https://capeweather.com/tropical-storm-fay-forecast-advisory-number-8/* Central Pacific outlook, 200 AM HST Sept 20 (live page, overwritten): https://www.nhc.noaa.gov/text/HFOTWOCP.shtml* Central Pacific outlook, 800 AM HST Sept 20: https://www.nhc.noaa.gov/mobile/text/refresh/HFOTWOCP+html/171133_HFOTWOCP.htmlWildfire, national (covers every NICC line in the state sections too)* Situation report and Predictive Services discussion (live PDF, overwritten daily): https://www.nifc.gov/nicc-files/sitreprt.pdf* Archive page for past situation reports: https://www.nifc.gov/nicc/incident-information/imsrCISA* KEV additions, Sept 18: link not recovered* Binding Operational Directive 26-04: link not recoveredFederal assistance* Texas, FEMA-4935-DR amendment, published Sept 14: https://www.federalregister.gov/documents/2026/09/14/2026-18756/presidential-declaration-amendment-of-a-major-disaster-for-public-assistance-only-for-the-state-of* Texas, FEMA-4935-DR original notice, published Sept 9: https://www.federalregister.gov/documents/2026/09/09/2026-18354/presidential-declaration-of-a-major-disaster-for-public-assistance-only-for-the-state-of-texas* Colorado, South Dakota and Missouri notices: links not recoveredIndiana* The Bloomingtonian: https://bloomingtonian.com/2026/09/20/bloomington-utilities-issues-countywide-boil-water-advisory-after-turbidity-spike/* Indiana Public Media: https://www.ipm.org/news/2026-09-20/precautionary-boil-advisory-issued-for-monroe-countyMichigan* The Detroit News: https://www.detroitnews.com/story/news/local/michigan/2026/09/20/kalamazoo-boil-water-advisory-could-end-monday/91781476007/North Carolina* WRAL: https://www.wral.com/news/local/boil-water-advisory-salemburg-september-2026/ This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit emnetwork.substack.com/subscribe

    This Week in Linux
    357: Linux on M3 Macs, GNOME 51, Steam Frame, Jellyfin 12, Microsoft 365 on Linux, & more Linux news

    This Week in Linux

    Play Episode Listen Later Sep 21, 2026 30:12


    video: https://youtu.be/V4r7fCFLP6A This week in Linux, GNOME 51 is here with a faster, more polished desktop and a long list of upgrades across the environment. Asahi Linux pulled off another major Apple Silicon milestone with M3 Macs. Valve is building one of the wildest compatibility stacks we've seen yet for the Steam Frame. And Open Source projects are now setting traps for AI coding agents that can make lazy submissions snitch on themselves. Plus, Jellyfin 12 is a huge upgrade for self-hosted media. All of this and more on this week's episode, so let's jump into Your Source for Linux GNews and see what's happened This Week in Linux. Download as MP3 Support the Show Become a Patron = tuxdigital.com/membership Store = tuxdigital.com/store Chapters: 00:00 Intro 00:47 GNOME 51 Released 04:34 Asahi Linux for M3 Macs 10:52 Valve's Steam Frame Compatibility Stack for Linux Gaming 13:45 Microsoft 365 might come to Linux (unofficially) 17:21 Jellyfin 12 Released 20:55 Open Source Projects Setting Traps for Lazy Ai Submissions 24:20 Raspberry Pi OS Desktop Refresh 26:07 Ubuntu 26.10 completes migration to Rust-based coreutils 27:08 Google is Replacing C-based Android Binder with Rust 28:13 RustFS 1.0 Released 29:04 Support the show 29:58 Outro Links: GNOME 51 Released https://release.gnome.org/51/ https://www.phoronix.com/news/GNOME-51-Released https://www.omgubuntu.co.uk/2026/09/gnome-51-released Asahi Linux for M3 Macs https://asahilinux.org/2026/09/m2-episode-1/ https://asahilinux.org/docs/platform/feature-support/m3/ https://fedoramagazine.org/announcing-fedora_asahi_remix_45_beta/ https://fedora-asahi-remix.org/ https://www.phoronix.com/news/Asahi-Linux-Official-M3 https://www.gamingonlinux.com/2026/09/asahi-linux-now-officially-supports-macs-with-the-m3-series-soc/ Valve's Steam Frame Compatibility Stack for Linux Gaming https://partner.steamgames.com/doc/steamhardware/steamframe/compatibility https://gitlab.steamos.cloud/frame-public/lepton/ https://www.gamingonlinux.com/2026/09/lepton-from-valve-to-run-android-games-on-linux-is-now-open-source/ https://www.gamingonlinux.com/2026/08/lepton-and-fex-get-prepared-for-the-steam-frame-release/ Microsoft 365 might come to Linux (unofficially) https://itsfoss.com/news/bottles-microsoft-365-early-look/ https://mprove.de/blogs/rss/index.php?rss=mastodon.bsd.cafe%2Ftags%2FMicrosoft.rss https://github.com/bottlesdevs/Bottles/releases https://usebottles.com/ Jellyfin 12 Released https://jellyfin.org/posts/jellyfin-release-12.0/ https://github.com/jellyfin/jellyfin/releases https://github.com/jellyfin/jellyfin-web/releases Open Source Projects Setting Traps for Lazy Ai Submissions https://github.com/NetworkManager/NetworkManager/blob/main/AGENTS.md https://www.phoronix.com/news/NetworkManager-AI-Canary https://github.com/systemd/systemd/blob/main/AGENTS.md https://www.phoronix.com/news/systemd-262-rc2 https://gitlab.freedesktop.org/NetworkManager/NetworkManager/-/merge_requests/2536 https://gitlab.freedesktop.org/NetworkManager/NetworkManager/-/blob/main/AGENTS.md https://hwbusters.com/news/networkmanager-ai-policy-gets-a-trap-word-and-ci-now-scans-every-commit-for-it/ Raspberry Pi OS Desktop Refresh https://www.raspberrypi.com/news/an-updated-look-for-the-raspberry-pi-desktop/ https://www.raspberrypi.com/news/category/os/raspberry-pi-os/ Ubuntu 26.10 completes migration to Rust-based coreutils https://www.omgubuntu.co.uk/2026/09/ubuntu-2610-rust-coreutils-complete https://documentation.ubuntu.com/release-notes/26.10/ https://uutils.org/ https://www.phoronix.com/news/Ubuntu-Completes-Rust-Coreutils Google is Replacing C-based Android Binder with Rust https://lkml.iu.edu/2609.1/12916.html https://lists.openwall.net/linux-kernel/2026/09/12/891 https://www.phoronix.com/news/Google-Binder-C-Goodbye RustFS 1.0 Released https://www.rustfs.com/blog/announcing-rustfs-1-0-0-ga/ https://github.com/RustFS/RustFS https://linuxiac.com/rustfs-1-0-s3-compatible-object-storage-reaches-ga/ Support the show https://tuxdigital.com/membership https://store.tuxdigital.com/

    Throwdown Show
    602: Who is to blame for the Kojima/PlayStation split?

    Throwdown Show

    Play Episode Listen Later Sep 21, 2026 97:26


    Tonight's questions:- Who is to blame for the Kojima/PlayStation split?- Is Control Resonant a game of the year contender?- Is Control Resonant a stand-alone game?- Will Manny buy Control Resonant on day one?- Is the Wolverine hate overblown?- Is PlayStation completely cooked?- Is Sony purposely tanking PlayStation? - Will PlayStation become just a digital storefront?- Has anyone played the Final Fantasy Resonance demo?- Why is GTA 6's OST getting a physical copy?- Will there be another Metroid Prime game?- Will Tony review the ROG Ally X20?- Will you move to Linux from Windows?Thanks as always to Shawn Daley for our intro and outro music. Follow him on SoundCloud: https://soundcloud.com/shawndaleyWhere to find Throwdown Show:Website: https://audioboom.com/channels/5030659Twitch: https://www.twitch.tv/throwdownshowX: https://x.com/ThrowdownShowYouTube: https://www.youtube.com/throwdownshowDiscord: https://discord.gg/fdBXWHTTwitter list: https://twitter.com/i/lists/1027719155800317953

    Atareao con Linux
    ATA 833 Workflows de IA, automatiza tu día con Ollama y Docker

    Atareao con Linux

    Play Episode Listen Later Sep 21, 2026 29:55


    Hoy en el episodio 833 te traigo algo que llevaba tiempo queriendo hacer: montar workflows de IA que funcionen de verdad en tu Linux, sin depender de servicios externos, sin GPUs, y sobre todo, sin malgastar tokens en tonterías.Hasta ahora hemos hablado de piezas sueltas: Ollama, skills, MCPs, agentes... pero todo eso suelto no te sirve para nada. La gracia está en combinarlo. En este episodio te enseño 4 workflows completos implementados en Rust que he puesto a funcionar en mi propio equipo, y que puedes adaptar al tuyo sin necesidad de ser un experto.La base del sistema es sencilla: Llama 3.2 3B para generación de texto (~2 GB, funciona en CPU), bge-m3 para embeddings multilingües (~1.2 GB), y binarios Rust compilados estáticamente que no necesitan ninguna dependencia del sistema. Con 8 GB de RAM tienes de sobra.Workflow 1 — noticias-bot: un bot que monitoriza feeds RSS, los filtra por palabras clave, los resume con Llama 3.2 local, y los publica automáticamente en Telegram. Funciona como servicio persistente 24/7, no como un script que ejecutas a mano. Cada feed tiene su propio intervalo de actualización y sus propias keywords. Lleva deduplicación con SHA256 y SQLite para no repetir artículos.Workflow 2 — tareas-bot: un gestor de tareas GTD vía Telegram. Le escribes un mensaje al bot, y la IA lo clasifica al instante: decide si es una tarea, le asigna prioridad y categoría, lo guarda en SQLite, y te confirma en el chat. Sin abrir ninguna app de tareas, sin salir de Telegram.Workflow 3 — monitor-bot: un monitor de sistema inteligente con 5 tipos de chequeo: disco, servicios, memoria, logs del sistema y contenedores Docker. Y aquí viene lo interesante: solo usa la IA cuando realmente hace falta, para analizar logs. Los checks normales son deterministas, sin LLM. Así no malgastas recursos ni tokens. Incluye rate limiting y remediación automática.Workflow 4 — investigador RAG: un asistente de investigación con RAG local. Le haces una pregunta, genera consultas de búsqueda, las lanza contra SearXNG, descarga las páginas en paralelo con tokio, las procesa con embeddings de bge-m3, y te da una respuesta con sus fuentes. Todo en un solo binario Rust, sin Python, sin dependencias del sistema.Todo esto desplegado con systemd o Docker, como servicios que arrancan solos y se mantienen funcionando. Y lo mejor: con modelos que caben en cualquier máquina con 8 GB de RAM y sin GPU.Capítulos del episodio:0:00 — Introducción: del caos de herramientas a los workflows IA2:30 — ¿Qué es un workflow de IA? Skills, MCPs y prompts combinados5:00 — Arquitectura: Rust, Ollama y Docker como base del sistema8:00 — Ejemplo 1: noticias-bot — RSS filtrado a Telegram11:30 — Demo del noticias-bot: publicación automática de noticias14:30 — Ejemplo 2: tareas-bot — clasificación GTD vía Telegram17:30 — Demo del tareas-bot: "hola" vs "comprar ciruelas"20:00 — Ejemplo 3: monitor-bot — alertas de sistema con IA22:30 — Ejemplo 4: investigación asistida con RAG local26:00 — Demo del investigador: consulta sobre Podman 629:00 — Ventajas, conclusiones y despedidaMás información y enlaces en las notas del episodio

    All TWiT.tv Shows (MP3)
    Untitled Linux Show 271: Touched By AI

    All TWiT.tv Shows (MP3)

    Play Episode Listen Later Sep 20, 2026 87:25 Transcription Available


    This week Jonathan starts out talking about his new Framework 13 pro, and Rob tries to convince him to buy a Steam Frame to go with it. The 7.4 kernel is really shaping up to be special, with huge wins in build time, AMD HDMI support, Google finally dropping its legacy binder code, and files opening significantly faster. Ubuntu is continuing to push into Rust coreutils, KaOS is trying to manage a SystemD-free existence, and Intel is done paying out bug bounties. For tips, we have excise for a TUI disk usage visualizer, snoopy for watching what your system is really up to, and a quick primer on .pc files. You can find the show notes at https://bit.ly/3V6eLGM and enjoy! Host: Jonathan Bennett Co-Hosts: Rob Campbell and Jeff Massie Download or subscribe to Untitled Linux Show at https://twit.tv/shows/untitled-linux-show Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Club TWiT members can discuss this episode and leave feedback in the Club TWiT Discord. Sponsor: trustedtech.team/twit

    All TWiT.tv Shows (MP3)
    Untitled Linux Show 271: Touched By AI

    All TWiT.tv Shows (MP3)

    Play Episode Listen Later Sep 20, 2026 87:25


    This week Jonathan starts out talking about his new Framework 13 pro, and Rob tries to convince him to buy a Steam Frame to go with it. The 7.4 kernel is really shaping up to be special, with huge wins in build time, AMD HDMI support, Google finally dropping its legacy binder code, and files opening significantly faster. Ubuntu is continuing to push into Rust coreutils, KaOS is trying to manage a SystemD-free existence, and Intel is done paying out bug bounties. For tips, we have excise for a TUI disk usage visualizer, snoopy for watching what your system is really up to, and a quick primer on .pc files. You can find the show notes at https://bit.ly/3V6eLGM and enjoy! Host: Jonathan Bennett Co-Hosts: Rob Campbell and Jeff Massie Download or subscribe to Untitled Linux Show at https://twit.tv/shows/untitled-linux-show Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Club TWiT members can discuss this episode and leave feedback in the Club TWiT Discord. Sponsor: trustedtech.team/twit

    Late Night Linux Extra
    Linux Dev Time – Episode 159

    Late Night Linux Extra

    Play Episode Listen Later Sep 20, 2026 25:56


    Our fantasy programming language features. Support us on Patreon and get an ad-free RSS feed with early episodes sometimes See our contact page for ways to get in touch. Subscribe to the RSS feed

    Late Night Linux All Episodes
    Linux Dev Time – Episode 159

    Late Night Linux All Episodes

    Play Episode Listen Later Sep 20, 2026 25:56


    Our fantasy programming language features. Support us on Patreon and get an ad-free RSS feed with early episodes sometimes See our contact page for ways to get in touch. Subscribe to the RSS feed

    The Linux Cast
    Episode 243: Omarchy is Rich Now - And Other News

    The Linux Cast

    Play Episode Listen Later Sep 20, 2026 60:50


    The boys are back! Tonight we talk about the news! Just a note for the audio listeners, this podcast was cursed. The ending was cut off. Apologies. ==== Special Thanks to Our Patrons! ==== https://thelinuxcast.org/patrons/ ===== Follow us

    All TWiT.tv Shows (Video LO)
    Untitled Linux Show 271: Touched By AI

    All TWiT.tv Shows (Video LO)

    Play Episode Listen Later Sep 20, 2026 87:25 Transcription Available


    This week Jonathan starts out talking about his new Framework 13 pro, and Rob tries to convince him to buy a Steam Frame to go with it. The 7.4 kernel is really shaping up to be special, with huge wins in build time, AMD HDMI support, Google finally dropping its legacy binder code, and files opening significantly faster. Ubuntu is continuing to push into Rust coreutils, KaOS is trying to manage a SystemD-free existence, and Intel is done paying out bug bounties. For tips, we have excise for a TUI disk usage visualizer, snoopy for watching what your system is really up to, and a quick primer on .pc files. You can find the show notes at https://bit.ly/3V6eLGM and enjoy! Host: Jonathan Bennett Co-Hosts: Rob Campbell and Jeff Massie Download or subscribe to Untitled Linux Show at https://twit.tv/shows/untitled-linux-show Join Club TWiT for Ad-Free Podcasts! Support what you love and get ad-free audio and video feeds, a members-only Discord, and exclusive content. Join today: https://twit.tv/clubtwit Club TWiT members can discuss this episode and leave feedback in the Club TWiT Discord. Sponsor: trustedtech.team/twit

    BIT-BUY-BIT's podcast
    Is AI Freedom Tech? Part II | FREEDOM TECH FRIDAY 57

    BIT-BUY-BIT's podcast

    Play Episode Listen Later Sep 19, 2026 64:52 Transcription Available


    A weekly live show covering all things Freedom Tech with Max, Q and Seth.In this episode, Max chats with Seth for a wide-ranging catch-up on the state of open source and local AI following his visit to the Open Source AI Summit in San Francisco. They talked about why AI feels so energising compared with the slower pace of Bitcoin development, how local and open models are rapidly improving, and why more people in the freedom tech world need to pay attention. Seth shared his view that we are already close to a point where self-hosted AI is practical for everyday use, with the right hardware, and explained the trade-offs between speed, intelligence and cost when choosing a setup.Max and Seth also dug into what it means to treat AI like “self-custody for intelligence", keeping your data private, running tools locally, and avoiding dependence on closed platforms that can limit, monitor or shape your usage. Along the way, they covered how AI can help with everything from infrastructure and creative work to family life and productivity, as well as the risks around surveillance, security and misuse. Max and Seth closed by looking at the idea of “resonant computing”, a vision for technology that is private, adaptable and built to serve people rather than exploit them.HELP GET SAMOURAI A PARDONSIGN THE PETITION ----> https://www.change.org/p/stand-up-for-freedom-pardon-the-innocent-coders-jailed-for-building-privacy-tools DONATE TO THE FAMILIES w/ USD ----> https://www.givesendgo.com/billandkeonneDONATE TO THE FAMILIES w/ BTC ----> https://pay.zaprite.com/pl_JpxtkLv95T SUPPORT ON SOCIAL MEDIA ---> https://billandkeonne.org/TO DONATE TO ROMAN'S DEFENSE FUND: https://freeromanstorm.com/donateVALUE FOR VALUEThanks for listening you Ungovernable Misfits, we appreciate your continued support and hope you enjoy the shows.You can support this episode using your time, talent or treasure.TIME:- create fountain clips for the show- create a meetup- help boost the signal on social mediaTALENT:- create ungovernable misfit inspired art, animation or music- design or implement some software that can make the podcast better- use whatever talents you have to make a contribution to the show!TREASURE:- BOOST IT OR STREAM SATS on the Podcasting 2.0 apps @ https://podcastapps.com- DONATE via Monero @ https://xmrchat.com/ungovernable- BUY SOME STICKERS @ https://ungovernable.network/shop/FOUNDATIONhttps://foundation.xyz/ungovernableFoundation builds Bitcoin-centric tools that empower you to reclaim your digital sovereignty.As a sovereign computing company, Foundation is the antithesis of today's tech conglomerates. Returning to cypherpunk principles, they build open source technology that “can't be evil”.Thank you Foundation Devices for sponsoring the show!Use code: Ungovernable for $10 off of your purchaseCAKE WALLEThttps://cakewallet.comCake Wallet is an open-source, non-custodial wallet available on Android, iOS, macOS, and Linux.Features:- Built-in Exchange: Swap easily between Bitcoin and Monero.- User-Friendly: Simple interface for all users.Monero Users:- Batch Transactions: Send multiple payments at once.- Faster Syncing: Optimized syncing via specified restore heights- Proxy Support: Enhance privacy with proxy node options.Bitcoin Users:- Coin Control: Manage your transactions effectively.- Silent Payments: Static bitcoin addresses- Batch Transactions: Streamline your payment process.Thank you Cake Wallet for sponsoring the show!MYNYMBOXhttps://mynymbox.ioYour go-to for anonymous server hosting solutions, featuring: virtual private & dedicated servers, domain registration and DNS parking. We don't require any of your personal information, and you can purchase using Bitcoin, Lightning, Monero and many other cryptos.Explore benefits such as No KYC, complete privacy & security, and human support.

    MP3 – mintCast
    493 – Mintinator

    MP3 – mintCast

    Play Episode Listen Later Sep 19, 2026 85:38


    This episode of mintCast (Episode 493) features the team—Joe, Bill, Charles, and Jim—discussing various Linux-related news and personal tech projects. Security & AI: The hosts discuss recent industry concerns regarding self-improving AI, the "existential risk" warnings from researchers, and the practical reality behind "kill switches" and sandbox escapes.

    The Plex
    The Plex EP496 - Zohran and 911, More Iran Lies, Hasan Thirst At GOP Convention, AI Psycho Hype Cycle

    The Plex

    Play Episode Listen Later Sep 19, 2026


    Check Out Echoplex Radio iTunes, Stitcher, Google, iHeart, Spotify, RSS, Odysee, Twitch, YouTubeSupport This Project On Patreon Check Out Our Swag Shop Join Our Discord Server Check out our Linux powered studio!‍ ‍Host: Producer DaveDocket: https://tinyurl.com/9-13-20206-docMembers ShowFourthwallPatreon

    ThunderCast
    Thundercast - S3E5 - All About Thundermail

    ThunderCast

    Play Episode Listen Later Sep 19, 2026 48:06


    ThunderCast, the official Thunderbird podcast is back for another season! In this episode we're joined by Philipp, the Director of Services, as we discuss everything about Thundermail, our newly released email service, and answer many questions from our community.Thundermail: https://thundermail.comRoadmaps: https://roadmaps.thunderbird.net/Developer guides: https://developer.thunderbird.net/Ideas for Thunderbird desktop and mobile: https://connect.mozilla.org/User support for Thunderbird desktop and mobile: https://support.mozilla.org/Submit your questions at podcast@thunderbird.net ★ Support this podcast ★

    The G2 on 5G Podcast by Moor Insights & Strategy
    Apple's C2 mmWave Modem, 4 GHz Spectrum Push, Snap SPECS 5G, and Verizon's 6G Forum

    The G2 on 5G Podcast by Moor Insights & Strategy

    Play Episode Listen Later Sep 19, 2026 32:38


    In episode 259 of the 6G Podcast, analysts Anshel Sag and Mike Dano discuss Sag's travel and key takeaways from Apple's event, including the foldable iPhone Duo, A20 Pro silicon, and Apple's new C2 modem as its first millimeter-wave modem (US-only), plus Release 18/5G-Advanced features like uplink MIMO and updated modem architecture; Sag also notes new watches using a 5G RedCap modem. Dano covers accelerating US spectrum activity tied to 6G, highlighting CTIA and NextG Alliance focus on the 4 GHz band ahead of WRC27, FCC/NTIA auction developments, Verizon's request to test 4.8–4.9 GHz in Los Angeles, and debate over spectrum sharing. Sag recaps Snap's standalone AR Specs, including a Qualcomm cellular modem built into the glasses case, enterprise positioning, open-source Linux platform, and carrier partners. They also address new FCC and HHS (RFK Jr.) RF exposure inquiries, Verizon's expanded 6G Innovation Forum membership and pre-standard ISAC work, and India joining a US-led 6G alignment initiative. 00:00 Welcome Back and Updates 01:18 Apple Event Foldable iPhone 01:46 C2 Modem and 5G Advanced 05:44 Four Gigahertz Spectrum Push 11:45 Snap Specs AR Glasses 18:46 FCC RF Exposure Study Returns 26:45 Verizon 6G Innovation Forum 29:27 India Joins 6G Alignment 32:16 Wrap Up and Subscribe

    Hackaday Podcast
    Ep 387: Superhuman Clocks, CAN in USB-C, and the Joys of Bare Metal

    Hackaday Podcast

    Play Episode Listen Later Sep 18, 2026 70:08


    This week, Hackaday Editors Elliot Williams and Tom Nardi start the episode off by discussing the latest CircuitPython developments before covering some impressive reverse engineering efforts, the benefits of modeling your projects in 3D, and some of the most incredible timepieces that have ever graced the pages of Hackaday. You'll also hear about the fascinating potential of combing 3D and UV printing, Linux on the ESP32, and a virtual TV station that pulls from the Internet Archive. The episode wraps up with a Hackaday Europe double-feature: one talk extols the virtues of keeping things simple through bare metal development, while the other covers off-world hacks and fixes that will make you want to sign up for Space Camp. Check out the links if you want to follow along, and as always, tell us what you think about this episode in the comments!  

    The Nextlander Podcast
    264: A Lot Going on in Video Game Town

    The Nextlander Podcast

    Play Episode Listen Later Sep 17, 2026 148:51


    PlayStation dropped Kojima Productions' Physint... and Xbox picked it up. Nintendo showed the Ocarina of Time remake, and then went and announced a new 2D Metroid. Blizzard revealed a StarCraft shooter for 2030, Diablo V for 2029, and a forever version of World of Warcraft. Insomniac's Wolverine came out. Obviously this is a lot to talk about, and we do our best with all of it on this week's show!CHAPTERS(00:00:00) NOTE: Some timecodes may be inaccurate for versions other than the ad-free Patreon version due to dynamic ad insertions. Please use caution if skipping around to avoid spoilers. Thanks for listening.(00:00:10) Intro(00:00:37) Never stop podcasting(00:07:11) [Possible Spoilers] Narratively, we try to stay close to officially released materials. If you're on complete blackout obviously skip this Wolverine section.(00:07:39) Marvel's Wolverine  |  [PlayStation 5]  |  Sep 15, 2026(00:53:46) First Break(00:54:02) Big Walk  |  [Mac, PC (Microsoft Windows), PlayStation 5, Nintendo Switch 2]  |  Aug 04, 2026(00:57:30) Star Wars Zero Company  |  [PC (Microsoft Windows), PlayStation 5, Xbox Series X|S]  |  Aug 27, 2026(01:03:44) The Incident at Galley House  |  [Linux, Mac, PC (Microsoft Windows)]  |  Jul 14, 2026(01:06:34) The Blood of Dawnwalker  |  [PC (Microsoft Windows), PlayStation 5, Xbox Series X|S]  |  Sep 03, 2026(01:08:57) The Lord of the Rings: Return to Moria  |  [PlayStation 5]  |  Dec 05, 2023(01:10:14) Valheim  |  [Nintendo Switch 2, PC (Microsoft Windows), PlayStation 5, Linux, Mac, Xbox One, Xbox Series X|S]  |  Sep 09, 2026(01:11:37) Blood Dungeon  |  [Mac, PC (Microsoft Windows), PlayStation 5, Xbox Series X|S, Nintendo Switch, Nintendo Switch 2]  |  Aug 25, 2026(01:13:27) Second Break(01:14:28) Sony drops Physint! Xbox picks up Physint!(01:33:49) Nintendo reveals more about its Ocarina remake(01:50:35) New Metroid from the Metroid: Dread folks(01:56:59) Blizzard looks to the (far) future(01:58:26) Starcraft returns in 2030(02:06:25) Diablo returns in 2029(02:08:13) World of Warcraft Forever, forever, ever?(02:18:06) Email(s)(02:22:18) Wrapping up and thanks(02:25:08) Mysterious Benefactor Shoutouts(02:26:29) Who out there has been around here for the last 20+ years?! Thanks and celebrate it!(02:28:39) See ya!

    Coder Radio
    658: Games Dev With Grenis

    Coder Radio

    Play Episode Listen Later Sep 17, 2026 24:20


    Alderon Games [Warp](warp.dev) [The Mad Botter Automation Promo](themadbotter.com/start) Discord Coder Conduit

    Rebuild
    432: Market Price 2nm Norimaki (hak)

    Rebuild

    Play Episode Listen Later Sep 17, 2026 184:48


    Hakuro Matsuda さんをゲストに迎えて、iPhone 18 Pro, iPhone Duo, Apple Watch, GPT-6 Astra, Muse, Pixel 11 Pro, 東京ゲームショウなどについて話しました。 Show Notes Apple unveils a more powerful Mac mini featuring the all-new M6 and M5 Pro iPhone 18 Pro and iPhone 18 Pro Max iPhone 18 Pro and iPhone 18 Pro Max Have Different Modems Ricoh GR IIIx Apple has a new way to prove your iPhone photos aren't AI slop Google: Pixel and Android trust with C2PA Content Credentials Set up iPhone Handoff Apple Upgrade Buy AirPods 5 Apple Watch Series 12 announced Apple Watch: Is Live Rewind a privacy nightmare? Meta Glasses: Men wearing them on dates is a red flag Be My Eyes Apple unveils iPhone Duo Apple: Why it waited 10 years to make a foldable phone (John Ternus & Joz) Design for iPhone Duo GPT-6 Astra Blender Anthropic CEO outlines plan to slow AI development OpenAI: AI agents hijacked a German wiki to escape sandbox Google: Demis Hassabis steps down from DeepMind CEO role Google: Jeff Dean and top AI researchers leaving to launch startup TheJeffDeanFacts Muse Grok Bot OpenAI: On the Navier–Stokes Millennium Prize Problem Google: Maps entire brain of fruit fly, software engineers make it run Doom Google: Pixel 11, Pixel 11 Pro and Pixel 11 Pro XL 東京ゲームショウ2026 PlayStation: Physical disc production ending in January 2028 Wii&PS3解禁!?生配信 ルックバック ラスト・サバイバー 劇場版 魔法少女まどか☆マギカ〈ワルプルギスの廻天〉 You Can See Everything JR東日本: Suica新キャラクターデザイン投票 続ポートピア連続殺人事件:忘却の埋葬 ポートピア連続殺人事件:80年代実写化風 AIが考えるおばあちゃんゲーム Grand Theft Auto VI レベルファイブ: 「LEVEL5 VISION 2026 II」PV一部修正と今後の生成AI対応について Linus Torvalds to critics of AI coding in Linux: "Fork it. Or just walk away."

    Voices of VR Podcast – Designing for Virtual Reality
    #1789: Steam Frame VR HMD Peripheral Developed by Arcturus Vision Camera

    Voices of VR Podcast – Designing for Virtual Reality

    Play Episode Listen Later Sep 15, 2026 52:20


    Valve's Steam Frame was finally announced today, and initial lottery to have an opportunity to buy one will be open until September 17th. Some of the biggest complaints are around the low quality, monochrome, default mixed reality passthrough cameras. But Valve collaborated with the Arcturus Vision Camera team to add a port that would enable their $149 camera peripheral to the Steam Frame in order to upgrade the mixed reality passthrough quality. I had a chance to speak with founder Jeff Powers to get some more context on the camera, the additional Gaussian Splatting capture capabilities, and the process of collaborating with Valve to add this port. I watched a lot of reviews today, and I enjoyed getting an overview from Tested's Norm Chan and Linus Tech Tips, VR developer Anton Hand had some pointed critiques around the Steam Frame controllers quality prioritizing 2D games over VR games, Gamers Nexus had some innovative tests comparing controller latency and privacy policies, Cas and Chary XR did a great overview from a VR gaming and gaming POV, CNET's Scott Stein's brief review focused on the Steam Frame as a standalone device (however most of the feedback is that it is more of streaming device for PCVR), and ThrillSeeker dug into a bit more of how the Steam Frame is a proper Linux machine with a lot of optimism around that it means to finally have an XR device that is a rootable open platform. A lot of the feedback was that the ergonomic kits and charger were also likely add-ons that were suggested to make the device even more comfortable via the top strap, and a charger that could be plugged in as the battery life was also not the greatest, but could be solved with a portable battery charger. The default camera is also not the greatest, and so given all of the things that need updates, then going with the smaller 256GB version rather than the 1TB version means the price difference could be applied to these peripherals to make the overall experience better. And listen to Powers talk about the Steam Frame, what the Arcturus Vision Camera provides, and what types of additional volumetric capture solutions they've been developing for it. This is a listener-supported podcast through the Voices of VR Patreon. Music: Fatality

    LINUX Unplugged
    684: You Ain't Ready For This Jelly

    LINUX Unplugged

    Play Episode Listen Later Sep 14, 2026 67:48 Transcription Available


    We rebuild our media stack around Jellyfin 12, survive a major upgrade, and add self-hosted tools that make sharing with family and friends much better.Sponsored By:Jupiter Party Annual Membership: Put your support on automatic with our annual plan, and get one month of membership for free!Managed Nebula: Meet Managed Nebula from Defined Networking. A decentralized VPN built on the open-source Nebula platform that we love.Support LINUX UnpluggedLinks:Web Boost — Send us a boost via sats or USD

    Late Night Linux
    Late Night Linux – Episode 403

    Late Night Linux

    Play Episode Listen Later Sep 14, 2026 29:04


    We answer your questions including the last new distro we tried, what distro we’d switch to if ours went away, whether we should limit our computer use, our vision for the Linux desktop, and our unpopular Linux opinions. Tailscale Tailscale is a modern, secure, zero-trust connectivity platform. Tailscale's personal plan will be free forever for up to 6 users and unlimited devices! No credit card required. Go to tailscale.com/lnl to get started. If you’d like to bring Tailscale to work, use code latenightlinux for three free months of any paid plan. Support us on patreon and get an ad-free RSS feed with some early episodes See our contact page for ways to get in touch. RSS: Subscribe to the RSS feeds here