Podcasts about aiops

  • 225PODCASTS
  • 580EPISODES
  • 33mAVG DURATION
  • 5WEEKLY NEW EPISODES
  • Sep 2, 2026LATEST

POPULARITY

20192020202120222023202420252026


Best podcasts about aiops

Latest podcast episodes about aiops

Packet Pushers - Full Podcast Feed
TCG083: Superintelligence for Everyone: Who Actually Holds the Power?

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Sep 2, 2026 53:22


Mark Zuckerberg argues that broadly distributed personal AI, or a “superintelligence” in his parlance, can increase prosperity and counter the risks of AI being controlled by a handful of government and corporate entities. Drew Conry-Murray joins Eyvonne and William to engage in a lively roundtable where they examine whether distributing access meaningfully distributes power when... Read more »

Packet Pushers - Fat Pipe
TCG083: Superintelligence for Everyone: Who Actually Holds the Power?

Packet Pushers - Fat Pipe

Play Episode Listen Later Sep 2, 2026 53:22


Mark Zuckerberg argues that broadly distributed personal AI, or a “superintelligence” in his parlance, can increase prosperity and counter the risks of AI being controlled by a handful of government and corporate entities. Drew Conry-Murray joins Eyvonne and William to engage in a lively roundtable where they examine whether distributing access meaningfully distributes power when... Read more »

Packet Pushers - Full Podcast Feed
TCG082: AI News Roundtable – Copyrights, AI Watermarks, and the Open Weight Debate

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Aug 19, 2026 45:52


William Collins and Eyvonne Sharp dig into the latest AI headlines, from the largest copyright settlement in American history to stolen AI models and invisible watermarks on Claude output. Plus, they discuss why so many companies have rallied around NVIDIA’s support for open weight AI models. Our hosts also examine the biggest questions arising from... Read more »

Packet Pushers - Fat Pipe
TCG082: AI News Roundtable – Copyrights, AI Watermarks, and the Open Weight Debate

Packet Pushers - Fat Pipe

Play Episode Listen Later Aug 19, 2026 45:52


William Collins and Eyvonne Sharp dig into the latest AI headlines, from the largest copyright settlement in American history to stolen AI models and invisible watermarks on Claude output. Plus, they discuss why so many companies have rallied around NVIDIA’s support for open weight AI models. Our hosts also examine the biggest questions arising from... Read more »

Packet Pushers - Full Podcast Feed
TCG081: Network Automation Forum: From Simple Survey to Global Community

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Aug 5, 2026 64:49


William Collins is joined by guest co-host Eric Chou as well as Network Automation Forum founders Scott Robohn and Chris Grundemann to discuss how their community emerged from a simple question: Why haven’t we seen full adoption of network automation, yet? They discuss the growth of AutoCon and how its practitioner-focused, vendor-neutral approach has fostered... Read more »

Packet Pushers - Fat Pipe
TCG081: Network Automation Forum: From Simple Survey to Global Community

Packet Pushers - Fat Pipe

Play Episode Listen Later Aug 5, 2026 64:49


William Collins is joined by guest co-host Eric Chou as well as Network Automation Forum founders Scott Robohn and Chris Grundemann to discuss how their community emerged from a simple question: Why haven’t we seen full adoption of network automation, yet? They discuss the growth of AutoCon and how its practitioner-focused, vendor-neutral approach has fostered... Read more »

UC Today - Out Loud
Is IT Evolving Fast Enough? Service Management, Connectivity & the AI Shift

UC Today - Out Loud

Play Episode Listen Later Aug 4, 2026 34:23


As organisations modernise their digital workplaces, service management and connectivity are no longer separate disciplines – they're converging into a single experience layer that shapes how people work, sell, and serve customers.In this UC Today roundtable, Christopher Carey is joined by Irwin Lazar, President and Principal Analyst at Metrigy; Chethan Visweswar, Chief Product Officer at Movius; John Finch, Global VP at RingCentral; and Mark Bunnell, Chief Operating Officer at Nuwave.Together they explore how the market is evolving, where organisations are falling behind, and what it takes to move from reactive IT to predictive, AI-driven operations.Topics covered include:

Telecom Reseller
ZPE Systems Brings AI-Driven Infrastructure Recovery to the Edge, Podcast

Telecom Reseller

Play Episode Listen Later Jul 28, 2026


By Doug Green “AI goes blind at exactly the moment when you need it most.” In this Technology Reseller News podcast, Vishal Gupta, Director of Product Management at ZPE Systems, explains why AI-driven infrastructure management needs an independent path to the devices it is expected to monitor, troubleshoot and recover. AIOps platforms have become increasingly effective at detecting problems, correlating events and automating routine infrastructure operations. The problem, Gupta says, is that these systems often run on the same production infrastructure they manage. When a network outage or hardware failure occurs, the AI platform can lose both its connection to the affected equipment and access to the telemetry it needs to diagnose the problem. “That is the gap out-of-band fills,” says Gupta. Out-of-band management provides an independent management plane that remains separate from the production network. Even when the primary infrastructure is unavailable, IT teams—and increasingly AI agents—can still reach devices through console access, examine system logs and kernel messages, and take corrective action. Gupta compares the architecture to an airport. Aircraft use the runway for normal operations, while emergency and service vehicles have separate roads and infrastructure. If the runway becomes unavailable, the service infrastructure can still reach the aircraft. The same principle applies to resilient IT operations. An isolated management environment should have its own connectivity, security, routing, switching, storage and compute capabilities. It may also include failover connectivity through 4G, 5G or satellite services such as Starlink. Building Out-of-Band for a Larger Edge ZPE Systems developed its Nodegrid Net Services Router 2U, or NSR 2U, in response to customers operating increasingly large and complex edge environments. These environments can include branch offices, remote facilities, ships, oil rigs, cell sites and other locations outside traditional data centers. They frequently contain more devices, require greater bandwidth and have fewer trained personnel available on-site. The NSR 2U was designed around three priorities: greater capacity, increased resiliency and support for AI workloads. The modular platform offers 10 expansion-card slots, allowing customers to configure the system around their particular deployment. It also includes redundant, field-serviceable power supplies and fans, two NVMe storage slots with RAID support, four native 10-gigabit SFP+ ports and an increased Power over Ethernet budget. ZPE has even addressed the possibility that the out-of-band device itself could fail. Two NSR 2U systems can be interconnected so that one system can provide remote console, power and reset control for the other—effectively providing out-of-band management for the out-of-band infrastructure. Taking NVIDIA Jetson AI to Remote Locations ZPE Systems has also developed an NVIDIA Jetson AI Expansion Card for the Nodegrid NSR family. The card supports NVIDIA Jetson Orin Nano and Orin NX modules, providing local AI processing within the isolated management environment. This allows organizations to deploy AI agents close to the infrastructure and data they manage, without relying entirely on a remote cloud connection. A key capability is remote lifecycle management. IT teams can remotely flash the Jetson operating system, deploy or update AI agents and models, access the console, and power the device on, off or into recovery mode. Ordinarily, updating or recovering an edge AI device may require someone to travel to the location and connect directly to the hardware. ZPE's approach is intended to reduce those truck rolls while allowing organizations to manage distributed AI infrastructure centrally. Potential applications extend beyond AIOps. The platform can support real-time video analytics, object detection, smart recording, manufacturing quality control, sensor-data aggregation and local automation. GPIO and I2C interfaces also allow sensors measuring conditions such as temperature, vibration or voltage to feed information directly into locally running AI models. Asking the Hard Infrastructure Questions Gupta says much of the AI conversation remains focused on models, software and the token economy. Those areas are important, but they can obscure fundamental infrastructure questions. Where will an AIOps platform run? Can it survive the outage it is expected to resolve? Will it still have a path to the affected equipment? Can it access sufficiently accurate data to diagnose the problem and select the right recovery action? “If you can't answer these questions, then there's a gap in your AIOps strategy,” Gupta says. “No software and no model will fix it for you.” As AI becomes more autonomous, infrastructure resilience will determine whether AI agents can move beyond identifying failures to actually recovering from them. ZPE Systems is positioning isolated out-of-band infrastructure, the NSR 2U and edge-based Jetson AI processing as the foundation for making that transition possible. More at Enterprise Network Management Solution | ZPE Systems

Packet Pushers - Full Podcast Feed
TCG080: Skills Over MCP and More

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jul 22, 2026 56:17


What if your MCP server shipped with its own manual? Angie Jones, VP of Developer Experience at the Agentic AI Foundation, joins William and Eyvonne to break down the Skills Over MCP working group effort, which delivers Agent Skills through MCP’s existing resources primitive (think voice over IP, not skills versus MCP). Angie shares her... Read more »

Packet Pushers - Fat Pipe
TCG080: Skills Over MCP and More

Packet Pushers - Fat Pipe

Play Episode Listen Later Jul 22, 2026 56:17


What if your MCP server shipped with its own manual? Angie Jones, VP of Developer Experience at the Agentic AI Foundation, joins William and Eyvonne to break down the Skills Over MCP working group effort, which delivers Agent Skills through MCP’s existing resources primitive (think voice over IP, not skills versus MCP). Angie shares her... Read more »

The New Stack Podcast
Meet Brain, the AI that decides when Azure is officially down

The New Stack Podcast

Play Episode Listen Later Jul 14, 2026 19:27


In this episode, Mark Russinovich, CTO of Microsoft Azure revealed Brain, the AI-powered AIOps system that continuously monitors Azure's health, detects incidents, identifies root causes, and increasingly automates responses such as pausing problematic deployments and notifying affected customers. Built on Azure Resource Graph, Brain creates a real-time digital twin of Azure, mapping dependencies across hundreds of services, data centers, and regions. Although Brain predates the generative AI boom, years of data engineering, standardized service-level indicators (SLIs), and machine learning laid the foundation for today's capabilities.  Brain combines standardized SLIs, service-specific monitoring, and third-party signals to detect anomalies, while ML models dynamically establish service baselines and correlate outages with software rollouts. Microsoft says automated notifications have reduced customer support tickets by four to six times, with 80–90% of Brain-covered services receiving notifications within 15 minutes, often in under five. The company is also layering LLM-powered agents, called Triangle, on top of Brain to streamline incident routing and eventually enable AI agents to autonomously troubleshoot and remediate outages.   Learn more from The New Stack around the latest in Microsoft Azure: Meet Brain, the AI that decides when Azure is officially down  Microsoft's pitch to enterprises: Ditch Azure Repos for GitHub, despite its rocky reliability record  Join our community of newsletter subscribers to stay on top of the news and at the top of your game. 

Packet Pushers - Full Podcast Feed
TCG079: Why Your State File is Actually a Distributed Systems Problem

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jul 1, 2026 47:39


Malcolm Matalka joins William and Eyvonne to challenge the narrative that Infrastructure as Code (IaC) is dead. Malcolm argues that the real value of IaC was never the syntax, but state and governance. Together they examine whether the state was a file problem at all, or a distributed systems problem in a JSON costume. Episode... Read more »

Packet Pushers - Fat Pipe
TCG079: Why Your State File is Actually a Distributed Systems Problem

Packet Pushers - Fat Pipe

Play Episode Listen Later Jul 1, 2026 47:39


Malcolm Matalka joins William and Eyvonne to challenge the narrative that Infrastructure as Code (IaC) is dead. Malcolm argues that the real value of IaC was never the syntax, but state and governance. Together they examine whether the state was a file problem at all, or a distributed systems problem in a JSON costume. Episode... Read more »

Packet Pushers - Full Podcast Feed
TCG078: The Pope's AI Encyclical: Navigating AI with Values

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jun 17, 2026 50:59


The Pope issued a recent encyclical on AI, urging developers to safeguard human agency in the age of artificial intelligence. Eyvonne and William explore this encyclical, moving beyond the headlines to the core message regarding human dignity. They examine how the document provides a values-based framework for evaluating technology and the need for a balanced... Read more »

Packet Pushers - Fat Pipe
TCG078: The Pope's AI Encyclical: Navigating AI with Values

Packet Pushers - Fat Pipe

Play Episode Listen Later Jun 17, 2026 50:59


The Pope issued a recent encyclical on AI, urging developers to safeguard human agency in the age of artificial intelligence. Eyvonne and William explore this encyclical, moving beyond the headlines to the core message regarding human dignity. They examine how the document provides a values-based framework for evaluating technology and the need for a balanced... Read more »

Telecom Reseller
Grokstream on L1 Agent and the Path to Autonomous Network Operations, Podcast

Telecom Reseller

Play Episode Listen Later Jun 15, 2026 7:06


By Doug Green “We're absolutely on the path, and we're not talking five, six, seven years. We're talking in the next 18 to 24 months.” In this episode of the Technology Reseller News podcast, Doug Green speaks with Josh Kindiger, COO and co-founder of Grokstream, about the company's new L1 Agent and what it means for the future of AI-driven network and IT operations. Grokstream is the company behind Grok, an AI-powered predictive agent platform for network and IT operations. The platform comes out of the event intelligence and AIOps space and is designed to help operations teams identify, triage, and resolve recurring issues more efficiently. Kindiger says Grokstream recently released its first role-based agent, the L1 Agent, in beta. The full production release is expected in Q2. The agent is already being used with customers to prove out real-world capabilities. Because many organizations remain cautious about AI-driven automation, Grokstream is starting with low-risk, repeatable use cases. In many operations centers, Kindiger notes, the same incidents occur repeatedly, sometimes accounting for as much as 70% of activity. The L1 Agent is designed to recognize those patterns and guide operators through triage and resolution. For example, if a recurring issue requires a service restart, the system can recommend or automate that step. If a pattern points to a commercial power outage at a site, the agent can help avoid unnecessary dispatches while monitoring backup power systems. Kindiger says the goal is not to remove human oversight immediately, but to build trust through guardrails, staged automation, and operator control. Low-risk automations can be handled end to end, while higher-risk actions may require human approval. The podcast also explores the broader opportunity for enterprises, MSPs, and CSPs. Kindiger says service providers and managed service providers face growing pressure to improve efficiency, reduce costs, and differentiate in competitive markets. AI-driven operations can help them respond faster, lower manual workload, and deliver better service outcomes. The long-term direction is clear: autonomous network operations are coming. Kindiger says companies should begin now because foundational work is needed before they can fully benefit from automation. For MSPs and CSPs, he says the urgency is even greater. Cost pressure is shaping renewals and new customer wins, and AI-powered operations may become a competitive advantage. Learn more at www.grokstream.com

Packet Pushers - Full Podcast Feed
TCG077: News Roundtable: Data Center Backlash and the AI Chip War

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jun 3, 2026 45:09


William and Eyvonne discuss recent tech news, including the growing political and community opposition to AI data centers driven by fears over power and water usage. They also analyze the “AI Chip War” as hyperscalers such as AWS and Google invest in specialized silicon for training and inference.  Episode Links: Amid backlash, O'Leary Digital CEO... Read more »

Packet Pushers - Fat Pipe
TCG077: News Roundtable: Data Center Backlash and the AI Chip War

Packet Pushers - Fat Pipe

Play Episode Listen Later Jun 3, 2026 45:09


William and Eyvonne discuss recent tech news, including the growing political and community opposition to AI data centers driven by fears over power and water usage. They also analyze the “AI Chip War” as hyperscalers such as AWS and Google invest in specialized silicon for training and inference.  Episode Links: Amid backlash, O'Leary Digital CEO... Read more »

Telecom Reseller
Grokstream: Predictive and Agentic AI Moves IT Operations Toward Self-Healing, Podcast

Telecom Reseller

Play Episode Listen Later Jun 1, 2026


Grokstream: Predictive and Agentic AI Moves IT Operations Toward Self-Healing, Podcast, Grokstream's platform is designed to operate from signals, not noise. The system fuses telemetry across domains, learns continuously from operational data and human feedback, and creates a unified source of truth for IT operations. That allows teams to move beyond correlation and toward understanding what is happening, why it is happening and what should be done next. By Doug Green Grokstream says the next generation of IT operations will not be built around more dashboards, more rules, or faster alert routing. It will be built around AI that can learn, reason, remember, recommend and eventually act with governed autonomy. “Agentic AI must be governed by design,” said Josh Kindiger, CEO of Grokstream. “Predictive intelligence is powerful, but safe, explainable autonomy is what drives real adoption.” In this Technology Reseller News podcast, Doug Green speaks with Josh Kindiger, Co-Founder and COO of Grokstream, about how the company is helping MSPs, CSPs and enterprise IT organizations move from reactive operations toward predictive, self-healing IT environments. The conversation comes as Grokstream advances its Grok L1 Agent, a new role-based agent designed for frontline IT operations teams. The L1 Agent is intended to reduce alert noise before incidents reach the queue, provide intelligent summaries, identify likely root causes, recommend next-best actions and trigger approved remediations inside tools such as Slack, Microsoft Teams and existing IT workflows. For service providers and enterprise operations teams, the problem is familiar. More tools often mean more alerts, but not necessarily more clarity. Traditional rules-based AIOps platforms can help with deduplication and routing, but they often stop short of true incident compression, causal reasoning and prevention. Grokstream is taking a different approach by combining classical machine learning, causal intelligence and generative AI into a single cognitive AI layer. Kindiger explains that Grokstream's platform is designed to operate from signals, not noise. The system fuses telemetry across domains, learns continuously from operational data and human feedback, and creates a unified source of truth for IT operations. That allows teams to move beyond correlation and toward understanding what is happening, why it is happening and what should be done next. A central theme of the podcast is the difference between AI that summarizes and AI that reasons. Grokstream argues that true agentic AI is not simply an LLM attached to a workflow. It requires memory, context, policy guardrails, procedural intelligence and the ability to improve over time. In Grokstream's model, agents begin as assisted tools, then move toward trusted operators and eventually toward predictive autonomous systems. The first practical on-ramp is the L1/NOC environment, where many organizations see the fastest measurable impact. Grokstream says its approach can deliver 2–3x more incident compression beyond traditional deduplication and rules-based correlation, while reducing L1 workload by more than 50% through noise compression, guided resolution and fewer unnecessary escalations. The timing is significant. Grokstream recently announced that Cirion Technologies selected the Cognitive Grok AI platform to support AI-driven predictive operations across Latin America's digital infrastructure. That deployment highlights the growing demand for systems that can detect emerging issues across network, transport and infrastructure layers before customer-facing impact occurs. For MSPs, CSPs and enterprise IT leaders, the message is clear: operational scale cannot be achieved simply by adding more people or more monitoring tools. The next step is an intelligence layer that can unify data, predict impact, explain cause and support governed automation. Grokstream is positioning Grok as that layer: a predictive and agentic AI platform that helps operations teams reduce noise, prevent incidents, improve engineer experience and move toward self-healing IT operations. Learn more at https://grokstream.com/ Related Grokstream Stories on Telecom Reseller Grokstream's Cognitive Grok® AI Platform Selected by Cirion Technologies to Power AI-Driven, Predictive Operations Across Latin America's Digital Infrastructure https://telecomreseller.com/2026/05/20/grokstreams-cognitive-grok-ai-platform-selected-by-cirion-technologies-to-power-ai-driven-predictive-operations-across-latin-americas-digital-infrastructure/ Grokstream Announces Grok® L1 Agent to Advance Predictive and Agentic AI for IT Operations https://telecomreseller.com/2026/04/06/grokstream-announces-grok-l1-agent-to-advance-predictive-and-agentic-ai-for-it-operations/ More Grokstream coverage on Telecom Reseller https://telecomreseller.com/?s=grokstream/

Packet Pushers - Full Podcast Feed
TCG076: Packet Pushers Assemble! Bridging the Telemetry Divide

Packet Pushers - Full Podcast Feed

Play Episode Listen Later May 20, 2026 56:25


Today our Packet Pushers team assembles to discuss whether the grass is greener on the NetOps or DevOps side of the telemetry fence. William of The Cloud Gambit, Scott of Total Network Operations, and Ned and Kyler of Day Two DevOps discuss the difficulties and differences of getting telemetry and state from devices across different... Read more »

Packet Pushers - Fat Pipe
TCG076: Packet Pushers Assemble! Bridging the Telemetry Divide

Packet Pushers - Fat Pipe

Play Episode Listen Later May 20, 2026 56:25


Today our Packet Pushers team assembles to discuss whether the grass is greener on the NetOps or DevOps side of the telemetry fence. William of The Cloud Gambit, Scott of Total Network Operations, and Ned and Kyler of Day Two DevOps discuss the difficulties and differences of getting telemetry and state from devices across different... Read more »

PurePerformance
Observability in the AI‑Native Era with Hilliary Lipsig and Rob Rati

PurePerformance

Play Episode Listen Later May 11, 2026 51:30


As the software world is transforming from cloud native to AI-native, observability must transform with it. But how exactly? How do we apply this in an existing enterprise with established processes and practices?In this PurePerformance episode, Andi Grabner hosts Hilliary Lipsig and Rob Rati to discuss their new book, Observability in the AI‑Native Era. The conversation explores how AIOps, automation, and modern observability must evolve as systems become cloud‑native, data‑heavy, and AI‑driven.We talk about why old alerting and SLO models no longer scale, how to balance AI with automation and human judgment, and why trust, security, and compliance matter more than ever when machines start making operational decisions. A must‑listen for SREs, platform engineers, and engineering leaders navigating the AI‑native future.Links we discussedBook on Amazon: https://www.amazon.com/Observability-AI-Native-Era-Artificial-Intelligence-ebook/dp/B0GHZH1YFLHilliary LinkedIn: https://www.linkedin.com/in/hilliary-lipsig-a5935245/Rob LinkedIn: https://www.linkedin.com/in/roberthrati/Andi LinkedIn: https://www.linkedin.com/in/grabnerandi/

Packet Pushers - Full Podcast Feed
TCG075: Say the Thing: How the Network Automation Conference Circuit Shaped One SP Operator's Voice

Packet Pushers - Full Podcast Feed

Play Episode Listen Later May 6, 2026 48:53


Eyvonne and William sit down with Joseph Nicholson, a Network Operations Engineer with NTT DATA, to share how public speaking transformed his career and technical experience. Joseph went from a terrifying ten minute lightning talk at AutoCon 2 to presenting 45-minute sessions at conferences like NANOG. Together they discuss how conversations in conference halls influenced... Read more »

Packet Pushers - Fat Pipe
TCG075: Say the Thing: How the Network Automation Conference Circuit Shaped One SP Operator's Voice

Packet Pushers - Fat Pipe

Play Episode Listen Later May 6, 2026 48:53


Eyvonne and William sit down with Joseph Nicholson, a Network Operations Engineer with NTT DATA, to share how public speaking transformed his career and technical experience. Joseph went from a terrifying ten minute lightning talk at AutoCon 2 to presenting 45-minute sessions at conferences like NANOG. Together they discuss how conversations in conference halls influenced... Read more »

Packet Pushers - Full Podcast Feed
TCG074: From SOAR to Agents: Why Practical Automation Has to Survive Contact with Real Infrastructure

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Apr 22, 2026 44:49


Eyvonne Sharp and William Collins speak with Sif Baksh, Principal Solutions Architect at Tines, to discuss the power of automation. Sif shares some personal stories of how he has been able to use automation to innovate and modernize networking operations. They also discuss the importance of learning AI and using it as a tool, how... Read more »

Packet Pushers - Fat Pipe
TCG074: From SOAR to Agents: Why Practical Automation Has to Survive Contact with Real Infrastructure

Packet Pushers - Fat Pipe

Play Episode Listen Later Apr 22, 2026 44:49


Eyvonne Sharp and William Collins speak with Sif Baksh, Principal Solutions Architect at Tines, to discuss the power of automation. Sif shares some personal stories of how he has been able to use automation to innovate and modernize networking operations. They also discuss the importance of learning AI and using it as a tool, how... Read more »

Business of Tech
Metered AI and Variable Output Are Shifting MSP Accountability and Margin Risks

Business of Tech

Play Episode Listen Later Apr 21, 2026 11:35


The episode identifies a structural shift in the integration of generative AI within organizational workflows: variable cost models, unpredictable output quality, and heightened accountability requirements are converging to reshape managed services operations. This shift is exemplified by Anthropic's move toward usage-based pricing for Claude Enterprise, combining compute consumption with per-user fees, and by reports of major enterprises and intelligence agencies piloting dedicated cybersecurity-focused generative AI models. These trends expose IT service providers, especially MSPs, to cost volatility, operational risk, and new governance challenges as generative AI transitions from experimental implementation to core workflow tooling. Primary evidence includes Anthropic's revised pricing strategy, which replaces predictable licensing with usage-based billing, introducing financial unpredictability for heavy users. The episode cites reporting from The Verge and The Guardian, noting that AI-generated outputs can create hidden labor through the need for manual review and corrections, while undetected errors escalate into operational disputes and rework. The implementation of generative AI in security-sensitive environments underscores the need to scrutinize how AI-driven processes are metered and governed. Supporting developments reinforce this shift: MSP platform providers such as Enable are embedding generative AI directly into operational workflows, connecting third-party tools to live data. This creates the need for controls over what AI systems can access, approve, and log, particularly in multi-tenant environments. Meanwhile, outcome-based service agreements—such as fixed response-time SLAs—set new client expectations for measurable performance and accountability in AI operations. The market is also rewarding those who wrap unmanaged technology surfaces, like BYOD or AI tooling, with enforceable policies and auditable evidence trails. Operational implications for MSPs include increased pressure on margins due to AI's variable usage costs colliding with fixed-fee contracts, the challenge of capturing and reporting hidden labor from AI output review, and the necessity for evidence-based governance. Service providers unable to implement and sell AI operations management (“AIOps”) as a billable, controlled service risk becoming de facto shock absorbers for unpriced spend, rework, and disputes. Those who standardize on enforceable budgets, approval gates, audit trails, and compliance-ready reporting stand to protect service margins and reduce liability exposure. 00:00 AI Cost Reckoning 02:39 AI Governance Gap 04:44 Govern or Lose 07:12 Why Do We Care?  Supported by:  TimeZest Zero Networks 

Packet Pushers - Full Podcast Feed
TCG073: From Vibes to Governed: What Building a Real Network Agent Reveals About Spec-Driven Development

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Apr 8, 2026 57:49


Vibe coding: give AI a description of what you want, the model writes the code, you ship it, and then you hope for the best. It works great for side projects, but it can fall apart the moment you point an AI agent at production infrastructure. Today, William and Eyvonne sit down with John Capobianco,... Read more »

Packet Pushers - Fat Pipe
TCG073: From Vibes to Governed: What Building a Real Network Agent Reveals About Spec-Driven Development

Packet Pushers - Fat Pipe

Play Episode Listen Later Apr 8, 2026 57:49


Vibe coding: give AI a description of what you want, the model writes the code, you ship it, and then you hope for the best. It works great for side projects, but it can fall apart the moment you point an AI agent at production infrastructure. Today, William and Eyvonne sit down with John Capobianco,... Read more »

Web and Mobile App Development (Language Agnostic, and Based on Real-life experience!)
AIOps and Modern IT Operations: Simplifying Multi-Cloud Operations (feat. Michael Nappi)

Web and Mobile App Development (Language Agnostic, and Based on Real-life experience!)

Play Episode Listen Later Apr 8, 2026 58:49


In this episode, Michael Nappi, Chief Product and Engineering Officer at ScienceLogic, shares insights into AI Ops, its role in modern IT management, and how it helps large enterprises and MSPs streamline their infrastructure monitoring and management. Discover how AI-driven automation and observability are transforming IT operations.

MLOps.community
Operationalizing AI Agents: From Experimentation to Production // Databricks Roundtable

MLOps.community

Play Episode Listen Later Mar 30, 2026 61:13


Databricks Roundtable episode: Operationalizing AI Agents: From Experimentation to Production. Join the Community: https://go.mlops.community/YTJoinInGet the newsletter: https://go.mlops.community/YTNewsletterMLOps GPU Guide: https://go.mlops.community/gpuguideBig shout-out to Databricks for the collaboration!// AbstractThis panel discusses the real-world challenges of deploying AI agents at scale. The conversation explores technical and operational barriers that slow production adoption, including reliability, cost, governance, and security.The panelists also examine how LLMOps, AIOps, and AgentOps differ from traditional MLOps, and why new approaches are required for generative and agent-based systems. Finally, experts define success criteria for GenAI frameworks, with a focus on robust evaluation, observability, and continuous monitoring across development and staging environments.// BioSamraj MoorjaniSamraj is a software engineer working on the Agent Quality team. Previously, Samraj worked at Meta on ads/product classification research and AppLovin on MLOps. Samraj graduated with a BS+MS in Computer Science from UIUC, advised by Professor Hari Sundaram, where he worked on controllable natural language generation to produce appealing, interpretable science to combat the spread of misinformation. He also worked with Professor Wen-mei Hwu on accelerating LLM inference through extreme sparsification.Apurva MisraApurva is an AI Consultant at Sentick, focusing on assisting startups with their AI strategy and building solutions. She leverages her extensive experience in machine learning and a Master's degree from the University of Waterloo, where her research bridged driving and machine learning, to offer valuable insights. Apurva's keen interest in the startup world fuels her passion for helping emerging companies incorporate AI effectively. In her free time, she is learning Spanish, and she also enjoys exploring hidden gem eateries, always eager to hear about new favourite spots!Ben EpsteinBen was the machine learning lead for Splice Machine, leading the development of their MLOps platform and Feature Store. He is now the Co-founder and CTO at GrottoAI, focused on supercharging multifamily teams and reducing vacancy loss with AI-powered guidance for leasing and renewals. Ben also works as an adjunct professor at Washington University in St. Louis, teaching concepts in cloud computing and big data analytics.Hosted by Adam Becker// Related LinksWebsite: https://www.databricks.com/https://mlflow.org/~~~~~~~~ ✌️Connect With Us ✌️ ~~~~~~~Catch all episodes, blogs, newsletters, and more: https://go.mlops.community/TYExploreJoin our Slack community [https://go.mlops.community/slack]Follow us on X/Twitter [@mlopscommunity](https://x.com/mlopscommunity) or [LinkedIn](https://go.mlops.community/linkedin)] Sign up for the next meetup: [https://go.mlops.community/register]MLOps Swag/Merch: [https://shop.mlops.community/]Connect with Demetrios on LinkedIn: /dpbrinkmConnect with Samraj on LinkedIn: /samrajmoorjani/Connect with Apurva on LinkedIn: /apurva-misra/Connect with Ben on LinkedIn: /ben-epstein/Connect with Adam on LinkedIn: /adamissimo/Timestamps:[00:00] Introduction[02:30] AI Agents in Operations[04:36] AI Strategy Consulting[05:30] Agent Quality Focus[06:17] AI Agent Expectations[11:44] AI Use Cases Evolution[15:25] Agent Expectations Adjustment[17:41] Agent Quality Monitoring[23:22] Trust in GenAI Systems[33:33] Data Prep vs Product Thinking[40:27] Quality Systems Distinction[44:54] Q & A[1:00:57] Wrap up

Packet Pushers - Full Podcast Feed
TCG072: AI and the Automation Engineer – When Your Scripts Start Writing Themselves

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Mar 25, 2026 49:39


William Collins and Eyvonne Sharp invite Skylar Sands, Senior Automation Engineer at World Wide Technology, to discuss what it means to integrate AI into the daily workflow in a meaningful way. Together they break down the shift in the automation engineer's role now that AI can instantly generate the “toolkit” of Python, Ansible, and Bash,... Read more »

Packet Pushers - Fat Pipe
TCG072: AI and the Automation Engineer – When Your Scripts Start Writing Themselves

Packet Pushers - Fat Pipe

Play Episode Listen Later Mar 25, 2026 49:39


William Collins and Eyvonne Sharp invite Skylar Sands, Senior Automation Engineer at World Wide Technology, to discuss what it means to integrate AI into the daily workflow in a meaningful way. Together they break down the shift in the automation engineer's role now that AI can instantly generate the “toolkit” of Python, Ansible, and Bash,... Read more »

Packet Pushers - Full Podcast Feed
TCG071: Cloud Cloning and Portability – Why Multi-Cloud Freedom Still Requires Translation (Sponsored)

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Mar 18, 2026 34:58


In this sponsored episode, FluidCloud co-founders Sharad Kumar and Harshit Omar sit down with William and Eyvonne to discuss how FluidCloud tackles multi-cloud portability. They detail how FluidCloud acts as a cloning platform that scans an existing cloud or VMware environment, extracts complex infrastructure configurations (including compute and storage, as well as firewall rules and... Read more »

Packet Pushers - Fat Pipe
TCG071: Cloud Cloning and Portability – Why Multi-Cloud Freedom Still Requires Translation (Sponsored)

Packet Pushers - Fat Pipe

Play Episode Listen Later Mar 18, 2026 34:58


In this sponsored episode, FluidCloud co-founders Sharad Kumar and Harshit Omar sit down with William and Eyvonne to discuss how FluidCloud tackles multi-cloud portability. They detail how FluidCloud acts as a cloning platform that scans an existing cloud or VMware environment, extracts complex infrastructure configurations (including compute and storage, as well as firewall rules and... Read more »

Packet Pushers - Full Podcast Feed
TCG070: The Effort Illusion: Why AI Tools Reward Expertise, Not Shortcuts

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Mar 11, 2026 48:27


The tech industry is split between two fantasies  – that AI writes production software while you get coffee, and that everything AI touches is slop. The reality is messier and more interesting: AI tools are force multipliers for people who already know what good looks like, and an expertise amplifier disguised as an easy button. ... Read more »

Packet Pushers - Fat Pipe
TCG070: The Effort Illusion: Why AI Tools Reward Expertise, Not Shortcuts

Packet Pushers - Fat Pipe

Play Episode Listen Later Mar 11, 2026 48:27


The tech industry is split between two fantasies  – that AI writes production software while you get coffee, and that everything AI touches is slop. The reality is messier and more interesting: AI tools are force multipliers for people who already know what good looks like, and an expertise amplifier disguised as an easy button. ... Read more »

Packet Pushers - Full Podcast Feed
TCG069: Viral Predictions, Waterfall's Comeback, and the SaaSpocalypse

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Feb 25, 2026 54:44


William and Eyvonne tackle the biggest AI stories of early 2026. They dissect Matt Schumer’s viral “Something Big is Happening” essay – agreeing professionals need to skill up now while pushing back on the doomsday framing with real-world examples from engineering disciplines. The conversation takes a fascinating turn as Eyvonne draws a parallel between AI-assisted... Read more »

Packet Pushers - Fat Pipe
TCG069: Viral Predictions, Waterfall's Comeback, and the SaaSpocalypse

Packet Pushers - Fat Pipe

Play Episode Listen Later Feb 25, 2026 54:44


William and Eyvonne tackle the biggest AI stories of early 2026. They dissect Matt Schumer’s viral “Something Big is Happening” essay – agreeing professionals need to skill up now while pushing back on the doomsday framing with real-world examples from engineering disciplines. The conversation takes a fascinating turn as Eyvonne draws a parallel between AI-assisted... Read more »

Packet Pushers - Full Podcast Feed
TCG068: Agents and Identity – Navigating What We Can't Predict

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Feb 11, 2026 56:48


We’ve spent a decade figuring out how to (more or less) securely authenticate humans. Now AI agents are crashing the party, and identity just got a whole lot more complicated. Today we sit down with Dan Moore, Senior Director of CIAM Strategy and Identity Standards at FusionAuth, to explore the collision course between artificial intelligence... Read more »

Packet Pushers - Fat Pipe
TCG068: Agents and Identity – Navigating What We Can't Predict

Packet Pushers - Fat Pipe

Play Episode Listen Later Feb 11, 2026 56:48


We’ve spent a decade figuring out how to (more or less) securely authenticate humans. Now AI agents are crashing the party, and identity just got a whole lot more complicated. Today we sit down with Dan Moore, Senior Director of CIAM Strategy and Identity Standards at FusionAuth, to explore the collision course between artificial intelligence... Read more »

The Tech Trek
Cloud Costs vs AI Workloads, The Storage Decisions That Decide Scale

The Tech Trek

Play Episode Listen Later Feb 9, 2026 26:26


Cloud bills are climbing, AI pipelines are exploding, and storage is quietly becoming the bottleneck nobody wants to own. Ugur Tigli, CTO at MinIO, breaks down what actually changes when AI workloads hit your infrastructure, and how teams can keep performance high without letting costs spiral. In this conversation, we get practical about object storage, S3 as the modern standard, what open source really means for security and speed, and why “cloud” is more of an operating model than a place. Key takeaways• AI multiplies data, not just compute, training and inference create more checkpoints, more versions, more storage pressure • Object storage and S3 are simplifying the persistence layer, even as the layers above it get more complex • Open source can improve security feedback loops because the community surfaces regressions fast, the real risk is running unsupported, outdated versions • Public cloud costs are often less about storage and more about variable charges like egress, many teams move data on prem to regain predictability • The bar for infrastructure teams is rising, Kubernetes, modern storage, and AI workflow literacy are becoming table stakes Timestamped highlights00:00 Why cloud and AI workloads force a fresh look at storage, operating models, and cost control 00:00 What MinIO is, and why high performance object storage sits at the center of modern data platforms 01:23 Why MinIO chose open source, and how they balance freedom with commercial reality 04:08 Open source and security, why faster feedback beats the closed source perception, plus the real risk factor 09:44 Cloud cost realities, egress, replication, and why “fixed costs” drive many teams back inside their own walls 15:04 The persistence layer is getting simpler, S3 becomes the standard, while the upper stack gets messier 18:00 Skills gap, why teams need DevOps plus AIOps thinking to run modern storage at scale 20:22 What happens to AI costs next, competition, software ecosystem maturity, and why data growth still wins A line worth keeping“Cloud is not a destination for us, it's more of an operating model.” Pro tips for builders and tech leaders• If your AI initiative is still a pilot, track egress and data movement early, that is where “surprise” costs tend to show up • Standardize around containerized deployment where possible, it reduces the gap between public and private environments, but plan for integration friction like identity and key management • Treat storage as a performance system, not a procurement line item, the right persistence layer can unblock training, inference, and downstream pipelines What's next:If you're building with AI, running data platforms, or trying to get your cloud costs under control, follow the show and subscribe so you do not miss upcoming episodes. Share this one with a teammate who owns infrastructure, data, or platform engineering.

Packet Pushers - Full Podcast Feed
TCG067: Progressive Delivery: Shipping Software is Just the Beginning with Adam Zimman

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jan 28, 2026 55:22


In this episode, we sit down with Adam Zimman, author and VC advisor, to explore the world of progressive delivery and why shipping software is only the beginning. Adam shares his fascinating journey through tech—from his early days as a fire juggler to leadership roles at EMC, VMware, GitHub, and LaunchDarkly – and how those... Read more »

Packet Pushers - Fat Pipe
TCG067: Progressive Delivery: Shipping Software is Just the Beginning with Adam Zimman

Packet Pushers - Fat Pipe

Play Episode Listen Later Jan 28, 2026 55:22


In this episode, we sit down with Adam Zimman, author and VC advisor, to explore the world of progressive delivery and why shipping software is only the beginning. Adam shares his fascinating journey through tech—from his early days as a fire juggler to leadership roles at EMC, VMware, GitHub, and LaunchDarkly – and how those... Read more »

Packet Pushers - Full Podcast Feed
TCG066: How Infrastructure Teams Can Scale Reasoning Without Losing Control with Chris Wade

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Jan 14, 2026 42:15


The industry has pivoted from scripting to automation to orchestration – and now to systems that can reason. Today we explore what AI agents mean for infrastructure with Chris Wade, Co-Founder and CTO of Itential. We also dive into the brownfield reality, the potential for vendor-specific LLMs trained on proprietary knowledge, and advice for the... Read more »

Packet Pushers - Fat Pipe
TCG066: How Infrastructure Teams Can Scale Reasoning Without Losing Control with Chris Wade

Packet Pushers - Fat Pipe

Play Episode Listen Later Jan 14, 2026 42:15


The industry has pivoted from scripting to automation to orchestration – and now to systems that can reason. Today we explore what AI agents mean for infrastructure with Chris Wade, Co-Founder and CTO of Itential. We also dive into the brownfield reality, the potential for vendor-specific LLMs trained on proprietary knowledge, and advice for the... Read more »

Packet Pushers - Full Podcast Feed
TCG065: AutoCon 4 Recap, AI Tools, MCP's First Birthday, and More

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Dec 10, 2025 41:49


In this year-end episode, William and Eyvonne recap their experiences at AutoCon 4 in Austin, Texas. They discuss the conference’s new multi-track format, including Eyvonne’s presentation in the leadership track on why technical projects fail. The conversation dives into how AI tools like Google Gemini can augment – not replace – human creativity, from research... Read more »

Packet Pushers - Full Podcast Feed
TCG064: Governing AI Agents for Real-World Infrastructure (Sponsored)

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Dec 3, 2025 39:12


In this sponsored episode recorded live at AutoCon 4 in Austin, we sit down with Peter Sprygada, Chief Architect at Itential, to discuss Itential’s on-stage announcement of FlowAI. Peter shares his journey from network engineering skeptic to AI advocate, explaining how Itential securely connects AI agents to infrastructure with enterprise-grade governance and traceability. We dive... Read more »

Packet Pushers - Full Podcast Feed
TCG063: Constraint Drives Innovation with John Capobianco

Packet Pushers - Full Podcast Feed

Play Episode Listen Later Nov 26, 2025 53:28


Recorded live at AutoCon4, William Collins and Eyvonne Sharp join forces with John Capobianco for some in the moment thoughts and reflections on the AutoCon experience – from the in-person connections to the workshops to the stage presentations.  John gives us the inside story on his very own workshop and the latest version releases in... Read more »

The CyberWire
Attack of the automated ops. [Research Saturday]

The CyberWire

Play Episode Listen Later Nov 1, 2025 19:40


Today we are joined by Dario Pasquini, Principal Researcher at RSAC, sharing the team's work on WhenAIOpsBecome “AI Oops”: Subverting LLM-driven IT Operations via Telemetry Manipulation. A first-of-its-kind security analysis showing that LLM-driven AIOps agents can be tricked by manipulated telemetry, turning automation itself into a new attack vector. The researchers introduce AIOpsDoom, an automated reconnaissance + fuzzing + LLM-driven telemetry-injection attack that performs “adversarial reward-hacking” to coerce agents into harmful remediations—even without prior knowledge of the target and even against some prompt-defense tools. They also present AIOpsShield, a telemetry-sanitization defense that reliably blocks these attacks without harming normal agent performance, underscoring the urgent need for security-aware AIOps design. The research can be found here: ⁠When AIOps Become “AI Oops”: Subverting LLM-driven IT Operations via Telemetry Manipulation Learn more about your ad choices. Visit megaphone.fm/adchoices