
AI agents are no longer just chatbots. They now execute code, access infrastructure, call APIs, and trigger real production actions. What works perfectly in a demo can quickly become a reliability and security nightmare at scale.
In this talk, Walid Mansia explores the hidden operational challenges behind autonomous agents, MCP servers, and LLM systems: infinite loops, hallucinated actions, exploding token costs, broken automations, and observability gaps traditional SRE tooling cannot explain.
Walid will also share practical patterns for safely operating AI systems in production, including sandboxing, execution limits, least-privilege access, and AI-native observability.
Walid Mansia is an AI Solutions Architect and Platform Engineering specialist focused on autonomous agents, MCP servers, and AI systems in production. He currently leads next-generation AI and data platform initiatives at CANAL+ Group, with previous experience across AWS, DevOps, cloud infrastructure, and large-scale platform engineering. Walid has spent over a decade designing reliable, scalable systems at the intersection of AI, cloud, and SRE.