I have spent the last few years at EDC designing agentic AI systems for banks, government entities and large enterprises. The thing I did not expect to learn is that the distance between a demo that works and a system running in production has almost nothing to do with the model.
Walk into most large organizations right now and you will find the same picture. Every department has an AI experiment running. Marketing has something drafting copy. Finance has something reading invoices. Someone in operations has quietly wired a public API into a process that touches customer records. The demos are genuinely good, and the enthusiasm is real.
Then a meeting happens, and someone from information security asks where exactly that customer record went.
Nobody can answer. Not because anything bad happened, but because nobody built the system that would let them answer. The initiative goes quiet, and few months later it is a slide in a lessons-learned deck.