Last month, I gave access to production to one of the AI agent to aspiring AI Engineer | Architect in my team. He was copying part of the code and accidentality remove portion of it. We figure few days later, when we had bigger load of incidents in queue. 🙃 This was my mistake. I gave keys to production to inexperienced person and it had huge impact on our operations. Also...I'm monitoring cost, usage, but...not how many incidents were investigated by each agent. Now I'm in the process of designing complete approach change how to use AI agents in production. The functionalities are great, the customisation brilliant. It's time to take it one step further. In the upcoming weeks I will spend some save guards to identify this kind of problems in the future. It should also cover testing different AI models and ranking them, how they score based on your unique data. My goal is to have AI orchestration agent, who would call the other ones specialized in specific tasks. This orchestration agent would monitor them all. At the same time, you could add new functionalities easily.