Given an exception and a set of tools, deciding autonomously which checks to run in what order — prior resolutions for this account, manager notes, anomaly score, predicted category — and proposing a resolution with its reasoning trail.
Tool-using LLM loop with read-only tools auto-executed and write actions gated behind human approval; every step logged
Native tool-use APIs; LangGraph or CrewAI for stateful multi-agent orchestration