What We’re Seeing

For the past several years, the dominant narrative in enterprise AI has been a race for scale. The prevailing wisdom suggested that larger, more powerful monolithic models were the inevitable path to solving complex business problems. We are now seeing clear, quantitative evidence from the field that this assumption is not just being challenged, but overturned. A new pattern is emerging: enterprises are achieving superior results by moving away from single, giant models and toward smaller, coordinated systems built on a principle of agentic orchestration.

Compelling proof of this shift comes from a recent paper published by researchers at a major accommodation marketplace, detailing their production AI for customer support. The paper, From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational Assistants at Scale, documents a migration from a large, all-in-one model to an agentic architecture. This new system uses a smaller orchestrator model to intelligently call a set of specialized, reliable tools. The business impact was not incremental; it was a step-change in performance and reliability.

The Number That Changes Everything

45%

The reduction in critical customer support escalations to human agents after replacing a large monolithic AI with a smaller, tool-using agentic system.


Who’s Ahead and Why

The market is bifurcating. On one side are organizations still focused on the raw capabilities of the largest foundation models, often getting stuck in pilot phases where unpredictable outputs and hallucinations prevent deployment in mission-critical workflows. On the other side are the leaders who have recognized that for most enterprise tasks, reliability and control trump raw conversational ability. These forward-thinking teams are winning because they are building systems of systems, not just bigger black boxes.

The agentic approach is gaining ground because it directly addresses the core weaknesses of monolithic models in a business context. Instead of asking one model to know everything and do everything, it asks a smaller, more efficient orchestrator model to do one thing well: understand user intent and route the task to the correct, deterministic tool. This could be an API call to a CRM, a database query for order status, or a function that processes a refund. Because each tool is specialized and verifiable, the system as a whole becomes more predictable, auditable, and less prone to factual invention. As documented in a recent McKinsey report on AI adoption, the value of AI is realized when it is embedded reliably into core business processes, a task for which agentic systems are far better suited.


The Gap Most Teams Miss

Many enterprise AI initiatives stall because they misdiagnose the central challenge. They treat the problem as a failure of the LLM—believing a better model or a more clever prompt will solve their issues. Consequently, they invest heavily in evaluating the next generation of foundation models, hoping for a silver bullet. The real gap, however, is not in the model but in the architecture. The fundamental challenge is one of systems engineering, not just applied AI.

The gap most teams miss is the distinction between a model’s capability and a system’s reliability. A large model might be capable of drafting an email, analyzing a contract, and looking up a customer record in a single prompt, but it cannot do so with the 99.9% reliability required for a core business workflow. The hard, often unglamorous work lies in building the robust, well-documented library of tools that an agent can use. This involves API design, data governance, access control, and observability—the foundational elements of enterprise software development that must be adapted for an AI-driven world.

This is why we see a significant difference between impressive demos and production-grade systems. The demo showcases the model; the production system showcases the architecture. Closing this gap requires a strategic pivot from model-centric exploration to system-centric engineering. Thinkia’s approach to Agentic AI Implementation focuses on building this robust tooling and orchestration layer, ensuring that AI systems deliver measurable business value safely and reliably.


How to Close the Gap

Transitioning from a monolithic to an agentic mindset requires a deliberate, structured approach. It’s a shift in both technology and strategy. First, teams must stop asking, “What can this model do?” and start asking, “What business process can we deconstruct into a series of reliable, tool-based steps?” This reframing is the critical first move.

Second, the focus of investment must shift. Instead of dedicating the majority of the budget to LLM inference costs and prompt engineering, resources should be reallocated to building and maintaining a library of high-quality, internal and external tools. This is the core intellectual property of an agentic system. An effective AI Strategy & Roadmap will prioritize the development of this tool ecosystem as a foundational enterprise asset.

Finally, success requires a new set of skills. Teams need not only AI specialists but also strong software architects, API designers, and DevOps engineers who can build and manage these complex, distributed systems. The goal is to create a resilient system where the orchestrator can fail, a tool can be updated, or a new capability can be added without bringing the entire process to a halt.

Maturity LevelCurrent StateNext ActionTimeline
ExploringUsing public web UIs (e.g., ChatGPT) for ad-hoc, isolated tasks.Identify a single, high-value, structured business process to target for a pilot.1-2 months
PilotingBuilding a proof-of-concept using a single, large monolithic model with complex prompts.Deconstruct the target process into discrete steps and map them to potential API/tool calls.3-6 months
ScalingA monolithic model is in production for a narrow use case but struggles with accuracy and cost.Begin developing a formal, version-controlled tool library and a basic orchestration layer.6-12 months
OptimisingAn agentic system is in production, delivering reliable results for a core process.Implement advanced observability to monitor tool performance and automate error recovery.Ongoing

Watch These Signals

As this trend accelerates, enterprise leaders should monitor several key indicators to stay ahead:

  • Rise of Tool-Use Benchmarks: Watch for a shift in industry benchmarks away from pure knowledge tests (like MMLU) toward evaluations that measure a model’s ability to reason and reliably use a complex set of tools to accomplish tasks (e.g., ToolBench).
  • Maturation of Orchestration Frameworks: Keep an eye on the enterprise-readiness of open-source and commercial frameworks designed to build, test, and deploy agentic systems. Increased focus on security, governance, and observability will signal maturity.
  • Performance of Smaller, Specialized Models: The utility of massive, generalist models for orchestration will decline. Track the function-calling and reasoning capabilities of smaller, fine-tuned models that offer a much better cost-performance ratio for the orchestrator role.

Our Take

We believe the evidence is now clear: for the vast majority of enterprise use cases, the future of AI is not bigger models, but smarter systems. The success of the agentic architecture detailed in the research is not an anomaly; it is a leading indicator of a fundamental shift in how high-performing organizations will build and deploy AI. The pursuit of monolithic, superhuman AI is a fascinating research endeavor, but the path to immediate and sustainable business value lies in the disciplined engineering of agentic orchestration.

This transition requires a new playbook, one that prioritizes control, reliability, and architectural rigor over raw model capability. Organizations that master the art of decomposing complex processes and building robust toolsets will create a durable competitive advantage. As our Enterprise AI Adoption Guide 2025 outlines, making this shift is central to moving from scattered pilots to a scalable, value-generating AI program. Thinkia helps enterprise leaders navigate this transition, building the strategy and technical foundations for the next generation of reliable, high-impact AI systems.