Skip to content
Tiatra, LLCTiatra, LLC
Tiatra, LLC
Information Technology Solutions for Washington, DC Government Agencies
  • Home
  • About Us
  • Services
    • IT Engineering and Support
    • Software Development
    • Information Assurance and Testing
    • Project and Program Management
  • Clients & Partners
  • Careers
  • News
  • Contact
 
  • Home
  • About Us
  • Services
    • IT Engineering and Support
    • Software Development
    • Information Assurance and Testing
    • Project and Program Management
  • Clients & Partners
  • Careers
  • News
  • Contact

When it comes to AI, bigger isn’t always better

There is growing concern about trust in AI as the technology is adopted by more people. Large language models (LLMs) continue to face persistent challenges with hallucinations and inaccurate outputs.

LLMs are probabilistic systems trained to give answers, even when the correct answer is unclear or unknowable with the context provided. Humans are more likely to admit that they do not know an answer, especially when there is a financial or reputational consequence at stake, while on the other hand, LLMs are designed to act confidently, no matter what.

Most leading models fall within a 20 to 27 percent range of hallucination rate, making it a persistent and unresolved challenge across current AI systems, because it’s not just an architectural problem, it’s a contextual one.

Enterprise AI can’t afford hallucinations

While consumer AI is typically optimized for scale and creativity, enterprise AI must optimize for consistency and precision.

There are many high-stakes industries where getting the answer wrong can have detrimental and long-lasting effects. For example, in healthcare, legal and finance, the margin for error is zero, and one single hallucination can lead to serious consequences. A wrong supplier name, a misread total, or a compliance misstep isn’t a quirky model behaviour, it’s a liability. Typically, in business, it’s not just the one issue that is the concern, it’s the compounding of issues and scale.

In enterprise AI tools, just bolting a general-purpose LLM onto a workflow and hoping for accuracy is a dangerous gamble. Frontier models can change overnight, resulting in a workflow that was 92 percent accurate on Monday but, by Tuesday, produces entirely different results, may be under export controls, or may refuse to process some items. When your product has dependencies on something not designed for the job, you may not get the accuracy you need or the cost you expect because you’re effectively renting your house, and the cost of the rent can change at any time. It’s fast to build with frontier AI models, but you can usually tell when there is no accuracy claim: ‘AI can make mistakes, we may or may not train on your data…’

Before generative AI, we lived in a world of deterministic code – there were bugs, but you could reason over the system. As we move into a world of generative AI and purely probabilistic systems, things are going to behave differently. Looking forward, enterprises are aiming to achieve a balance between the two. A blend of deterministic logic, specialised models, frontier systems and the correct grounding context with human supervision. Building and maintaining this orchestrated symphony at scale—while ensuring absolute trust—is a non-trivial challenge.

Enter SLMs: Faster, cheaper and more accurate

We need to move away from a one-size-fits-all approach to AI, or even a one-model system, and this is where small language models (SLMs) come into play. SLMs are specialised, domain-specific language models designed for a purpose.

By training on narrow, high-quality datasets, these models operate in a world focused on accuracy. Because they leverage highly targeted, niche datasets, compared to the internet and world knowledge which LLMs are trained on, SLMs are inherently leaner, faster and cheaper to run. Consequently, their logic is easier to reason about, allowing them to guarantee a much higher degree of accuracy in their output. Unlike LLMs, they aren’t trying to be clever; they’re trying to be correct.

A report from Gartner found that LLM response accuracy declines when tasks require specific business context. As a result, they predict that by 2027, smaller, context-specific models will see usage volumes at least three times greater than those of general-purpose LLMs. It’s critical to state that the difference lies in the focus and the distribution of data these models have seen.

Early adoption of AI was driven by experimentation and productivity gains. The next step is for AI systems to influence operational decisions, and that shifts the standards required for trust. For CIOs and technology leaders, we need to build systems that operate consistently under real-world conditions – ones that maintain performance over time and will withstand regulatory and customer scrutiny.

Building trusted systems

The most effective enterprise AI tools will combine both LLM and SLM models – using LLMs for orchestration and SLMs for deterministic verification. By using multiple specialist models built for precision, organizations can build AI systems with accuracy they can stand behind and that can run efficiently from a cost perspective.

1. Getting work done right requires AI and humans to work together

AI adoption isn’t a binary choice between models; true operational resilience comes from utilizing the best elements of different architectures. When large and small models are combined, large thinking models can dispatch and reason while smaller models can act and verify, and the whole system can adapt to improve itself.

Models only get better when there are feedback loops through signals and context. On the surface things should look simple and feel like magic, but under the covers, the complexity is often many layers deep.

In financial document processing, for example, large thinking models can handle complex reasoning and user/company/accounting preferences, while the smaller models handle values, tax and line items. The key is that it is not a one-size-fits-all, and it’s not just the models; it’s the embeddings of similarity for processing and the context and guidance that shapes the outcome.

2. Creating best in class

To build a best-in-class system, you need data, compute and a feedback loop. We achieve this by training in-house models on our billions of documents for fast and accurate extraction, then leverage frontier-level models for reasoning. Relying solely on frontier-level models for extraction would tank our accuracy and skyrocket costs. This balance of accuracy comes at the price of generability and reasoning, but leveraging reasoning post-processing provides a top-tier solution, especially when that reasoning is looking to mimic the user’s preferences.

3. The human feedback loop

Human and AI oversight must be built into frontier systems from the beginning, not treated as a fallback when something goes wrong. There are always edge cases that have “it depends” answers – sometimes the answer is we do not know, but without any auditing or sampling, it’s impossible to know how well you did. AI and complex systems can drift, and accuracy is heavily dependent on the distribution of data, so continuous sampling and monitoring are critical components to these systems.

Human users and reviewers in the accounting world are those who are accountable for outcomes, and they provide the signal of feedback to the models and weights. On our systems, we sample over 100,000 documents every month to ensure accuracy is a measurable metric, not a hope.

In our system, when a bookkeeper accepts, edits, or rejects AI suggestions, that data is fed into an evaluation loop too. The result is a continuously improving system that understands the unique financial context of each of the 700,000 SMBs in our system. Through these feedback loops and rigorous evaluations, outputs remain accurate and dependable over time.

The AI hype

The market is flooded with AI tools and demos that trivialise complex systems. Yet many of these crumble when they encounter the complexity and nuances of real-world workflows, leading to a new tagline: “AI can make mistakes”.

This era of AI will not be defined by who is building the biggest model or who has the best demos. It will be defined by those delivering the biggest sustainable impact – the true overnight success stories, 15 years in the making.

This article is published as part of the Foundry Expert Contributor Network.
Want to join?


Read More from This Article: When it comes to AI, bigger isn’t always better
Source: News

Category: NewsJuly 28, 2026
Tags: art

Post navigation

PreviousPrevious post:Security at the speed of AI: How to protect against sophisticated cyberattacksNextNext post:Earnings from SAP, ServiceNow, and IBM challenge the SaaSpocalypse narrative

Related posts

Compassion is not a control: What veterinary practices reveal about AI governance
July 31, 2026
The blueprint for innovation: 3 ways regulatory readiness is a competitive advantage
July 31, 2026
The gen AI helping Aetna review millions of medical records
July 31, 2026
How AI helps the US Senate Federal Credit Union better manage risk
July 31, 2026
Your AI model isn’t the problem. Your data was never ready for it
July 31, 2026
Microsoft doubles down on multi-model AI as it builds a Copilot super app
July 31, 2026
Recent Posts
  • Compassion is not a control: What veterinary practices reveal about AI governance
  • How AI helps the US Senate Federal Credit Union better manage risk
  • The gen AI helping Aetna review millions of medical records
  • The blueprint for innovation: 3 ways regulatory readiness is a competitive advantage
  • Your AI model isn’t the problem. Your data was never ready for it
Recent Comments
    Archives
    • July 2026
    • June 2026
    • May 2026
    • April 2026
    • March 2026
    • February 2026
    • January 2026
    • December 2025
    • November 2025
    • October 2025
    • September 2025
    • August 2025
    • July 2025
    • June 2025
    • May 2025
    • April 2025
    • March 2025
    • February 2025
    • January 2025
    • December 2024
    • November 2024
    • October 2024
    • September 2024
    • August 2024
    • July 2024
    • June 2024
    • May 2024
    • April 2024
    • March 2024
    • February 2024
    • January 2024
    • December 2023
    • November 2023
    • October 2023
    • September 2023
    • August 2023
    • July 2023
    • June 2023
    • May 2023
    • April 2023
    • March 2023
    • February 2023
    • January 2023
    • December 2022
    • November 2022
    • October 2022
    • September 2022
    • August 2022
    • July 2022
    • June 2022
    • May 2022
    • April 2022
    • March 2022
    • February 2022
    • January 2022
    • December 2021
    • November 2021
    • October 2021
    • September 2021
    • August 2021
    • July 2021
    • June 2021
    • May 2021
    • April 2021
    • March 2021
    • February 2021
    • January 2021
    • December 2020
    • November 2020
    • October 2020
    • September 2020
    • August 2020
    • July 2020
    • June 2020
    • May 2020
    • April 2020
    • January 2020
    • December 2019
    • November 2019
    • October 2019
    • September 2019
    • August 2019
    • July 2019
    • June 2019
    • May 2019
    • April 2019
    • March 2019
    • February 2019
    • January 2019
    • December 2018
    • November 2018
    • October 2018
    • September 2018
    • August 2018
    • July 2018
    • June 2018
    • May 2018
    • April 2018
    • March 2018
    • February 2018
    • January 2018
    • December 2017
    • November 2017
    • October 2017
    • September 2017
    • August 2017
    • July 2017
    • June 2017
    • May 2017
    • April 2017
    • March 2017
    • February 2017
    • January 2017
    Categories
    • News
    Meta
    • Log in
    • Entries feed
    • Comments feed
    • WordPress.org
    Tiatra LLC.

    Tiatra, LLC, based in the Washington, DC metropolitan area, proudly serves federal government agencies, organizations that work with the government and other commercial businesses and organizations. Tiatra specializes in a broad range of information technology (IT) development and management services incorporating solid engineering, attention to client needs, and meeting or exceeding any security parameters required. Our small yet innovative company is structured with a full complement of the necessary technical experts, working with hands-on management, to provide a high level of service and competitive pricing for your systems and engineering requirements.

    Find us on:

    FacebookTwitterLinkedin

    Submitclear

    Tiatra, LLC
    Copyright 2016. All rights reserved.