ibl.ai Agentic AI Blog

Insights on building and deploying agentic AI systems. Our blog covers AI agent architectures, LLM infrastructure, MCP servers, enterprise deployment strategies, and real-world implementation guides. Whether you are a developer building AI agents, a CTO evaluating agentic platforms, or a technical leader driving AI adoption, you will find practical guidance here.

Topics We Cover

Featured Research and Reports

We analyze key research from leading institutions and labs including Google DeepMind, Anthropic, OpenAI, Meta AI, McKinsey, and the World Economic Forum. Our content includes detailed analysis of reports on AI agents, foundation models, and enterprise AI strategy.

For Technical Leaders

CTOs, engineering leads, and AI architects turn to our blog for guidance on agent orchestration, model evaluation, infrastructure planning, and building production-ready AI systems. We provide frameworks for responsible AI deployment that balance capability with safety and reliability.

Back to Blog

You Cannot Govern a Clinical Model You Cannot Observe

ibl.ai EngineeringSeptember 29, 2026
Premium

Anthropic reported that ~950 agents surfaced a novel enzyme system in 21 hours, and the first FDA-approved AI margin-assessment device reached its first operating rooms in August. The capability question is closing. The governance one is not β€” and even that approved device ships AI updates under a change-control plan the hospital does not hold.

The Short Answer

Clinical AI is arriving faster than hospitals can govern it: 950 agents found a novel enzyme system in 21 hours, and the FDA approved the first AI margin-assessment device in March. Ban or license, both assume you can observe the model. On ibl.ai you own all the code and the data.

The debate about clinical AI has been about whether the models are good enough. That question is closing. The one underneath it β€” whether a hospital can see what the model is doing β€” has barely been asked.

How fast is clinical capability actually moving?

Fast enough that the recent examples are qualitatively different from the last decade's.

On 23 September 2026 Anthropic released a preprint reporting that roughly 950 Claude agent sessions, running for 21 hours and consuming 210 million tokens, surveyed 1.9 billion protein clusters, recovered about 200,000 reverse transcriptases, and surfaced a previously unknown enzyme system it calls ART β€” array-associated reverse transcriptases, found mostly in bacteriophages.

Anthropic presents it as early results from one of its first research programs, and says the work would have taken an expert scientist weeks or months.

Worth stating the limits as clearly as the result. The function of the system itself is not yet known β€” not merely the accessory protein's β€” and whether it is useful for gene editing is undetermined.

Dario Amodei has acknowledged that a Stanford team previously described a system similar in some ways to this one, so "previously unknown" is doing careful work.

It is not purely computational, though: the preprint reports that ART arrays are highly expressed and appear as discrete units during Staphylococcus phage infection.

Expression was confirmed at the bench; function was not. A strong demonstration of search at scale, not a validated therapy.

On 3 March 2026 the FDA granted premarket approval to Claire, from Perimeter Medical Imaging AI β€” the first AI-enabled imaging device approved in the United States for intraoperative breast cancer margin assessment.

It images the excised lumpectomy specimen during surgery, not tissue in the patient.

Its pivotal trial reported 88.1% margin accuracy and a statistically significant reduction in patients with residual cancer against standard of care.

Repeat surgery occurs in about 20% of breast-conserving surgeries in the US, against roughly 300,000 breast cancer surgeries a year.

In August 2026 Intermountain Health became the first US health system to deploy it commercially, at LDS and American Fork hospitals. Five months from approval to an operating room is fast for a class III device.

So what is the bottleneck now?

Deployment, and then something harder than deployment.

The deployment part is documented. In a February 2026 survey of 120 US health systems, 75% were using or planning to use at least one AI application.

The share implementing or planning three or more AI solutions rose from 30% to 59% β€” a 67% year-on-year increase. Respondents named slow implementation timelines among their challenges.

That is a real drag, and it is the one most people name. But a slow procurement cycle is a solvable problem β€” budget, staffing, integration work. The constraint underneath it is not.

What is the constraint underneath it?

You cannot govern a model you cannot observe.

When clinical AI runs on a vendor's cloud, four things are outside the hospital's control at once:

Question a quality committee must answer Vendor cloud Self-hosted
Which model version produced this recommendation, on this date?Knowable if you pinned it β€” several platforms let youA version you pinned
When did the weights last change?On the vendor's schedule, with notice β€” until the version retiresWhen you decided
Has behaviour drifted since validation?Measurable only from outputsMeasurable against a fixed artefact
Can you reproduce the decision a year later?Only if the vendor retained itYes

Be fair to the managed platforms first, because the absolute version of this claim is wrong.

Azure lets a deployment opt out of automatic model upgrades, exposes the current version through the portal and API, gives at least two weeks' notice before a default moves, and keeps the previous major version until its retirement date.

ONC's HTI-1 rule goes further, requiring certified health IT to disclose the update and validation schedule for predictive decision support.

So versions can be pinned and changes can be announced. What ends is the pinning: a retired version is gone, and with it the ability to re-run a case as it ran.

And the regulated case is not the clean exception either. Claire's own PMA authorises a predetermined change control plan covering "planned AI enhancements that can be implemented without additional FDA interaction." That is good regulation β€” the changes were reviewed in advance and are bounded β€” but it means an approved class III AI device also updates without a fresh submission, on a schedule the hospital does not set.

That is the sense in which "ban it" and "license it" answer a narrower question than they appear to. Both assume you can see what the model is doing well enough to decide whether it is acceptable.

Prohibition does not need observability because nothing is deployed. Licensure absolutely does β€” and licensing a service whose behaviour changes on a schedule you do not hold licenses a moving target.

Isn't this an argument against clinical AI generally?

No, and it would be a bad one. Repeat surgery after one in five breast-conserving operations is a real harm, and Claire's trial showed a statistically significant reduction in patients left with residual cancer.

The company puts the re-operation benefit as potential rather than measured. Either way, a hospital that refuses AI on principle is choosing the status quo.

The argument is about where the model runs, not whether it runs. An institution can adopt aggressively and still insist on three things: a version it controls, logs it holds, and the ability to reproduce a decision for a morbidity and mortality review or a malpractice claim.

Those are ordinary clinical governance expectations. They are only difficult when the model is somebody else's service.

What does a hospital need in place before it scales?

Four things, and none of them is a model choice:

  1. A pinned version β€” the artefact you validated is the artefact in production, and changing it is a decision with a date and a signature.
  2. Logs you hold β€” inputs, outputs, model version and the clinician who reviewed it, in your systems.
  3. Drift measurement against a fixed baseline β€” not vibes, and not the vendor's dashboard.
  4. Reproducibility β€” the ability to re-run a case as it ran then, years later, because that is the window in which claims arrive.

Why does ownership decide all four?

Because every one of them is a property of where the model lives.

On ibl.ai you own all the code and the data. The platform runs under a perpetual licence inside the hospital's own perimeter, so the model version, the inference logs and the audit trail are institutional records rather than a vendor's telemetry.

It is model-agnostic, which is what makes the pinning real: a health system can run an open-weight model entirely inside its own network, upgrade on its own schedule, and keep the prior version available for reproduction.

Pricing is usage-based with no per-seat pricing, and you can deploy anywhere: your cloud, your VPC, on-premise, or fully air-gapped.

1.6M+ users across 400+ organizations run the platform this way, including NVIDIA, MIT, and Syracuse University.

ibl.ai is family-owned and operated from New York, NY β€” a U.S.-headquartered, domestically-owned long-term partner, not a vendor that sells licenses and moves on.

A stethoscope does not get a firmware update. Clinical models will, including the approved ones β€” so the governance question was never really about the model. It is about who holds the version, the logs, and the ability to reconstruct the day.

Sources: the enzyme discovery from Anthropic's announcement, its preprint and TechCrunch; Claire's approval, PCCP, indications and the re-excision figure from Perimeter's approval release, and the first deployment from its Intermountain release; the change-control pathway from FDA's PCCP guidance; version pinning from Azure's model-version documentation; adoption figures from Eliciting Insights' 2026 AI Adoption Survey.

Related: AI Agents Do Licensed Work. Liability Law Doesn't Fit. β€” the ban-versus-license split in detail, and why three liability doctrines all miss.

Why does owning the AI stack matter?

ibl.ai is the agentic AI platform where you own all the code and the data. You self-host the entire stack inside your own perimeter, run it model-agnostic across any LLM and switch anytime, and pay by usage with no per-seat pricing β€” so you can deploy anywhere: your cloud, on-premise, GovCloud, or fully air-gapped.

  • You own all the code and the data

    Full source code under a perpetual license, running on your infrastructure. Not API access to someone else's platform β€” the stack itself is yours.

  • Model-agnostic

    Run any LLM β€” Claude, GPT, Gemini, Llama, Command, or your own fine-tune β€” and switch providers without rewriting the platform.

  • No per-seat pricing

    Usage-based billing against a budget cap you set. Cost tracks what your organization actually uses, not how many people you employ.

  • Deploy anywhere

    Your cloud, your VPC, on-premise, GovCloud, or a fully air-gapped network with no outbound connectivity.

1.6M+ users across 400+ organizations run the platform this way, including NVIDIA, MIT, and Syracuse University.

ibl.ai is family-owned and operated from New York, NY β€” a U.S.-headquartered, domestically-owned long-term partner, not a vendor that sells licenses and moves on.

Related Articles

Beyond LLMs: What Reasoning Limits Mean for Clinical AI

A widely-shared DeepMind position paper argues LLMs cannot make the abductive leap that produces new scientific theories. It is a narrower claim than the headlines suggest, and it is not the reason clinical AI fails today β€” but it does explain why a health system should build for model replacement rather than model selection.

Miguel AmigotAugust 17, 2026

Self-Hosted AI Agents for Healthcare: PHI Never Leaves

Self-hosted AI agents for healthcare are autonomous clinical and administrative agents that run entirely inside your HIPAA-covered environment β€” reading from and writing to your EHR through connectors, with PHI never leaving the boundary. The agents, the architecture, the cost math, and why owning the stack is the defensible posture.

Mikel AmigotJune 8, 2026

Self-Hosted AI for Hospitals and Health Systems: The Deployment That Survives Audit

Self-hosted AI for hospitals and health systems means the runtime executes inside your existing HIPAA-covered environment β€” PHI never traverses a third-party cloud. The deployment options, the workloads, the cost math, and why this becomes the default endpoint for any serious clinical AI program.

Mikel AmigotJune 1, 2026

AI Agents Do Licensed Work. Liability Law Doesn't Fit.

Agents now draft motions and reason across clinical documents β€” work that is licensed when a human does it. Product liability assumes a defect, professional liability assumes a licensed practitioner, and agency law reaches non-human agents only partway. At least seven states have answered by prohibiting AI therapy; the Cicero Institute proposes licensing the service instead.

ibl.ai EngineeringSeptember 28, 2026

See the ibl.ai AI Operating System in Action

Discover how leading universities and organizations are transforming education with the ibl.ai AI Operating System. Explore real-world implementations from Harvard, MIT, Stanford, and users from 400+ institutions worldwide.

View Case Studies
Work with our team

Pilots, deployment, and full ownership

Most enterprise engagements are one-time, not subscriptions. You integrate ibl.ai with your own data, deploy it on your own infrastructure, and the engineering hours scale with the work β€” so the price tracks the scope, not your headcount.

Start here

Pilot

from $15K

fixed scope Β· fixed timeline

A time-boxed proof of value on your real data β€” not a slide deck.

Best for: Teams that want to see ibl.ai working before committing.

  • Deployed on your infrastructure or our cloud
  • 1–2 production agents wired to a slice of your data
  • One integration (LMS / SIS / SSO / data source)
  • Weekly working sessions with our engineers
  • Pilot fee credits toward a full engagement
Scope a pilot
Most common

Integration & Deployment

$25K – $80K

one-time Β· not a subscription

Full deployment integrated with your data and systems. Engineering hours scale with scope.

Best for: Organizations rolling ibl.ai out across a department, campus, or business unit.

  • Platform deployed in your VPC, on-prem, or air-gapped
  • Integrated with your data + identity (SSO / SAML)
  • Multiple custom agents built to your workflows
  • Engineering hours proportional to scope
  • You own the data Β· run any LLM you choose
Plan a deployment
Full ownership

Codebase Transfer + Custom AI Engineering

Custom quote

perpetual license Β· you own the stack

We transfer the full source code. You own and self-host the entire platform β€” outright.

Best for: Organizations and enterprises that benefit from perpetual ownership and sovereignty.

  • Complete source-code transfer + perpetual license
  • Dedicated AI engineering team on your roadmap
  • Custom agents, models, and integrations to spec
  • Air-gapped capable Β· zero vendor lock-in
  • Family-owned, New York–based long-term partner
Talk about ownership
You own the code and data Run any LLM β€” Claude, GPT, Gemini, Llama Family-owned & operated from New York, NY