Introducing MAVIS

Monitoring Assertions and Verifying Identifiers in Science

AI tools now draft a growing share of what research teams read: reports, literature reviews, summaries and data notes. MAVIS keeps a standing watch on the claims in those files. It has fred check each claim against public sources, checks again when something changes, and tells the person responsible when a claim cannot be traced, conflicts with its source, or stops holding.

Figure 1. How MAVIS helps a life-sciences team. AI tools write quickly and are sometimes wrong. MAVIS watches the team's files and has fred check every claim, then checks again when files, the schedule or sources change. The team sees what holds and what needs attention, with the evidence, and decides what to do. Click the cartoon to enlarge it.

Why MAVIS is needed

Pre-clinical decisions rest on claims: this compound binds that target, this paper showed that effect, this protein has that sequence. When people wrote every claim, a reviewer could ask where it came from. AI tools now write many of them, quickly, and they write wrong claims in the same confident style as right ones.

A one-off check at the moment of writing is not enough. The claims a program relies on need a standing watch.

What MAVIS is

MAVIS is an application that sits beside a research program's files and watches the claims in them. It does not write content and it does not make decisions. It makes sure that every claim it can check has been checked, keeps checking it, and shows people the result with its evidence.

It watches your files
Your program's notes, reports and data, as Markdown and JSON files with a README that says what each file is and where it came from. Files drafted by AI are marked as such.
It has fred check every claim
Any statement that names an identifier (a PubMed ID, DOI, UniProt or PubChem entry) or a value with a unit is sent to fred, which looks it up in public databases.
It alerts the owner
When a claim cannot be traced, disagrees with its source, or stops holding, the person responsible for it is told, once, with the reason.
It shows the provenance
For every claim: the file and line it came from, the check that tested it, each database lookup behind the result, and the record it found.

People stay in charge. The claim's owner reads the alert and its evidence, then corrects the file, accepts the claim with a note, or withdraws it. MAVIS records the decision.

How it works

MAVIS runs continuously, in a loop of four steps.

  1. Probe the files. MAVIS looks for new or changed files and picks out every claim in them that names an identifier or a value with a unit.
  2. Ask fred. It sends the claims to fred, which checks each one against public databases. A claim never counts as its own evidence: MAVIS does not send a claim's own file as the material to check it against.
  3. Read fred's result. For each claim, the verdict and every lookup behind it.
  4. Compare and alert. MAVIS compares the result with the last check. A claim counts as holding only when a public source confirmed it. A change raises an alert only when a second check agrees, because AI-based checks can vary from one run to the next.

When it checks

When a file is added or edited, when a scheduled re-check is due, and when a public source may have changed, such as a paper being retracted.

What the owner sees

Everything appears on the MAVIS dashboard: what was checked and when, open alerts, claims labeled by where their support comes from, and the provenance of each claim.

Figure 2. The MAVIS workflow, from your data files through the loop around fred to the dashboard and the people who decide. Click the diagram to enlarge it.

How MAVIS uses fred

fred is a reasoning engine for pre-clinical drug discovery that runs on your own machine with a free open-weight model. It looks things up in public databases and traces every identifier in its answers to where it came from. MAVIS does not change fred. It calls fred the way any other application does, and adds the watching, the memory of past checks, the alerts and the dashboard.

fred does

  • Reasons with a local open-weight model.
  • Looks values up in UniProt, PubChem, PubMed, Open Targets, Reactome and the web.
  • Traces each identifier to the lookup or file it came from, and flags what it cannot trace.
  • Keeps a record of every run.

MAVIS does

  • Watches your files and finds the claims in them.
  • Decides what to check and when, and asks fred.
  • Remembers every past check, so it can see change.
  • Raises alerts, shows the dashboard and records decisions.

MAVIS gives fred your program's own Markdown and JSON files, with a README that labels each one, and asks fred for a verdict on every claim: confirmed by a public source, in conflict with it, not found, or supported only by your own files. fred also reports when a paper a claim cites has been retracted or corrected.

Because both run on your own machine, your files and fred's reasoning never go to a model provider. This is private, not air-gapped: fred's searches of public databases do leave the machine.

What MAVIS checks, and what it does not

MAVIS checks whether each claim's identifiers and values trace to a public source and still agree with it. It does not check whether the science is right; that judgment stays with people.

Sources

  1. Topaz M, et al. Fabricated citations: an audit across 2.5 million biomedical papers. The Lancet 407 (2026): 1779-81. Summary: casrai.org.
  2. Biomedical reference generation remains unreliable across 26 large language models. arXiv 2609.14988 (2026): arxiv.org/pdf/2609.14988.
  3. FDA and EMA. Guiding principles of good AI practice in drug development, January 2026: fda.gov/media/189581/download.
Figure 1. How MAVIS helps a life-sciences team
The MAVIS cartoon, full size.
Figure 2. The MAVIS workflow
The MAVIS workflow diagram, full size.