CiteAbility Labs

Frontier LLM Citation Integrity Benchmark

See what changes when the same frontier AI model answers the same question with and without the CiteAbility Citation Integration Layer.

Method

Same question, same model, one variable.

Each model answers the identical question twice: once natively, and once with the CiteAbility Citation Integration Layer applied during source selection. Wording, model version, generation settings and answer constraints are held the same.

The native answer receives no CiteAbility input while it is being written. After both answers are finished, the sources each one used are assessed independently, so both columns are compared on the same footing.

This is an experiment, not a guarantee. A model saying it would use a source is not evidence that it has cited, or will cite, that source in live search.

Results

What the runs show.

Loading results…