Skip to main content
v2026.11,858 entries · CC-BY 4.0
NIKOLAI elementN4 · Claims and argumentProposednikolai-v0.1

Limitation

NIKOLAI proposes Limitation as a stated reason, attached to an evaluation or a safety case, why that assessment's conclusions may not hold or may not generalise -- e.g. elicitation strength, evaluation awareness, incomplete data capture, or transcript integrity concerns. This is a NIKOLAI editorial proposal. It must not be confused with a developer's stated restrictions or conditions on end-user use of a model, which is a different concept some sources label with the same word (see the false-friend note below).

This is CASRAI's own proposed definition, not a definition any named organisation has agreed to. See what NIKOLAI is and is not.

Source of record

Where this definition comes from

Crosswalk

How named organisations use this concept

Every row below is a shadow mapping. It is CASRAI's own reading of a published document. No lab, evaluator or regulator named here has declared, endorsed, or been consulted on this mapping. That will change only when an organisation files its own Mapping Declaration — see the non-endorsement policy.
OrganisationTheir term, as publishedMatchSource
Anthropic
Anthropic Risk Report, August 2026
Explicit limitations section (§2.16); "Evaluation awareness is conceded to 'partially undermine' confidence in the alignment assessment"; "we have not provided clear evidence that this elicitation is sufficiently strong" (§2.16.1).exact
confidence: high
Anthropic Risk Report, August 2026
OpenAI
OpenAI Preparedness Framework v2
Safeguards Report contents include "Any notable limitations with the information provided" (§4.2).exact
confidence: high
OpenAI Preparedness Framework v2
Meta
Meta Advanced AI Scaling Framework v2
"Preparedness reports will also disclose any known issues that could hinder generalizing our safety testing to realworld risks, including changes to the training process that reduce interpretability" (§2.2.1).exact
confidence: high
Meta Advanced AI Scaling Framework v2
California SB 53
California SB 53
Transparency report "(G) Any generally applicable restrictions or conditions on uses of the frontier model" (22757.12(c)(1)). This is a use restriction on the deployed model, not a stated limitation of the assessment's conclusions.
FALSE FRIEND, explicitly coded FF in the source table: SB 53's transparency-report item labelled 'restrictions or conditions on uses' is a product/usage-policy disclosure, not an assessment-limitation disclosure. See the element-level divergence_note.
none
confidence: high
California SB 53
STREAM (discovery sweep)
Discovery sweep (evaluator ecosystem)
Evaluation reporting template (content not read at time of the working table) [UV].
STREAM's evaluation reporting template (arXiv 2508.09853) was subsequently confirmed open-access and fetched directly (ALIGNMENT-MATRIX.md §6 item 6, resolved fourth pass 16 Sep 2026): title, scope (a pilot limited to ChemBio benchmarks, 23 contributing experts), format (three-page template with worked 'gold standard' examples), and purpose (comparable, item-level evaluation disclosure) are confirmed from the primary source. Still open: the template's exact field names -- needed to know whether it has a distinct 'limitations' field matching this element -- require reading the PDF body, not yet done. Match type left as 'none' rather than guessed; citation URL corrected to the resolved arXiv page (the sources table's {DEV} tag names 'STREAM=arXiv 2508.09853' but has no single URL of its own, so the arXiv abstract-page URL is used here rather than the compound {DEV} tag).
none
confidence: low
STREAM evaluation reporting template (arXiv 2508.09853; per {DEV} sources-table entry)
Anthropic
Anthropic Advanced AI Framework
AAF system card contents: "Model capabilities and limitations, as well as intended and observed model behaviors" (p.6).close
confidence: high
Anthropic Advanced AI Framework

Related, not mapped

Pointers that are not crosswalk claims

These sources mention this concept but do not define or map it clearly enough to count as a crosswalk row — noted here so the research is visible without overstating it as a mapping.

Divergence

Where sources materially disagree

REQUIRED false-friend note (an SB 53 crosswalk row above is coded FF): California SB 53's transparency-report item '(G) Any generally applicable restrictions or conditions on uses of the frontier model' (22757.12(c)(1)) uses limitation-adjacent language but names a product/usage-policy restriction imposed BY the developer ON users of the deployed model -- e.g. acceptable-use terms -- not a stated reason why the developer's own safety assessment might not hold or generalise, which is what every other source in this cluster means by 'limitation'. A naive string or keyword match on 'restrictions'/'limitations' between SB 53 and the lab frameworks would wrongly equate a use-policy disclosure with an assessment-caveat disclosure; NIKOLAI's crosswalk keeps them separate.

Referenced across the research world

University of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logoUniversity of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logo
  • University of Cambridge logo
  • Columbia University logo
  • Crossref logo
  • University of Edinburgh logo
  • Harvard University logo
  • University of Oxford logo
  • Princeton University logo
  • Stanford School of Medicine logo
  • University College London logo
  • ORCID logo

View CASRAI adoption →