SEARCH · ENGLISH EDITION
Find an AI concept.
Search titles, summaries, article text, source-language titles, and known aliases.
Benchmark Data Contamination
The presence of evaluation examples or closely related material in training data, potentially inflating benchmark results.
Datasheets for Datasets
A documentation framework for describing how datasets were created, composed, maintained, and intended to be used.
Foundation Model
A broadly trained model that can be adapted to many downstream tasks and applications.
Hallucination in AI
A generated statement that is unsupported, fabricated, or inconsistent with the available evidence or context.
Red Teaming
Adversarial testing intended to discover failure modes, unsafe behavior, or exploitable weaknesses in an AI system.
Safety Evaluation
Testing designed to measure risks, harmful behaviors, and the effectiveness of safeguards in an AI system.
System Card
Documentation describing a deployed AI system, including capabilities, evaluations, mitigations, and limitations.