{
  "@context": {
    "dcterms": "http://purl.org/dc/terms/",
    "dcmitype": "http://purl.org/dc/dcmitype/"
  },
  "dcterms:identifier": [
    "tag:aim-pro.eu,2026:oer/8b4ef701d9f2",
    "https://arxiv.org/abs/2306.05949",
    "arxiv:2306.05949",
    "doi:10.1093/oxfordhb/9780198940272.013.0025"
  ],
  "dcterms:title": [
    "Evaluating the Social Impact of Generative AI Systems in Systems and Society"
  ],
  "dcterms:type": [
    {
      "@id": "dcmitype:Text"
    },
    "preprint"
  ],
  "dcterms:creator": [
    "Irene Solaiman",
    "Zeerak Talat",
    "William Agnew",
    "Lama Ahmad",
    "Dylan Baker",
    "Su Lin Blodgett",
    "Canyu Chen",
    "Hal Daumé",
    "Jesse Dodge",
    "Isabella Duan",
    "Ellie Evans",
    "Felix Friedrich",
    "Avijit Ghosh",
    "Usman Gohar",
    "Sara Hooker",
    "Yacine Jernite",
    "Ria Kalluri",
    "Alberto Lusoli",
    "Alina Leidinger",
    "Michelle Lin",
    "Xiuzhu Lin",
    "Sasha Luccioni",
    "Jennifer Mickel",
    "Margaret Mitchell",
    "Jessica Newman",
    "Anaelia Ovalle",
    "Marie-Therese Png",
    "Shubham Singh",
    "Andrew Strait",
    "Lukas Struppek",
    "Arjun Subramonian"
  ],
  "dcterms:description": [
    "Generative AI systems across modalities, ranging from text (including code), image, audio, and video, have broad social impacts, but there is no official standard for means of evaluating those impacts or for which impacts should be evaluated. In this paper, we present a guide that moves toward a standard approach in evaluating a base generative AI system for any modality in two overarching categories: what can be evaluated in a base system independent of context and what can be evaluated in a societal context. Importantly, this refers to base systems that have no predetermined application or deployment context, including a model itself, as well as system components, such as training data. Our framework for a base system defines seven categories of social impact: bias, stereotypes, and representational harms; cultural values and sensitive content; disparate performance; privacy and data protection; financial costs; environmental costs; and data and content moderation labor costs. Suggested methods for evaluation apply to listed generative modalities and analyses of the limitations of existing evaluations serve as a starting point for necessary investment in future evaluations. We offer five overarching categories for what can be evaluated in a broader societal context, each with its own subcategories: trustworthiness and autonomy; inequality, marginalization, and violence; concentration of authority; labor and creativity; and ecosystem and environment. Each subcategory includes recommendations for mitigating harm."
  ],
  "dcterms:subject": [
    "cs.CY (arxiv)",
    "cs.AI (arxiv)"
  ],
  "dcterms:language": [
    "en"
  ],
  "dcterms:license": [
    {
      "@id": "http://creativecommons.org/licenses/by-sa/4.0/"
    }
  ],
  "dcterms:rights": [
    "CC-BY-SA-4.0 — assessed as SHARE_ALIKE by the harvester's licence gate. Conditions: attribution required, adaptations must carry the same licence. open but ShareAlike/copyleft — derivatives keep the license"
  ],
  "dcterms:accessRights": [
    "open"
  ],
  "dcterms:format": [
    "text/markdown"
  ],
  "dcterms:extent": [
    "2373 bytes (extracted text)"
  ],
  "dcterms:issued": [
    "2023-06-09"
  ],
  "dcterms:modified": [
    "2024-06-28"
  ],
  "dcterms:bibliographicCitation": [
    "The Oxford Handbook of the Foundations and Regulation of Generative AI, 18 December 2025"
  ],
  "dcterms:provenance": [
    "Retrieved from arXiv on 2026-10-09 in response to the search string “(all:\"artificial intelligence\" OR all:\"machine learning\" OR all:\"generative AI\" OR all:\"deep learning\" OR all:\"reinforcement learning\" OR all:\"large language model\") AND (all:\"AI concepts\" OR all:\"types of AI\" OR all:\"AI fundamentals\" OR all:\"recognizing AI\" OR all:\"recognising AI\" OR all:\"general versus narrow AI\" OR all:\"narrow AI\" OR all:\"general AI\" OR all:\"machine intelligence\" OR all:\"AI strengths and weaknesses\" OR all:\"traditional software\" OR all:\"rule-based systems\" OR all:\"introduction to AI\" OR all:\"introduction to artificial intelligence\" OR all:\"artificial intelligence introduction\" OR all:\"AI primer\" OR all:\"foundations of artificial intelligence\" OR all:\"overview of AI\" OR all:\"understanding AI\" OR all:\"history of AI\" OR all:\"AI essentials\" OR all:\"AI terminology\" OR all:\"metaphors for AI\" OR all:\"AI fundamental concepts\" OR all:\"AI key concepts\" OR all:\"philosophy of AI\" OR all:\"critical AI literacy\")”. arXiv served the resource and is not asserted to be its publisher or author.",
    "Text extracted from abstract to Markdown by abstract; the original is retained unchanged beside it."
  ],
  "inferred": {
    "dcterms:isReferencedBy": [
      {
        "value": {
          "@id": "https://arxiv.org/abs/2410.23704",
          "title": "Using Scenario-Writing for Identifying and Mitigating Impacts of Generative AI"
        },
        "method": "inferred",
        "rule": "citation",
        "rule_version": "links/1",
        "evidence": {
          "links": [
            {
              "how": "citation",
              "line": 76,
              "basis": "reference",
              "literal": "arXiv:2306.05949"
            }
          ],
          "kind": "citation"
        }
      }
    ]
  }
}