The AI Safety Measurement Problem

The AI Safety Measurement Problem

Nimble Books LLC

33,78 €
IVA incluido
Disponible
Editorial:
W. Frederick Zimmerman
Año de edición:
2026
Materia
Influencia de la ciencia y la tecnología sobre la sociedad
ISBN:
9798259503489
33,78 €
IVA incluido
Disponible
Añadir a favoritos

A safety claim about an AI system is only as credible as the measurement that backs it. Everything else in this book follows from that sentence. If a developer says a model is 'safe enough to deploy,' a regulator says it is 'low-risk,' or a buyer says it is 'fit for purpose,' each of those statements is a load-bearing claim about behavior under conditions that have not yet happened. Without a measurement program - without tests that were specified before the system was built, executed by people who can be honest about the results, and reported in a form that outsiders can interrogate - those claims are aspirations dressed up as findings. This book is about how to tell the difference. AI governance discussions in 2026 are crowded with rules, principles, and voluntary commitments. The substance of those commitments, however, almost always reduces to a claim about what an AI system will or will not do. 'We will not deploy a model that meaningfully uplifts the creation of biological weapons.' 'Our system does not exhibit unacceptable bias in hiring contexts.' 'The model refuses to generate child sexual abuse material.' Each of these claims is a measurement claim. Each requires that someone - the developer, an independent lab, a government body - define what the dangerous behavior looks like, design a probe that can elicit it if it is present, run that probe under conditions representative of real use, and report the result. The NIST AI Risk Management Framework treats this measurement function as one of four continuous functions - Govern, Map, Measure, and Manage (National Institute of Standards and Technology, 2023). The framework does not tell organizations what to measure; it tells them that measurement is a precondition for trustworthy AI rather than a downstream activity. The 2024 Generative AI Profile sharpened this for foundation-model systems, calling out evaluation, documentation, and disclosure controls that have to be in place before generative systems are deployed at scale (National Institute of Standards and Technology, 2024). The framework’s posture, read carefully, is that without measurement infrastructure there is no risk management - only assertion. The case for treating measurement as infrastructure, not as ornament, has three parts. First, the systems themselves are now too capable and too widely deployed for narrative assurance to substitute for evidence. External evaluators including METR and its ARC Evals predecessor program have published public materials on frontier capability evaluation and dangerous-capability framing (METR, 2023; METR, 2025); the existence of those public evaluation efforts, whatever one thinks of their methodology, implies that the relevant questions cannot be answered by inspection alone. Second, governments have started to build state evaluation capacity - most visibly the UK AI Security Institute, which has positioned itself as a public-facing evaluation body for advanced AI safety and security (UK AI Security Institute, 2025). Third, the academic and open-source community has produced standing benchmark infrastructure, with Stanford’s Holistic Evaluation of Language Models (HELM) project running multi-dimensional evaluations across many models and a broad range of scenarios on a continuing basis (Stanford Center for Research on Foundation Models, 2025). Each of these efforts is partial. Each is contested. Together they constitute the first generation of what it would mean to have actual measurement infrastructure for AI - and they make visible how far that infrastructure still has to go.

Artículos relacionados

  • Interactive Media Use and Youth
    Modern advancements in technology have changed the way that young people use interactive media. Learning from such methods was not even considered until recently. It is now slowly defining the landscape of contemporary pedagogical practices. Interactive Media Use and Youth: Learning, Knowledge Exchange and Behavior provides a comprehensive collection of knowledge based on diffe...
  • Social and Economic Effects of Community Wireless Networks and Infrastructures
    Abdelaal
    It is surprising to think that in today’s rapidly evolving world of technology, over half of the globe still does not have access to high speed internet. Creating community wireless networks has in the past been a way to provide remote communities with internet and network access. Social and Economic Effects of Community Wireless Networks and Infrastructures highlights the succ...
  • The New Atlantis
    Francis Bacon
    In this work, Bacon portrayed a vision of the future of human discovery and knowledge, expressing his aspirations and ideals for humankind. The novel depicts the creation of a utopian land where 'generosity and enlightenment, dignity and splendour, piety and public spirit' are the commonly held qualities of the inhabitants of 'Bensalem'. ...
  • Sport 2.0
    Andy Miah
    ...
    Disponible

    28,86 €

  • Productivity Machines
    Corinna Schlombs
    ...
    Disponible

    48,52 €

  • Open Space
    Mariel Borowitz
    ...
    Disponible

    38,25 €

Otros libros del autor

  • The Urban Birdsong Decoder
    Nimble Books LLC
    A city bird is easiest to misidentify when the listener starts with a name instead of a pattern. In Opening Matter: The City Has a Sound Map, the controlling focus is color guide framing, real-audio spectrogram discipline, and differential diagnosis. The Cornell source packet supports species descriptions, occurrence checks, audio-archive use, and spectrogram workflow, but it d...
    Disponible

    35,27 €

  • Meridian CSM Exam Prep
    Nimble Books LLC
    Meridian CSM Exam Prep is a Meridian Certification Press study guide built for candidates who need a structured, print-friendly path from orientation to final review. The book emphasizes exam-domain coverage, practical vocabulary, scenario recognition, and disciplined self-assessment rather than unsupported promises of certification success. Readers get a coherent sequence of r...
    Disponible

    33,68 €

  • The Congress of Vienna Was a Dinner Party That Got Out of Hand
    Nimble Books LLC
    This book is what happens when you take a serious diplomatic event and notice that it was also a months-long house party with consequences. The Congress of Vienna, convened in late 1814 and concluded with its Final Act on 9 June 1815, redrew the map of Europe after twenty-three years of revolutionary and Napoleonic warfare (The Avalon Project at Yale Law School 1815; Encyclopae...
    Disponible

    33,86 €

  • Meridian CNA Exam Prep
    Nimble Books LLC
    Meridian CNA Exam Prep is an independent, source-grounded study guide for candidates preparing for the certified nursing assistant knowledge and skills exam. It explains the CNA role, resident rights, safety, infection control, basic nursing skills, activities of daily living, communication, and skills-test reasoning in plain exam-focused language. The guide is built around fed...
    Disponible

    33,69 €

  • The Coast Guard Index
    Nimble Books LLC
    It is a dangerous mistake to view the Coast Guard as just a medium-sized navy, maritime police force, emergency service, or one environmental agency. A mission index is more useful than a single institutional label. The official mission statement gives the service a broad public vocabulary, while federal law supplies the primary-duty frame and the departmental rule that explain...
    Disponible

    33,83 €

  • The Governance Gap Index
    Nimble Books LLC
    This book is the framework volume for an index. The Governance Gap Index is a structured way to compare what 195 nations have written down about artificial intelligence - laws, strategies, policies, ministerial declarations, agency guidance - against what those nations have actually built in the way of enforceable capacity to govern AI systems used inside their borders. It is n...
    Disponible

    33,78 €