8 found
Order:
  1. Taking AI Welfare Seriously.Robert Long, Jeff Sebo, Patrick Butlin, Kathleen Finlinson, Kyle Fish, Jacqueline Harding, Jacob Pfau, Toni Sims, Jonathan Birch & David Chalmers - manuscript
    In this report, we argue that there is a realistic possibility that some AI systems will be conscious and/or robustly agentic in the near future. That means that the prospect of AI welfare and moral patienthood — of AI systems with their own interests and moral significance — is no longer an issue only for sci-fi or the distant future. It is an issue for the near future, and AI companies and other actors have a responsibility to start taking it (...)
    Direct download (2 more)  
     
    Export citation  
     
    Bookmark   109 citations  
  2. Operationalising Representation in Natural Language Processing.Jacqueline Harding - 2023 - British Journal for the Philosophy of Science.
    Despite its centrality in the philosophy of cognitive science, there has been little prior philosophical work engaging with the notion of representation in contemporary NLP practice. This paper attempts to fill that lacuna: drawing on ideas from cognitive science, I introduce a framework for evaluating the representational claims made about components of neural NLP models, proposing three criteria with which to evaluate whether a component of a model represents a property and operationalising these criteria using probing classifiers, a popular analysis (...)
    Direct download (4 more)  
     
    Export citation  
     
    Bookmark   34 citations  
  3. What is AI safety? What do we want it to be?Jacqueline Harding & Cameron Domenico Kirk-Giannini - 2025 - Philosophical Studies 182 (7):1495-1518.
    The field of AI safety seeks to prevent or reduce the harms caused by AI systems. A simple and appealing account of what is distinctive of AI safety as a field holds that this feature is constitutive: a research project falls within the purview of AI safety just in case it aims to prevent or reduce the harms caused by AI systems. Call this appealingly simple account The Safety Conception of AI safety. Despite its simplicity and appeal, we argue that (...)
    Direct download (4 more)  
     
    Export citation  
     
    Bookmark   3 citations  
  4. What is it for a Machine Learning Model to Have a Capability?Jacqueline Harding & Nathaniel Sharadin - forthcoming - British Journal for the Philosophy of Science.
    What can contemporary machine learning (ML) models do? Given the proliferation of ML models in society, answering this question matters to a variety of stakeholders, both public and private. The evaluation of models' capabilities is rapidly emerging as a key subfield of modern ML, buoyed by regulatory attention and government grants. Despite this, the notion of an ML model possessing a capability has not been interrogated: what are we saying when we say that a model is able to do something? (...)
    Direct download (4 more)  
     
    Export citation  
     
    Bookmark   10 citations  
  5. AI language models cannot replace human research participants.Jacqueline Harding, William D’Alessandro, N. G. Laskowski & Robert Long - 2024 - AI and Society 39 (5):2603-2605.
    In a recent letter, Dillion et. al (2023) make various suggestions regarding the idea of artificially intelligent systems, such as large language models, replacing human subjects in empirical moral psychology. We argue that human subjects are in various ways indispensable.
    Direct download (4 more)  
     
    Export citation  
     
    Bookmark   6 citations  
  6. A Communication-First Account of Explanation.Jacqueline Harding, Tobias Gerstenberg & Thomas F. Icard - forthcoming - Noûs.
    This paper develops a formal account of causal explanation, grounded in a theory of conversational pragmatics, and inspired by the interventionist idea that explanation is about asking and answering what-if-things-had-been-different questions. We illustrate the fruitfulness of the account, relative to previous accounts, by showing that widely recognised "explanatory virtues" emerge naturally, as do subtle empirical patterns concerning the impact of norms on causal judgments. This shows the value of a "communication-first" approach to explanation: getting clear on explanation’s communicative dimension is (...)
    Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  7. Everettian Quantum Mechanics and the Metaphysics of Modality.Jacqueline Harding - 2021 - British Journal for the Philosophy of Science 72 (4):939-964.
    This article sits at a point of intersection between the philosophy of physics and the metaphysics of modality. There are clear similarities between Everettian quantum mechanics and various modal metaphysical theories, but there have hitherto been few attempts at exploring how the two topics relate. In this article, I build on a series of recent papers by Wilson ([2011], [2012], [2013]), who argues that Everettian quantum mechanics’ connections with traditional modal metaphysics are vital in defending it against objections. I show (...)
    Direct download (3 more)  
     
    Export citation  
     
    Bookmark   3 citations  
  8.  71
    Proxy Selection in Transitive Proxy Voting.Jacqueline Harding - 2022 - Social Choice and Welfare 58:69-99.
    Transitive proxy voting (or "liquid democracy") is a novel form of collective decision making, often framed as an attractive hybrid of direct and representative democracy. Although the ideas behind liquid democracy have garnered widespread support, there have been relatively few attempts to model it formally. This paper makes three main contributions. First, it proposes a new social choice-theoretic model of liquid democracy, which is distinguished by taking a richer formal perspective on the process by which a voter chooses a proxy. (...)
    Direct download (2 more)  
     
    Export citation  
     
    Bookmark