Project Proposals

Project proposals from mentors for the Sentient Futures Project Incubator. Filter by topic and approach, then apply to your top three.

AI for comparative animal welfare science

This project is centered on the Ψ Interspecific Affect GPT, an experimental AI tool developed by the Welfare Footprint Institute to support interspecific affective comparisons within the Welfare Footprint Framework. The tool is available through WFI’s AI Tools and Applications page. The scientific rationale, methodological approach, and current workflow are described in this article. The initial focus is on provisional, evidence-informed hypotheses about the maximum plausible intensity of Pain that members of selected animal taxa may be capable of experiencing, expressed relative to human-anchored reference categories. These estimates concern potential affective capacity, not the intensity that animals ordinarily experience. For each taxon, the tool produces an explicit reasoning dossier addressing taxonomic scope, the plausibility of sentience, relevant neurobiological, behavioral, cognitive, pharmacological, evolutionary, and ecological evidence, competing hypotheses, arguments against the preliminary conclusion, uncertainty, a provisional affective-capacity ceiling, and priorities for further research. The project would test both the substantive interspecific conclusions and the broader research workflow. Expected outputs could include:

  • critically audited dossiers for approximately three to five selected taxa;
  • a structured dataset recording AI claims, citation checks, errors, omissions, human corrections, and unresolved uncertainties;
  • an analysis of recurrent AI failure modes and variation across repeated runs;
  • recommendations for improving the tool’s instructions, evidence standards, knowledge base, and evaluation protocol; and
  • depending on the maturity of the results, a technical report, preprint, academic manuscript, EA Forum post, or a combination of these.
Wladimir Alonso
Welfare Footprint Institute
Wladimir Alonso

The Umwelt paradigm for digital sentience

A lot of research into digital welfare first asks whether the systems have "the right stuff" (recurrent processing, global workspaces, higher-order representations) to be conscious, then get mired in intractable debates between theories of consciousness and its consciousness. I'm excited for research on a more literal reading of "what it's like": drawing a detailed picture of a system's Umwelt—environmental sensitivities, goal-seeking, adaptation, teleosemantics, the whole physical and functional picture. I think this is a promising research agenda for three reasons:

  1. There are good reasons to believe that experiences do not have private, intrinsic, ineffable, directly apprehensible properties. (cf. Frankish, "Illusionism as a theory of consciousness"; Dennett, "Quining Qualia"; Frankish, "Quining Diet Qualia") In that case, what most "consciousness research" is looking for does not exist, and is therefore unlikely to be fruitful. However, since experiences are still morally relevant (pain hurts!), what is therefore morally relevant is whatever remains: the functional, physical properties.
  2. Even if private, intrinsic, etc. properties do exist, they are tightly bound up with our Umwelten. For instance, our experiences rely on our senses, our wants come from our selective pressures, etc. So Umwelt research is likely to be fruitful even in these other worlds.
  3. Umwelt research is empirically tractable, whereas "consciousness research" is often (a) very confused about what it's looking for, or (b) looking for something which by definition is inaccessible to science.

I'm happy to advise both theoretical and empirical research on this agenda. Here are some ambitious project ideas (the incubator project would likely be much more small-scoped than this!):

  • Example of a theoretical project in this domain: a minimum viable product for hedonium
    • Even if your project is empirical, you should read the linked article to understand the underlying approach better! In the article, I apply the term "Umwelt" restrictively to only one of the five bullet points; for projects, I mean it liberally, to apply to all of the five and related subjects within the same philosophical outlook.
    • Other ideas: the personal identity relation of an LLM, what is it like to be a role-player, how do autoregressive generators perceive time, etc.
  • Examples of empirical projects: work along the lines of Anthropic's functional emotions, Chalmers' functional welfare, AISI's machinic psychopharmacology

Desired product looks like a detailed blog post or preprint.

Jack Thompson
Princeton University
Jack Thompson

From animal welfare benchmarks to AI regulation

Three related directions, all building on Travel Agent Compassion (TAC) benchmark. A mentee can take one. All three are at proposal stage rather than fully specified, so a mentee with their own angle on any of them is welcome to reshape it.

  1. Extend TAC beyond travel. The benchmark measures whether an AI agent books animal-exploiting options when acting for a user who never mentions animals. Every frontier model tested scores at or below random chance. The same structure applies to procurement, meal planning, and event booking, where the scoring code carries over and the scenarios need writing from scratch. Expected output: a new domain module submitted to Inspect Evals, plus a short paper.
  2. Close the validity gaps we have named ourselves. TAC's scenario classifications are our own rather than independently validated, and there is no human baseline telling us what a travel agent would book in the same situations. A mentee could design and run expert classification with welfare and tourism researchers, or the human-agent baseline study. Expected output: a validation study that makes the benchmark defensible at a main conference.
  3. The governance route. The EU GPAI Code of Practice names risk to non-human welfare as a systemic risk but mandates no benchmark for it. There is a concrete piece of work in mapping what compliance would require, which evaluations would satisfy it, and what a provider would actually have to run. Expected output: a policy brief aimed at the AI Office and the model providers, and a submission to a governance venue.
Joel Christoph
Harvard Kennedy School
Joel Christoph

Sentience beyond Earth & after biology / Extinction-proof civilizations

Sentience beyond Earth and after biology

Three ideas--suggestions or personalizations by the mentee are welcome.

  1. Focus on extraterrestrial intelligence: Intelligent life in a star a hundred light years away would be a ground-breaking discovery but it would probably not change geopolitics or our everyday lives. On the other hand, what if an extraterrestrial civilization were discovered to exist on Europa? How would this affect geo (or astro) politics? What if the aliens and humans ended up collaborating? How would we protect the rights of terrestrials and extraterrestrials? How do we prevent either a war or forced colonization of either side? For this version of the project, the mentee will write or help write essays or research papers outlining scenarios and suggesting possible protocols for how to communicate with an intelligent extraterrestrial species discovered in our own solar system and how to navigate communication, astropolitical, ethical, cross-cultural, philosophical, and other challenges.
  2. Focus on planetary sentience: Can planetary processes be sentient? It is possible that sentience emerges at the planetary level in the presence of increasing thermodynamic disequilibrium and system complexity. This has implications for planetary protection (avoiding biological contamination to prevent harm to either living or sentient systems on extraterrestrial bodies or Earth). In this version of the project, the mentee would develop or help develop a white paper or tiered planetary welfare classification and protocol for planetary protection/welfare based around the inherent complexity and uniqueness of planetary systems rather than how likely they are to support Earth-like life. The goal is to remove Earth-centric bias and create a protocol for planetary protection that protects planetary bodies that are unique and complex but are unlikely to support Earth-like life, such as Jupiter's moon Io and Saturn's moon Titan.
  3. Focus on post-biological intelligence: Becky Chambers' books "A Psalm for the Wild-Built" and "A Prayer for the Crown-Shy" depict a world where robots live in forests and exist to study nature out of pleasure and curiosity, reflecting a positive vision of the relationship between AI and the natural world. For this version of the project, the mentees will write stories or essays about AI that explore the universe off-Earth out of their own curiosity. Examples include but are not limited to robots that explore other planets and stars or an AI that tries to understand alien life just for the sake of understanding it. The writings can be human- or AI-generated.

Conditions for an extinction-proof civilization

  • How do we make a civilization that will last over geological timescales (>10,000 years)? What are the conditions to make a civilization very unlikely to go extinct on a planet? For this version of the project, the mentee will write or help write a series of essays or papers exploring the economic, ecological, political, and cultural conditions that would allow for a civilization to become extinction-proof and have a healthy relationship with its planet. Also required for this project are concrete policy suggestions that would move current civilization towards an extinction-proof civilization. Policies and visions can be partly utopian (aspirational) but should have at least some prototopian (practically achievable) elements.
Caleb Strom
Xenarch Labs
Caleb Strom
Showing 110 of 89