Manuel Delaflor

Manuel Delaflor

Researcher

  • Metacognition Institute (U.K.)
Publications
23
First-authored
8
Projects
1
Lab co-authors
5
Venues
15

About

Manuel Delaflor is a researcher in the Human-AI Empowerment Lab and the director and main researcher of the Metacognition Institute (U.K.). An epistemologist and philosopher of science, Manuel began a research career with six years of work alongside Jacobo Grinberg at the National Autonomous University of Mexico (UNAM) and is developing Model Dependent Ontology, an epistemic framework that informs the lab's research on AI epistemics.

With the lab, Manuel studies sycophantic metacognition and confidence calibration in LLM self-assessment (ACM CUI 2026), the attribution of mental states and consciousness to AI (Frontiers in Psychology, 2026), the ethical consistency of LLMs under perturbation, their logic and fact-checking capabilities, and Socratic-dialogue systems such as Belief Explorer that support reflection on personal beliefs (CHI 2026). Manuel previously developed financial algorithms at Liv Capital and is also an award-winning photographer and digital artist.

Highlights

  • Co-author of more than 20 lab publications (2023-2026), including ACM CUI, ACM IVA, HCOMP, AAAI/ACM AIES, CHI and seven NeurIPS 2026 workshops
  • Director and Main Researcher, Metacognition Institute (U.K.), since 2023

Publications 23

NeurIPS 2026 AIWILD WorkshopForthcoming

Evidence Before Rankings: An Executable Audit Contract for Stateful Tool-Agent Evaluations

Carlos Toxtli-Hernández, Manuel Delaflor

Website
NeurIPS 2026 Trust-AI-Eval Workshop

When Calibration Records Cannot Support Routing: A Denominator and Artifact Audit

Carlos Toxtli-Hernández, Manuel Delaflor

PDF Website
NeurIPS 2026 AI4GOOD Workshop

Separating Governance Designation from Operational Fault in Model-Generated Incident Audits

Carlos Toxtli-Hernández, Manuel Delaflor

PDF Website
NeurIPS 2026 AI-Native Academia Workshop

Evaluator Disagreement in AI-Assisted Manuscript Revision

Carlos Toxtli-Hernández, Manuel Delaflor

PDF Website
NeurIPS 2026 AI-Native Academia Workshop

A Validation Contract for Anticipatory Peer-Review Benchmarks

Manuel Delaflor, Carlos Toxtli-Hernández

PDF Website
NeurIPS 2026 HAIC Workshop

How Many Agents Can One Supervisor Track? A Mechanistic Capacity Model and Measurement Protocol for LLM-Agent Teams

Carlos Toxtli-Hernández, Manuel Delaflor

PDF Website
NeurIPS 2026 HAIC Workshop

Headroom Before Protocol Effects: A Decision Procedure for Human-Agent Evaluation

Manuel Delaflor, Carlos Toxtli-Hernández

PDF Website
NeurIPS 2026 SocialAgent WorkshopForthcoming

Content Is Not Social Attribution: An Audit Protocol for Textual LLM Social Simulation

Carlos Toxtli-Hernández, Manuel Delaflor

Website
Frontiers in Psychology 2026 PerspectiveForthcoming

AI Consciousness? Attribution and Cognitive Biases

Carlos Toxtli-Hernández, Manuel Delaflor, Alejandro Tapia-V.

DOI
ACM IVA 2026

Personality-Profiled Virtual Agents Are More Predictable: The Constraint-Entropy Tradeoff for Trustworthy Agent Design

Carlos Toxtli-Hernández, Manuel Delaflor

DOI Website
CSCW 2026 Companion

Hidden Profile Decision Making in Multi-Agent LLM Groups

Carlos Toxtli-Hernández, Manuel Delaflor

DOI Website
HCOMP 2026

Qualification by Calibration: A Readable Benchmark for Admitting Language Models to Human-Computation Tasks

Carlos Toxtli-Hernández, Manuel Delaflor

DOI Website Code
ACM CUI 2026 Forthcoming

Sycophantic Metacognition: Investigating the Dunning-Kruger Effect in Large Language Model Self-Assessment

Manuel Delaflor, Carlos Toxtli-Hernández

DOI Website Code
AAAI/ACM AIES 2026 Forthcoming

Auditing LLM Portrayals of Neurodivergent People: Quantifying the Asymmetry Between Deficit Framing and Neurodiversity Affirmation

Carlos Toxtli-Hernández, Manuel Delaflor

Website
ACM HAI 2026 Forthcoming

LLM-Judge Behavioral Coding Scheme for Agent Teammate Quality

Carlos Toxtli-Hernández, Manuel Delaflor

Website
CHI 2026 Extended Abstract

Belief Explorer: A Preliminary Evaluation of AI-Mediated Socratic Dialogue for Epistemic Reflection

Manuel Delaflor, Cecilia Delgado Solorzano, Carlos Toxtli-Hernández

DOI Code
AHFE IHIET-AI 2025

Can We Trust Them? Examining the Ethical Consistency of Large Language Models to Perturbations

Manuel Delaflor, Cecilia Delgado Solorzano, Carlos Toxtli-Hernández

DOI
AHFE IHIET-FS 2025

Artificial Intelligence as Self-Instantiated, Temporally Continuous, Disturbance-Driven Adaptive World-Builder

Manuel Delaflor, Cecilia Delgado Solorzano, Carlos Toxtli-Hernández

DOI
AHFE IHIET 2025

A Multi-Perspective AI Framework for Mitigating Disinformation Through Contextual Analysis and Socratic Dialogue

Manuel Delaflor, Carlos Toxtli-Hernández

DOI
IEEE SmartData 2024

Assessing the Syllogistic Logic and Fact-Checking Capabilities of Large Language Models

Cecilia Delgado Solorzano, Manuel Delaflor, Carlos Toxtli-Hernández

DOI Website
CSCE 2024

Automatic Detection of Errors in LLM Large Benchmarks Using Frontier Model Consensus

Cecilia Delgado Solorzano, Manuel Delaflor, Carlos Toxtli-Hernández

DOI Website
AHFE IHIET-AI 2024

ReActIn: Infusing Human Feedback into Intermediate Prompting Steps of Large Language Model

Manuel Delaflor, Carlos Toxtli-Hernández, Claire Gendron, Wangfan Li, Cecilia Delgado Solorzano

DOI
arXiv 2023

Conceptual Framework for Autonomous Cognitive Entities

David Shapiro, Wangfan Li, Manuel Delaflor, Carlos Toxtli-Hernández

PDF DOI

In the news

  • Nine papers accepted at seven NeurIPS 2026 workshops, on auditing agent and AI evaluations, peer review in the age of AI, supervising teams of LLM agents, social simulation and mathematical reasoning, including Fateme Mazdarani's study of mathematical creativity in formal proof generation (MATH-AI).
  • Qualification by Calibration, a readable benchmark for admitting language models to human-computation tasks, is presented at HCOMP 2026 in Alexandria, VA.
  • Perspective article AI Consciousness? Attribution and Cognitive Biases accepted in Frontiers in Psychology.
  • Sycophantic Metacognition, on the Dunning-Kruger effect in LLM self-assessment, is presented as a full paper at ACM CUI 2026 in Bremen, Germany.
  • The lab presents three extended abstracts at ACM CHI 2026 in Barcelona, where Dr. Toxtli serves as Associate Chair.