Order:
  1. Reverse Turing Tests for Human-Machine Task Suitability Assessments Should be Profile-Driven.Jonathan Prunty, Marko Tešić, John Burden, Ben Slater, Zachary Tidler, Paul Clothier, Luning Sun, Katherine Collins, Bernardo Gonçalves, Giulio Corsi, Seán Ó hÉigeartaigh, Lucy Cheke & Jose Hernandez-Orallo - manuscript
    As AI is integrated into the workplace, organisations increasingly face allocation decisions between human and machine workers. These decisions are increasingly made or assisted by algorithms, creating a Reverse Turing Test dynamic wherein the machine is now the judge. In addition, human and machine workers may ``compete'' for a given task, reproducing aspects of adversarial games. This raises new methodological questions about assessing task suitability between humans and machines. The criteria often used to assess people (e.g., education, experience, references) cannot (...)
    No categories
    Direct download  
     
    Export citation  
     
    Bookmark  
  2. Your Prompt is my command: On Assessing the Human-Centred Generality of Multimodal Models.Wout Schellaert, Fernando Martínez-Plumed, Karina Vold, John Burden, Pablo A. M. Casares, Bao Sheng Loe, Roi Reichart, Sean Ó hÉigeartaigh, Anna Korhonen & José Hernández-Orallo - 2023 - Journal of Artificial Intelligence Research 77.
    Even with obvious deficiencies, large prompt-commanded multimodal models are proving to be flexible cognitive tools representing an unprecedented generality. But the directness, diversity, and degree of user interaction create a distinctive “human-centred generality” (HCG), rather than a fully autonomous one. HCG implies that —for a specific user— a system is only as general as it is effective for the user’s relevant tasks and their prevalent ways of prompting. A human-centred evaluation of general-purpose AI systems therefore needs to reflect the personal (...)
    No categories
    Direct download  
     
    Export citation  
     
    Bookmark   2 citations  
  3.  12
    Predictable artificial intelligence.Lexin Zhou, Pablo A. M. Casares, Fernando Martínez-Plumed, John Burden, Ryan Burnell, Lucy Cheke, Cèsar Ferri, Alexandru Marcoci, Behzad Mehrbakhsh, Yael Moros-Daval, Seán Ó hÉigeartaigh, Danaja Rutar, Wout Schellaert, Konstantinos Voudouris & José Hernández-Orallo - 2026 - Artificial Intelligence 353 (C):104491.
    Direct download (2 more)  
     
    Export citation  
     
    Bookmark