Jos Hernndez-Orallo
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
56
Venues
18
Active years
1999–2026
Best venue rank
A*
Where they publish
Papers
56 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | TRACE: A Corpus of Team Creative Discussions. | Yixuan Jiang, Tiancheng Hu, Jos Hernndez-Orallo, David Stillwell, Luning Sun |
| 2025 | ACL | PredictaBoard: Benchmarking LLM Score Predictability. | Lorenzo Pacchiardi, Konstantinos Voudouris, Ben Slater, Fernando Martnez-Plumed, Jos Hernndez-Orallo, Lexin Zhou, Wout Schellaert |
| 2025 | ECAI | Relative Drawing Identification Complexity Is Invariant to Modality in Vision-Language Models. | Diogo Freitas, Brigt Hvardstun, Daro Garigliotti, Jan Arne Telle, Csar Ferri, Jos Hernndez-Orallo |
| 2025 | IJCAI | Paradigms of AI Evaluation: Mapping Goals, Methodologies and Culture. | John Burden, Marko Tesic, Lorenzo Pacchiardi, Jos Hernndez-Orallo |
| 2025 | IJCAI | Contamination Budget: Trade-offs Between Breadth, Depth and Difficulty. | Behzad Mehrbakhsh, Fernando Martnez-Plumed, Jos Hernndez-Orallo |
| 2024 | AAAI | Your Prompt Is My Command: On Assessing the Human-Centred Generality of Multimodal Models (Abstract Reprint). | Wout Schellaert, Fernando Martnez-Plumed, Karina Vold, John Burden, Pablo A. M. Casares, Bao Sheng Loe, Roi Reichart, Sen higeartaigh, Anna Korhonen, Jos Hernndez-Orallo |
| 2024 | ECAI | Caveats and Solutions for Characterising General-Purpose AI. | Jos Hernndez-Orallo |
| 2024 | ECAI | Distilling the Effects of Language Model Contamination. | Behzad Mehrbakhsh, Fernando Martnez-Plumed, Jos Hernndez-Orallo |
| 2024 | ECAI | Language Task Difficulty Prediction Through LLM-Annotated Meta-Features. | Yael Moros-Daval, Fernando Martnez-Plumed, Jos Hernndez-Orallo |
| 2024 | IDEAL | How Resilient are Language Models to Text Perturbations? | Daniel Romero-Alvarado, Jos Hernndez-Orallo, Fernando Martnez-Plumed |
| 2023 | ECAI | Adversarial Benchmark Evaluation Rectified by Controlling for Difficulty. | Behzad Mehrbakhsh, Fernando Martnez-Plumed, Jos Hernndez-Orallo |
| 2022 | AAAI | How General-Purpose Is a Language Model? Usefulness and Safety with Human Prompters in the Wild. | Pablo Antonio Moreno Casares, Bao Sheng Loe, John Burden, Sen higeartaigh, Jos Hernndez-Orallo |
| 2022 | AAAI | Training on the Test Set: Mapping the System-Problem Space in AI. | Jos Hernndez-Orallo, Wout Schellaert, Fernando Martnez-Plumed |
| 2022 | AAAI | When AI Difficulty Is Easy: The Explanatory Power of Predicting IRT Difficulty. | Fernando Martnez-Plumed, David Castellano Falcn, Carlos Monserrat Aranda, Jos Hernndez-Orallo |
| 2022 | IJCAI | Not a Number: Identifying Instance Features for Capability-Oriented Evaluation. | Ryan Burnell, John Burden, Danaja Rutar, Konstantinos Voudouris, Lucy Cheke, Jos Hernndez-Orallo |
| 2022 | IJCAI | A Framework for Categorising AI Evaluation Instruments. | Anthony G. Cohn, Jos Hernndez-Orallo, Julius Sechang Mboli, Yael Moros-Daval, Zhiliang Xiang, Lexin Zhou |
| 2022 | IJCAI | Non-Cheating Teaching Revisited: A New Probabilistic Machine Teaching Model. | Csar Ferri, Jos Hernndez-Orallo, Jan Arne Telle |
| 2022 | IJCAI | Measuring the Occupational Impact of AI: Tasks, Cognitive Abilities and AI Benchmarks (Extended Abstract). | Songl Tolan, Annarosa Pesole, Fernando Martnez-Plumed, Enrique Fernndez-Macas, Jos Hernndez-Orallo, Emilia Gmez |
| 2022 | IJCAI | Evaluating Object Permanence in Embodied Agents using the Animal-AI Environment. | Konstantinos Voudouris, Niall Donnelly, Danaja Rutar, Ryan Burnell, John Burden, Jos Hernndez-Orallo, Lucy Cheke |
| 2022 | IJCAI | Reject Before You Run: Small Assessors Anticipate Big Language Models. | Lexin Zhou, Fernando Martnez-Plumed, Jos Hernndez-Orallo, Csar Ferri, Wout Schellaert |
| 2021 | AAAI | Negative Side Effects and AI Agent Indicators: Experiments in SafeLife. | John Burden, Jos Hernndez-Orallo, Sen higeartaigh |
| 2021 | IDA | Muppets: Multipurpose Table Segmentation. | Gust Verbruggen, Lidia Contreras Ochando, Csar Ferri, Jos Hernndez-Orallo, Luc De Raedt |
| 2020 | AAAI | Exploring AI Safety in Degrees: Generality, Capability and Control. | John Burden, Jos Hernndez-Orallo |
| 2020 | AIES | Does AI Qualify for the Job?: A Bidirectional Model Mapping Labour and AI Intensities. | Fernando Martnez-Plumed, Songl Tolan, Annarosa Pesole, Jos Hernndez-Orallo, Enrique Fernndez-Macas, Emilia Gmez |
| 2020 | DSN | AI Safety Landscape From short-term specific system engineering to long-term artificial general intelligence. | Jos Hernndez-Orallo |
| 2020 | ECAI | Family and Prejudice: A Behavioural Taxonomy of Machine Learning Techniques. | Ral Fabra-Boluda, Csar Ferri, Fernando Martnez-Plumed, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2020 | ECAI | Finite and Confident Teaching in Expectation: Sampling from Infinite Concept Classes. | Jos Hernndez-Orallo, Jan Arne Telle |
| 2020 | ECAI | AI Paradigms and AI Safety: Mapping Artefacts and Techniques to Safety Issues. | Jos Hernndez-Orallo, Fernando Martnez-Plumed, Shahar Avin, Jess Whittlestone, Sen higeartaigh |
| 2020 | ECAI | Tracking AI: The Capability Is (Not) Near. | Fernando Martnez-Plumed, Jos Hernndez-Orallo, Emilia Gmez |
| 2019 | AAAI | Surveying Safety-relevant AI Characteristics. | Jos Hernndez-Orallo, Fernando Martnez-Plumed, Shahar Avin, Sen higeartaigh |
| 2019 | AIES | AI Extenders: The Ethical and Societal Implications of Humans Cognitively Extended by AI. | Jos Hernndez-Orallo, Karina Vold |
| 2018 | IJCAI | The Facets of Artificial Intelligence: A Framework to Track the Evolution of AI. | Fernando Martnez-Plumed, Bao Sheng Loe, Peter A. Flach, Sen higeartaigh, Karina Vold, Jos Hernndez-Orallo |
| 2017 | IJCAI | Computer Models Solving Intelligence Test Problems: Progress and Implications (Extended Abstract). | Jos Hernndez-Orallo, Fernando Martnez-Plumed, Ute Schmid, Michael Siebers, David L. Dowe |
| 2016 | ECAI | Is Spearman's Law of Diminishing Returns (SLODR) Meaningful for Artificial Agents? | Jos Hernndez-Orallo |
| 2016 | ECAI | Making Sense of Item Response Theory in Machine Learning. | Fernando Martnez-Plumed, Ricardo B. C. Prudncio, Adolfo Martnez Us, Jos Hernndez-Orallo |
| 2014 | ICMLA | A Knowledge Growth and Consolidation Framework for Lifelong Machine Learning Systems. | Fernando Martnez-Plumed, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2011 | ICML | A Coherent Interpretation of AUC as a Measure of Aggregated Classification Performance. | Peter A. Flach, Jos Hernndez-Orallo, Csar Ferri Ramirez |
| 2011 | ICML | Brier Curves: a New Cost-Based Visualisation of Classifier Performance. | Jos Hernndez-Orallo, Peter A. Flach, Csar Ferri Ramirez |
| 2010 | FLOPS | An Integrated Distance for Atoms. | Vicent Estruch, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2010 | ICDM | Quantification via Probability Estimators. | Antonio Bella, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2009 | IDEAL | Similarity-Binning Averaging: A Generalisation of Binning Calibration. | Antonio Bella, Csar Ferri Ramirez, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2009 | PAKDD | An Instantiation of Hierarchical Distance-Based Conceptual Clustering for Propositional Learning. | Ana Funes, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2007 | IDEAL | Joint Cutoff Probabilistic Estimation Using Simulation: A Mailing Campaign Application. | Antonio Bella, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2006 | ILP | Minimal Distance-Based Generalisation Operators for First-Order Objects. | Vicent Estruch, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2005 | ICMLA | Knowledge acquisition through machine learning: minimising expert's effort. | Ricardo Blanco-Vega, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2005 | ILP | Distance Based Generalisation. | Vicent Estruch, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2004 | DIS | Analysing the Trade-Off Between Comprehensibility and Accuracy in Mimetic Models. | Ricardo Blanco-Vega, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2004 | ICML | Delegating classifiers. | Csar Ferri, Peter A. Flach, Jos Hernndez-Orallo |
| 2002 | DIS | From Ensemble Methods to Comprehensible Models. | Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2002 | ICCS | Induction of Decision Multi-trees Using Levin Search. | Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2002 | ICML | Learning Decision Trees Using the Area Under the ROC Curve. | Csar Ferri, Peter A. Flach, Jos Hernndez-Orallo |
| 2002 | JELIA | SMILES: A Multi-purpose Learning System. | Vicent Estruch, Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2001 | FLOPS | Incremental Learning of Functional Logic Programs. | Csar Ferri, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2000 | FASE | Software as Learning: Quality Factors and Life-Cycle Revised. | Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 2000 | ILP | Universal Learning of Classes from Sparse and Non-uniform Evidence. | Csar Ferri Ramirez, Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |
| 1999 | ILP | A Strong Complete Schmema for Inductive Functional Logic Programming. | Jos Hernndez-Orallo, M. Jos Ramrez-Quintana |