Publications

publications by categories in reversed chronological order. generated by jekyll-scholar.

2026

  1. ICML
    Calibrated Preference Learning: The Case of Label Ranking
    Santo M. A. R. Thies, Viktor Bengs, Timo Kaufmann, and 2 more authors
    In Proceedings of the International Conference on Machine Learning (ICML), 2026

2025

  1. NeurIPS
    ResponseRank: Data-Efficient Reward Modeling through Preference Strength Learning
    Timo Kaufmann, Yannick Metz, Daniel Keim, and 1 more author
    In The Annual Conference on Neural Information Processing Systems (NeurIPS), 2025
  2. TMLR
    A Survey of Reinforcement Learning from Human Feedback
    Timo Kaufmann, Paul Weng, Viktor Bengs, and 1 more author
    Transactions on Machine Learning Research, 2025
  3. arXiv
    Feedback Forensics: A Toolkit to Measure AI Personality
    Arduin Findeis, Timo Kaufmann, Eyke Hüllermeier, and 1 more author
    2025
  4. ICLR
    Inverse Constitutional AI: Compressing Preferences into Principles
    Arduin Findeis, Timo Kaufmann, Eyke Hüllermeier, and 2 more authors
    In Proceedings of the International Conference on Learning Representations (ICLR), 2025
  5. ICML
    Comparing Comparisons: Informative and Easy Human Feedback with Distinguishability Queries
    Xuening Feng, Zhaohui Jiang, Timo Kaufmann, and 3 more authors
    In Proceedings of the International Conference on Machine Learning (ICML), 2025
  6. AAAI
    DUO: Diverse, Uncertain, On-Policy Query Generation and Selection for Reinforcement Learning from Human Feedback
    Xuening Feng, Zhaohui Jiang, Timo Kaufmann, and 4 more authors
    In Proceedings of the AAAI Conference on Artificial Intelligence, 2025
  7. CL
    Problem Solving Through Human-AI Preference-Based Cooperation
    Subhabrata Dutta, Timo Kaufmann, Goran Glavaš, and 7 more authors
    Computational Linguistics, 2025

2024

  1. MHFAIA
    Comparing Comparisons: Informative and Easy Human Feedback with Distinguishability Queries
    Xuening Feng, Zhaohui Jiang, Timo Kaufmann, and 3 more authors
    In ICML 2024 Workshop on Models of Human Feedback for AI Alignment (MHFAIA), 2024
  2. MHFAIA
    Relatively Rational: Learning Utilities and Rationalities Jointly from Pairwise Preferences
    Taku Yamagata, Tobias Oberkofler, Timo Kaufmann, and 3 more authors
    In ICML 2024 Workshop on Models of Human Feedback for AI Alignment (MHFAIA), 2024
  3. RLBRew
    OCALM: Object-Centric Assessment with Language Models
    Timo Kaufmann, Jannis Blüml, Antonia Wüst, and 3 more authors
    In RLC 2024 Workshop on Reinforcement Learning Beyond Rewards (RLBRew), 2024

2023

  1. HLDM
    On the Challenges and Practices of Reinforcement Learning from Real Human Feedback
    Timo Kaufmann, Sarah Ball, Jacob Beck, and 2 more authors
    In Machine Learning and Principles and Practice of Knowledge Discovery in Databases, 2023
  2. Reinforcement Learning from Human Feedback for Cyber-Physical Systems: On the Potential of Self-Supervised Pretraining
    Timo Kaufmann, Viktor Bengs, and Eyke Hüllermeier
    In Proceedings of the International Conference on Machine Learning for Cyber-Physical Systems (ML4CPS), 2023