AutoRecLab: An Autonomous Recommender Systems Lab

Abstract. Recommender systems (RecSys) research depends on extensive empirical evaluation, yet translating experimental designs into executable code remains a manual, error-prone process. This paper presents AutoRecLab, a Python-based autonomous RecSys lab that automates RecSys experimentation from natural-language prompts. Starting from a research idea, AutoRecLab derives explicit experiment requirements, develops and Read more…

From AutoRecSys to AutoRecLab: A Call to Build, Evaluate, and Govern Autonomous Recommender-Systems Research Labs

Joeran Beel (University of Siegen, Germany) Bela Gipp (University of Göttingen, Germany) Tobias Vente (University of Antwerp, Belgium) Moritz Baumgart (University of Siegen, Germany) Philipp Meister (University of Göttingen, Germany) Pre-Print @misc{beel2025autorecsysautoreclabbuildevaluate, title={From AutoRecSys to AutoRecLab: A Call to Build, Evaluate, and Govern Autonomous Recommender-Systems Research Labs}, author={Joeran Beel and Read more…

Min-Yen Kan and Joeran Beel speak with ‘Nature’ about the potential (and threats) of AI Scientists

Following my recent visit to Prof. Min-Yen Kan at the National University of Singapore, Moritz Baumgart, Min-Yen Kan and I co-authored a paper critically evaluating Sakana’s AI Scientist (to be published in SIGIR Forum). Our review attracted the attention of Nature, which was preparing an editorial feature on the opportunities Read more…

Reproducing Machine Learning: Seven Bachelor Projects That Hold Science Accountable

Over the past semester, several Bachelor students at the University of Siegen undertook a bold challenge in our Machine Learning Praktikum: they didn’t just learn algorithms—they tried to reproduce published machine learning or recommender systems research. Each student or team selected a recent paper, rebuilt the experimental pipeline, validated (or Read more…

Versorgungsauskunft für neue Beamte/Professoren: Mein Fall, mein Widerspruch, mein Ergebnis (8 Jahre mehr = 100.000€ mehr Pension)

In diesem Blogpost dokumentiere ich, wie das Landesamt für Besoldung und Versorgung NRW (LBV) versucht hat, mir eine Versorgungsauskunft faktisch zu verwehren, wie anschließend zu meinen Ungunsten entschieden wurde – und wie ich am Ende dennoch rund acht Jahre mehr ruhegehaltsfähige Zeit anerkannt bekommen habe, als zunächst vorgesehen. Das entspricht Read more…

Evaluating Sakana’s AI Scientist: Bold Claims, Mixed Results, and a Promising Future?

Abstract. Recently, Sakana.ai introduced the AI Scientist, which claims to automate the entire research lifecycle and conduct research autonomously. This is a concept we call Artificial Research Intelligence (ARI). Achieving ARI would be a major milestone toward Artificial General Intelligence (AGI) and a prerequisite to achieving Super Intelligence. The AI Read more…

Checky, the Paper-Submission Checklist Generator for Authors, Reviewers and LLMs

Following our proposal for evidence-based best-practices for recommender systems evaluation and our Dagstuhl manuscript about Best-Practices for Offline Evaluations of Recommender Systems, we are glad to announce Checky, a tool for conference chairs and journal editors to create and manage submission checklists. Abstract. Submission checklists have become increasingly prevalent for Read more…

ChatGPT (und ich) verlieren gegen die LVM vor dem Amtsgericht – ein Praxistest für den KI-Anwalt

Vor etwa einem halben Jahr endete mein erster Versuch, mit Unterstützung von ChatGPT einen Zivilprozess selbst zu führen. Das Ergebnis vorweg: Wir haben verloren, auf eine ziemlich unbefriedigende Art. Das Amtsgericht hat nicht entschieden, ob meine tatsächliche Position gegenüber der Beklagten richtig oder falsch war. Es hat meine Klage als Read more…

8 Recommender Systems Illustration

ISG will present 8 papers and posters at the ACM Recommender Systems Conference and Workshops

We are thrilled that 8 of our 11 submissions to the 18th ACM Recommender-Systems Conference and Workshops (RobustRecSys and RecSoGood) were accepted for publication. Our research was conducted jointly with partners from the University of Gothenburg (Alan Said), the University of Antwerpen (Lien Michiels), and some excellent Bachelor and Master Read more…

Our use of AI-tools for writing research papers

The German Research Association (Deutsche Forschungsgemeinschaft, DFG) recently issued guidelines for using artificial intelligence, including tools like ChatGPT, to write research papers and grant applications. The DFG supports using AI for these purposes, excluding reviews but emphasizes the need for transparency. Therefore, “Scientists should disclose whether, for what purpose, and Read more…

e-fold cross-validation: A computing and energy-efficient alternative to k-fold cross-validation with adaptive folds [Proposal]

This proposal is also available as pre-print (PDF) on OSF.io. If you want to cite this proposal, please cite: Introduction K-fold cross-validation is widely regarded as a robust method for model evaluation in machine learning and related fields, including recommender systems. Unlike a simple hold-out split, k-fold cross-validation ensures that Read more…

Deutsche und Indische Fachkräfte bzw. Studenten im Vergleich: Von vergleichbar bis katastrophal

Im Mai 2023 veröffentlichte ich einen Tweet auf Twitter Beitrag auf X. Dort zeigte und kommentierte ich Statistiken, dass eine große Anzahl vorwiegend indischer Studenten in einer meiner Vorlesungen plagiierte. Einigen Personen gefiel dieser Beitrag nicht und sie beschwerten sich. Zugegebener Maßen ist es schwierig, sich auf X in kurzen Read more…