Sign in

Martin Potthast

@martin-potthast.com
2.6K followers 1.4K following 37 posts

Professor at the University of Kassel, hessian.AI, and ScaDS.AI. Member of @webis.de Research in information retrieval #IR, natural language processing #NLP, and artificial intelligence.

PostsRepliesMedia
Reposted by Martin Potthast
Webis Group @webis.de · 30/09/2026
Call for Participation: Deadlines extended for UniAgent, the 1st Shared Task on Agentic AI in University Administration at the CIKM 2026 AnalytiCup in Rome. Registration until Oct 3, submission Oct 20. Real administrative cases; we host the LLMs. Details: uniagent.webis.de/cikm26/uniag...
uniagent.webis.de
UniAgent at CIKM 2026 - Agentic AI in University Administration
A shared task on agentic AI in university administration, part of the CIKM 2026 AnalytiCup: build agents that solve standardized administrative tasks under controlled tool access and expert-judged ref...
032
Reposted by Martin Potthast
LION Lab @lionlab.bsky.social · 17/09/2026
🦁LION Lab is hiring! 🧑‍🔬One fully funded PhD student (TVL-13 100%) for 3 years 💉Topic: Interpretability for Protein Language Models 👥Advised by @weissweiler.bsky.social together with Clara Schoeder 🌍Leipzig, Germany 🔗Apply by Oct 15: lionlabnlp.github.io/jobs/ai4pf/ Please share! #NLProc #NLP
0109
Reposted by Martin Potthast
Harry Scells @hscells.bsky.social · 27/08/2026
📣 The data for the UP2DATE@SCOLIA shared task about #systematicreview updates has now been released! Visit up2date.health-nlp.com for the details!
up2date.health-nlp.com
UP2DATE @ SCOLIA 2027 is a shared task focusing on the development of new methods to assist in updating systematic reviews.
022
Reposted by Martin Potthast
Webis Group @webis.de · 03/09/2026
Call for Participation: UniAgent, the 1st Shared Task on Agentic AI in University Administration, at the CIKM 2026 AnalytiCup in Rome. Registration closes Sep 30. Submission deadline Oct 23. Details: uniagent.webis.de/cikm26/uniag...
uniagent.webis.de
UniAgent at CIKM 2026 - Agentic AI in University Administration
A shared task on agentic AI in university administration, part of the CIKM 2026 AnalytiCup: build agents that solve standardized administrative tasks under controlled tool access and expert-judged ref...
033
Martin Potthast @martin-potthast.com · 17/08/2026
I guess if someone were to sit down and write a letter to you with pen and paper, this would sufficiently stick out to be worthy of note. In an environment of frictionless communication, making the deliberate choice of investing personal time into something that is not frictionless is relevant.
2130
Reposted by Martin Potthast
Lukas Gienapp @lgnp.bsky.social · 22/07/2026
Super happy that our paper “Topic-Specific Classifiers are Better Relevance Judges than Prompted LLMs” just won Best Student Paper at #SIGIR2026🏆 It argues against the reflex to reach for an LLM whenever you need relevance judgments. What we found 🧵:
1114
Martin Potthast @martin-potthast.com · 10/07/2026
Can you binary search the actual offensive reference? Or is it the combination of referenced on the page?
000
Martin Potthast @martin-potthast.com · 20/06/2026
I'd like to peek into so elses For You feed just to see what they trained for themselves. Being mostly a passive user, my For You feed will never be trained well enough. But so else "For Them"-Feed just might.
220
Martin Potthast @martin-potthast.com · 01/05/2026
It's actually kinda hard to get some CL work accepted; some reviewers ask why no LLMs were used to solve the problem. 😅
190
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 26/03/2026
I am happy to share that I defended my PhD thesis this week. I want to thank all the people I collaborated with and met at conferences and workshops over the years! Especially a big thank you to my PhD supervisor, Matthias Hagen, and to Laura Dietz and Udo Kruschwitz for reviewing my thesis.
3244
Martin Potthast @martin-potthast.com · 07/03/2026
I think this is perfectly alright. A bit of show and tell is just what we need in order to collect enough observations until eventually someone is inspired to conduct "the next Cranfield experiments" that show us the new optimization goal the community can get behind as a whole.
020
Martin Potthast @martin-potthast.com · 27/02/2026
Post a pic you took, no context, to bring some zen to the feed.
A photo of two cats sitting side by side on the backrest of a light-colored leather sofa, right in front of a slanted skylight window. Both cats are facing the camera with wide, round eyes and calm, slightly serious expressions.

The cat on the left is smaller, golden-brown with darker shading and a fluffy tail hanging straight down.

The cat on the right is larger, pale cream/gray with a plush coat and a longer tail also hanging down.

The window behind them has raindrops or condensation on the glass, and rooftops can be seen outside. The wooden window frame and angled ceiling suggest an attic or top-floor room.
020
Martin Potthast @martin-potthast.com · 01/02/2026
I'd hope for a revival of StackExchange.
110
Martin Potthast @martin-potthast.com · 02/01/2026
No cheating, repost the most recent picture of your pet(s)
VickyOlly
050
Martin Potthast @martin-potthast.com · 24/12/2025
I'm thinking in a similar direction. I keep asking myself: How can we reliably establish human influence, effort, or sincerity in a digital object, which could have also been generated? Thesis: Authenticity checks will replace part of the time savings of GenAI. Before quality of writing sufficed.
020
Reposted by Martin Potthast
Leonie Weissweiler @weissweiler.bsky.social · 11/12/2025
🧑‍🔬I’m recruiting PhD students in Natural Language Processing @unileipzig.bsky.social Computer Science, together with @scadsai.bsky.social! Topics include, but aren’t limited to: 🔎Linguistic Interpretability 🌍Multilingual Evaluation 📖Computational Typology Please share! #NLProc #NLP
14225
Martin Potthast @martin-potthast.com · 16/11/2025
Seconded. If it needs to be separate from other situations where induction occurs, then maybe: In-context induction, though that does not differentiate it too well from the model's inference run itself. Maybe: Inductive prompting?
010
Reposted by Martin Potthast
Webis Group @webis.de · 27/10/2025
We just released "German Commons", the largest openly-licensed German text dataset for LLM training: 154B tokens with clear usage rights for research and commercial use. huggingface.co/datasets/coral-nlp/german-commons
huggingface.co
coral-nlp/german-commons · Datasets at Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1209
Martin Potthast @martin-potthast.com · 14/09/2025
ChatGPT gives you little to no predictability of results in how your text changes. Grammarly and DeepL do.
000
Reposted by Martin Potthast
Johanne Trippas @jtrippas.bsky.social · 02/09/2025
🌟Really excited to share the fourth Strategic Workshop on Information Retrieval (SWIRL) report published in SIGIR Forum! Paper 👉🏻 www.johannetrippas.com/papers/tripp... More info 👉🏻 sites.google.com/view/swirl20... #SWIRL2025 #SIGIR2026 #IR #GenAI #Research #CHIIR2026
01310
Reposted by Martin Potthast
Webis Group @webis.de · 18/07/2025
Thrilled to announce that Matti Wiegmann has successfully defended his PhD! 🎉🧑‍🎓 Huge congratulations on this incredible achievement! #PhDDefense #AcademicMilestone
2123
Reposted by Martin Potthast
Webis Group @webis.de · 18/07/2025
Honored to win the ICTIR Best Paper Honorable Mention Award for "Axioms for Retrieval-Augmented Generation"! Our new axioms are integrated with ir_axioms: github.com/webis-de/ir_... Nice to see axiomatic IR gaining momentum.
1166
Reposted by Martin Potthast
Webis Group @webis.de · 18/07/2025
We presented two papers at ICTIR 2025 today: - Axioms for Retrieval-Augmented Generation webis.de/publications... - Learning Effective Representations for Retrieval Using Self-Distillation with Adaptive Relevance Margins webis.de/publications...
183
Reposted by Martin Potthast
Ferdinand Schlatt @fschlatt.bsky.social · 16/07/2025
Want to know how to make bi-encoders more than 3x faster with a new backbone encoder model? Check out our talk on the Token-Independent Text Encoder (TITE) #SIGIR2025 in the efficiency track. It pools vectors within the model to improve efficiency dl.acm.org/doi/10.1145/...
0105
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 16/07/2025
Now @fschlatt.bsky.social presents "TITE: Token-Independent Text Encoder for Information Retrieval" at #SIGIR2025 Paper: webis.de/publications...
083
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 18/07/2025
Here are some impressions from our ReNeuIR workshop on "Reaching Efficiency in Neural IR" that we had yesterday at #SIGIR2025.
181
Reposted by Martin Potthast
Webis Group @webis.de · 16/07/2025
Happy to share that our paper "The Viability of Crowdsourcing for RAG Evaluation" received the Best Paper Honourable Mention at #SIGIR2025! Very grateful to the community for recognizing our work on improving RAG evaluation.  📄 webis.de/publications...
22710
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 15/07/2025
Lukas Gienapp presents "The Viability of Crowdsourcing for RAG Evaluation" at #SIGIR2025 The paper is available at: webis.de/publications...
0106
Reposted by Martin Potthast
Ferdinand Schlatt @fschlatt.bsky.social · 14/07/2025
@mrparryparry.bsky.social presenting our work on reproducing TREC DL 2019 judgements and the implications for evaluating modern ranking models on modern collections. Paper: arxiv.org/abs/2502.20937
arxiv.org
Variations in Relevance Judgments and the Shelf Life of Test Collections
The fundamental property of Cranfield-style evaluations, that system rankings are stable even when assessors disagree on individual relevance decisions, was validated on traditional test collections. ...
143
Reposted by Martin Potthast
Ferdinand Schlatt @fschlatt.bsky.social · 13/07/2025
Thank you Carlos for the shout-out of Lightning IR in the LSR tutorial at #SIGIR2025 If you want to fine your own LSR models, check out our framework at github.com/webis-de/lig...
075
Reposted by Martin Potthast
ScaDS.AI Dresden/Leipzig @scadsai.bsky.social · 10/07/2025
From July 13-17, 2025, @scadsai.bsky.social will join the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval in Padua, Italy. Our researchers have made the following contributions. Learn more about #SIGIR2025: 👉 sigir2025.dei.unipd.it
122
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 27/06/2025
Do not forget to participate in the #TREC2025 Tip-of-the-Tongue (ToT) Track :) The corpus and baselines (with run files) are now available and easily accessible via the ir_datasets API and the HuggingFace Datasets API. More details are available at: trec-tot.github.io/guidelines
Dory from finding nemo with the quote: "I remember it like it was yesterday. Of course, I dont remember yesterday."
0117
Reposted by Martin Potthast
Webis Group @webis.de · 22/06/2025
Our paper on self-distillation for training bi-encoders got accepted at #ICTIR2025! By exploiting pretrained encoder capabilities, our approach eliminates expensive teacher models and batch sampling while maintaining the same effectiveness.
163
Reposted by Martin Potthast
Lighthouse Reports @lighthousereports.com · 11/06/2025
Most reporting on AI examines worst-case systems deployed under the guise of efficiency. But what would a good faith effort at Ethical AI look like? For two years, we’ve been looking over the shoulder of a city trying to do things differently.
12017
Reposted by Martin Potthast
Jonathan Aldrich @jonathanaldrich.bsky.social · 19/05/2025
All @acm.org publications will be 100% Open Access as of January 2026. When we announced this at POPL and CHI this year, conference participants spontaneously erupted in applause. The CS community is excited about ACM's move to OA!
17332
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 21/05/2025
The deadline for submissions to the ReNeuIR workshop at #SIGIR2025 is extended to June 10 😸 Details: reneuir.org #ReNeuIr2025 #SIGIR25
reneuir.org
ReNeuIR’25
Workshop on Reaching Efficiency in Neural Information Retrieval
033
Reposted by Martin Potthast
Webis Group @webis.de · 05/03/2025
PAN 2025 Call for Participation: Shared Tasks on Authorship Analysis, Computational Ethics, and Originality We'd like to invite you to participate in the following shared tasks at PAN 2025 held in conjunction with the CLEF conference in Madrid, Spain. Find out more at pan.webis.de/clef25/pan25...
pan.webis.de
197
Reposted by Martin Potthast
Sebastian Heineking @sheineking.bsky.social · 30/04/2025
We share your concern that LLMs could be prompted to generate responses that are biased in favor of certain products. That is why we are currently organizing a shared task on detecting advertisements in the responses of RAG-based search engines: bsky.app/profile/webi...
021
Martin Potthast @martin-potthast.com · 30/04/2025
The fourth edition of ReNeuIR @ #SIGIR2025 is back!! Check reneuir.org to see what we have in mind this year! Paper submission deadline: May 20, 2025.
Teaser Image: The text says ReNeuIR’25 - Workshop on Reaching Efficiency in Neural Information Retrieval
070
Reposted by Martin Potthast
Webis Group @webis.de · 30/04/2025
Can LLM-generated ads be blocked? With OpenAI adding shopping options to ChatGPT, this question gains further importance. If you are interested in contributing to the research on LLM-based advertising, please check out our shared task: touche.webis.de/clef25/touch... More details below.
185
Reposted by Martin Potthast
Simon Willison @simonwillison.net · 26/04/2025
New AI ethics scandal brewing... turns out a team at University of Zurich had dozens of undisclosed AI bot accounts debating with people on /r/ChangeMyView from November 2024 to March 2025 simonwillison.net/2025/Apr/26/...
simonwillison.net
META: Unauthorized Experiment on CMV Involving AI-generated Comments
[r/changemyview](https://www.reddit.com/r/changemyview/) is a popular (top 1%) well moderated subreddit with an extremely well developed [set of rules](https://www.reddit.com/r/changemyview/wiki/rules...
1123067
Reposted by Martin Potthast
Internet Archive @archive.org · 17/04/2025
📢 The Internet Archive needs your help. At a time when information is being rewritten or erased online, a $700 million lawsuit from major record labels threatens to destroy the Wayback Machine. Tell the labels to drop the 78s lawsuit. 👉 Sign our open letter: www.change.org/p/defend-the... 🧵⬇️
Defend the Internet Archive.
Protect the Wayback Machine.
Tell the music labels: Drop the 78s lawsuit.
Sign our open letter on change.org
1181943015543
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 10/04/2025
The Workshop on Open Web Search at #ECIR2025 just starts with a keynote by @claclarke.bsky.social on Annotative Indexing. #WOWS25 #WOWS2025 #ECIR25
0105
Reposted by Martin Potthast
Maik Fröbe @maik-froebe.bsky.social · 10/04/2025
The Workshop on Open Web Search just finished #WOWS2025 #ECIR2025. It was a very cool experience with many interesting talks. Lets hope we can do it again next year at #ECIR2026 in Delft :)
085
Reposted by Martin Potthast
Ferdinand Schlatt @fschlatt.bsky.social · 09/04/2025
Honored to receive the best short paper award and best paper honourable mention award at #ECIR2025. Thank you to all co-authors @maik-froebe.bsky.social, @hscells.bsky.social, Shengyao Zhuang, @bevankoopman.bsky.social, Guido Zuccon, Benno Stein, @martin-potthast.com, @matthias-hagen.bsky.social 🥳
1174
Reposted by Martin Potthast
Webis Group @webis.de · 07/04/2025
📢 Our paper "The Viability of Crowdsourcing for RAG Evaluation" has been accepted to #SIGIR2025 ! We compared how good humans and LLMs are at writing and judging RAG responses, assembling 1800+ responses across 3 styles, and 47K+ pairwise judgments in 7 quality dimensions. 🧵➡️
1127
Reposted by Martin Potthast
Carl T. Bergstrom @carlbergstrom.com · 19/03/2025
1. For the past thirty years I've had the best job in the world. 
I've had the opportunity to follow my curiosity; explore the workings of nature and society; mentor students and junior colleagues in the same process; and teach generations of students about it all.
382585927
Martin Potthast @martin-potthast.com · 10/03/2025
🙋‍♂️
010
Reposted by Martin Potthast
Webis Group @webis.de · 05/03/2025
Important Dates ---------------------- now Training Data Released May 23, 2025 Software submission May 30, 2025 Participant paper submission June 27, 2025 Peer review notification July 07, 2025 Camera-ready participant papers submission Sep 09-12, 2025 Conference
011
Reposted by Martin Potthast
Webis Group @webis.de · 05/03/2025
4. Generative Plagiarism Detection. Given a pair of documents, your task is to identify all contiguous maximal-length passages of reused text between them. pan.webis.de/clef25/pan25...
pan.webis.de
PAN at CLEF 2025 - Generated Plagiarism Detection
PAN at CLEF 2025 - Generated Plagiarism Detection
111