Sign in

Ingo Frommholz

@ingo.idf.social.ap.brid.gy
18 followers 6 following 79 posts

Professor of Applied Data Science at Modul University Vienna, Austria. Interested in all things Information Retrieval, AI, Data Science, Digital Libraries. Also FC Schalke […] 🌉 bridged from ⁂ idf.social/@ingo, follow @ap.brid.gy to interact

PostsRepliesMedia
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 28/09/2026
RE: idf.social/@djoerd/1173476250200419… “The real test of an AI-mediated tutor is not what a child can do with AI beside them, but what the child can still do when the AI is gone.” That holds true for students also. Children and students are not guinea pigs for a technology that […]
idf.social
Original post on idf.social
004
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 25/09/2026
Information Retrieval folk, the #ECIR2027 short paper abstract submission deadline is approaching – 5 October 2026. Full papers a week later. www.ecir2027.co.uk/call-for-short-p… @ecir2027.bsky.social
ecir2027.co.uk
Call for Short Papers
000
Reposted by Ingo Frommholz
ECIR 2027 @ecir2027.bsky.social · 25/09/2026
📣 #ECIR2027 deadline reminder! Coming up on **5 October 2026**: 📄 Full paper submissions 📝 Short paper abstracts 🌍 IR for Good paper submissions We’re looking forward to your submissions! 🔗 www.ecir2027.co.uk/calls-home
ecir2027.co.uk
Calls
see links below to the various calls for submissions:
024
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 21/09/2026
macOS Golden Gate think #ChatGPT contains malware? Seems legit.
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 18/09/2026
Great pleasure to be in Dagstuhl again, this time for the Frontiers of Technology Assisted Review Systems seminar. Had wonderful conversations and thought-provoking discussions about TAR with so many brilliant colleagues and friends. Dagstuhl is always a memorable experience. And this time I […]
idf.social
Original post on idf.social
001
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 13/09/2026
I had the great pleasure to spend three weeks on an OMINO staff exchange secondment in Erik Cambria’s group at Nanyang Technological University Singapore. I also had the pleasure of meeting Min-Yen Kan and his fantastic group at National University of Singapore, where I gave a talk about […]
idf.social
Original post on idf.social
010
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 08/09/2026
ECIR 2027 short paper submission is open now! www.ecir2027.co.uk/call-for-short-p… @ecir2027.bsky.social #ECIR2027
ecir2027.co.uk
Call for Short Papers
002
Reposted by Ingo Frommholz
ECIR 2027 @ecir2027.bsky.social · 07/09/2026
📣 Submitting a full paper to #ECIR2027? ⏰ Abstract deadline: 21 Sept 2026 📝 Full paper deadline: 5 Oct 2026 Topics include Information Retrieval, RAG/LLMs, recommender systems, conversational search, evaluation, trustworthy IR & more. More info: www.ecir2027.co.uk/call-for-ful...
ecir2027.co.uk
Call for Full Papers
ECIR 2027 invites the submission of high-quality and original papers in the broad field of Information Retrieval. Relevant topics include, but are not limited to:
087
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/08/2026
What a tragedy on so many levels. My condolences to his family. www.lbc.co.uk/article/jason-arday-c…
lbc.co.uk
Former Cambridge professor Jason Arday found dead aged 41 | LBC
Mr Arday is understood to have been found "unresponsive" at an address in Battersea on Friday afternoon
011
Reposted by Ingo Frommholz
ECIR 2027 @ecir2027.bsky.social · 10/08/2026
📢 ECIR 2027 Call for Papers 👉 www.ecir2027.co.uk/calls-home #ECIR2027 is the prime European forum for the presentation of original research in the field of #IR, and will take place as a physical (in-person) conference from 21-25 March 2027 at Southampton Football Club, Southampton, England, UK.
ecir2027.co.uk
Calls
see links below to the various calls for submissions:
196
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 28/07/2026
Still time to submit a proposal for a thought-provoking panel at JCDL 2026 in Dallas, Texas. Proposals should address pressing or emerging issues in digital libraries, archives, and related domains, offering fresh perspectives and fostering meaningful dialogue. Deadline July 31 […]
idf.social
Original post on idf.social
021
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 10/07/2026
Graduation ceremony at Modul University Vienna – congratulations to our students and all the best on your journey. A degree is not the end, it’s the beginning! #academicLife #academia #students
021
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 30/06/2026
RE: idf.social/@bcs_irsg/11683862870416… Please nominate for the KSJ Award!
idf.social
101
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 12/06/2026
As AI systems grow more capable, so does the pressure they place on people and organisations. Our research in the OMINO project proposes a taxonomy of seven types of AI overload. We explore how increasing AI agency is creating entirely new challenges for human–AI collaboration. The paper is […]
idf.social
Original post on idf.social
010
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 04/06/2026
As impressive and impactful (in a good way and bad way) as they are, any anthropomorphisation of "AI" or LLMs and discussion about consciousness or them having feelings is pure marketing. We shouldn't have to waste time on this […]
idf.social
Original post on idf.social
002
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 02/06/2026
RE: idf.social/@bcs_irsg/11668233622209… Please apply if you want to give a tutorial at Search Solutions 2026!
idf.social
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/05/2026
Cranfield University, the place where Cyril Cleverdon worked on early Information Retrieval evaluation (when it was the College of Aeronautics) and from which the "Cranfield paradigm" is derived, will merge with King's College London. www.bbc.com/news/articles/cnvp7gpyg…
001
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 03/04/2026
At our SCOLIA workshop at #ECIR2026 we introduced shared tasks. UP2DATE@SCOLIA focusses on the development of new methods to assist in updating systematic reviews. Interested? Check out up2date.health-nlp.com! @hscells @potthast @carsten @ecir2026.eu
up2date.health-nlp.com
up2date.health-nlp.com
031
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 31/03/2026
@ecir2026.eu in full swing in Delft, and next year ECIR will come home to the UK. But where will it be in 2028? It could be in your city! The BCS Information Retrieval Specialist Group (IRSG) is calling for bids for ECIR 2028 – […]
idf.social
Original post on idf.social
001
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 30/03/2026
Hello #ECIR2026, the proceedings of our SCOLIA workshop on Thursday are online – ceur-ws.org/Vol-4187. Looking forward to (hopefully) seeing you. Happy reading!
111
Reposted by Ingo Frommholz
Djoerd Hiemstra 🍉 @djoerd.idf.social.ap.brid.gy · 19/03/2026
I will be giving an invited talk at the #ECIR2026 IR4Good track about #OpenWebSearchEU: "Towards a shared infrastructure for assembling web search engines" djoerdhiemstra.com/2026/towards-a-s…
djoerdhiemstra.com
Towards a shared infrastructure for assembling web search engines
Web search engines are essential for navigating the web. Suppose we look at the web as a service that is provided by public utility companies, a service similar to electricity, water or telephone. To make sure that everyone has access to the web, public utility companies have to be subject to public control and regulation. Without regulation, a single firm may abuse their natural monopoly, for instance by raising prices, by deteriorating the service, by delivering unequal quality to different groups, or by pushing advertisements and propaganda. Equity requires that all citizens can access the web at a fair price, and at a sufficient level of quality, via transparent, well-regulated, community-based or government-based control. OpenWebSearch.eu is a European Union funded project that researches what a transparent, well-regulated, community-based web search engine would look like. The project builds the index for a web search engine on open infrastructure that is distributed over four data centers in four different European countries. The data centers cooperatively crawl the web, cooperatively preprocess and enrich the web data, and cooperatively build an inverted index that is shared with the world. We envision a future where a search engine is “assembled” from parts provided by many different companies, based on public standards. I will discuss public standards for search engine indexes, such as the common index file format (CIFF) and approaches based on open data formats like Parquet and open cloud object storage like S3. Furthermore, I will show how researchers can query the Open Web Index remotely using a low-cost local machine, without the need to download the full index, even though it currently consists of more than 10 billion web pages. _To be presented at the European Conference on Information Retrieval (ECIR 2026)IR 4Good track on 30 March 2026 in Delft, the Netherlands_
051
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 21/03/2026
Going to @ecir2026.eu in Delft? Or need another reason to go there? Please drop by SCOLIA 2026, our Second International Workshop on Scholarly Information Access to enjoy some exciting presentations! Our programme is online now at sites.google.com/view/bir-ws/scolia…. See you in Delft!
011
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/03/2026
Self-retraction when you made an honest mistake is a credible thing that should be celebrated, because that shows you care about scientific integrity. Keep calm and be transparent: advice from scientists who retracted their papers www.nature.com/articles/d41586-026-…
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/03/2026
Nice and intense days this week at our OMINO workshop in Wrocław (including visiting an actual 5-bit quantum computer at the impressive Wrocław Centre for Networking and Supercomputing).
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 11/02/2026
Great day yesterday at the University of Southampton where my PhD student Libo Ren and I gave a talk at a mini workshop in the context of our EU staff exchange project OMINO. It was great to see so many friends and current and former colleagues!
000
Reposted by Ingo Frommholz
IRRJ @irrj.sigmoid.social.ap.brid.gy · 05/02/2026
Published at #IRRJ: "CRAWLDoc: A System for Contextual Ranking and Bibliographic Metadata Extraction from Web Resources" by Fabian Karl and Ansgar Scherp. #DocumentRanking, #BibliographicMetadataExtraction, #ScholarlyDataset doi.org/10.54195/irrj.23861
irrj.org
CRAWLDoc: A System for Contextual Ranking and Bibliographic Metadata Extraction from Web Resources | Information Retrieval Research
012
Reposted by Ingo Frommholz
Djoerd Hiemstra 🍉 @djoerd.idf.social.ap.brid.gy · 03/02/2026
Generative AI and the Threat to Thinking by Martha Larson and Zhuoran Liu: "The proliferation of content created using generative artificial intelligence can overwhelm the ability of people to process information." Larson and Liu consider a model of human thinking to unpack this threat […]
idf.social
Original post on idf.social
213
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 03/02/2026
The growing crisis of “AI slop” that endangers scientific publishing. #FakeScience #InformationOverload #AIEthics #ScientificIntegrity #ResearchQuality #PeerReview #AcademicPublishing #OpenScience #TrustInScience www.theatlantic.com/science/2026/01…
theatlantic.com
Science Is Drowning in AI Slop
Peer review has met its match.
008
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 28/01/2026
ArXiv preprint server clamps down on AI slop – first-time posters to venerable platform now need an endorsement from an established author. www.science.org/content/article/arx…
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 26/01/2026
This is an interesting policy change regarding author attendance. Much more inclusive, but will authors struggle now to justify their travel expenses? Would be interesting to see how this affects author participation. From the ICML 2026 CfP […] [Original post on idf.social]
ICML policy change regarding author attendance, requiring authors only to register virtually, but not to attend in-person any more. Introduced proceedings-only papers.
001
Reposted by Ingo Frommholz
IRRJ @irrj.sigmoid.social.ap.brid.gy · 24/01/2026
There will be an #IRRJ special issue on #FIRE: Participants of the Forum for Information Retrieval Evaluation (FIRE 2025), which took place from 17 to 20 December 2025 at the Indian Institute of Technology in Varanasi, India, are invited to submit an extended version of their FIRE 2025 paper.
122
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 19/01/2026
Deadline extended to January 31! idf.social/@ingo/115893305306305111
idf.social
Ingo Frommholz (@ingo@idf.social)
Still a few days to go to submit to our SCOLIA workshop at ECIR 2026! If you're interested in scholarly search and recommendation, scholarly document processing, the intersection of bibliometrics/scientometrics and information retrieval, this is your place! https://sites.google.com/view/bir-ws/scolia-2026
000
Reposted by Ingo Frommholz
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/01/2026
Still a few days to go to submit to our SCOLIA workshop at ECIR 2026! If you're interested in scholarly search and recommendation, scholarly document processing, the intersection of bibliometrics/scientometrics and information retrieval, this is your place! […]
idf.social
Original post on idf.social
103
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 14/01/2026
Still a few days to go to submit to our SCOLIA workshop at ECIR 2026! If you're interested in scholarly search and recommendation, scholarly document processing, the intersection of bibliometrics/scientometrics and information retrieval, this is your place! […]
idf.social
Original post on idf.social
103
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 24/12/2025
I was honoured to give a keynote on Information Overload in Academia, and the role of Information Retrieval as a remedy, at FIRE 2025 in Varanasi, India. We need to look at the quality, not just topical relevance, of scientific publications. And think out of the box! Met lots of great people […]
idf.social
Original post on idf.social
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 24/12/2025
I was honoured to give a keynote on Information Overload in Academia, and the role of Information Retrieval as a remedy, at FIRE 2025 in Varanasi, India. We need to look at the quality, not just topical relevance, of scientific publications. And think out of the […] [Original post on idf.social]
101
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 06/12/2025
Academic #InformationOverload hitting the news. www.theguardian.com/technology/2025…
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 02/12/2025
This is terrible for students and staff alike and, unfortunately, not the first or the last of such cases in #UKAcademia. brignews.com/2025/12/01/students-fe…
brignews.com
_Disclaimer: This has been amended from the print edition. Gerry McCormac’s pay is £438,000 including pension contributions and £414,000 without. His pay increase over two years was £119,000._ University of Stirling students are fearing the impact of the recent voluntary severance scheme which has seen 175 staff members leave the institution. 112 professional services staff and 63 academics were granted the severance package. The university said the scheme was launched to support their “financial stability” as the higher education sector continues to experience uncertainty. However, in the financial accounts dated 2023-2024, the University recorded 5.6 per cent total income growth, turning over £179.2 million with a surplus of almost £7 million. Vice Chancellor Gerry McCormac’s salary also increased again, growing to £414,000 or £438,000, including pension contributions. In the past two years, his salary has increased by £119,000. This makes him Scotland’s best-paid higher education chief. Students at the University of Stirling say they find the situation concerning. Ali, a mature Journalism Studies student, said: “There is so much anxiety amongst students about the severance scheme. “Modules are being cancelled, people are losing project or dissertation supervisors, it’s really stressful. “On top of that, the staff that are left are seeing their workloads increase massively, it hardly seems fair. “When you look at the numbers, Greedy Gerry could pay the salary of 5-6 academic staff or 10-12 professional services staff and still be left earning £100,000 per year, an amount that most of us can’t even aspire to. “It’s disgusting.” Three French lecturers, including the programme director, are leaving the university leaving students concerned about future modules. One French student said: “The whole department will be reorganised and we are losing probably some of the greatest teachers that I’ve met during my entire education. “The fact that they can’t see out the semester, and are having to leave before our assignments are due to be submitted, has given my class a lack of faith in the university, that they would decide something that does affect our education. “I think a lot of people are frustrated… there has been no acknowledgement of the fact that it will be affecting us and our education. “It is also sad to see these teachers, who have been at the uni for such a long time, upset at the fact they can’t speak to us and are being put in this position, where they can’t see out the rest of the semester at least.” Two Journalism Studies lecturers are also leaving. Students in their final year have expressed worries about their final project, as supervisors are allegedly moving from having two or three students each to seven or eight. A 21-year-old second year student, who wishes to remain anonymous, said she is worried about the changes, and worried about who is left to support them: “It is quite unnerving to be honest. “We have no clue what this means for students further on. How are lecturers going to maintain the same quality of teaching with less staff? “I am also a bit concerned about other services the university provides, such as counselling. “It all feels very unclear, which doesn’t help when uni is stressful already.” A UCU spokesperson said: “Staff are very concerned about the impacts the recently closed voluntary severance scheme is going to have on the student experience this year. “A recent survey of UCU Stirling members revealed widespread concerns about workload, including how this will impact on teaching (66 per cent think there will be significant impacts, 19 per cent some impacts, a further 13 per cent are unsure/too early to say).” One member of staff commented: “I have never seen colleagues so de-moralised. The ‘business as usual’ message is misleading. Students are already recognising that staff are struggling. We have seen no leadership. “The constant uncertainty is having a huge impact on morale as is the loss of key academic and professional services staff.” The Students’ Union Officers said: “We understand why the university has resorted to this scheme and we are in constant and regular communication with the Trade Unions on campus to ensure that staff are being treated fairly. “Whilst the university assures us that this will not have an impact on staff workload, teaching staff have raised concerns of increased workload due to this scheme and limited resources. “We want to assure our students that we are working hard to ensure that there is as little disruption to your studies as possible. “We do ask you to demonstrate understanding and kindness towards staff during this period. “Please do not hesitate to seek support if you need it either from us or Student Support Services.” A University of Stirling spokesperson said: “The University’s Voluntary Severance Scheme was one of several strategic measures designed to support long-term financial sustainability. “In an increasingly challenging and unpredictable external environment – where income generation and cost pressures are continuing to impact the entire UK higher education sector – strong financial stewardship and good governance remain essential. “The Scheme attracted significant interest, and applicants have now received their outcomes. “The savings achieved – through this Scheme and wider organisational efficiencies – are intended to support our strategic priorities and improve operational effectiveness.” _Featured image credit: University of Stirling_ __ ##### Alex Paterson + postsBio Editor-in-Chief. Twitter/X and BlueSky: @AlexPaterson01 * Alex Paterson https://brignews.com/author/alexpaterson01/ __LIVE: Rachel Reeves outlines tax and other changes in 2025 budget * Alex Paterson https://brignews.com/author/alexpaterson01/ __Stirling club pays lucky winner’s full year rent in unique giveaway * Alex Paterson https://brignews.com/author/alexpaterson01/ __Stirling Uni Student saves life with first aid * Alex Paterson https://brignews.com/author/alexpaterson01/ __Support offered after student found dead in University of Stirling accommodation * Click to share on X (Opens in new window) X * Click to share on Facebook (Opens in new window) Facebook * Click to share on LinkedIn (Opens in new window) LinkedIn * Click to email a link to a friend (Opens in new window) Email * Click to share on Reddit (Opens in new window) Reddit * Click to share on Tumblr (Opens in new window) Tumblr * Click to share on Pinterest (Opens in new window) Pinterest * Click to share on WhatsApp (Opens in new window) WhatsApp * ### Like this: Like Loading... ### _Related_
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 22/11/2025
We’re inviting submissions to the 2nd International Workshop on Scholarly Information Access (SCOLIA 2026 @ ECIR) — April 2, 2026 in Delft. Focus: IR, NLP, bibliometrics, GenAI, RAG, academic search, integrity of the scientific record. More info & CFP […]
idf.social
Original post on idf.social
001
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 03/11/2025
Simple conclusion: don’t use it unless you want to be deliberately misled. www.theguardian.com/technology/2025… #grok #grokipedia
theguardian.com
In Grok we don’t trust: academics assess Elon Musk’s AI-powered encyclopedia
From publishing falsehoods to pushing far-right ideology, Grokipedia gives chatroom comments equal status to research
010
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 02/11/2025
Probably everyone I message with has received a scrambled message from me once in a while. Seems I’m not the only one who has to constantly autocorrect their autocorrection, and it’s not always because I keep hitting the wrong keys with my “Wurstfinger” […]
idf.social
Original post on idf.social
000
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 02/11/2025
For the sake of arXiv not just being a pool of research articles, but also to provide basic quality assurance, this decision is understandable. I don’t think just accepting everything and letting the reader sort it out is feasible, given we are already overloaded […]
idf.social
Original post on idf.social
011
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 20/10/2025
Uncertainty is an integral part of scientific research. How can we automatically detect it? blog.gesis.org/behind-every-discove…
blog.gesis.org
Behind Every Discovery Lies a Question: How UnScientify Maps the Invisible Web of Scientific Uncertainty
_The text explains that uncertainty is a fundamental part of scientific writing, yet difficult to detect automatically. The new system “UnScientify”, developed by researchers from France and Germany, uses transparent, rule-based methods to identify and categorize different forms of scientific uncertainty. It even distinguishes whether the uncertainty is expressed by the authors themselves or by others. In benchmark tests, UnScientify outperformed advanced AI models like GPT-4 in both accuracy and explainability. The tool aims to help science and society better identify uncertainty within research texts, promoting greater transparency and more effective knowledge transfer._ _Der Text beschreibt, wie Unsicherheit ein integraler Bestandteil wissenschaftlicher Texte ist und bisher schwer automatisch erfasst werden konnte. Das neue System “UnScientify”, entwickelt von Forschern in Frankreich und Deutschland, erkennt und kategorisiert verschiedene Ausdrucksformen wissenschaftlicher Unsicherheit durch regelbasierte, transparente Methoden. Es unterscheidet dabei sogar, wer die Unsicherheit äußert – die Autoren selbst oder andere Forschende. In Vergleichstests schlägt UnScientify fortschrittliche KI-Modelle wie GPT-4 in der Zuverlässigkeit und Nachvollziehbarkeit seiner Ergebnisse. Das System soll Wissenschaft und Gesellschaft helfen, Unsicherheiten in Forschungstexten besser zu erkennen, was zu mehr Transparenz und besserem Wissenstransfer führt._ DOI: 10.34879/gesisblog.2025.106 * * * ### The Hidden Patterns of Doubt – And Why They Drive Real Progress Open an academic journal, and you’ll find more than facts and figures. There are also subtle pauses, careful “maybes,” even blunt admissions of ignorance—woven seamlessly into the text. In science, uncertainty is not a weakness; it’s a driving force. Yet until recently, the art of tracing and understanding these signals of doubt remained impossible for machines—and tricky even for humans. Now, researchers at the Université Marie et Louis Pasteur in France and GESIS – Leibniz Institute for the Social Sciences in Germany are changing that, putting uncertainty itself under the microscope. ### Setting the Scene: Reading between the lines of scientific uncertainty Picture a university café on a rainy afternoon. Groups of scholars argue animatedly over results and methods, forever balancing certainty and speculation. Their articles reflect this same reality. Scientific texts are not just collections of data and conclusions—they are rich landscapes of cautious optimism, measured skepticism, and, crucially, hundreds of ways to admit, “We don’t fully know.” In this world, words like “possible,” “remains unclear,” “we hypothesize,” or “the evidence is inconclusive” become as significant as any hard number or equation. Reading between the lines is not a luxury, but a necessity: only by identifying the contours of certainty and uncertainty can science progress authentically. ### Why Does Mapping Uncertainty Matter? Uncertainty marks the edges of knowledge. When researchers hedge their statements or highlight the limits of a finding, they are—often unwittingly—placing flags for the next expedition into the unknown. These signals are crucial for other scientists, for policymakers sifting through complex reports, and for anyone seeking to distinguish established facts from open questions. But language is slippery, and context is everything. A “may” can express a minor doubt or substantial knowledge gap depending on where and how it’s used. Traditional computational tools have often failed here. Even the most advanced AI struggles to consistently recognize the nuanced signals of scientific doubt, and the reasons behind AI decisions often remain opaque—a “black box” that few are comfortable trusting in critical contexts. ### UnScientify: Letting Machines Read Between the Lines Enter UnScientify. Instead of treating language as a bag of words or pouring everything into a deep-learning engine, this system adopts a rule-based, transparent approach. Drawing from 12 distinct patterns of uncertainty—from explicit statements (“remains unresolved”), to modal verbs (“might affect”), conditional reasoning (“if x, then possibly y”), indirect questions, and even scholarly disagreement—UnScientify annotates scientific texts sentence by sentence. Crucially, it doesn’t stop there: it also discerns who is expressing the uncertainty. Is this the author’s personal doubt, or a recitation of others’ skepticism? This matters enormously. Imagine two sentences: “We believe further research is needed,” and “Some studies suggest the results are unreliable.” They both communicate limits, but the source (and thus the weight) of uncertainty is different. UnScientify uses linguistic analysis (with tools such as spaCy and custom pattern-matching) to make these distinctions clear, providing not just a label but a rationale for each annotation. ### When Human-Designed Rules Beat Super AIs What good is all this pattern-hunting? To test UnScientify, the team pitted it against the most sophisticated AI models of our time: GPT-4, RoBERTa, SciBERT and more, using a carefully constructed, multidisciplinary dataset of almost 1,000 annotated sentences. UnScientify came out ahead, scoring an impressive 80.8% accuracy—surpassing all tested machine-learning and large language models. But the story doesn’t end at raw scores. UnScientify’s real triumph is its explainability. Where “black-box” AIs often flip-flop or leave humans guessing about their decisions, UnScientify shows its work: every detection is justified by its language rulebook. If a new kind of uncertainty emerges, domain experts can update the patterns—no retraining on enormous datasets required. ### Unlocking Better Communication in Science and Society Imagine a future where, with a click, every sentence carrying academic doubt is highlighted in a research paper; where policymakers can instantly see which findings are robust and which rest on shaky ground; where journalists and educators can model scientific caution with ease. By charting uncertainty, UnScientify offers a tool not just for text mining, but for transparency, trust, and better conversation between science, decision-makers, and the public. The implications ripple outwards. Science does not advance in leaps of absolute certainty, but in careful steps—each one marked by open questions. UnScientify stands to help make these questions visible, opening new paths for interdisciplinary research and genuine dialogue. ### Looking Further: Charting New Territories of Doubt While UnScientify already covers vital ground, the journey has only begun. The team aims to broaden its understanding, expand its training to new fields, and perhaps even hybridize linguistic clarity with the deep, abstruse capacities of modern AI. The project’s publicly available dataset is already fueling further research, encouraging scientists everywhere to refine the art of charting uncertainty. As we look to a future brimming with data—and questions—the real heroes may not be those who declare the answers, but those who help reveal the borders between what is known and what is still waiting to be explored. UnScientify is helping science learn to read its own hesitations—and in those hesitations, to find the seeds of its next revolutions. Original scientific publication: **Panggih Kusuma Ningrum, Philipp Mayr, Nina Smirnova, Iana Atanassova: Annotating scientific uncertainty: A comprehensive model using linguistic patterns and comparison with existing approaches, Journal of Informetrics, Volume 19, Issue 2, 2025: ****https://doi.org/10.1016/j.joi.2025.101661** This article was written by Christian Kolle with the support of ChatGPT 4.1 based on the original scientific publication and reviewed by one of the researchers involved, Dr. Philipp Mayr.
000
Reposted by Ingo Frommholz
petersuber @petersuber.fediscience.org.ap.brid.gy · 23/07/2025
#AI tools seem to be generating a large swath of low-quality, formulaic biomedical articles drawn from #OpenAccess biomedical databases. For example, since the rise of #LLMs about three years ago, the number of new biomedical articles is about 5k larger than previous moving average would have […]
fediscience.org
Original post on fediscience.org
023
Reposted by Ingo Frommholz
petersuber @petersuber.fediscience.org.ap.brid.gy · 19/10/2025
Update. In response to this problem (previous post, this thread), some publishers are desk-rejecting papers based on open health datasets. The problem is not the quality of the data, but the absence of additional work to validate the findings. Two reports: 1. "Journals and publishers crack […]
fediscience.org
Original post on fediscience.org
022
Reposted by Ingo Frommholz
Ingo Frommholz @frommholz.org · 19/10/2025
‘[…]LLMs have utility, but the absurd way they've been over-hyped, the fact they're being forced on everyone, and the insistence on ignoring the many valid critiques about them make it very difficult to focus on legitimate uses where they might add value.’ www.anildash.com//2025/10/17/...
anildash.com
The Majority AI View - Anil Dash
A blog about making culture. Since 1999.
021
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 28/09/2025
Time Berners-Lee: Why I gave the world wide web away for free www.theguardian.com/technology/2025…
001
Reposted by Ingo Frommholz
Ingo Frommholz @frommholz.org · 10/09/2025
When I was a PhD student in Duisburg, we had our own mail server and managed our own machines including laptops – all Linux-based. We didn’t miss anything! Interesting move by @djoerd.idf.social.ap.brid.gy. Perhaps time to reconsider the relationship to Microsoft? www.voxweb.nl/en/professor...
voxweb.nl
Professor closes down Radboud University email address in protest against Microsoft - Vox magazine
Professor of Data Science Djoerd Hiemstra can no longer be reached through his standard Radboud University email address. Anyone emailing him at this address is referred to a page explaining that he b...
013
Reposted by Ingo Frommholz
Ingo Frommholz @frommholz.org · 05/09/2025
The Proceedings of SCOLIA 2025 are now available online. SCOLIA is the first workshop on Scholarly Information Access and the successor to our BIR workshop series --> ceur-ws.org/Vol-4022/. Happy reading!
ceur-ws.org
CEUR-WS.org/Vol-4022 - First International Workshop on Scholarly Information Access (SCOLIA)
031
Ingo Frommholz @ingo.idf.social.ap.brid.gy · 24/08/2025
“UKAI, a trade body representing the UK’s artificial intelligence industry, has argued repeatedly that the government’s approach is focused too narrowly on big tech at the expense of smaller players.” Exactly my thought when I read this […]
idf.social
Original post on idf.social
000