An important component of their proposed 'Scientist AI' architecture is the presentation of the underlying data. Training directly on the internet (as is done now for LLMs) would just lead to the AI to adopt (some average of) human claims and beliefs, regardless of whether they are true or false […]
fediscience.org
Original post on fediscience.org