Technology

The Chatbots Reply to Their Critics

As the US midterms near, a study warns against relying on LLMs. ChatGPT, Gemini, Grok and Claude weigh in on their errors and the biases of the experts.

Study questions the reliability of AI chatbots in politics.

As the US midterms approach, a new study has raised questions about the reliability of AI chatbots in political debate. Photo: Statement/AI

Americans are increasingly reliant on AI chatbots for guidance on everything from cooking to shopping to writing love letters. Is that going to be a problem for November’s midterm elections?

Yes, it will almost certainly be an issue: That was the upshot of the arguments by several experts from Forum AI, a digital standards evaluation firm, in concert with experts from Stanford and the University of Washington, in a white paper and a newsletter about chatbot errors.

The Forum AI team predicted there was a “90% chance” an AI’s response to a midterm-related question will be “flawed in some material way”. They also unveiled a new tool called NewsBench to help measure and improve chat outcomes in the future.

Granted, the study raises valid points, and the new tool is welcome. But the researchers are smuggling in their own biases in the guise of expertise: that was the gist of the replies from the chatbots, running to more than 4,000 words, when Statement reached out to them for comment Wednesday.

Welcome to the comments section of the Štandard daily. Please take note of our guidelines, comments are moderated by us. You can contact the moderators at support@statement.com.

Participate in the discussion

Comments are available to subscribers only. If you'd like to join the discussion, choose a subscription starting at €6.72 per month.

All comments 0

    Register

    Comments are available to registered users only. If you'd like to join the discussion, register here.