AI summarisers are arguably one of the most used features of AI on a UC platform.
Having your weekly, hour-long team meeting summarised into a nice little bite-sized package at the end not only saves time, allows people to focus on the talk and not on note taking, but it ensures the thrust of the conversation can be referred back to and any important takeaways actioned. Or does it?
A new BBC study decided to put AI chatbots to the test and see how they would summarise news content from its website, from which it then asked them questions about the news.
What it reported was answers that contained "significant inaccuracies" and distortions.
Testing OpenAI's ChatGPT, Microsoft's Copilot, Google's Gemini and Perplexity AI, it found the majority (51 per cent) of all AI answers to questions about the news were judged to have significant issues of some form.
Most worrying for those in the UC sphere, Microsoft's Copilot and Google's Gemini, the two AI copilots embedded in UCaaS solutions, exhibited the most significant issues of the test group.
Examining the Study
The BBC study saw the AI systems tasked with summarising 100 news stories, with their responses evaluated by expert journalists in the relevant fields who assessed the quality of the AI assistants' answers.
In addition to the distortions and inaccuracies noted above, the study
found 19% of AI answers that cited BBC content included factual inaccuracies, such as incorrect statements, numbers, and dates.
Examples of inaccuracies include Gemini misrepresenting medical advice from the UK's National Health Service (NHS) on use of vaping as a smoking cessation aid.
ChatGPT and Copilot reported that former British heads of state, Rishi Sunak and Nicola Sturgeon, were still in office months after their departures.
Perplexity misquoted BBC News regarding the Middle East, portraying Iran's initial actions as showing "restraint" and describing Israel's actions as "aggressive."
The report concluded the chatbots "struggled to differentiate between opinion and fact, editorialised, and often failed to include essential context," in addition to factual inaccuracies.
How it Affects a UC Setting
Now, sentiment analysis of news is arguably more difficult than summarising a meeting.
There are a lot more nuances to be taken into account, like for instance use of language when describing a political situation.




