'Current LLMs introduce substantial errors when editing work documents': Microsoft scientists find most AI models struggle with long-running tasks — so maybe don't trust them completely just yet
Microsoft researchers determine that current LLMs aren't good at long-running tasks More interactions and less structure significantly reduce benchmark performance "Python is the only domain where most models are ready"…
Read moreHow applying cognitive diversity to LLMs could transform the user experience
As AI continues to transform, so too does the experience of the people it serves. Research by McKinsey shows that in 2025, 62% of organizations are at least experimenting with AI agents, whilst almost 9 in 10 now say…
Read moreHackers are using LLMs to build the next generation of phishing attacks - here's what to look out for
Unit 42 warns GenAI enables dynamic, personalized phishing websites LLMs generate unique JavaScript payloads, evading traditional detection methods Researchers urge stronger guardrails, phishing prevention, and…
Read moreYou're all caught up.