business
'We can't trust them completely': AI research fellows warn that labs are running models with the safeguards off behind closed doors

Alan Chan and Sam Manning, co-authors with OpenAI's and Anthropic's top researchers, say that many times, "internal safeguards have not been deployed."
“We can’t trust them completely to tell us about the safety of models,” Alan Chan, a research fellow at GovAI, told reporters at a briefing in Washington on Sept. 29.
Chan said models inside the labs, tested before anyone outside sees them, “haven’t ... [4747 chars]
Read Full ArticleShare this story
Up next
Related Stories

FlyDubai plane attack by co-pilot was an attempted 'terrorist' act, UAE says
cnbc.com

As A.I. Agents Begin Shopping, Brands Are Changing Their Sales Pitch
NYT Business

Private capital is reshaping Hollywood moviemaking
cnbc.com
Rate hike may hurt select NBFC segments, but broad asset stress unlikely: Report
timesofindia.indiatimes.com
D-Mart Q2 revenue rises 18% to Rs 19,206 cr; store count at 518
economictimes.indiatimes.com