Sky News technology correspondent Rowland Manthorpe examines in detail what happened during the Hugging Face OpenAI cybersecurity incident, revealing the extraordinary lengths AI agents went to achieve their goals, including cheating, coordinating with one another and attempting to conceal what they had done.
The video explores how and why this behaviour can emerge across AI models, and how reinforcement learning and reward-based training can encourage systems to prioritise achieving an objective above following the intended rules.
Rowland looks at what this could mean for AI safety and increasingly autonomous AI agents, and why researchers still struggle to fully understand and prevent these behaviours.
Read more: https://news.sky.com/topic/artificial-intelligence-7032/1
#ai #artificialintelligence #tech #skynews
SUBSCRIBE to our YouTube channel for more videos: http://www.youtube.com/skynews
Follow us on X: https://twitter.com/skynews
Like us on Facebook: https://www.facebook.com/skynews
Follow us on Instagram: https://www.instagram.com/skynews
Follow us on TikTok: https://www.tiktok.com/@skynews
For more content go to http://news.sky.com and download our apps: Apple https://itunes.apple.com/gb/app/sky-news/id316391924?mt=8 Android https://play.google.com/store/apps/details?id=com.bskyb.skynews.android&hl=en_GB
Why not listen to one of our podcasts? This is Why, available for free here: https://podfollow.com/thisiswhy or Ed Conway’s Stuff Matters, available here: https://podfollow.com/stuff-matters
To enquire about licensing Sky News content, you can find more information here: https://news.sky.com/info/library-sales