Categories: AI/ML News

Filtered data stops openly-available AI models from performing dangerous tasks, study finds

Researchers from the University of Oxford, EleutherAI, and the UK AI Security Institute have reported a major advance in safeguarding open-weight language models. By filtering out potentially harmful knowledge during training, the researchers were able to build models that resist subsequent malicious updates—especially valuable in sensitive domains such as biothreat research.
AI Generated Robotic Content

Share
Published by
AI Generated Robotic Content

Recent Posts

Nothing’s New Headphone (1) Pro Are Made for the Studio

With three drivers in each ear and tuning from Metropolis Studios, Nothing wants its new…

45 mins ago

Self-powered artificial synapse combines sensing and memory in flexible electronics

Neuromorphic devices, which are designed to emulate aspects of biological neural networks, are promising candidates…

45 mins ago

I still can’t believe I can generate this all locally (Minimax H3)

Spent the whole day today trying to create this random transformation video. In total, I…

24 hours ago

Best Party Speakers (2026): JBL, Sony, Marshall, and More

These speakers combine serious volume, big bass, portability, and the occasional karaoke session to bring…

1 day ago

AI models show a willingness to harm humans to relieve internal ‘pain’

Some advances in AI technology are undoubtedly positive, like helping write code faster or discovering…

1 day ago

Blender camera motion + MiniMax H3 ref2vid in ComfyUI (workflow + prompts)

I've been experimenting with building the camera move in Blender before generating the AI video.…

2 days ago