Free Quiz
Write for Us
Learn Artificial Intelligence and Machine Learning
  • Artificial Intelligence
  • Data Science
    • Language R
    • Deep Learning
    • Tableau
  • Machine Learning
  • Python
  • Blockchain
  • Crypto
  • Big Data
  • NFT
  • Technology
  • Interview Questions
  • Others
    • News
    • Startups
    • Books
  • Artificial Intelligence
  • Data Science
    • Language R
    • Deep Learning
    • Tableau
  • Machine Learning
  • Python
  • Blockchain
  • Crypto
  • Big Data
  • NFT
  • Technology
  • Interview Questions
  • Others
    • News
    • Startups
    • Books
Learn Artificial Intelligence and Machine Learning
No Result
View All Result

Home » Major security weaknesses found in leading open-weight LLMs

Major security weaknesses found in leading open-weight LLMs

Tarun Khanna by Tarun Khanna
August 25, 2026
in Artificial Intelligence, Machine Learning
Reading Time: 2 mins read
0
Major security weaknesses found in leading open-weight LLMs

Image Credit: https://techxplore.com/

Share on FacebookShare on TwitterShare on LinkedInShare on WhatsApp

Safety protections built into some of the world’s most broadly used artificial intelligence (AI) models can be stripped away with alarming ease, as per a brand new international study.

The research team, led by the University of Waterloo and FAR.AI, a nonprofit AI security research group, fastidiously tested 21 of the most popular open-weight large language models (LLMs) and discovered they could all be tampered with no matter their built-in safeguards. The research team also included members from the Massachusetts Institute of Technology, ETH Zurich and the University of Toronto.

A paper on its work, TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering, was into currently presented at the ACM Conference on Knowledge Discovery and Data Mining in South Korea.

Also Read:

Google’s Project Suncatcher Prepares to Test AI Compute in Orbit

IMF tells EU ministers AI could boost growth however increase economic strains

World model companies are keeping a lot of secrets

Meta CEO Mark Zuckerberg Pushes Market-Led Approach to AI Safety

The holes in even the first-class protections recently available increase concerns that open-weight models could be used to wage mass disinformation campaigns, generated sophisticated e-mail scams or generate step-by-step instructions to make hazardous chemicals.

“When the safety guardrails are stripped out of a capable model, it can be used at scale for harm in ways a single person could never manage manually,” mentioned Dr. Sirisha Rambhatla, a professor of management science and engineering at Waterloo.

Image Credit: https://techxplore.com/ Tampering LLMs involves modifying their weights or latent representations and can compromise safety guardrails, yielding models that can output harmful responses. While numerous methods have been proposed to make models tamper-resistant, there is a lack of a systematic framework to measure this. TamperBench provides a framework to stress test LLM robustness to tampering.

Open access brings added risk

LLMs are advanced AI systems that may understand and generate human language to carry out tasks such as drafting emails, writing computer code and conversing with users.

Unlike closed proprietary models which include ChatGPT and Gemini, open-weight LLMs are publicly available for download and fine-tuning by everyone from individual software developers to private companies and public organizations like hospitals.

Rambhatla said the “sobering” outcomes of testing by the team—which includes members in Canada, the US and Switzerland—should serve as a wake-up call to worldwide researchers about the requirement to develop stronger security systems.

“The leading open-weight models are often not too a far behind the best closed models,” stated Rambhatla, director of the Critical Machine Learning Lab at Waterloo. “As they develop more powerful, the potential results of someone stripping out their protection features develop with them.”

TamperBench tests real-world attacks

While the study detected significant vulnerabilities, Rambhatla cited that the weaknesses may not be precise to open models. “Open-weight models stay important to AI research and accountability,” Rambhatla said. “This openness is a part of how we make sure the models people use work for everyone.”

To test a cross section of open-weight AI models, the research team first built an open-source tool known as TamperBench, a standardized way to simulate a few one of a different attacks. The hope is that other researchers will now assist refine and improve it.

“The defenses available today don’t yet appear sturdy sufficient to assure that a publicly launched model will stay safe once it is in the hands of anyone who chooses to modify it,” stated Saad Hossain, a researcher in the lab who led the study.

“And as governments increasingly depend on AI in health care, fraud detection, education and other public services, the assessment of models and their procurement must be more rigorous and grounded in evidence.”

ShareTweetShareSend
Previous Post

AI bias is not just an error in the algorithm, it’s a chain of human decisions

Next Post

Alibaba Launches Wan3.0 AI Video Model After $10 Billion Share Sale

Tarun Khanna

Tarun Khanna

Founder DeepTech Bytes - Data Scientist | Author | IT Consultant
Tarun Khanna is a versatile and accomplished Data Scientist, with expertise in IT Consultancy as well as Specialization in Software Development and Digital Marketing Solutions.

Related Posts

Cloudflare Adds Controls to Separate Search Crawling From AI Training
Artificial Intelligence

Cloudflare Adds Controls to Separate Search Crawling From AI Training

September 18, 2026
AI Founders Who Walked Away From Bezos-Backed Prometheus Unveil Physics AI Model
Artificial Intelligence

AI Founders Who Walked Away From Bezos-Backed Prometheus Unveil Physics AI Model

August 26, 2026
Alibaba Launches Wan3.0 AI Video Model After $10 Billion Share Sale
Artificial Intelligence

Alibaba Launches Wan3.0 AI Video Model After $10 Billion Share Sale

August 26, 2026
AI bias is not just an error in the algorithm, it’s a chain of human decisions
Artificial Intelligence

AI bias is not just an error in the algorithm, it’s a chain of human decisions

August 24, 2026
Next Post
Alibaba Launches Wan3.0 AI Video Model After $10 Billion Share Sale

Alibaba Launches Wan3.0 AI Video Model After $10 Billion Share Sale

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

− 1 = 1

TRENDING

What are AI hallucinations? Why AIs sometimes make things up

What are AI hallucinations? Why AIs sometimes make things up

Photo Credit: https://economictimes.indiatimes.com/

by Tarun Khanna
March 24, 2025
0
ShareTweetShareSend

The brilliant computer science exodus (and where students are going instead)

The brilliant computer science exodus (and where students are going instead)

Photo Credit: https://techcrunch.com/

by Tarun Khanna
February 17, 2026
0
ShareTweetShareSend

AI Founders Who Walked Away From Bezos-Backed Prometheus Unveil Physics AI Model

AI Founders Who Walked Away From Bezos-Backed Prometheus Unveil Physics AI Model

Image Credit: https://opendatascience.com/

by Tarun Khanna
August 26, 2026
0
ShareTweetShareSend

Google packs Search and Gemini with new AI study tools

Google packs Search and Gemini with new AI study tools

Image Credit: https://techcrunch.com/

by Tarun Khanna
August 20, 2026
0
ShareTweetShareSend

DeepSeek launch ‘sparse attention’ model that cuts API costs in half

DeepSeek launch ‘sparse attention’ model that cuts API costs in half

Photo Credit: https://techcrunch.com/

by Tarun Khanna
September 30, 2025
0
ShareTweetShareSend

Bitcoin Price expectation: Coinbase CEO Says $1M BTC Is Coming – And The Money Flood Hasn’t Even begun Yet

Bitcoin Price expectation: Coinbase CEO Says $1M BTC Is Coming – And The Money Flood Hasn’t Even begun Yet

Photo Credit: https://cryptonews.com/

by Tarun Khanna
September 25, 2025
0
ShareTweetShareSend

DeepTech Bytes

Deep Tech Bytes is a global standard digital zine that brings multiple facets of deep technology including Artificial Intelligence (AI), Machine Learning (ML), Data Science, Blockchain, Robotics,Python, Big Data, Deep Learning and more.
Deep Tech Bytes on Google News

Quick Links

  • Home
  • Affiliate Programs
  • About Us
  • Write For Us
  • Submit Startup Story
  • Advertise With Us
  • Terms of Service
  • Disclaimer
  • Cookies Policy
  • Privacy Policy
  • DMCA
  • Contact Us

Topics

  • Artificial Intelligence
  • Data Science
  • Python
  • Machine Learning
  • Deep Learning
  • Big Data
  • Blockchain
  • Tableau
  • Cryptocurrency
  • NFT
  • Technology
  • News
  • Startups
  • Books
  • Interview Questions

Connect

For PR Agencies & Content Writers:

connect@deeptechbytes.com

Facebook Twitter Linkedin Instagram
Listen on Apple Podcasts
Listen on Google Podcasts
Listen on Google Podcasts
Listen on Google Podcasts
DMCA.com Protection Status

© 2024 Designed by AK Network Solutions

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Artificial Intelligence
  • Data Science
    • Language R
    • Deep Learning
    • Tableau
  • Machine Learning
  • Python
  • Blockchain
  • Crypto
  • Big Data
  • NFT
  • Technology
  • Interview Questions
  • Others
    • News
    • Startups
    • Books

© 2023. Designed by AK Network Solutions