Skip to content
September 15, 2026
  • Bluesky
  • Facebook
  • Linkedin
  • Mastodon
  • RSS
  • Twitter
  • Youtube

Daily CyberSecurity

Zero-hour alerts. Unmatched analysis.

Primary Menu
  • Home
  • CVE Data
    • CVE Watchtower
    • Top Exploited CVEs
    • CVE Stats by Vendor
    • Q2 2026 Report
    • CVE Alerts
    • CVE Alert Settings
    • Pricing
  • Cyber Criminals
  • Data Leak
  • Free Tools
    • CVSS 3.1 Calculator
    • Certificate Viewer
    • DNS Lookup
    • Encoder & Hash Generator
    • IP / Subnet Calculator
    • Whois Lookup
  • Linux
  • Malware
  • Vulnerability
  • Submit Press Release
  • Weekly Recap
Light/Dark Button
  • Home
  • News
  • Technology
  • Prevention of AI misguided, IBM launches open source Adversarial Robustness Toolbox
  • Technology

Prevention of AI misguided, IBM launches open source Adversarial Robustness Toolbox

Do Son April 21, 2018 3 minutes read
Add Daily CyberSecurity as a preferred source on Google

In order to prevent AI models from being misinformed and producing erroneous judgments, researchers need to go through continuous simulation attacks to ensure that AI models are not deceived. The IBM research team recently opened the detection model and the Adversarial Robustness Toolbox, a toolkit to combat attacks, to help developers strengthen their defensiveness against deep neural network attacks and make AI systems more secure.

Adversarial Robustness Toolbox is a library dedicated to adversarial machine learning. Its purpose is to allow rapid crafting and analysis of attacks and defense methods for machine learning models. It provides an implementation for many state-of-the-art methods for attacking and defending classifiers.

The Adversarial Robustness Toolbox contains implementations of the following attacks:

  • Deep Fool (Moosavi-Dezfooli et al., 2015)
  • Fast Gradient Method (Goodfellow et al., 2014)
  • Jacobian Saliency Map (Papernot et al., 2016)
  • Universal Perturbation (Moosavi-Dezfooli et al., 2016)
  • Virtual Adversarial Method (Moosavi-Dezfooli et al., 2015)
  • C&W Attack (Carlini and Wagner, 2016)
  • NewtonFool (Jang et al., 2017)

The following defense methods are also supported:

  • Feature squeezing (Xu et al., 2017)
  • Spatial smoothing (Xu et al., 2017)
  • Label smoothing (Warde-Farley and Goodfellow, 2016)
  • Adversarial training (Szegedy et al., 2013)
  • Virtual adversarial training (Miyato et al., 2017)

Adversarial Robustness Toolbox is available on Github.

In recent years, AI has made many breakthroughs in cognitive issues. Many tasks in life have begun to incorporate AI technology, such as identifying objects in images and videos, transliterating text, and machine translation. However, if the deep learning network is influenced by a designed interference signal, it is easy to produce erroneous judgments. This type of interference is difficult for humans to perceive, and interested people may use such weaknesses to mislead the judgment of the AI model and use it improperly. behavior.

The Adversarial Robustness toolkit currently provides enhanced defensibility for computer vision, provides developers with new defense technologies, and defends against malicious misleading attacks when deploying AI models. The toolbox was written in Python because Python It is the most commonly used language for building, testing, and deploying deep neural networks. It contains methods for combating and defending against attacks.

Firstly, developers can use this toolbox to detect the robustness of deep neural networks. They mainly record the output of the model for different interferences. Then, they reinforce the AI model through the attack data set and finally mark the attack mode and signal to prevent the model. Causes wrong results due to interference signals.

The Adversarial Robustness toolkit currently supports TensorFlow and Keras. In the future, it is expected to support more frameworks, such as PyTorch or MXNet. At this stage, the defense is mainly to provide image recognition. In the future, more versions will be added, such as speech recognition and text. Identify or time series and so on.

Related coverage

  • GPT-4.5 Released: Enhanced Accuracy, Reduced Hallucinations, and Expanded Knowledge
  • UK uses artificial intelligence to deal with ISIS extremist propaganda
  • OpenAI Codex Unleashed: Internet Access & Pro Features for Developers
Track all actively exploited CVEs →

Support Our Threat Intelligence

Find our tech and OS security coverage helpful? Support our work today and unlock a 100% ad-free reading experience!

Buy Me a Coffee Logo Buy Me a Coffee
Select your plan
Free Pro Team

Hover over a plan to see its benefits.

Stay Ahead of the Threat

Join security professionals receiving zero-hour CVE alerts, PoC updates, and threat analysis directly to their inbox.

No spam. One actionable email per week. Unsubscribe anytime.

SHARE
Share on FacebookShare on XShare on LinkedInShare on TelegramShare on BlueskyShare on Mastodon
Written by
@DdoS · Security Researcher

Do Son

Do Son is the Founder and Editor of SecurityOnline.info. Working in cybersecurity since 2013, he reports on vulnerabilities, malware, and emerging threats, providing timely analysis to help organizations and individuals stay ahead of evolving risks.

Tags: Adversarial Robustness Toolbox

Search

Translation

CVE ALERTS
📧

Email Delivery
Get threat intel straight to your inbox.

♾️

Unlimited Vendors
Track every technology in your stack.

🚨

All New CVE Alerts
Be the first to know about new flaws.

⚙️

Custom EPSS Threshold
Filter noise, focus on real risks.

💬

Slack & Teams Webhook
Integrate directly into your SecOps.

🚫

100% Ad-Free
Enjoy an uninterrupted reading experience.

$7/mo
Subscribe Now

🚨 Active Exploits in the Wild

  • CVE-2026-87827CVSS 10.0
    Certain KGUARD DVR devices running vulnerable firmware expose a system command execution service on all network interfaces without...
    Admin intel📅 Updated: Sep 15, 2026
  • CVE-2026-78006CVSS 9.8
    The The Events Calendar plugin for WordPress is vulnerable to Remote Code Execution in all versions up to,...
    Admin intel📅 Updated: Sep 15, 2026
  • CVE-2026-39364
    Vite is a frontend tooling framework for JavaScript. From 7.1.0 to before 7.3.2 and 8.0.5, on the Vite...
    Admin intel📅 Updated: Sep 15, 2026
  • CVE-2026-27540CVSS 9.0
    Unrestricted Upload of File with Dangerous Type vulnerability in Rymera Web Co Pty Ltd. Woocommerce Wholesale Lead Capture...
    Admin intel📅 Updated: Sep 15, 2026
  • CVE-2026-76461CVSS 9.8
    A vulnerability in the email parsing of Cisco AsyncOS Software for Cisco Secure Email Gateway could allow an...
    CISA KEV📅 Added to KEV: Sep 14, 2026
  • CVE-2026-51990
    A critical remote code execution vulnerability in Sogou Input Method, one of the most widely used Chinese-language input...
    Admin intel📅 Updated: Sep 12, 2026
  • CVE-2026-85706CVSS 10.0
    GitLab has remediated an issue that, under certain conditions, an unauthenticated user could have read arbitrary files from...
    Admin intelCISA KEV📅 Added to KEV: Sep 11, 2026📅 Updated: Sep 11, 2026
  • CVE-2026-42016CVSS 8.1
    JFrog Artifactory (Self Hosted) versions before 7.133.11 are vulnerable to a privilege escalation attack due to a validation...
    Admin intelCISA KEV📅 Added to KEV: Sep 11, 2026📅 Updated: Sep 11, 2026
Powered by CVE Watchtower

🔴 Live Critical Threats

  • CVE-2026-91949CVSS 9.3
    FreeRDP server versions before 3.31.0 contain a protocol negotiation bypass vulnerability that...
  • CVE-2026-63696CVSS 9.1
    Dell SmartFabric OS10 Software, versions prior to 10.6.1.3, contains a Download of...
  • CVE-2026-63695CVSS 9.8
    Dell SmartFabric OS10 Software, versions prior to 10.6.1.3, contains a Session Fixation...
  • CVE-2026-39919CVSS 9.8
    Ghostscript before 10.08.0 contains a heap-based buffer overflow vulnerability in the JPEG...
  • CVE-2026-91998CVSS 9.9
    Casdoor through 4.4.0 contains an authorization bypass vulnerability in the /api/mcp endpoint...
  • CVE-2026-91995CVSS 9.1
    pig before 4.1.0 contains an authentication bypass vulnerability in the /register/password endpoint...
  • CVE-2026-90711CVSS 9.1
    proxy-addr is a Node.js module that determines a request's client address behind...
  • CVE-2026-91003CVSS 9.1
    A flaw has been found in D-Link DI-8300 16.07. The affected element...
  • CVE-2026-91001CVSS 9.9
    A security flaw has been discovered in D-Link DI-8400 16.07. This affects...
  • CVE-2026-90847CVSS 9.1
    A vulnerability was determined in EFM ipTIME C200E 1.094. The impacted element...
Powered by CVE WATCHTOWER

Daily CyberSecurity

  • About SecurityOnline.info
  • Advertise with us
  • Announcement
  • Contact
  • Contributor Register
  • Login
  • Disclaimer
  • DCMA
  • Privacy Policy
  • About SecurityOnline.info
  • Advertise on SecurityOnline.info
  • Contact Us

When you purchase through links on our site, we may earn an affiliate commission. Here’s how it works

  • CVE Watchtower
  • CVE Statistics by Vendor 2026
  • Q2 2026 Report
  • Top Exploited CVEs
  • Bluesky
  • Facebook
  • Linkedin
  • Mastodon
  • RSS
  • Twitter
  • Youtube
© 2017 - 2026 Daily CyberSecurity. All Rights Reserved.