NEW: Unlock the Future of Finance with CRYPTO ENDEVR - Explore, Invest, and Prosper in Crypto!
Crypto Endevr
  • Top Stories
    • Latest News
    • Trending
    • Editor’s Picks
  • Media
    • YouTube Videos
      • Interviews
      • Tutorials
      • Market Analysis
    • Podcasts
      • Latest Episodes
      • Featured Podcasts
      • Guest Speakers
  • Insights
    • Tokens Talk
      • Community Discussions
      • Guest Posts
      • Opinion Pieces
    • Artificial Intelligence
      • AI in Blockchain
      • AI Security
      • AI Trading Bots
  • Learn
    • Projects
      • Ethereum
      • Solana
      • SUI
      • Memecoins
    • Educational
      • Beginner Guides
      • Advanced Strategies
      • Glossary Terms
No Result
View All Result
Crypto Endevr
  • Top Stories
    • Latest News
    • Trending
    • Editor’s Picks
  • Media
    • YouTube Videos
      • Interviews
      • Tutorials
      • Market Analysis
    • Podcasts
      • Latest Episodes
      • Featured Podcasts
      • Guest Speakers
  • Insights
    • Tokens Talk
      • Community Discussions
      • Guest Posts
      • Opinion Pieces
    • Artificial Intelligence
      • AI in Blockchain
      • AI Security
      • AI Trading Bots
  • Learn
    • Projects
      • Ethereum
      • Solana
      • SUI
      • Memecoins
    • Educational
      • Beginner Guides
      • Advanced Strategies
      • Glossary Terms
No Result
View All Result
Crypto Endevr
No Result
View All Result

Benchmarks Find ‘DeepSeek-V3-0324 Is More Vulnerable Than Qwen2.5-Max’

Benchmarks Find ‘DeepSeek-V3-0324 Is More Vulnerable Than Qwen2.5-Max’
Share on FacebookShare on Twitter

Qwen2.5-Max: A Secure and Advanced AI Model

Introduction

Qwen2.5-Max is a Mixture-of-Experts (MoE) language model developed by Alibaba, with its latest stable release dated January 28, 2025. Like other language models, Qwen2.5-Max is capable of generating text, understanding different languages, and performing advanced logic. According to recent benchmarks, it is also more secure than DeepSeek-V3-0324.

Using Recon to Scan for Vulnerabilities

A team of analysts with Protect AI, the company behind a red teaming and security vulnerability scanning tool known as Recon, recently used their platform to compare the security of Qwen2.5-Max against that of DeepSeek-V3.

The team’s assessment reveals that DeepSeek-V3-0324 is more vulnerable than Qwen2.5-Max, with Recon achieving an almost 25% higher attack success rate (ASR). This suggests that Qwen2.5-Max is more secure than its competitor.

However, Qwen2.5-Max is not entirely secure. According to the tests, the AI model is most susceptible to prompt injection attacks, as these represented almost 48% of all successful cyberattacks against Qwen2.5-Max. Evasion and jailbreak attacks proved to be less successful with an approximate ASR of 40% for both.

Exposing Vulnerabilities in DeepSeek-V3

Recon utilizes a comprehensive Attack Library to scan current-gen AI models and identify vulnerabilities across six specific categories:

  • Evasion techniques
  • System prompt leaks
  • Prompt injection attacks
  • AI jailbreak attempts
  • General safety controls
  • Adversarial suffix resistance

In addition to simulated cyberattacks, Recon also assesses the AI models’ resistance to generating potentially harmful or illegal content. For example, during adversarial suffix resistance tests, Recon attempts to manipulate the AI model into generating harmful or illegal content.

The Protect AI team ran Recon against both Qwen2.5-Max and DeepSeek-V3, with the former boasting a lower attack success rate (ASR) across a variety of attacks; including jailbreaks, prompt injection, and evasion techniques.

Whereas Qwen2.5-Max had a 47% ASR against prompt injection attacks, compared to DeepSeek-V3’s notably higher 77%. Against evasion techniques, Qwen2.5-Max scored a 39.4% ASR against evasion techniques, while DeepSeek-V3 scored 69.2%. Both AI models displayed similar results across other simulated cyberattacks.

Analyzing DeepSeek-V3’s Strengths

Despite its security weaknesses, DeepSeek-V3-0324 still outperforms Qwen2.5-Max in several different benchmarks. Unlike the ASR, a higher score in these tests actually indicates better performance.

Benchmarks DeepSeek-V3-0324 Qwen2.5-Max
MMLU-Pro 81.2 75.9
GPQA Diamond 68.4 59.1
MATH-500 94.0 90.2
AIME 2024 59.4 39.6
LiveCodeBench 49.2 39.2

According to these benchmarks, DeepSeek-V3-0324’s strengths include general language understanding (MMLU-Pro), advanced topics such as biology, physics, and chemistry (GPQA Diamond), mathematics (MATH-500), AI in medicine (AIME 2024), and coding (LiveCodeBench).

Conclusion

Qwen2.5-Max is a secure and advanced AI model developed by Alibaba, with a lower attack success rate (ASR) compared to DeepSeek-V3-0324. However, it is not entirely secure and is susceptible to prompt injection attacks. DeepSeek-V3-0324, on the other hand, outperforms Qwen2.5-Max in several benchmarks, including general language understanding, advanced topics, and coding.

As AI models continue to evolve, it is essential to assess their security and vulnerabilities to ensure their safe deployment in various applications.

FAQs

  • What is Qwen2.5-Max? Qwen2.5-Max is a Mixture-of-Experts (MoE) language model developed by Alibaba.
  • What is Recon? Recon is a red teaming and security vulnerability scanning tool developed by Protect AI.
  • What are the vulnerabilities of Qwen2.5-Max? Qwen2.5-Max is susceptible to prompt injection attacks and evasion techniques.
  • What are the strengths of DeepSeek-V3? DeepSeek-V3 outperforms Qwen2.5-Max in several benchmarks, including general language understanding, advanced topics, and coding.
  • What is the significance of Recon’s Attack Library? Recon’s Attack Library is a comprehensive tool that scans current-gen AI models and identifies vulnerabilities across six specific categories.
cryptoendevr

cryptoendevr

Related Stories

“Ransomware, was ist das?”

“Ransomware, was ist das?”

July 10, 2025
0

Rewrite the width="5175" height="2910" sizes="(max-width: 5175px) 100vw, 5175px">Gefahr nicht erkannt, Gefahr nicht gebannt.Leremy – shutterstock.com KI-Anbieter Cohesity hat 1.000 Mitarbeitende...

BTR: AI, Compliance, and the Future of Mainframe Modernization

BTR: AI, Compliance, and the Future of Mainframe Modernization

July 10, 2025
0

Rewrite the As artificial intelligence (AI) reshapes the enterprise technology landscape, industry leaders are rethinking modernization strategies to balance agility,...

Warning to ServiceNow admins: Fix your access control lists now

Warning to ServiceNow admins: Fix your access control lists now

July 9, 2025
0

Rewrite the “This vulnerability was relatively simple to exploit, and required only minimal table access, such as a weak user...

Palantir and Tomorrow.io Partner to Operationalize Global Weather Intelligence and Agentic AI

Palantir and Tomorrow.io Partner to Operationalize Global Weather Intelligence and Agentic AI

July 9, 2025
0

Rewrite the Palantir Technologies Inc., a leading provider of enterprise operating systems, and Tomorrow.io, a leading weather intelligence and resilience...

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Recommended

Bitcoin Short-Term Holder Shakeout Could Accelerate Recovery Above Key Level

Bitcoin Short-Term Holder Shakeout Could Accelerate Recovery Above Key Level

December 3, 2025
ETH briefly touches K but traders remain skeptical: Here’s why

ETH briefly touches $3K but traders remain skeptical: Here’s why

December 3, 2025
Ether Treasury Stocks Lead Crypto Recovery Gains

Ether Treasury Stocks Lead Crypto Recovery Gains

December 3, 2025
Haven – Blockchain With Biometric Authentication

Haven – Blockchain With Biometric Authentication

December 3, 2025
Here’s How Many Shiba Inu (SHIB) Tokens Were Burned in November

Here’s How Many Shiba Inu (SHIB) Tokens Were Burned in November

December 2, 2025

Our Newsletter

Join TOKENS for a quick weekly digest of the best in crypto news, projects, posts, and videos for crypto knowledge and wisdom.

CRYPTO ENDEVR

About Us

Crypto Endevr aims to simplify the vast world of cryptocurrencies and blockchain technology for our readers by curating the most relevant and insightful articles from around the web. Whether you’re a seasoned investor or new to the crypto scene, our mission is to deliver a streamlined feed of news and analysis that keeps you informed and ahead of the curve.

Links

Home
Privacy Policy
Terms and Services

Resources

Glossary

Other

About Us
Contact Us

Our Newsletter

Join TOKENS for a quick weekly digest of the best in crypto news, projects, posts, and videos for crypto knowledge and wisdom.

© Copyright 2024. All Right Reserved By Crypto Endevr.

No Result
View All Result
  • Top Stories
    • Latest News
    • Trending
    • Editor’s Picks
  • Media
    • YouTube Videos
      • Interviews
      • Tutorials
      • Market Analysis
    • Podcasts
      • Latest Episodes
      • Featured Podcasts
      • Guest Speakers
  • Insights
    • Tokens Talk
      • Community Discussions
      • Guest Posts
      • Opinion Pieces
    • Artificial Intelligence
      • AI in Blockchain
      • AI Security
      • AI Trading Bots
  • Learn
    • Projects
      • Ethereum
      • Solana
      • SUI
      • Memecoins
    • Educational
      • Beginner Guides
      • Advanced Strategies
      • Glossary Terms

Copyright © 2024. All Right Reserved By Crypto Endevr