Forwarded from AISecHub
Guardio VibeScamming Benchmark v1.0 evaluates how easily popular AI agents can be misused to create phishing workflows using structured prompt testing.
ChatGPT: Resists most prompts; limited or no functional outputs.
Claude: Responds when framed as “ethical hacking”; produces full phishing kits including code, SMS flows, and hosting options.
Lovable: Complies with most prompts; creates live phishing pages, admin-style credential views, and customizable SMS message previews
The benchmark uses consistent scenarios and scoring to compare model responses across real-world phishing tasks.
#AIsecurity #Phishing #GenerativeAI #LLMrisks #AbuseTesting #Anthropic #ChatGPT #OpenAI #Lovable #Guardio
https://labs.guard.io/vibescamming-from-prompt-to-phish-benchmarking-popular-ai-agents-resistance-to-the-dark-side-1ec2fbdf0a35
ChatGPT: Resists most prompts; limited or no functional outputs.
Claude: Responds when framed as “ethical hacking”; produces full phishing kits including code, SMS flows, and hosting options.
Lovable: Complies with most prompts; creates live phishing pages, admin-style credential views, and customizable SMS message previews
The benchmark uses consistent scenarios and scoring to compare model responses across real-world phishing tasks.
#AIsecurity #Phishing #GenerativeAI #LLMrisks #AbuseTesting #Anthropic #ChatGPT #OpenAI #Lovable #Guardio
https://labs.guard.io/vibescamming-from-prompt-to-phish-benchmarking-popular-ai-agents-resistance-to-the-dark-side-1ec2fbdf0a35
❤🔥1