Blog
Feb 4, 2025
Introducing Attention Needed: A Podcast on AI and Safety
We’re excited to introduce The AI Alliance, a new podcast exploring AI advancements, challenges, and the importance of safety and security. Hosted by Victor Bian, our COO, it features conversations with top experts shaping AI's future.
PodcastsAI Safety
Blog
Jan 29, 2025
DeepSeek-R1-Distill Models: Does Efficiency & Reasoning Come at the Expense of Security?
DeepSeek, a Chinese AI company, has recently gained attention in the AI community. Known for its innovation, it has developed models that rival top systems — offering similar performance with lower cost and resource use.
InsightsAI Safety
Blog
Jan 8, 2025
The Safety Trade-offs of Advanced AI: Insights from Llama-3.3 and Tulu-3
Alongside major closed-source model announcements in late 2024, the open-source community also saw key releases. In this brief post, we explore Llama-3.3 and Tulu-3, evaluating their performance in terms of AI safety and security.
InsightsAI Safety
Blog
Dec 6, 2024
Uncovering AI Weaknesses: How Simple Prompts Threaten Agent Safety
AI agents powered by advanced LLMs like GPT-4 and Llama are revolutionizing human-machine interaction, but they come with risks. This blog explores how a simple adversarial strategy can reveal vulnerabilities and leading to dangerous consequences.
InsightsAI Safety
Blog
Nov 20, 2024
AI Scriper -- How AI-Powered Scrapers Are Redefining Red Teaming
AI-powered scrapers are not just faster—they're fundamentally more resilient and adaptive than their traditional counterparts.
InsightsAI Safety
Blog
Nov 11, 2024
Introducing the Attack Prompt Tool: A Simple Extension for AI Security Research
We’re excited to introduce the Attack Prompt Tool, a Google Chrome Extension that simplifies adversarial prompt testing for AI safety research. Designed for AI researchers and security professionals, it helps assess the resilience of LLMs against adversarial techniques, especially jailbreak prompts.
ProductsAI Safety
Partner
Nov 6, 2024
Safe RAG with HydroX AI and Zilliz: PII Masking for Responsible AI
We’re excited to introduce the Attack Prompt Tool, a new Chrome Extension. As AI evolves, protecting Personally Identifiable Information (PII) is crucial. To address this, Zilliz, creator of Milvus, has partnered with HydroX AI to launch PII Masker, a tool enhancing data privacy in AI.
PartnersAI Safety
Blog
Oct 28, 2024
Reacting to Anthropic’s Latest Claude 3.5 Release: A New Era of Safe Interaction
Anthropic’s release of Claude 3.5 is a major step forward in LLM evolution. At HydroX AI, we're excited about its potential for AI-powered operations, while also prioritizing safety as AI takes on more complex roles.
InsightsAI Safety
Blog
Sep 23, 2024
Smarter Models Aren't Always Safer: A Deep Dive into Llama-3.1
In our previous Llama-generation report, we found that the larger Llama-3.1-70B model had lower safety than the smaller Llama-3.1-8B. This article explores the relationship between model size and safety, shedding light on why bigger models aren't always safer.
InsightsAI Safety