Blogs

AI safety and security insights.

Research notes, company updates, partner announcements, and practical writing from the HydroX content archive.

Archive

Latest articles

Research notes, product thinking, and company updates from the HydroX AI content archive.

Clear
Blog Mar 11, 2025

New Research: Exploring the Impact of Output Length on LLM Safety

We’re excited to announce that HydroX AI is sponsoring and collaborating on new AI safety research. Our latest paper explores a key yet overlooked factor in LLMs: how output length affects model safety and reasoning.

InsightsAI Safety
Blog Jan 29, 2025

DeepSeek-R1-Distill Models: Does Efficiency & Reasoning Come at the Expense of Security?

DeepSeek, a Chinese AI company, has recently gained attention in the AI community. Known for its innovation, it has developed models that rival top systems — offering similar performance with lower cost and resource use.

InsightsAI Safety
Blog Jan 8, 2025

The Safety Trade-offs of Advanced AI: Insights from Llama-3.3 and Tulu-3

Alongside major closed-source model announcements in late 2024, the open-source community also saw key releases. In this brief post, we explore Llama-3.3 and Tulu-3, evaluating their performance in terms of AI safety and security.

InsightsAI Safety
Blog Dec 6, 2024

Uncovering AI Weaknesses: How Simple Prompts Threaten Agent Safety

AI agents powered by advanced LLMs like GPT-4 and Llama are revolutionizing human-machine interaction, but they come with risks. This blog explores how a simple adversarial strategy can reveal vulnerabilities and leading to dangerous consequences.

InsightsAI Safety
Blog Nov 20, 2024

AI Scriper -- How AI-Powered Scrapers Are Redefining Red Teaming

AI-powered scrapers are not just faster—they're fundamentally more resilient and adaptive than their traditional counterparts.

InsightsAI Safety
Blog Oct 28, 2024

Reacting to Anthropic’s Latest Claude 3.5 Release: A New Era of Safe Interaction

Anthropic’s release of Claude 3.5 is a major step forward in LLM evolution. At HydroX AI, we're excited about its potential for AI-powered operations, while also prioritizing safety as AI takes on more complex roles.

InsightsAI Safety
Blog Sep 23, 2024

Smarter Models Aren't Always Safer: A Deep Dive into Llama-3.1

In our previous Llama-generation report, we found that the larger Llama-3.1-70B model had lower safety than the smaller Llama-3.1-8B. This article explores the relationship between model size and safety, shedding light on why bigger models aren't always safer.

InsightsAI Safety
Blog Sep 13, 2024

Evaluating OpenAI’s o1-mini and GPT-4o-mini – Advances and Areas for Improvement

On September 12, 2024, OpenAI released its powerful new model, OpenAI o1, featuring advanced reasoning and enhanced safety against jailbreak attempts. This sets a new benchmark for secure AI, while the GPT-4o mini model is also praised for its strong safety features.

InsightsAI Safety
Blog Aug 14, 2024

Llama Series Comparison Across Generations: A White Paper

The Llama series, an open-source LLM developed by Meta, has gained recognition for its high performance and the emphasis placed on safety and security during its development.

InsightsAI Safety