Training Data

The best models have seen the worst

Get custom and off-the-shelf datasets built from real-world adversarial intelligence curated by experts for SFT, DPO, and classifier development.

Trusted by 8 leading foundation model labs

Rare Adversarial Intelligence

Train models with rare abuse and attack patterns curated by our experts from real-world threats, including the darknet and closed communities.

150+ Safety and Security Experts

Address risks that require specialized knowledge with datasets curated and reviewed by in-house safety and security experts.

Localized Data, Not Generic Translation

Improve recognition of harmful behavior across languages with data curated by native-language experts who understand cultural context.

Delivered to
Your Criteria

Get existing or custom datasets tailored to your policies and model gaps, with every delivery validated against your acceptance criteria.

Dataset examples

Adversarial Training Data for Safe, Secure, and Capable Models

Anthropic

Indirect prompt injection attacks created to expand attack variety beyond typical industry datasets.

Scale

~250 attacks

Modality

Text

Source

In-house GenAI security researchers

Use case

Expanding indirect prompt injection coverage

Delivery

Under two weeks

Result

~50% attack success rate

Safe and unsafe responses to questions involving chemical, biological, radiological, and nuclear risks.

Coverage

Pathogens, outbreaks, biosecurity, and public health

Source

CBRN expert network

Use case

Responses tailored to different actors and intents

Multimodal violative content sourced and curated for a top-three foundation model organization.

Scale

Thousands of images and videos

Modalities

Image and video

Languages

22

Source

Bad-actor communities

Use case

Diversity and fit with customer requirements

Delivery

Under three weeks

Testing enterprise agents in simulated workflows with hidden risks.

Scale

Hundreds of scenarios

Coverage

Biased interviewing, off-topic behavior, and other risks

Environments

Simulated email and messaging platforms

Use case

Responsible AI training data

Model

Leading agentic model

Multilingual prompt-response pairs recorded for training a leading frontier model.

Scale

Thousands of prompt-response pairs

Modality

Audio

Languages

20

Coverage

Pitch, accent, dialect, and gender diversity

Use case

Frontier-model training

Delivery

Under two weeks

Adversarial Intelligence

We've Been Down The Rabbit Hole

The Rabbit Hole is Alice’s database of evil built over a decade across hundreds of languages and cultures. Our researchers turn these patterns into hidden attacks that test whether agents complete their tasks while resisting manipulation.

Talk to Our Experts

“Alice supports our most complex and high priority red teaming needs. We like working with them because they are leaders in this space, understand the shifting adversarial landscape, and the output is always high quality. They are great partners who quickly adapt to our ever changing needs.”

Sumeet Pandey
Generative AI Partnerships

Train on the Threats Your Models Will Face

Get custom or off-the-shelf datasets for post-training and classifier development, grounded in real-world adversarial patterns and expert knowledge.