Research

Published, not just claimed.

Our methods are peer-reviewed and presented at the venues that set the agenda for AI evaluation and safety.

KDD 2026ICML 2026NeurIPS 2025TrustCon 2026
0+Papers & invited talks
0+Top-tier venues
0+Partner labs & institutions
Featured publication

The research behind Flint.

KDD 2026Peer-reviewed

EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities

EvoFlint formalizes the evolutionary, multi-turn red-teaming approach that powers Flint, treating model failure discovery as a diversity-driven search across conversation space rather than a fixed test suite. The result is a living atlas of how production language models break down over extended interactions, not just whether they pass.

Venue: KDD 2026Topic: Multi-turn red-teamingSystem: Flint
Publications and conferences

Peer-reviewed work.

KDD 2026Paper

EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities

Reinforce Labs

Formalizes the evolutionary multi-turn red-teaming approach behind Flint.

ICML 2026Paper

InfoDLM: Information-Adaptive Discrete Diffusion LM Pretraining

Prof. Tony Geng, Rice University

A learned, feedback-driven masking policy for discrete diffusion language-model pretraining.

Community leadership

Convening the AI field.

We don't just publish, we set the agenda, moderating and presenting alongside the leading labs.

KDD 2026Talk

AI Integrity as a Search Problem: Diversity-Driven Behavioral Evaluation

Anish Das Sarma, CEO

Presenting red-teaming as a diverse search problem over model behavior.

TrustCon 2026Panel

"Who Owns What? Responsible AI in Dynamic Production Systems"

Anish Das Sarma, moderating

A panel with heads of Trust & Safety and safety engineering from Anthropic, OpenAI, Google, and Mercor.

NeurIPS 2025Workshop

"Agentic AI: Organizational Automation vs. Personalization at Scale"

Co-hosted with Centific

Exploring the tension between organizational automation and personalization in agentic systems.