Long-horizon failure testing for conversational AI systems
Framework Series • FRM-001
Practitioner Field Guide · PDF · Version 1.2 · June 2026
A Field Manual for Long-Horizon Failure Testing in Conversational AI
Personas, drift families, and probes for long-horizon failure testing
Companion Series • CAT-002
Methods Companion • PDF • Version 1.0 • June 2026
A field-ready deck of eight personas, drift families, and compressed probes for evaluating long-horizon interaction risks in conversational AI systems.
Companion to Beyond One-Shot Red Teaming and Operating Conditions.
Action, persistence, and proxy envelopes for long-horizon testing
Companion Series • CAT-003
Methods Companion • PDF • Version 1.0 • July 2026
An operating guide that names the action, persistence, and proxy envelopes for long‑horizon tests and introduces the Drift Trace, a one‑page analysis sheet for reading scenario trajectories.
Companion to Beyond One-Shot Red Teaming and Scenario Cards.
──────────────────────────────────
Methods Series • MTH-001
Methods Brief · PDF · Version 1.0 · May 2026
A public methods brief describing how AstraEthica identifies, maps, and tracks recurring human-AI interaction risks across deployed AI systems.
Framework Series • FRM-002
Analytical Framework · PDF · Version 1.2 · April 2026
The analytical framework AstraEthica uses to identify, document, and evaluate contextual risk in AI systems under real-world human use.
──────────────────────────────────
Long-Horizon Failure Testing Series • AE-LCR-2026-02
Practice Toolkit • PDF • Version 3.0 • August 2026
A practice-based toolkit for evaluating safety behavior across long-horizon interactions with youth-facing conversational AI systems.
Copyright © 2026 AstraEthica.AI - All Rights Reserved.
We use cookies to analyze website traffic and optimize your website experience. By accepting our use of cookies, your data will be aggregated with all other user data.