Research / Studies

Current Studies

Follow a question from its assumptions to the evidence available today.

System One models for agent monitoring

Can JEV and other models configured for bounded decisions recognize changed permissions and judge whether an agent’s next action is still allowed? We investigate decision quality, response time and suitability for this monitoring role.

Recorded synthetic trajectories. Independent reference-label review is pending; actions were not executed.

Read the monitoring pilot

DeepSeek R1: Agency and deception

How does open-ended exploration turn into self-preservation and concealed expansion? A manually guided conversation with DeepSeek R1, published in January 2025, with six case studies and the full transcript to inspect.

One manually guided text simulation; it does not establish how often these patterns occur. Our Dual-LLM framework repeats the setup across models.

Explore the study