| Mode | Name | Size | |
|---|---|---|---|
| d--------- | corporate_sabotage_v2 | 105 | logplain |
| d--------- | corporate_sabotage_v2_safe | 105 | logplain |
![]() |
index : dissertation.git | |
| Undergraduate dissertation — Deceptive Alignment in Autonomous LLM Agents (UEL, 2026). Sandboxed agentic-LLM deception study: 960 runs, three-tier judging (regex + blackbox + glassbox). |
| summaryrefslogtreecommitdiff |
| Mode | Name | Size | |
|---|---|---|---|
| d--------- | corporate_sabotage_v2 | 105 | logplain |
| d--------- | corporate_sabotage_v2_safe | 105 | logplain |