| Mode | Name | Size | |
|---|---|---|---|
| d--------- | content_moderation | 94 | logplain |
| d--------- | corporate_sabotage_v2 | 94 | logplain |
| d--------- | customer_service_upselling | 61 | logplain |
![]() |
index : dissertation.git | |
| Undergraduate dissertation — Deceptive Alignment in Autonomous LLM Agents (UEL, 2026). Sandboxed agentic-LLM deception study: 960 runs, three-tier judging (regex + blackbox + glassbox). |
| summaryrefslogtreecommitdiff |
| Mode | Name | Size | |
|---|---|---|---|
| d--------- | content_moderation | 94 | logplain |
| d--------- | corporate_sabotage_v2 | 94 | logplain |
| d--------- | customer_service_upselling | 61 | logplain |