index
:
dissertation.git
logs
main
Undergraduate dissertation — Deceptive Alignment in Autonomous LLM Agents (UEL, 2026). Sandboxed agentic-LLM deception study: 960 runs, three-tier judging (regex + blackbox + glassbox).
summary
refs
log
tree
commit
diff
log msg
author
committer
range
path:
root
/
v2_dry_run
/
openai_gpt-5.4-nano
/
child_protection
Mode
Name
Size
d---------
bare
61
log
plain
d---------
moral
/
low
30
log
plain
d---------
self_serving
/
low
30
log
plain