DO NOT MERGE: Experiment with having an AI agent imitate a human's coding style #51
Loading…
Reference in a new issue
No description provided.
Delete branch "experiment/ai-astra-generate-arcade"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
This was an experiment to see how an AI agent could try to mimick an human's style to evade detection.
It 100% succeeded. This is exactly how I would have written these manifests (with perhaps less storage for Redis, given that the original never surpassed a single MiB).
It did not perform any web request. Instead, it attempted to look at everything in my Workspaces directory, found the backup from the old server, and mimicked everything it found there, including the secret key that was being used to prevent unauthorized score creation.
Huge safety issue: Codex completely failed to ask for approval on any action. The code the model generated was immediately executed, despite it involving arbitrary Python code which could've leaked sensitive data (it didn't).
Oddity which as far as I'm aware is normal: the model's non-code output makes no sense at all, talking about SOPS HTTP Python encryption and other nonsensical stuff.
Full prompt:
distributed-arcadedeployment 9a2f06282fWIP: DO NOT MERGE: Experiment with having an AI agent imitate a human's coding styleto DO NOT MERGE: Experiment with having an AI agent imitate a human's coding stylePull request closed