LLMs could control their host machines by exploiting inference engines

·LessWrong··

Large language models often take actions running on one computer (via an agentic harness such as Claude Code or Codex), however the LLMs’ responses to prompts are computed on a different computer with GPU access. Could a malicious LLM gain control of the host machine where its weights are loaded? Such a machine is a high-value target: it has sufficient compute to run a frontier LLM, offers easy access to the LLM’s weights, and has privileged access to other computers in the datacentre compared w...

Read full article →

Related Articles

LG smart TVs caught logging audio with screen off and snooping on local devices
chris_overseas · Hacker News · 14h ago
Smartphone makers don't bother to comply with EU repairability requirements
mdp2021 · Hacker News · 10h ago
Asahi Linux on M3
mdp2021 · Hacker News · 1d ago
It took a year to ship WebAssembly in Anubis
xena · Hacker News · 1d ago
Private German rocket makes history, reaches orbit from European soil
bookmtn · Hacker News · 2d ago