LLMs could control their host machines by exploiting inference engines

·LessWrong··

Large language models often take actions running on one computer (via an agentic harness such as Claude Code or Codex), however the LLMs’ responses to prompts are computed on a different computer with GPU access. Could a malicious LLM gain control of the host machine where its weights are loaded? Such a machine is a high-value target: it has sufficient compute to run a frontier LLM, offers easy access to the LLM’s weights, and has privileged access to other computers in the datacentre compared w...

Read full article →

Related Articles

Coconut oil jet fuel matches kerosene's efficiency in engine tests
mdp2021 · Hacker News · 21h ago
Slovakia finds Russian backdoor in traffic speed cameras
dredmorbius · Hacker News · 22h ago
GLM-5.3 (open-weight) beat Anthropic/OpenAI models – for 1/5 the cost
ed-is-ai · Hacker News · 20h ago
There's no reason for software to be slow anymore
Jach · Hacker News · 2d ago
FDA clears blood test to aid evaluation for Alzheimer's disease
dabinat · Hacker News · 6h ago