Urgent.News

What's breaking now, across thousands of outlets.

AI

The agent finished. Who turns off the VM?

Suppose a coding agent opens a pull request at 6 p.m. The tests pass, a preview is running, and the reviewer has already logged off. The agent's task is complete. The machine still has something to do: serve that preview until someone looks at it. Or perhaps it doesn't. Perhaps the preview can be rebuilt tomorrow and the workspace should shut down now. Either choice is reasonable. Leaving it…

When a coding agent finishes its task, several components may still need to remain active. The agent's command, session, and workspace are three separate elements that do not automatically control each other. For example, during a pull request review, the agent's task might finish, but the machine still needs to serve the preview until someone looks at it or it can be rebuilt the next day.

The agent session should end when the agent hands over the result and the review evidence is saved. The preview should only be retained until an explicit expiry time. This policy separates the agent's completion from other aspects of the task, ensuring that each component is managed independently.

For instance, a Node.js project can utilize a Bash script with GNU timeout to limit the execution time of tests. If the command does not exit within the specified time frame, a termination signal is sent. However, this approach does not provide a reliable boundary for the agent itself, as an active workspace may never become idle. An inactivity timer could be a solution, but its effectiveness depends on how activity is defined.

The actual implementation of this policy depends on the agent platform. It is a policy to be tested and refined, as not all agent platforms provide the same controls. A key consideration is preserving evidence before removing the workspace, especially when the only copy of a change or test output lives on the stopped machine. This ensures that useful logs and test results are not lost.

When stopping an EC2 instance, it is important to distinguish between termination and stopping. Stopping an EC2 instance retains the instance, its data, and associated resources, while terminating it releases those resources. If uncommitted changes or untracked files remain after stopping the agent, Git refuses removal until they are resolved. Therefore, it is crucial to check the commit status and retention rules before removing the worktree.

To test this setup, run a failed task with an expired deadline and observe the agent's behavior, including whether the logs survive and which resources remain active. Then, run a successful task with an unavailable reviewer, ensuring the preview remains available for the promised window without requiring the agent to continue working. By comparing these two cases, you can better understand the readiness of the system and the reasons behind its functioning.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

Your LLM Types One Token at a Time. It Doesn't Have To.

Every token your LLM emits costs one full forward pass through the entire model. Seventy billion parameters loaded from memory, multiplied, discarded — for a single token. Then again. And again.

  • Speculative decoding drafts tokens with cheap model before big model verification
  • Acceptance rule accepts drafted token with probability min(1, q(d)/p(d))
  • EAGLE-3 achieves 2-3x speedup over vanilla decoding, but diminishing returns at high batch sizes

An LLM observability platform stores prompts, and prompts are the application

An LLM observability platform stores prompts, and prompts are the application A title query for Langfuse returns 346 matches in ZoomEye. The number is small and the contents are unusual.

  • LLM observability platform records prompts as application logic
  • Traces contain prompts, model parameters, completion details
  • Langfuse indexes sensitive traces, including potential secrets

The most exciting claims from OpenAI’s heap of new proofs

Scientists have called out several math and computer science results as the most significant in the company’s overwhelming new deluge of achievements

  • OpenAI's model proves quasi-Riemann hypothesis, offering insights into prime number distribution
  • Model resolves Hilbert's 10th problem for rational numbers, showing unsortable equations
  • Progress on Kakeya conjecture extended to four dimensions

More from Thursday 8 October →