TraceMachina / TraceMachina/nativelink

how to kill worker after the remote action execution?

Open
#815 3 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Rust
Stars
1.6k
Forks
244
Avg merge
1d 16h
Merged PRs (30d)
54

Description

Thank you for the awesome project!

Let's say that we have the following setup:

  • k8s deployment with fixed (for simplicity) number of pods
  • each pod is running nativelink worker

It would be very useful to allow work to exit nativelink binary after a single execution -- this way I can kill the pod and it would be re-created proving a clean environment for the next action execution.

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No source file, test, or entry point is named. Start by locating the worker lifecycle and configuration handling, then determine where a single-execution exit mode belongs. Done means a worker can exit after one remote action so Kubernetes recreates the pod with a clean environment.

Written by the indexing model from the issue text.

Assessment

Tech stack
kubernetes, rust
Domain
backend, build-system, devops
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.