Skip to main content
The tracebloc CLI is how you manage your secure environment and its data from your own machine: sign in, connect this machine to tracebloc, ingest datasets, list and delete them, check the connection, and set how much of the machine training may use. tb is the short alias — every command on this page works as tb … too.
You never touch Helm, edit YAML, or run kubectl. tracebloc data ingest finds your secure environment, checks your data locally, copies it in, runs the ingestion, and streams the progress for you.

Install

If you deployed with the Quick Start one-liner, the CLI is already installed. Otherwise, install the CLI on its own:
Binaries are cosign-signed and multi-arch; the installer also creates the tb alias (unless an unrelated tb is already installed). Pin a version with sh -s -- --version vX.Y.Z (or $env:RELEASE_VERSION='vX.Y.Z' on Windows). Verify:
Run tracebloc with no arguments for a welcome screen that shows who you’re signed in as, whether your secure environment is running, and the next commands to try.

Sign in and connect this machine

login works on a headless or SSH machine — the browser doesn’t have to be on the same machine — and is safe to re-run (pass --force to switch accounts). client create takes its name from --name or TRACEBLOC_CLIENT_NAME, otherwise it generates one; on a cluster that already runs a client, it adopts that client instead of creating a second one.

Check your secure environment

doctor exits 0 when healthy, 2 when it found a problem, 3 when it can’t read your local config. --diagnose writes a redacted support bundle you can send to tracebloc.
If a command reports that no tracebloc client was found (exit code 4), your client runs in another namespace — add -n <your-namespace>. cluster info, doctor, resources, and the data commands all accept -n, --context, and --kubeconfig.

Ingest a dataset

tracebloc data ingest copies a local dataset into your secure environment’s storage, runs the ingestion, and follows it to completion. Your data never leaves your infrastructure.
Omit the flags on a terminal to run guided — the CLI asks for the task, name, and label column. Add --dry-run to check your data and your secure environment without creating anything.
tracebloc data push (and tracebloc dataset push) is a deprecated alias of tracebloc data ingest and will be removed in a future release. Use data ingest.

Tasks

Pick one with --task:

Dataset layout

What you pass as <dataset> depends on the task family. A bare .csv file is accepted only for tabular and time-series tasks; image and text datasets must be a folder.
A single CSV — pass the file itself, or a folder holding exactly one .csv:
Time-series classification adds fixed sequence_id and timestamp columns — see its template.
Each task’s dataset template shows its exact columns and an example — for example causal language modeling, sequence-to-sequence, token classification, sentence-pair classification, embeddings, and semantic segmentation.
One data ingest is capped at 1 GiB total and 500 MiB per file.

Key flags

Run tracebloc data ingest --help for the full list.

Exit codes

Useful when you script data ingest:

List and delete datasets

data delete is destructive and can’t be undone; it asks for confirmation (--yes skips it, --dry-run shows what would be deleted). The dataset’s catalog entry on tracebloc is kept as a record and marked unavailable, so experiments that used it keep their history. tracebloc data rm is an alias. Both commands accept --output-json.

Validate a config locally

Checks an ingest.yaml against the embedded v1 schema in milliseconds — no cluster required. Exits 0 when valid, 2 on schema violations, 3 when the file can’t be read or parsed. The declarative form (also accepted by the Helm ingestor chart):

Command reference

data also answers to dataset. Add --help to any command for the full flag list, --verbose for step-by-step detail, or --plain to turn off color. For lifecycle tasks (upgrade, stop/start, uninstall), see Operations.

CLI or Helm chart?

Both submit to the same ingestion protocol — pick the one that fits your workflow:
  • CLI — local data on your workstation; the everyday choice. Handles copying and submission for you.
  • Helm ingestor chart — Kubernetes-native / GitOps, when your data is already staged on the cluster.