Enterprise Data Lineage for LoRA Policy Fleets
Enterprise reinforcement learning creates more than a training-data problem. It produces a chain of sensitive artifacts: task interactions, tool responses, evaluator judgments, rewards, checkpoints, LoRA weights, reports, and serving traces. When many policies share one foundation model, those artifacts may belong to different customers, workflows, or authorization boundaries even though they depend on the same base deployment.