With our local directory junction surgery complete and 350GB of duplicated GGUF cache bloat thoroughly expunged from the NVMe storage array, the infrastructure bounds were finally behaving themselves. But what good is a lean, sovereign local inference registry if the models you run inside it still act like corporate customer-service