Boyd-depth duality and cone programs meet second-order, distributed, and sharpness-aware training — the bridge from convex structure to modern large-scale ML optimization.