Verify inference output precision for vllm-plugin-FL on any hardware backend. Compares token-level outputs against NVIDIA + upstream vLLM ground truth. Use this skill whenever a task requires a correctness gate: model porting, hardware adaptation, plugin version upgrades, or regression detection after any code change. Trigger when the user says "check precision", "verify correctness", "compare outputs", "run E2E precision test", "precision alignment", or "does the output match NVIDIA?". Works for text-only and multimodal models; greedy decoding (temperature=0) is the standard mode.