- New experiments/dsv4_h200_sglang_vs_vllm/ with TP=8 on all 8 H200 cards - Phase1 short-context throughput + Phase2 long context up to 200k - Unified scenario matrix, warmup, parsing, and side-by-side comparison report - H200 SGLang vs vLLM entry added to root README