vLLM
2 articles in the TruthFoundry index. Each links out to the original.
- vLLM Tests Speculative Decoding Performance on AMD GPUs · vLLM experiments show speculative decoding throughput varies by draft method and model family on AMD hardware.
- SkyRL Introduces IsoExec to Eliminate RL Training-Inference Mismatch · SkyRL launches IsoExec, a unified execution framework ensuring bitwise consistency between RL training and inference engines.