@remi_or_
Synthetic data generation is now native in transformers 🔥 Last week, transformers continuous batching (CB) hit 84% of vLLM throughput. This week, we tuned torch.compile: now we are at 95% for 8K generation length 🦾 The gap isn't closing anymore. It's gone.💀 https://t.co/t5J9TOA5zM