Apple has announced a collaboration with Nvidia to accelerate large language model inference using its open source technology, Recurrent Drafter .It aims to address the computational challenges of auto-regressive token generation, which is crucial for improving efficiency and reducing latency in real-time LLM applications.Through its integration into Nvidia’s TensorRT-LLM framework, ReDrafter extends its impact by enabling faster LLM inference on Nvidia GPUs widely used in production environments. Are this collaborations among Big tech companies killing creativity ?

