Projects with this topic
Sort by:
-
Experimental text-to-speech utility via keyboard shortcut, powered by Large Language Models (LLM).
Updated -
A utility for instant grammar correction in any text field via keyboard shortcut, powered by Large Language Models (LLM).
Updated -
Extreme KV Cache Compression for LLM Inference — C++17/CUDA implementation of TurboQuant (arXiv 2504.19874). 7.5x compression, <2% quality loss.
Updated