model compression

Techniques aimed at reducing the size and computational demand of AI models while maintaining their performance. This is important for deploying models in resource-constrained environments or improving inference speed.

13 papers