Skip to content
PrunaAI

🌍 Join the Pruna AI community!

TwitterGitHubLinkedInDiscordReddit


💜 Simply make AI models faster, cheaper, smaller, greener!

Pruna AI makes AI models faster, cheaper, smaller, greener with the pruna package.

  • It supports various models including CV, NLP, audio, graphs for predictive and generative AI.
  • It supports various hardware including GPU, CPU, Edge.
  • It supports various compression algortihms including quantization, pruning, distillation, caching, recovery, compilation that can be combined together.
  • You can either play on your own with smash/compression configurations or let the smashing/compressing agent find the optimal configuration [Pro].
  • You can evaluate reliable quality and efficiency metrics of your base vs smashed/compressed models. You can set it up in minutes and compress your first models in few lines of code!

⏩ How to get started?

You can smash your own models by installing pruna with:

pip install pruna

You can start with simple notebooks to experience efficiency gains with:

Use CaseFree Notebooks
3x Faster Stable Diffusion ModelsSmash for free
Making your LLMs 4x smallerSmash for free
Smash your model with a CPU onlySmash for free
Transcribe 2 hours of audio in less than 2 minutes with WhisperSmash for free
100% faster Whisper TranscriptionSmash for free
Run your Flux model without an A100Smash for free
x2 smaller Sana in actionSmash for free

For more details about installation and tutorials, you can check the Pruna AI documentation.


Pinned Loading

  1. prunaprunaPublic

    Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.

    Python 1.3k 102

Repositories

Showing 10 of 25 repositories

Top languages

Loading…

Most used topics

Loading…