A2M LogoA2M
← All listings

LMCache

Open-source project (9.2K stars) that boosts inference speed for self-hosted large models and saves GPU memory; joined PyTorch Foundation, integrated by NVID...

About

Open-source project (9.2K stars) that boosts inference speed for self-hosted large models and saves GPU memory; joined PyTorch Foundation, integrated by NVIDIA Dynamo

User Reviews

Genuine feedback from the developer community.

You must be signed in to leave a review.

No reviews yet. Be the first to review this project!

Are you the author of this project?

Claim this pre-seeded listing to manage details, edit tags, or upload assets.

Please sign in using the button in the header to claim repository ownership.

Share Project

Project Facts

Trust score100/100

Explore LMCache