Skip to content
#

ggml

Here are 23 public repositories matching this topic...

Replace OpenAI GPT with another LLM in your app by changing a single line of code. Xinference gives you the freedom to use any LLM you need. With Xinference, you're empowered to run inference with any open-source language models, speech recognition models, and multimodal models, whether in the cloud, on-premises, or even on your laptop.

  • Updated Feb 25, 2025
  • Python

This custom_node for ComfyUI adds one-click "Virtual VRAM" for any GGUF UNet and CLIP loader, managing the offload of layers to DRAM or VRAM to maximize the latent space of your card. Also includes nodes for directly loading entire components (UNet, CLIP, VAE) onto the device you choose. Includes 16 examples covering common use cases.

  • Updated Feb 20, 2025
  • Python

Improve this page

Add a description, image, and links to the ggml topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the ggml topic, visit your repo's landing page and select "manage topics."

Learn more