graphragzen.llm.llama_cpp_models

Description

Functions

suppress_prompter_output(output_structure)

Classes

BaseLlamaCpp

Loads a GGUF model using llama cpp python and it's corresponding tokenizer from HF

Gemma2GGUF

Loads the GGUF version of a gemma2 model using llama-cpp-python

Phi35MiniGGUF

Loads the GGUF version of a Phi 3.5 Mini model using llama-cpp-python