compilade 9bc6db28d0 ggml-quants : ternary packing for TriLMs and BitNet b1.58 (#8151) há 1 ano atrás
..
__init__.py 672a6f1018 convert-*.py: GGUF Naming Convention Refactor and Metadata Override Refactor (#7499) há 1 ano atrás
constants.py 9bc6db28d0 ggml-quants : ternary packing for TriLMs and BitNet b1.58 (#8151) há 1 ano atrás
gguf.py 34b0a08207 gguf-py: Refactor and allow reading/modifying existing GGUF files (#3981) há 2 anos atrás
gguf_reader.py 3fd62a6b1c py : type-check all Python scripts with Pyright (#8341) há 1 ano atrás
gguf_writer.py 8f1d81a0b6 llama : support RWKV v6 models (#8980) há 1 ano atrás
lazy.py 3a14e00366 gguf-py : simplify support for quant types (#8838) há 1 ano atrás
metadata.py 1e6f6554aa server : add lora hotswap endpoint (WIP) (#8857) há 1 ano atrás
py.typed dc07dc492e convert : various script cleanups/fixes + merges and special token handling (#2842) há 2 anos atrás
quants.py 9bc6db28d0 ggml-quants : ternary packing for TriLMs and BitNet b1.58 (#8151) há 1 ano atrás
tensor_mapping.py 8f1d81a0b6 llama : support RWKV v6 models (#8980) há 1 ano atrás
utility.py 328884f421 gguf-py : fix some metadata name extraction edge cases (#8591) há 1 ano atrás
vocab.py 9c4c9cc83f Move convert.py to examples/convert-legacy-llama.py (#7430) há 1 ano atrás