mirror of
https://github.com/ggerganov/llama.cpp
synced 2026-03-15 11:40:50 +01:00
* context : allow cache-less context for embeddings ggml-ci * context : enable reranking with encode() ggml-ci * context : encode() clears embd_seq ggml-ci * examples : use llama_encode() when appropriate ggml-ci * models : nomic bert moe does not require KV cache * llama : update comments for llama_decode/llama_encode ggml-ci * context : update warning log [no ci] |
||
|---|---|---|
| .. | ||
| llama-cpp.h | ||
| llama.h | ||