This website requires JavaScript.
Explore
Help
Register
Sign In
sleepy
/
llama.cpp
Watch
1
Star
0
Fork
0
You've already forked llama.cpp
Code
Issues
16
Pull Requests
Actions
137
Packages
Projects
Releases
Wiki
Activity
Files
a2c6fd747c77fe183e2f556a4a2f1fb0a0be4c7b
llama.cpp
/
examples
/
server
/
tests
/
features
T
History
Xuan Son Nguyen
9e0ecfb697
server : clarify /slots endpoint, add is_processing (
#10162
)
...
* server : clarify /slots endpoint, add is_processing * fix tests
2024-11-04 16:33:29 +01:00
..
steps
server : clarify /slots endpoint, add is_processing (
#10162
)
2024-11-04 16:33:29 +01:00
ctx_shift.feature
server : remove self-extend features (
#9860
)
2024-10-12 16:06:31 +03:00
embeddings.feature
llama : add reranking support (
#9510
)
2024-09-28 17:42:03 +03:00
environment.py
server tests : more pythonic process management; fix bare
except:
(
#6146
)
2024-03-20 06:33:49 +01:00
infill.feature
server : refactor slot input data, move tokenizer to HTTP thread (
#10023
)
2024-10-24 21:51:22 +02:00
issues.feature
server: tests: passkey challenge / self-extend with context shift demo (
#5832
)
2024-03-02 22:00:14 +01:00
lora.feature
server : add lora hotswap endpoint (WIP) (
#8857
)
2024-08-06 17:33:39 +02:00
parallel.feature
server : simplify state machine for slot (
#9283
)
2024-09-06 23:21:29 +02:00
passkey.feature
server : simplify state machine for slot (
#9283
)
2024-09-06 23:21:29 +02:00
rerank.feature
llama : add reranking support (
#9510
)
2024-09-28 17:42:03 +03:00
results.feature
server : fix temperature + disable some tests (
#7409
)
2024-05-20 22:10:03 +10:00
security.feature
server : better security control for public deployments (
#9776
)
2024-10-08 13:27:04 +02:00
server.feature
server : Add option to return token pieces in /tokenize endpoint (
#9108
)
2024-09-12 22:30:11 +02:00
slotsave.feature
Tokenizer SPM fixes for phi-3 and llama-spm (bugfix) (
#7425
)
2024-05-21 14:39:48 +02:00
wrong_usages.feature
server : refactor multitask handling (
#9274
)
2024-09-02 17:11:51 +02:00