mags0ft
|
8a2982faa4
|
adjust incorrect llama-server start command log
|
2025-12-27 23:51:00 +01:00 |
|
mags0ft
|
d4ac091735
|
add timeout for llama-server startup to prevent indefinite waiting
|
2025-12-27 18:00:58 +01:00 |
|
mags0ft
|
3d442d9123
|
add error handling for unexpected llama-server process exit
|
2025-12-26 14:51:06 +01:00 |
|
mags0ft
|
0c814177fe
|
add caching support for models in start.sh and update documentation
|
2025-12-26 14:49:03 +01:00 |
|
mags0ft
|
0a59d57025
|
rename tests.json temporarily to skip apparently buggy RunPod CI
The RunPod support told me to do this until the issues are resolved.
|
2025-12-26 11:37:23 +01:00 |
|
mags0ft
|
398b59c0a1
|
update test timeout and default LLAMA_SERVER_CMD_ARGS for improved performance
|
2025-12-25 19:17:31 +01:00 |
|
mags0ft
|
889d3c5c96
|
move find_cached.py to src
|
2025-12-18 10:57:10 +01:00 |
|
mags0ft
|
41b3b3d02b
|
add helper script to use cached models
|
2025-12-18 10:50:33 +01:00 |
|
mags0ft
|
403d318ffc
|
update default args to include -ngl 99, improve startup sleep duration
|
2025-11-19 19:16:18 +01:00 |
|
mags0ft
|
381ed9ffff
|
prepare for production release
|
2025-11-19 18:40:00 +01:00 |
|
mags0ft
|
c42b80ebbd
|
add default behavior for undefined LLAMA_SERVER_CMD_ARGS
|
2025-11-19 18:11:36 +01:00 |
|
mags0ft
|
40d4097799
|
fix several major bugs in startup script
|
2025-11-19 17:47:06 +01:00 |
|
mags0ft
|
36a8b32d1e
|
fix: ensure script fails on error by setting 'set -e'
|
2025-11-15 16:07:54 +01:00 |
|
mags0ft
|
5f6c099504
|
further fixes for start.sh script
|
2025-11-15 15:15:17 +01:00 |
|
mags0ft
|
be3de61c52
|
fix: correct logic for port validation in start.sh
|
2025-11-15 15:02:05 +01:00 |
|
mags0ft
|
558da755c3
|
add empty handler.py file to mark repository as Runpod-compatible endpoint
|
2025-11-15 14:11:32 +01:00 |
|
mags0ft
|
9388fae14b
|
add badge to README
|
2025-11-15 14:07:00 +01:00 |
|
mags0ft
|
3a8d2decfd
|
change project to work with llama.cpp instead of Ollama
|
2025-11-15 14:03:10 +01:00 |
|