Commit Graph
15 Commits
Author SHA1 Message Date
mags0ft 0c814177fe add caching support for models in start.sh and update documentation 2025-12-26 14:49:03 +01:00
mags0ft 0a59d57025 rename tests.json temporarily to skip apparently buggy RunPod CI
The RunPod support told me to do this until the issues are resolved.
2025-12-26 11:37:23 +01:00
mags0ft 398b59c0a1 update test timeout and default LLAMA_SERVER_CMD_ARGS for improved performance 2025-12-25 19:17:31 +01:00
mags0ft 889d3c5c96 move find_cached.py to src 2025-12-18 10:57:10 +01:00
mags0ft 41b3b3d02b add helper script to use cached models 2025-12-18 10:50:33 +01:00
mags0ft 403d318ffc update default args to include -ngl 99, improve startup sleep duration 2025-11-19 19:16:18 +01:00
mags0ft 381ed9ffff prepare for production release 2025-11-19 18:40:00 +01:00
mags0ft c42b80ebbd add default behavior for undefined LLAMA_SERVER_CMD_ARGS 2025-11-19 18:11:36 +01:00
mags0ft 40d4097799 fix several major bugs in startup script 2025-11-19 17:47:06 +01:00
mags0ft 36a8b32d1e fix: ensure script fails on error by setting 'set -e' 2025-11-15 16:07:54 +01:00
mags0ft 5f6c099504 further fixes for start.sh script 2025-11-15 15:15:17 +01:00
mags0ft be3de61c52 fix: correct logic for port validation in start.sh 2025-11-15 15:02:05 +01:00
mags0ft 558da755c3 add empty handler.py file to mark repository as Runpod-compatible endpoint 2025-11-15 14:11:32 +01:00
mags0ft 9388fae14b add badge to README 2025-11-15 14:07:00 +01:00
mags0ft 3a8d2decfd change project to work with llama.cpp instead of Ollama 2025-11-15 14:03:10 +01:00