mags0ft
|
0c814177fe
|
add caching support for models in start.sh and update documentation
|
2025-12-26 14:49:03 +01:00 |
|
mags0ft
|
403d318ffc
|
update default args to include -ngl 99, improve startup sleep duration
|
2025-11-19 19:16:18 +01:00 |
|
mags0ft
|
381ed9ffff
|
prepare for production release
|
2025-11-19 18:40:00 +01:00 |
|
mags0ft
|
40d4097799
|
fix several major bugs in startup script
|
2025-11-19 17:47:06 +01:00 |
|
mags0ft
|
3a8d2decfd
|
change project to work with llama.cpp instead of Ollama
|
2025-11-15 14:03:10 +01:00 |
|
SvenBrnn
|
3920146e31
|
change MODEL_NAME to OLLAMA_MODEL_NAME to prevent runpod from blocking deploy
|
2025-10-06 08:10:12 +02:00 |
|
Nicholas
|
51d38adcf3
|
Borrow vLLM concurrenccy logic a little.
|
2025-09-11 00:31:09 -05:00 |
|
SvenBrnn
|
26384db5d1
|
remove tests for now and make model configurable
|
2025-05-23 06:53:59 +00:00 |
|
SvenBrnn
|
94e1abaa9e
|
Add runpod configs for hub
|
2025-05-22 04:53:14 +00:00 |
|