• Joined on 2026-07-31
The RunPod worker template for serving our large language model endpoints. Powered by vLLM.
Updated 2026-08-01 11:04:44 -04:00
Updated 2026-08-01 10:56:50 -04:00
A serverless worker to run LLMs in the cloud - using llama.cpp!
Updated 2026-08-01 10:56:33 -04:00
Not a serious project. Just mucking around with GPT 5.5 to see what it can do.
Updated 2026-08-01 10:55:48 -04:00
Updated 2026-08-01 10:55:33 -04:00
Updated 2026-08-01 10:52:02 -04:00