Logo
Explore Help
Sign In
haylan/LLM-Server
Watch 1
Star 0
Fork 0
Code Issues 9 Pull Requests Actions Packages Projects Releases Wiki Activity
Labels Milestones New Pull Request
0 Open 14 Closed
Label
Use alt + click/enter to exclude labels
All labels No label
wayfinder:grilling

wayfinder:map

wayfinder:prototype

wayfinder:research

wayfinder:task

Milestone
All milestones No milestones
Project
All projects No project
Author
All users
Assignee
Assigned to nobody Assigned to anybody
haylan (Arthur Erlich)
Sort
Newest Oldest Most recently updated Least recently updated Most commented Least commented Nearest due date Farthest due date
0 Open 14 Closed
Label
Clear labels
wayfinder:grilling
wayfinder:map
wayfinder:prototype
wayfinder:research
wayfinder:task
Milestone
No milestone
Projects
Clear projects
Assignee
Clear assignees
haylan
Fix/omniroute pr agent timeout
#54
by haylan was merged 2026-09-09 14:29:13 +00:00
main
fix/omniroute-pr-agent-timeout
fix(update.sh): stop config sync loop from dying silently on a missing key
#53
by haylan was merged 2026-09-07 18:10:23 +00:00
main
feat-rag-databases
Add Qdrant/Neo4j RAG storage + update.sh gum fixes
#52
by haylan was merged 2026-09-07 18:07:48 +00:00
main
feat-rag-databases
feat: remove llama-server-fast (Qwen3-4B classifier model)
#51
by haylan was merged 2026-09-07 17:56:27 +00:00
main
remove-fast-model
feat: add qdrant and neo4j for RAG vector/graph storage
#50
by haylan was merged 2026-09-07 17:56:49 +00:00
main
feat-rag-databases
Fix llama-server-fast context-size exhaustion breaking Auto Mode
#49
by haylan was merged 2026-09-06 20:21:53 +00:00
main
fix-fastmodel-context-size
Run git pull first in update.sh, not mid-script
#48
by haylan was merged 2026-09-06 19:40:10 +00:00
main
fix-update-sh-pull-order
Fix GPU pinned at 100% with two containers, flaky render group
#47
by haylan was merged 2026-09-06 19:35:32 +00:00
main
fix-gpu-pin-and-render-group
Downloader for Qwen-Image weights, switch-model.sh
#46
by haylan was merged 2026-09-06 19:01:02 +00:00
main
comfyui-model-and-switch-script
Add llama-server-fast: small non-thinking classifier/fast model
#45
by haylan was merged 2026-09-06 18:39:31 +00:00
main
add-fast-model
feat(llama.cpp): raise default context to 128K, document RAM/SSD offload knobs
#30
by haylan was merged 2026-09-03 04:44:34 +00:00
main
ctx-size-128k
fix(litellm): default max_tokens=4096 for the reasoning model
#20
by haylan was merged 2026-09-02 18:58:12 +00:00
main
fix/litellm-reasoning-max-tokens
feat(litellm): add UI_USERNAME/UI_PASSWORD for the admin UI login
#19
by haylan was merged 2026-09-02 18:29:30 +00:00
main
fix/litellm-ui-admin-creds
fix(downloader): run as root to fix permission denied on models volume
#18
by haylan was merged 2026-09-02 18:03:59 +00:00
main
fix/downloader-permission-denied
Powered by Gitea Version: 1.27.3 Page: 25ms Template: 4ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API