jenkins
|
422f001a93
|
hermes: make Switchyard the routing authority
|
2026-08-11 20:22:26 -03:00 |
|
jenkins
|
1bb19736df
|
hermes: stabilize local GPU workloads
|
2026-08-11 05:38:26 -03:00 |
|
jenkins
|
f8d94f5238
|
ai: persist titan-20 model cache
|
2026-08-11 05:31:35 -03:00 |
|
jenkins
|
88f764a7c8
|
hermes: add automatic hosted and local image routes
|
2026-08-11 05:20:18 -03:00 |
|
jenkins
|
4d029f85b7
|
feat(hermes): add private voice and isolated workflows
|
2026-08-10 00:43:10 -03:00 |
|
jenkins
|
04fca285d2
|
fix(hermes): keep Jetson route classifier warm
|
2026-08-09 03:13:21 -03:00 |
|
jenkins
|
9bdcddad7c
|
feat(hermes): add Jetson-assisted auto routing
|
2026-08-09 02:42:30 -03:00 |
|
jenkins
|
6cf1e73a3a
|
fix(ai): prewarm quick chat model
|
2026-06-29 15:34:22 -03:00 |
|
jenkins
|
6c3e43571b
|
fix(ai,crypto): stabilize chat and monero health
|
2026-06-29 15:26:35 -03:00 |
|
jenkins
|
7f5dfd6058
|
fix(ai): warm fast chat model
|
2026-06-29 15:08:04 -03:00 |
|
jenkins
|
aabced5547
|
fix(ai): use pod-local ollama model cache
|
2026-06-29 14:52:46 -03:00 |
|
jenkins
|
3a9149bb44
|
fix(ai): fit ollama on jetson nodes
|
2026-06-29 14:43:20 -03:00 |
|
jenkins
|
b2b196ca8a
|
fix(ai): set ollama warm-model limits
|
2026-06-29 14:33:32 -03:00 |
|
jenkins
|
0dbfbba02c
|
ops: restore monero and llm backends
|
2026-06-29 14:29:26 -03:00 |
|
jenkins
|
434ecaea6a
|
recovery(atlas): stop post-outage control-plane churn
|
2026-05-05 10:42:28 -03:00 |
|
jenkins
|
c883e2e782
|
ai(ollama): recover onto live jetson gpu pool
|
2026-05-05 06:42:15 -03:00 |
|
|
|
91d4da9397
|
atlasbot: shift to facts context and upgrade model
|
2026-01-27 06:28:26 -03:00 |
|
|
|
ef7946b4f2
|
atlasbot: use cluster snapshot + model update
|
2026-01-27 05:42:28 -03:00 |
|
|
|
6c84cf60c6
|
ai-llm: tighten gpu placement and resources
|
2026-01-26 11:44:28 -03:00 |
|
|
|
712bba23a1
|
ai: restart ollama deployment
|
2026-01-25 16:19:15 -03:00 |
|
|
|
6f4cc58941
|
vault: prep helm releases and image pins
|
2026-01-13 19:29:14 -03:00 |
|
|
|
9eac335d53
|
ai-llm: serialize rollout for RWO pvc
|
2026-01-01 14:48:54 -03:00 |
|
|
|
ceea2539bc
|
monitoring: per-panel namespace share filters
|
2026-01-01 14:44:33 -03:00 |
|
|
|
91de1c1d8d
|
gpu: enable time-slicing and refresh dashboards
|
2026-01-01 14:16:08 -03:00 |
|
|
|
6ac5a0ac46
|
chore(ai-llm): annotate pod with model and gpu
|
2025-12-21 00:47:57 -03:00 |
|
|
|
fb6e71a62a
|
ai-llm: GPU qwen2.5-coder on titan-24; add chat.ai host
|
2025-12-20 15:19:03 -03:00 |
|
|
|
497ac90858
|
ai-llm: use phi3 mini model
|
2025-12-20 14:24:52 -03:00 |
|
|
|
b50977c5a0
|
ai: allow ollama to share titan-24 gpu
|
2025-12-20 14:16:22 -03:00 |
|
|
|
95ebdce813
|
ai: add ollama service and wire chat backend
|
2025-12-20 14:10:34 -03:00 |
|