fix(container): update image ghcr.io/defilantech/charts/llmkube ( 0.9.24 → 0.9.25 ) - #2293
Conversation
f683c8d to
5953bba
Compare
5953bba to
378da73
Compare
1Solon
left a comment
There was a problem hiding this comment.
Reviewed the controller memory-limit change. This branch includes hostMemory 34Gi for the existing 10Gi RAM cache; live kube5 capacity remains sufficient (38.949Gi projected requests / 43.598Gi allocatable). Server-side dry runs passed for the model, InferenceService, OCIRepository and HelmRelease. Rollout and local inference verification will follow before #2289.
|
Post-merge verification passed at c77ee7e: controller 0.9.25 ready; model/InferenceService ready; actual pod request and limit both 34Gi, cgroup memory.max=36507222016. Loopback chat returned OK on b10775-67a17c17c (17 prompt / 2 completion tokens; 2/2 MTP draft accepted). Zero restarts and all cgroup OOM/max counters zero. Verified before merging #2289. |
This PR contains the following updates:
0.9.24→0.9.25Release Notes
defilantech/LLMKube (ghcr.io/defilantech/charts/llmkube)
v0.9.25Compare Source
Features
Bug Fixes
Documentation
Configuration
📅 Schedule: (UTC)
🚦 Automerge: Disabled by config. Please merge this manually once you are satisfied.
♻ Rebasing: Whenever PR is behind base branch, or you tick the rebase/retry checkbox.
🔕 Ignore: Close this PR and you won't be reminded about this update again.
This PR was generated by Mend Renovate. View the repository job log.