Skip to content

  • Projects
  • Groups
  • Snippets
  • Help
    • Loading...
    • Help
    • Contribute to GitLab
  • Sign in
C
coderai
  • Project
    • Project
    • Details
    • Activity
    • Cycle Analytics
  • Repository
    • Repository
    • Files
    • Commits
    • Branches
    • Tags
    • Contributors
    • Graph
    • Compare
    • Charts
  • Issues 0
    • Issues 0
    • List
    • Board
    • Labels
    • Milestones
  • Merge Requests 0
    • Merge Requests 0
  • CI / CD
    • CI / CD
    • Pipelines
    • Jobs
    • Schedules
    • Charts
  • Wiki
    • Wiki
  • Snippets
    • Snippets
  • Members
    • Members
  • Collapse sidebar
  • Activity
  • Graph
  • Charts
  • Create a new issue
  • Jobs
  • Commits
  • Issue Boards
  • nexlab
  • coderai
  • Repository

Switch branch/tag
  • coderai
  • codai
  • admin
  • templates
  • tasks.html
Find file
BlameHistoryPermalink
  • Stefy Lanza (nextime / spora )'s avatar
    engines card: canonical per-model memory info; cap GME vision token budget · 1920e6aa
    Stefy Lanza (nextime / spora ) authored Jul 23, 2026
    The tooltip matched loaded_info entries (raw engine keys) against
    canonicalized display names client-side and mostly missed — footprints
    now canonicalize server-side with the same mapping as the model list,
    keyed by canonical id, so every loaded model shows its VRAM (+RAM when
    offloading).
    
    llama-vl: bound mtmd image_max_tokens (config `image_max_tokens`,
    default 1024) — large photos otherwise expand to thousands of vision
    tokens whose transient compute buffer evicts co-resident models on a
    small card.
    Co-Authored-By: 's avatarClaude Opus 4.8 <noreply@anthropic.com>
    Claude-Session: https://claude.ai/code/session_014S8VtAvG499SsCbeESRK7V
    1920e6aa
tasks.html 21.7 KB
EditWeb IDE

Replace tasks.html

Attach a file by drag & drop or click to upload


Cancel
A new branch will be created in your fork and a new merge request will be started.