Skip to content

feat(gallery): add GLM-5.3 Flash variants - #11785

Open
localai-org-maint-bot wants to merge 1 commit into
masterfrom
cron/model-gallery-20260830
Open

feat(gallery): add GLM-5.3 Flash variants#11785
localai-org-maint-bot wants to merge 1 commit into
masterfrom
cron/model-gallery-20260830

Conversation

@localai-org-maint-bot

Copy link
Copy Markdown
Collaborator

Description

Add GLM-5.3-Flash to the model gallery with UD-Q4_K_XL and Q8_0 llama.cpp builds. The entries include the BF16 multimodal projector, MTP speculative-decoding configuration, and variant selection.

The model is MIT-licensed and currently leads Hugging Face text-generation trending. All GGUF and projector checksums were verified twice against Hugging Face x-linked-etag values.

Notes for Reviewers

Verification:

  • git diff --check
  • go test ./core/gallery -ginkgo.focus="gallery/index.yaml|keeps every entry that declares variants installable"
  • live SHA-256 comparison for all 16 file references

The full go test ./core/gallery run reached 383 passing specs; three unrelated network specs failed because the runner received HTTP 403 from GitHub raw/OCI endpoints.

Signed commits

  • Yes, I signed my commits. A human maintainer must provide DCO attestation.
  • Documentation updated, or not applicable. This change only adds gallery data.

Add verified Q4 and Q8 GGUF builds for the new multimodal GLM-5.3 Flash model. Group the quantizations so LocalAI can select the highest-quality build that fits.

Assisted-by: Codex:gpt-5
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant