@alvarobartt published a step-by-step guide to deploy zai-org/GLM-5.2 on AMD GPUs, using newly released features in Microsoft Foundry. The FP8 model fits on a single node of MI300X GPUs, which cuts the bill in half vs. H100.
Released lafzyn , built over Qwen, an Urdu language model that converts Urdu text into IPA phonetic transcription, with GGUF builds for local inference.
Release contents: - mahwizzzz/lafzyn: full weights - mahwizzzz/lafzyn-gguf: quantized builds
@retrain-pipelines v0.2.0 is out ! I'm at Station F at My booth with GOSIM Paris 2026 today & tomorrow. Come meet me for a live in-person demo and a chat !