- Added GLM-5.2 FP8, NVIDIA Nemotron‑Nano‑12B‑v2, and GLM‑OCR models to SageMaker JumpStart for one‑click deployment.
- Highlights: GLM‑5.2 FP8 offers a 1 M‑token context for long‑horizon engineering; Nemotron‑Nano‑12B‑v2 provides 128K context with hybrid Mamba‑2/Transformer architecture and up to 6× inference throughput; GLM‑OCR delivers fast multimodal ...
- Users can launch any of these models via the SageMaker console or Python SDK to power AI use‑cases such as agentic workflows, high‑throughput reasoning, and real‑time document processing.