Dokumentationerreichbar
GLM-OCR
huggingface.co (externe Seite)
Modellkarte des Erkennungsmodells mit 0,9 Milliarden Parametern, das Druck und Tabellen liest. Der Anbieter gibt 94,62 Punkte auf OmniDocBench in der Fassung V1.5 an und nennt Formelerkennung, Tabellenerkennung und Informationsextraktion als Stärken. Die Karte sagt nebenbei, dass die vollständige Verarbeitung dieses Anbieters selbst zweistufig aufgebaut ist und dafür PP-DocLayoutV3 einsetzt.
geprüft 24.09.2026
Worauf sich diese Seite beruft, wörtlich, abgerufen am 07.09.2026:
GLM-OCR is a multimodal OCR model for complex document understanding
bestätigt 24.09.2026Achieves a score of 94.62 on OmniDocBench V1.5, ranking #1 overall
bestätigt 24.09.2026delivers state-of-the-art results across major document understanding benchmarks, including formula recognition, table recognition, and information extraction
bestätigt 24.09.2026With only 0.9B parameters, GLM-OCR supports deployment via vLLM, SGLang, and Ollama, significantly reducing inference latency and compute cost
bestätigt 24.09.2026The complete OCR pipeline integrates PP-DocLayoutV3 for document layout analysis, which is licensed under the Apache License 2.0.
bestätigt 24.09.2026
Wie eine Maschine ein Dokument zerlegtWo in der eigenen Texterkennung das Modell sitzt