glm-bfb6f02e·3 events·first seen Aliases: GLM
A new arXiv paper evaluates GPT, Claude Opus, Gemini, and GLM on automated grading of 1,200 real student Linux/bash command responses, benchmarked against three expert instructors. Using a four-level cognitive taxonomy, Gemini 3.0 Pro with rubric-guided prompting achieved the highest human-AI agreement (ICC=0.888, MAE=0.10). Key findings: rubric quality mattered more than model choice, and grading accuracy declined consistently at higher cognitive complexity levels. The study proposes a taxonomy-based framework for deciding which exam questions are suitable for AI-assisted grading.
Zhipu AI, the organization behind the GLM model family, has released ZCode, a coding agent tool positioned as analogous to Anthropic's Claude Code. The announcement is generating notable community discussion on Hacker News with 260 points and 140 comments. This represents another entrant in the agentic coding assistant space, this time from a prominent Chinese AI lab.
GLM-OCR is an open-source OCR project from zai-org built on the GLM model family, positioning itself as accurate, fast, and comprehensive. The repository has accumulated 6,787 GitHub stars with 82 added today, indicating notable community traction. It represents an application of large language/vision models to document understanding and text recognition tasks.