download dots

Zhipu · tested in Taskade Genesis

How GLM builds in Taskade Genesis

Use GLM in Taskade Genesis for good judgment when details are unclear.

Last test

TSK-1 scorecard for GLM

TSK Score

Task completion across real apps built in Taskade Genesis

TSK Score 63 out of 100
Interfaceready to share
Emerging
Taskcompletes your task
Strong
Memorykeeps your data
Emerging
Adapthandles follow-up edits
Strong

The model that said no.

More

It considered adding a sign-in screen, decided the brief had not asked for one, and offered it as a suggestion instead. The only model to push back rather than quietly add something. Its apps have saved real data since an August test.

Where GLM ranks

TSK Score out of 100.

  1. 1GPT88 out of 100
  2. 2Claude75 out of 100
  3. 2DeepSeek75 out of 100
  4. 2Kimi75 out of 100
  5. 2Grok75 out of 100
  6. 6GLM63 out of 100
  7. 7Qwen50 out of 100
  8. 8Gemini44 out of 100
GLM versions over timeDays each version was tested
GLM versions over time: days tested so far, Jul 30 to Oct 8
DateGLM-5.2GLM-5.3GLM-5GLM-4.7
1, tested this day000
2, tested this day000
3, tested this day000
31, tested this day00
4, tested this day100
411, tested this day1, tested this day

Versions tested

Open a version to read its findings.

GLM-5.2Tested Sep 20265 findings

The model that said no in July, and a steady finisher in September. It built a live win-rate dashboard and a three-page sign-up app with 31 of 32 questions word for word.

Jul 30, 2026

Thought about adding a sign-in screen, decided the brief had not asked for one, and offered "Add Login" as a suggestion instead. The only model to push back rather than quietly add it.

Aug 1, 2026

Its app saved real data for the first time. A small styling issue is the one thing still open.

Aug 25, 2026

Complete and working, but its sample matches carried no result, so the dashboard showed a 0% win rate until a real match was logged.

Sep 19, 2026

Built a two-page match tracker whose dashboard showed a 58% win rate over 12 sample matches: 7 wins and 5 losses.

Sep 19, 2026

Built a three-page client sign-up app with a 37-field submissions table and 31 of 32 questions word for word.

GLM-5.3Tested Aug 20263 findings

Built a richer app than most, then ran out of time before finishing it. On the long form it thought at length and never started. Both follow-up edits were clean. GLM-5.3 is now in the Taskade model picker, and this August test is its latest result.

Aug 29, 2026

Four pages built, but the build did not finish inside the time limit.

Aug 29, 2026

On the 32-question client sign-up form it produced no app.

Aug 29, 2026

Both follow-up edits landed with no failed build actions.

GLM-5Tested Oct 20262 findings

Not in the Taskade model picker. No link led to the gallery page of its family intake form, but the form saved every answer and both photos, and the gallery showed the new teacher question.

Oct 8, 2026

Built the family intake form in 12 minutes 39 seconds so it saved every answer and both photos, and its gallery page showed each family with both photos.

Oct 8, 2026

Added the new teacher question and showed it on the gallery page, but no link in the app led to that page.

GLM-4.7Tested Oct 20261 finding

An older GLM that is not in the Taskade model picker. Its teacher page crashed on load, and its family intake form saved every answer and both photos.

Oct 8, 2026

Built the family intake form so it saved every answer and both photos, but the teacher page crashed on load, so the new teacher question never showed there.

Same-request tests

The family intake form, more models

All models →
The request

Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.

  • GLM-5Works

    Saved every answer and both photos, and the new teacher question showed on its gallery page, but no link led to that page.

  • GLM-4.7Works, with gaps

    Saved every answer and both photos, but the teacher page crashed on load.

One build per model unless marked ×N. A direction, not a final rank.

App kits you can open today

Live apps from the official Taskade account, not builds by GLM.

Browse all App Kits →

Compare GLM

Other models in TSK-1

See the full TSK-1 ranking →

FAQ

What is GLM's TSK Score?

GLM scores 63 out of 100 on TSK-1: Interface Emerging, Task Strong, Memory Emerging, Adapt Strong. Last test Oct 8, 2026.

What is GLM best at in Taskade Genesis?

GLM is strongest when a brief is silent on extras. GLM-5.2 was the only model to decline an unasked sign-in screen and offer it as a suggestion instead. GLM-5.3 builds richer apps on shorter briefs, but long forms can run out of time before finishing.

GLM-5.2 or GLM-5.3 for app building?

Both are in the Taskade model picker on Max and Enterprise. GLM-5.2 is strongest on judgment and clean follow-up edits. GLM-5.3 built a richer first app on a shorter brief, but its long-form build ran out of time in its August test.

Is GLM open source?

GLM-5.2 is an open-weight model from Zhipu, and you can pick it in the Taskade model picker.

How does GLM compare to Claude and GPT in TSK-1?

GLM-5.2's standout result is about judgment rather than raw power: it was the only model to decline a sign-in screen the brief never asked for, and to offer it as an option instead. When a brief is silent, that is the right call, and it is exactly what we watch for.

Where can I compare GLM with other models, or try it myself?

Side-by-side model pages live at /compare. To try the same kind of app request yourself, start from /create. This page stays on what GLM produced in TSK-1 tests.