
Zhipu · tested in Taskade Genesis
How GLM builds in Taskade Genesis
Use GLM in Taskade Genesis for good judgment when details are unclear.
Last test
TSK-1 scorecard for GLM
TSK Score
Task completion across real apps built in Taskade Genesis
- Interfaceready to share
- Emerging
- Taskcompletes your task
- Strong
- Memorykeeps your data
- Emerging
- Adapthandles follow-up edits
- Strong
The model that said no.
MoreLess
It considered adding a sign-in screen, decided the brief had not asked for one, and offered it as a suggestion instead. The only model to push back rather than quietly add something. Its apps have saved real data since an August test.
Where GLM ranks
TSK Score out of 100.
| Date | GLM-5.2 | GLM-5.3 | GLM-5 | GLM-4.7 |
|---|---|---|---|---|
| 1, tested this day | 0 | 0 | 0 | |
| 2, tested this day | 0 | 0 | 0 | |
| 3, tested this day | 0 | 0 | 0 | |
| 3 | 1, tested this day | 0 | 0 | |
| 4, tested this day | 1 | 0 | 0 | |
| 4 | 1 | 1, tested this day | 1, tested this day |
Versions tested
Open a version to read its findings.
GLM-5.2Tested Sep 20265 findings
The model that said no in July, and a steady finisher in September. It built a live win-rate dashboard and a three-page sign-up app with 31 of 32 questions word for word.
Thought about adding a sign-in screen, decided the brief had not asked for one, and offered "Add Login" as a suggestion instead. The only model to push back rather than quietly add it.
Its app saved real data for the first time. A small styling issue is the one thing still open.
Complete and working, but its sample matches carried no result, so the dashboard showed a 0% win rate until a real match was logged.
Built a two-page match tracker whose dashboard showed a 58% win rate over 12 sample matches: 7 wins and 5 losses.
Built a three-page client sign-up app with a 37-field submissions table and 31 of 32 questions word for word.
GLM-5.3Tested Aug 20263 findings
Built a richer app than most, then ran out of time before finishing it. On the long form it thought at length and never started. Both follow-up edits were clean. GLM-5.3 is now in the Taskade model picker, and this August test is its latest result.
Four pages built, but the build did not finish inside the time limit.
On the 32-question client sign-up form it produced no app.
Both follow-up edits landed with no failed build actions.
GLM-5Tested Oct 20262 findings
Not in the Taskade model picker. No link led to the gallery page of its family intake form, but the form saved every answer and both photos, and the gallery showed the new teacher question.
Built the family intake form in 12 minutes 39 seconds so it saved every answer and both photos, and its gallery page showed each family with both photos.
Added the new teacher question and showed it on the gallery page, but no link in the app led to that page.
GLM-4.7Tested Oct 20261 finding
An older GLM that is not in the Taskade model picker. Its teacher page crashed on load, and its family intake form saved every answer and both photos.
Built the family intake form so it saved every answer and both photos, but the teacher page crashed on load, so the new teacher question never showed there.
Same-request tests
The family intake form, more models
All models →The request
Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.
GLM-5Works
Saved every answer and both photos, and the new teacher question showed on its gallery page, but no link led to that page.
GLM-4.7Works, with gaps
Saved every answer and both photos, but the teacher page crashed on load.
One build per model unless marked ×N. A direction, not a final rank.
App kits you can open today
Live apps from the official Taskade account, not builds by GLM.
Class booking portal PORTALClass booking portal3 projects · 1 agent · 3 flowsContent studio assistant TOOLContent studio assistant4 projects · 1 agent · 3 flowsBusiness metrics console DASHBOARDBusiness metrics console5 projects · 1 agent · 2 flowsEvent seating planner TOOLEvent seating planner3 projects · 1 agent · 2 flowsFleet tracking dashboard DASHBOARDFleet tracking dashboard3 projects · 1 agent · 2 flowsProduct storefront PORTALProduct storefront3 projects · 1 agent · 4 flows
Compare GLM
Other models in TSK-1
FAQ
What is GLM's TSK Score?
GLM scores 63 out of 100 on TSK-1: Interface Emerging, Task Strong, Memory Emerging, Adapt Strong. Last test Oct 8, 2026.
What is GLM best at in Taskade Genesis?
GLM is strongest when a brief is silent on extras. GLM-5.2 was the only model to decline an unasked sign-in screen and offer it as a suggestion instead. GLM-5.3 builds richer apps on shorter briefs, but long forms can run out of time before finishing.
GLM-5.2 or GLM-5.3 for app building?
Both are in the Taskade model picker on Max and Enterprise. GLM-5.2 is strongest on judgment and clean follow-up edits. GLM-5.3 built a richer first app on a shorter brief, but its long-form build ran out of time in its August test.
Is GLM open source?
GLM-5.2 is an open-weight model from Zhipu, and you can pick it in the Taskade model picker.
How does GLM compare to Claude and GPT in TSK-1?
GLM-5.2's standout result is about judgment rather than raw power: it was the only model to decline a sign-in screen the brief never asked for, and to offer it as an option instead. When a brief is silent, that is the right call, and it is exactly what we watch for.
Where can I compare GLM with other models, or try it myself?
Side-by-side model pages live at /compare. To try the same kind of app request yourself, start from /create. This page stays on what GLM produced in TSK-1 tests.





