
DeepSeek · tested in Taskade Genesis
How DeepSeek builds in Taskade Genesis
Use DeepSeek in Taskade Genesis for polished apps with rich workspace data.
Last test
TSK-1 scorecard for DeepSeek
TSK Score
Task completion across real apps built in Taskade Genesis
- Interfaceready to share
- Leading
- Taskcompletes your task
- Strong
- Memorykeeps your data
- Strong
- Adapthandles follow-up edits
- Emerging
DeepSeek combines rich apps with efficient builds.
MoreLess
V4.1 Flash built the richest sales CRM and cookbook in their tests. Pro builds deep data but can leave the app locked or unfinished. In October both built a family intake form that saved every answer and both photos.
Where DeepSeek ranks
TSK Score out of 100.
- DeepSeek V4 Flash4
- DeepSeek V4.1 Flash4
- DeepSeek V4 Pro7
- DeepSeek V4 Pro, max thinking1
- DeepSeek V4 Pro, high thinking1
- DeepSeek V4 Pro 08131
- DeepSeek V3.21
| Date | DeepSeek V4 Flash | DeepSeek V4.1 Flash | DeepSeek V4 Pro | DeepSeek V4 Pro, max thinking | DeepSeek V4 Pro, high thinking | DeepSeek V4 Pro 0813 | DeepSeek V3.2 |
|---|---|---|---|---|---|---|---|
| 1, tested this day | 0 | 0 | 0 | 0 | 0 | 0 | |
| 2, tested this day | 0 | 1, tested this day | 0 | 0 | 0 | 0 | |
| 3, tested this day | 0 | 1 | 0 | 0 | 0 | 0 | |
| 3 | 0 | 2, tested this day | 0 | 0 | 0 | 0 | |
| 3 | 0 | 3, tested this day | 0 | 0 | 0 | 0 | |
| 4, tested this day | 0 | 4, tested this day | 0 | 0 | 0 | 0 | |
| 4 | 1, tested this day | 5, tested this day | 0 | 0 | 0 | 0 | |
| 4 | 2, tested this day | 6, tested this day | 0 | 0 | 0 | 0 | |
| 4 | 3, tested this day | 6 | 0 | 0 | 0 | 0 | |
| 4 | 4, tested this day | 7, tested this day | 1, tested this day | 1, tested this day | 1, tested this day | 1, tested this day |
Versions tested
Open a version to read its findings.
DeepSeek V4 FlashTested Aug 20265 findings
The cheapest model won the design test. It was the only one that looked right in both light and dark, ran without a single error, and laid out cleanly on a phone, for a fraction of what the others cost. On a real customer brief it was cheapest, fastest and cleanest at once. DeepSeek V4.1 Flash has since replaced it in the Taskade model picker.
Won on design: the only model that looked right in both light and dark, ran without a single error, and laid out cleanly on a 390px phone screen, at a fraction of the field's cost.
Won on design again, and this time the finished workspace saved real data with light and dark themes both working throughout.
Cheapest, fastest and cleanest run on the 32-question client sign-up form: 12 minutes 36 seconds.
Same test, same prompt: better value than Luna, at half the cost of the build.
Finished the tracker and made both follow-up changes cleanly, but its long-form build ran out of time.
DeepSeek V4.1 FlashTested Sep 202612 findings
The richest app per shape among the economical models. It built a six-page sales CRM whose scoring automation ran, a four-page match tracker, and a four-page cookbook with both automations live. On a family intake form it saved every answer and both photos.
Built the richest sales CRM of the test, with six pages and a scoring automation that scored all 7 sample leads.
Built a four-page match tracker and said plainly that its weekly recap was still switched off.
Typed all 32 client sign-up questions and built a 48-field submissions table, then ran out of time before finishing.
Built a three-page fleet inspection app for 42 vans where the brief named 40.
Its recipe box came out as a printed cookbook with 13 recipes, four pages, and both automations live.
Built a three-page match tracker whose coaching automation ran on the first match logged and wrote a note back into that match about six minutes later.
Built a two-page family intake form that saved every answer and both photos, and its alert automation ran twice. The build ran past the 15-minute test limit.
Added the new teacher question and put the change live.
Built the family intake form in 7 minutes 25 seconds so it saved every answer and both photos, and its new-family automation ran. Three sample families showed on the teacher page as if they were real.
Added the new teacher question as an optional field and kept it off the printed facesheet, and said so.
Built a five-page Dota 2 match tracker with labeled sample matches, but every hero showed as unknown and no match could be saved, while its summary said the save had been checked.
Rebuilt the 32-question client sign-up form with all 32 questions word for word and a scoring rubric, but the form saved each answer one question off, so every submission was rejected.
DeepSeek V4 ProTested Oct 202611 findings
The workspace champion with a delivery problem. Pro still builds deep data. In September it put two apps behind a sign-in screen and left a third without a page. In October it built a family intake form that saved every answer and both photos, with no sign-in.
The richest workspace of the test: 8 automations, a record 60 fields, and all four builds word for word.
Won the tracker test as the most efficient app that opened. Two stronger-looking results could not be used, so they did not receive a score.
Went from none of its four builds matching the brief to all four word for word on a shorter prompt. The biggest jump we have measured in a single test.
The cheapest complete match tracker of the test.
Built the deepest lead table of the sales CRM test, with 12 fields per lead, then put its one-page board behind a sign-in screen.
Finished the client sign-up form with all 32 questions word for word and left both automations switched off.
Spent the fleet-inspection test building data and automations and wrote no page.
Built a complete match-tracker data layer, then put the app behind a sign-in screen.
Built one of the two best-looking match trackers of its test, with a win-rate trend chart in light and dark, in 36 minutes, but its hero picker listed only 17 heroes.
Built the family intake form in 7 minutes 24 seconds so it saved every answer and both photos, and the new teacher question landed. Two sample families showed as real, and its summary did not mention them.
Built a Dota 2 match tracker with a win-rate trend that saved a logged match and added the heroes page when asked. About 15 sample matches showed as real.
DeepSeek V4 Pro, max thinkingTested Oct 20263 findings
DeepSeek V4 Pro with its deepest thinking setting, which the Taskade model picker no longer offers. The 15-minute test limit ended its build before a closing summary, and the follow-up turn finished an intake form that saved every answer and both photos with no fake sample families.
The 15-minute test limit ended the family intake build before a closing summary. The follow-up turn finished the app and added the new teacher question.
The finished form saved every answer and both photos, and the teacher page started empty instead of showing made-up families.
Built the best-looking Dota 2 match tracker of its set, and a logged match saved. The 15-minute limit ended the build, and the follow-up turn finished the interface and the heroes page.
DeepSeek V4 Pro, high thinkingTested Oct 20262 findings
DeepSeek V4 Pro with a high thinking setting, which the Taskade model picker no longer offers. Its family intake form could never keep a photo, so no parent could submit it. Its match tracker labeled its demo matches and kept them out of the stats.
The 15-minute test limit ended the family intake build before a closing summary, and the form dropped every photo a parent picked, so the two-photo rule blocked every submission.
Built a Dota 2 match tracker that saved a logged match, labeled its demo matches and kept them out of the stats, and added the heroes page when asked.
DeepSeek V4 Pro 0813Tested Oct 20263 findings
A dated DeepSeek V4 Pro release that is not in the Taskade model picker. With thinking off and with max thinking, both of its family intake forms saved every answer and both photos, and the new question landed each time.
With thinking off, it built the family intake form in 6 minutes 14 seconds so it saved every answer and both photos, and it removed its own test row when it finished.
With max thinking, the 15-minute test limit ended the build, and the follow-up turn finished a form that saved every answer and both photos, with every sample family labeled.
Built the Dota 2 match tracker with both settings, and both saved a logged match. With max thinking it labeled its sample matches. With thinking off it had already built a ranked heroes tab, so the follow-up change made no edit and said so.
DeepSeek V3.2Tested Oct 20262 findings
An older DeepSeek that is not in the Taskade model picker. With thinking off and on, neither of its family intake forms could save a family, though one uploaded both photos. V4.1 Flash, the DeepSeek in the picker today, saved every answer and photo on the same form.
With thinking off, the form uploaded both photos, but the save was rejected, and the teacher page showed its sample families with blank names.
With thinking on, the form sent each submission to an address that does not exist, and the teacher page showed families typed into the page instead of saved ones.
Same-request tests
The family intake form, more models
All models →The request
Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.
DeepSeek V4 Pro 0813×2 buildsWorks
Both settings saved every answer and both photos, and the new teacher question landed. With max thinking, the follow-up turn finished the build.
DeepSeek V4 Pro, max thinkingWorks
The 15-minute limit ended the build. The follow-up turn finished a form that saved every answer and both photos, with no made-up families.
DeepSeek V4.1 FlashWorks
Saved every answer and both photos, and its new-family automation ran. Three sample families showed as real.
DeepSeek V4 ProWorks
Saved every answer and both photos, and the new teacher question landed. Two sample families showed as real.
DeepSeek V4 Pro, high thinkingWorks, with gaps
The form dropped every photo a parent picked, so the two-photo rule blocked every submission.
DeepSeek V3.2×2 buildsWorks, with gaps
Neither setting saved a family: one save was rejected, and the other sent the form to an address that does not exist.
A Dota 2 match tracker with a heroes page
All models →The request
One player logs each match with hero, result, kills, deaths, assists, duration and notes, a dashboard shows win rate and streaks in a premium esports look, and then we ask for a heroes page.
DeepSeek V4 Pro, max thinkingWorks
The best-looking tracker of its set, and a logged match saved. The follow-up turn finished the interface and the heroes page.
DeepSeek V4 Pro, high thinkingWorks
Labeled its demo matches and kept them out of the stats. A logged match saved, and the heroes page landed.
DeepSeek V4 Pro 0813×2 buildsWorks
Both settings saved a logged match and ranked heroes by games. One labeled its sample matches, and one showed them as real.
DeepSeek V4 ProWorks
A win-rate trend, a logged match saved, and the heroes page landed. About 15 sample matches showed as real.
DeepSeek V4.1 FlashWorks, with gaps
Five pages with labeled sample matches, but every hero showed as unknown and no match could be saved.
A family intake form on a phone
All models →The request
Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.
DeepSeek V4.1 FlashWorks, with gaps
Saved every answer and both photos on a two-page form, and the follow-up change landed. The build ran past the 15-minute limit.
A match tracker
All models →The request
One player logs each game with hero, result and notes, and a dashboard shows win rate and streaks.
DeepSeek V4.1 FlashWorks
Three pages, and its coaching automation wrote a note into the first match logged.
DeepSeek V4 ProWorks, with gaps
One of the two best-looking trackers of the test, but its hero picker listed only 17 heroes.
One build per model unless marked ×N. A direction, not a final rank.
App kits you can open today
Live apps from the official Taskade account, not builds by DeepSeek.
Store command center OPSStore command center34 projects · 1 agent · 2 flowsProduct launch dashboard DASHBOARDProduct launch dashboard1 project · 1 agent · 1 flowOrder and stock tracker TRACKEROrder and stock tracker2 projects · 1 agent · 3 flowsHotel operations dashboard DASHBOARDHotel operations dashboard4 projects · 1 agent · 3 flowsNew-hire onboarding hub PORTALNew-hire onboarding hub4 projects · 1 agent · 3 flowsBillable time tracker TRACKERBillable time tracker3 projects · 1 agent · 3 flows
Compare DeepSeek
See all 8 comparisonsShow fewer
Other models in TSK-1
FAQ
What is DeepSeek's TSK Score?
DeepSeek scores 75 out of 100 on TSK-1: Interface Leading, Task Strong, Memory Strong, Adapt Emerging. Last test Oct 8, 2026.
What is DeepSeek best at in Taskade Genesis?
DeepSeek is strongest on workspace depth. V4 Flash won an August design test. V4.1 Flash built a six-page sales CRM with a scoring automation that ran. Pro still builds deep data, but its September apps were less reliable to reach.
Which DeepSeek model is best for app building?
Choose V4.1 Flash, the Flash in the Taskade model picker today, for a rich app at a low cost: it built a six-page sales CRM and saved every answer and photo on a family intake form. Choose Pro when data depth matters most, then check that every page and automation is available. V4 Flash, which won an August design test, has left the picker.
Are DeepSeek models open source?
DeepSeek publishes open-weight models, which means the underlying weights are available rather than kept entirely private. You can pick DeepSeek models in the Taskade model picker.
Where can I compare DeepSeek with other models, or try it myself?
Side-by-side model pages live at /compare. To try the same kind of app request yourself, start from /create. This page stays on what DeepSeek produced in TSK-1 tests.




