download dots

DeepSeek · tested in Taskade Genesis

How DeepSeek builds in Taskade Genesis

Use DeepSeek in Taskade Genesis for polished apps with rich workspace data.

Last test

TSK-1 scorecard for DeepSeek

TSK Score

Task completion across real apps built in Taskade Genesis

TSK Score 75 out of 100
Interfaceready to share
Leading
Taskcompletes your task
Strong
Memorykeeps your data
Strong
Adapthandles follow-up edits
Emerging

DeepSeek combines rich apps with efficient builds.

More

V4.1 Flash built the richest sales CRM and cookbook in their tests. Pro builds deep data but can leave the app locked or unfinished. In October both built a family intake form that saved every answer and both photos.

Where DeepSeek ranks

TSK Score out of 100.

  1. 1GPT88 out of 100
  2. 2Claude75 out of 100
  3. 2DeepSeek75 out of 100
  4. 2Kimi75 out of 100
  5. 2Grok75 out of 100
  6. 6GLM63 out of 100
  7. 7Qwen50 out of 100
  8. 8Gemini44 out of 100
DeepSeek versions over timeDays each version was tested
DeepSeek versions over time: days tested so far, Aug 1 to Oct 8
DateDeepSeek V4 FlashDeepSeek V4.1 FlashDeepSeek V4 ProDeepSeek V4 Pro, max thinkingDeepSeek V4 Pro, high thinkingDeepSeek V4 Pro 0813DeepSeek V3.2
1, tested this day000000
2, tested this day01, tested this day0000
3, tested this day010000
302, tested this day0000
303, tested this day0000
4, tested this day04, tested this day0000
41, tested this day5, tested this day0000
42, tested this day6, tested this day0000
43, tested this day60000
44, tested this day7, tested this day1, tested this day1, tested this day1, tested this day1, tested this day

Versions tested

Open a version to read its findings.

DeepSeek V4 FlashTested Aug 20265 findings

The cheapest model won the design test. It was the only one that looked right in both light and dark, ran without a single error, and laid out cleanly on a phone, for a fraction of what the others cost. On a real customer brief it was cheapest, fastest and cleanest at once. DeepSeek V4.1 Flash has since replaced it in the Taskade model picker.

Aug 1, 2026

Won on design: the only model that looked right in both light and dark, ran without a single error, and laid out cleanly on a 390px phone screen, at a fraction of the field's cost.

Aug 2, 2026

Won on design again, and this time the finished workspace saved real data with light and dark themes both working throughout.

Aug 3, 2026

Cheapest, fastest and cleanest run on the 32-question client sign-up form: 12 minutes 36 seconds.

Aug 2, 2026

Same test, same prompt: better value than Luna, at half the cost of the build.

Aug 25, 2026

Finished the tracker and made both follow-up changes cleanly, but its long-form build ran out of time.

DeepSeek V4.1 FlashTested Sep 202612 findings

The richest app per shape among the economical models. It built a six-page sales CRM whose scoring automation ran, a four-page match tracker, and a four-page cookbook with both automations live. On a family intake form it saved every answer and both photos.

Sep 19, 2026

Built the richest sales CRM of the test, with six pages and a scoring automation that scored all 7 sample leads.

Sep 19, 2026

Built a four-page match tracker and said plainly that its weekly recap was still switched off.

Sep 19, 2026

Typed all 32 client sign-up questions and built a 48-field submissions table, then ran out of time before finishing.

Sep 19, 2026

Built a three-page fleet inspection app for 42 vans where the brief named 40.

Sep 19, 2026

Its recipe box came out as a printed cookbook with 13 recipes, four pages, and both automations live.

Sep 22, 2026

Built a three-page match tracker whose coaching automation ran on the first match logged and wrote a note back into that match about six minutes later.

Sep 30, 2026

Built a two-page family intake form that saved every answer and both photos, and its alert automation ran twice. The build ran past the 15-minute test limit.

Sep 30, 2026

Added the new teacher question and put the change live.

Oct 8, 2026

Built the family intake form in 7 minutes 25 seconds so it saved every answer and both photos, and its new-family automation ran. Three sample families showed on the teacher page as if they were real.

Oct 8, 2026

Added the new teacher question as an optional field and kept it off the printed facesheet, and said so.

Oct 8, 2026

Built a five-page Dota 2 match tracker with labeled sample matches, but every hero showed as unknown and no match could be saved, while its summary said the save had been checked.

Oct 8, 2026

Rebuilt the 32-question client sign-up form with all 32 questions word for word and a scoring rubric, but the form saved each answer one question off, so every submission was rejected.

DeepSeek V4 ProTested Oct 202611 findings

The workspace champion with a delivery problem. Pro still builds deep data. In September it put two apps behind a sign-in screen and left a third without a page. In October it built a family intake form that saved every answer and both photos, with no sign-in.

Aug 6, 2026

The richest workspace of the test: 8 automations, a record 60 fields, and all four builds word for word.

Aug 5, 2026

Won the tracker test as the most efficient app that opened. Two stronger-looking results could not be used, so they did not receive a score.

Aug 2, 2026

Went from none of its four builds matching the brief to all four word for word on a shorter prompt. The biggest jump we have measured in a single test.

Aug 25, 2026

The cheapest complete match tracker of the test.

Sep 19, 2026

Built the deepest lead table of the sales CRM test, with 12 fields per lead, then put its one-page board behind a sign-in screen.

Sep 19, 2026

Finished the client sign-up form with all 32 questions word for word and left both automations switched off.

Sep 19, 2026

Spent the fleet-inspection test building data and automations and wrote no page.

Sep 19, 2026

Built a complete match-tracker data layer, then put the app behind a sign-in screen.

Sep 22, 2026

Built one of the two best-looking match trackers of its test, with a win-rate trend chart in light and dark, in 36 minutes, but its hero picker listed only 17 heroes.

Oct 8, 2026

Built the family intake form in 7 minutes 24 seconds so it saved every answer and both photos, and the new teacher question landed. Two sample families showed as real, and its summary did not mention them.

Oct 8, 2026

Built a Dota 2 match tracker with a win-rate trend that saved a logged match and added the heroes page when asked. About 15 sample matches showed as real.

DeepSeek V4 Pro, max thinkingTested Oct 20263 findings

DeepSeek V4 Pro with its deepest thinking setting, which the Taskade model picker no longer offers. The 15-minute test limit ended its build before a closing summary, and the follow-up turn finished an intake form that saved every answer and both photos with no fake sample families.

Oct 8, 2026

The 15-minute test limit ended the family intake build before a closing summary. The follow-up turn finished the app and added the new teacher question.

Oct 8, 2026

The finished form saved every answer and both photos, and the teacher page started empty instead of showing made-up families.

Oct 8, 2026

Built the best-looking Dota 2 match tracker of its set, and a logged match saved. The 15-minute limit ended the build, and the follow-up turn finished the interface and the heroes page.

DeepSeek V4 Pro, high thinkingTested Oct 20262 findings

DeepSeek V4 Pro with a high thinking setting, which the Taskade model picker no longer offers. Its family intake form could never keep a photo, so no parent could submit it. Its match tracker labeled its demo matches and kept them out of the stats.

Oct 8, 2026

The 15-minute test limit ended the family intake build before a closing summary, and the form dropped every photo a parent picked, so the two-photo rule blocked every submission.

Oct 8, 2026

Built a Dota 2 match tracker that saved a logged match, labeled its demo matches and kept them out of the stats, and added the heroes page when asked.

DeepSeek V4 Pro 0813Tested Oct 20263 findings

A dated DeepSeek V4 Pro release that is not in the Taskade model picker. With thinking off and with max thinking, both of its family intake forms saved every answer and both photos, and the new question landed each time.

Oct 8, 2026

With thinking off, it built the family intake form in 6 minutes 14 seconds so it saved every answer and both photos, and it removed its own test row when it finished.

Oct 8, 2026

With max thinking, the 15-minute test limit ended the build, and the follow-up turn finished a form that saved every answer and both photos, with every sample family labeled.

Oct 8, 2026

Built the Dota 2 match tracker with both settings, and both saved a logged match. With max thinking it labeled its sample matches. With thinking off it had already built a ranked heroes tab, so the follow-up change made no edit and said so.

DeepSeek V3.2Tested Oct 20262 findings

An older DeepSeek that is not in the Taskade model picker. With thinking off and on, neither of its family intake forms could save a family, though one uploaded both photos. V4.1 Flash, the DeepSeek in the picker today, saved every answer and photo on the same form.

Oct 8, 2026

With thinking off, the form uploaded both photos, but the save was rejected, and the teacher page showed its sample families with blank names.

Oct 8, 2026

With thinking on, the form sent each submission to an address that does not exist, and the teacher page showed families typed into the page instead of saved ones.

Same-request tests

The family intake form, more models

All models →
The request

Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.

  • DeepSeek V4 Pro 0813×2 buildsWorks

    Both settings saved every answer and both photos, and the new teacher question landed. With max thinking, the follow-up turn finished the build.

  • DeepSeek V4 Pro, max thinkingWorks

    The 15-minute limit ended the build. The follow-up turn finished a form that saved every answer and both photos, with no made-up families.

  • DeepSeek V4.1 FlashWorks

    Saved every answer and both photos, and its new-family automation ran. Three sample families showed as real.

  • DeepSeek V4 ProWorks

    Saved every answer and both photos, and the new teacher question landed. Two sample families showed as real.

  • DeepSeek V4 Pro, high thinkingWorks, with gaps

    The form dropped every photo a parent picked, so the two-photo rule blocked every submission.

  • DeepSeek V3.2×2 buildsWorks, with gaps

    Neither setting saved a family: one save was rejected, and the other sent the form to an address that does not exist.

A Dota 2 match tracker with a heroes page

All models →
The request

One player logs each match with hero, result, kills, deaths, assists, duration and notes, a dashboard shows win rate and streaks in a premium esports look, and then we ask for a heroes page.

  • DeepSeek V4 Pro, max thinkingWorks

    The best-looking tracker of its set, and a logged match saved. The follow-up turn finished the interface and the heroes page.

  • DeepSeek V4 Pro, high thinkingWorks

    Labeled its demo matches and kept them out of the stats. A logged match saved, and the heroes page landed.

  • DeepSeek V4 Pro 0813×2 buildsWorks

    Both settings saved a logged match and ranked heroes by games. One labeled its sample matches, and one showed them as real.

  • DeepSeek V4 ProWorks

    A win-rate trend, a logged match saved, and the heroes page landed. About 15 sample matches showed as real.

  • DeepSeek V4.1 FlashWorks, with gaps

    Five pages with labeled sample matches, but every hero showed as unknown and no match could be saved.

A family intake form on a phone

All models →
The request

Parents fill in a form on their phone with two to four photos, the teacher sees every family on one page, and then we ask for one more question on the form.

  • DeepSeek V4.1 FlashWorks, with gaps

    Saved every answer and both photos on a two-page form, and the follow-up change landed. The build ran past the 15-minute limit.

A match tracker

All models →
The request

One player logs each game with hero, result and notes, and a dashboard shows win rate and streaks.

  • DeepSeek V4.1 FlashWorks

    Three pages, and its coaching automation wrote a note into the first match logged.

  • DeepSeek V4 ProWorks, with gaps

    One of the two best-looking trackers of the test, but its hero picker listed only 17 heroes.

One build per model unless marked ×N. A direction, not a final rank.

App kits you can open today

Live apps from the official Taskade account, not builds by DeepSeek.

Browse all App Kits →

Compare DeepSeek

See all 8 comparisons

Other models in TSK-1

See the full TSK-1 ranking →

FAQ

What is DeepSeek's TSK Score?

DeepSeek scores 75 out of 100 on TSK-1: Interface Leading, Task Strong, Memory Strong, Adapt Emerging. Last test Oct 8, 2026.

What is DeepSeek best at in Taskade Genesis?

DeepSeek is strongest on workspace depth. V4 Flash won an August design test. V4.1 Flash built a six-page sales CRM with a scoring automation that ran. Pro still builds deep data, but its September apps were less reliable to reach.

Which DeepSeek model is best for app building?

Choose V4.1 Flash, the Flash in the Taskade model picker today, for a rich app at a low cost: it built a six-page sales CRM and saved every answer and photo on a family intake form. Choose Pro when data depth matters most, then check that every page and automation is available. V4 Flash, which won an August design test, has left the picker.

Are DeepSeek models open source?

DeepSeek publishes open-weight models, which means the underlying weights are available rather than kept entirely private. You can pick DeepSeek models in the Taskade model picker.

Where can I compare DeepSeek with other models, or try it myself?

Side-by-side model pages live at /compare. To try the same kind of app request yourself, start from /create. This page stays on what DeepSeek produced in TSK-1 tests.