Skip to main content
Introducing TSK-1Introducing TSK-1—Taskade's intelligence layer.
taskade
PricingHelpDashboard →Dashboard →
PricingLoginSign up for free →Sign up for free →
Dashboard →Dashboard →
Sign up →Sign up →
Loved by 1M+ users·Hosting 100K+ apps·Deploying 500K+ AI agents·Running 1M+ automations·Backed by Y Combinator·Powered by TSK-1
TaskadePricingFeaturesContact usIntegrationsMCP ServerDeveloper APIChangelogPressLearnAbout
ConnectProductivityKitsVideosReviewsFAQ
VibeVibe AppsVibe AgentsVibe CodingVibe WorkflowsVibe Marketing
Vibe DashboardsVibe CRMVibe AutomationVibe PaymentsVibe DesignVibe SEOVibe Tracking
Community
FeaturedQuick AppsToolsDashboardsWebsites
WorkflowsProjectsFormsCreators
DownloadsAndroidiOSMacWindows
ChromeFirefoxEdge
Compare
vs Cursorvs Boltvs Lovablevs V0vs Windsurf
vs Replitvs Emergentvs Devinvs Claude Codevs ChatGPTvs Claudevs Perplexityvs GitHub Copilotvs Figma AIvs Notionvs ClickUpvs Asanavs Mondayvs Trellovs Jiravs Linearvs Todoistvs Evernotevs Obsidianvs Airtablevs Basecampvs Mirovs Slackvs Bubblevs Retoolvs Webflowvs Framervs Softrvs Glidevs FlutterFlowvs Base44vs Adalovs Durablevs Gammavs Squarespacevs WordPressvs UI Bakeryvs Zapiervs Makevs n8nvs Jaspervs Copy.aivs Writervs Rytrvs Manusvs Crewvs Lindyvs Relevance AIvs Wrikevs Smartsheetvs Monday Magicvs Codavs TickTickvs Any.dovs Thingsvs OmniFocusvs MeisterTaskvs Teamworkvs Workfrontvs Bitrix24vs Process Streetvs Toggl Planvs Motionvs Momentumvs Habiticavs Zenkitvs Google Docsvs Google Keepvs Google Tasksvs Microsoft Teamsvs Dropbox Papervs Quipvs Roam Researchvs Logseqvs Memvs WorkFlowyvs Dynalistvs XMindvs Whimsicalvs Zoomvs Remember The Milkvs Wunderlist
Taskade GenesisVideo GuideApp BuilderVibe CodingAgent BuilderDashboard Builder
CRM BuilderWebsite BuilderForm BuilderWorkflow AutomationWorkflow BuilderBusiness-in-a-BoxAI for MarketingAI for Developers
AI Agents
FeaturedProject ManagementProductivityMarketingTranslator
ContentWorkflowResearchPersonalSalesSocial MediaTo-Do ListCRMTask AutomationCoachingCreativityTask ManagementBrandingFinanceLearning and DevelopmentBusinessCommunity ManagementMeetingsAnalyticsDigital AdvertisingContent CurationKnowledge ManagementProduct DevelopmentPublic RelationsProgrammingHuman ResourcesE-CommerceEducationLegalEmailSEODeveloperVideo ProductionDesignFlowchartDataPromptNonprofitAssistantsTeamsCustomer ServiceTrainingTravel PlanningUML DiagramER DiagramMath TutorLanguage LearningCode ReviewerLogo DesignerUI WireframeFitness CoachLead EnrichmentFounder OSSales DevelopmentBookkeepingRecruitingWebsite MonitoringAll Categories
Automations
FeaturedBusiness-in-a-BoxInvestor OperationsEducation & LearningHealthcare & Clinics
Real EstateStripeSalesE-commerceContentMarketingEmailCustomer SupportHubSpotProject ManagementAgentic WorkflowsBooking & SchedulingCalendarReportsSlackWebsiteFormTaskWeb ScrapingWeb SearchChatGPTText to ActionYoutubeLinkedInTwitterGitHubDiscordMicrosoft TeamsWebflowRSS & Content FeedsGoogle WorkspaceManufacturing & OperationsAI Agent TeamsMulti-Agent AutomationNotion AutomationsAgentic AutomationProposalBookkeeping & ExpensesClient OnboardingAll Categories
Wiki
Taskade GenesisAI AgentsAutomation
ProjectsLiving DNAAutonomous Workspaces, Agents & AppsQuantum AI & Taskade Genesis QuantumPlatformIntegrationsProductivityMethodsProject ManagementAgileScrumAI ConceptsCommunityTerminologyFeatures
Templates
FeaturedChatGPTTablePersonalProject Management
SalesFlowchartTask ManagementEngineeringEducationDesignTo-Do ListMarketingMind MapGantt ChartOrganizationalPlanningMeetingsTeam ManagementStrategyGamingProductionProduct ManagementStartupRemote WorkY CombinatorRoadmapCustomer ServiceLegalEmailBudgetsContentConsultingE-CommerceStandard Operating Procedure (SOP)Human ResourcesProgrammingMaintenanceCoachingSocial MediaHow-TosResearchMusicTrip PlanningCRMClient OnboardingEmployee OnboardingSOPBug TrackerRecruitment TrackerFormSales PipelineContent CalendarMarketing PlanProduct RoadmapBusiness PlanSWOT Analysis30-60-90 Day PlanInterviewNotion AlternativeKPIStrategic PlanMeeting AgendaInvoiceRisk RegisterIT Asset ManagementKanban BoardChange ManagementCommunication PlanRFPScope of WorkStatement of WorkHelpdeskKnowledge BaseCreative BriefGoal SettingExecutive SummaryGap AnalysisBooking SystemEvent ManagementPortfolio TrackerCustomer Onboarding PortalsClient PortalAgency OperationsFinance TrackingAll Categories
Generators
AI SoftwareNo-Code AI AppAI AppAI WebsiteAI Dashboard
AI FormAI AgentAI Client Portal BuilderAI WorkspaceAI ProductivityAI To-Do ListAI WorkflowsAI EducationAI Mind MapsAI FlowchartAI Scrum Project ManagementAI Agile Project ManagementAI MarketingAI Project ManagementAI Social Media ManagementAI BloggingAI Agency WorkflowsAI ContentAI Software DevelopmentAI MeetingAI PersonasAI OutlineAI SalesAI ProgrammingAI DesignAI FreelancingAI ResumeAI Human ResourceAI SOPAI E-CommerceAI EmailAI Public RelationsAI InfluencersAI Content CreatorsAI Customer ServiceAI BusinessAI PromptsAI Tool BuilderAI SEOAI Gantt ChartAI CalendarsAI BoardAI TableAI ResearchAI LegalAI ProposalAI Video ProductionAI Health and WellnessAI WritingAI PublishingAI NonprofitAI DataAI Event PlanningAI Game DevelopmentAI Project Management AgentAI Productivity AgentAI Marketing AgentAI Personal AgentAI Business and Work AgentAI Education and Learning AgentAI Task Management AgentAI Customer Relations AgentAI Programming AgentAI SchemaAI Business PlanAI Pitch DeckAI InvoiceAI Lesson PlanAI Social Media CalendarAI API DocumentationAI Database SchemaAI Marketing PlanAI Sales Pipeline GeneratorAI Course BuilderInternal ToolsBooking SystemReal Estate CRMInventory ManagementAI CRM BuilderAI TimesheetAI DispatchAI NewsletterAI Clinic OperationsAI Directory BuilderAll Categories
Converters
AI Featured ConvertersAI PDF ConvertersAI CSV ConvertersAI Markdown ConvertersAI Prompt to App Converters
AI Data to Dashboard ConvertersAI Workflow to App ConvertersAI Idea to App ConvertersAI Flowcharts ConvertersAI Mind Map ConvertersAI Text ConvertersAI Youtube ConvertersAI Knowledge ConvertersAI Spreadsheet ConvertersAI Email ConvertersAI Web Page ConvertersAI Video ConvertersAI Coding ConvertersAI Task ConvertersAI Kanban Board ConvertersAI Notes ConvertersAI Education ConvertersAI Language TranslatorsAI Business → Backend App ConvertersAI File → App ConvertersAI SOP → Workflow App ConvertersAI Portal → App ConvertersAI Form → App ConvertersAI Schedule → Booking App ConvertersAI Metrics → Dashboard ConvertersAI Game → Playable App ConvertersAI Catalog → Directory App ConvertersAI Creative → Studio App ConvertersAI Agent → Agent App ConvertersAI Audio ConvertersAI DOCX ConvertersAI EPUB ConvertersAI Image ConvertersAI Resume & Career ConvertersAI Presentation ConvertersAI PDF to Spreadsheet ConvertersAI PDF to Database ConvertersAI PDF to Quiz ConvertersAI Image to Notes ConvertersAI Audio to Notes ConvertersAI Email to Tasks ConvertersAI CSV to Dashboard ConvertersAI YouTube to Flashcards ConvertersURL to NotesVideo → SummaryAI Receipts to Expense Tracker ConvertersAI Docs to Knowledge Base ConvertersAI Form to Client Portal ConvertersSpreadsheet to CRMAll Categories
Prompts
Blog WritingBrandingPersonal Finance
Human ResourcesPublic RelationsTeam CollaborationProduct ManagementSupportAgencyReal EstateMarketingCodingResearchSalesAdvertisingSocial MediaCopywritingContentProject ManagementWebsite CreationDesignStrategyE-commerceEngineeringSEOEducationEmail MarketingUX/UIProductivityInfluencer MarketingAnalyticsEntrepreneurshipLegalVibe CodingCRMCustomer SupportRecruitingAll Categories
Blog
Introducing Taskade TSK-1: The System Kernel Behind Every App (2026)How to Automate 99% of Data Entry with AI Agents (Full Guide, 2026)How to Run Your Whole Day on AI Autopilot with Taskade (Full Guide, 2026)
How to Automate 99% of Reporting and Dashboards with AI (2026)How to Automate Google Workspace (Drive, Sheets, Gmail) with AI (2026)How to Use Taskade to Automate 99% of Your Small Business (Full Guide, 2026)How to Automate 99% of Your Email with AI Agents (Full Guide, 2026)How to Build Custom AI Agents Without Code (Step-by-Step, 2026)How to Use Taskade to Automate 99% of Your Busywork (Full Guide, 2026)The State of AI App Building 2026: Market Map, Funding League Table & the Operation TurnHow to Automate 99% of Your Social Media with AI Agents (2026)Why AI-Generated Apps Break: The Complete Failure Taxonomy (2026)Build an Internal App With No Engineers: A 2026 GuideAI Workflow Generator: Build Automations From a Single Prompt (2026)How to Automate 99% of Business Operations with AI Agents (2026)The 8 Best AI App Builders With Memory, Ranked (2026)Agentic Process Automation (APA): How AI Agents Run Business Processes (2026)AI Content Pipelines: Automate Your Distribution in 2026How to Automate 99% of HR and Recruiting with AI Agents (2026)Run Your Whole Business in One App with Taskade Genesis (June 2026)
AIAutomationProductivityProject ManagementRemote WorkStartupsKnowledge ManagementCollaborative WorkUpdates
Changelog
MCP Connections Do More & Signed Webhooks (Jul 21, 2026)Flexible Form Triggers & Refreshed Plans (Jul 19, 2026)Taskade Genesis Reviews Its Own Visuals (Jul 17, 2026)
Free Thinking Lane & Personal MCP Tools (Jul 16, 2026)Clearer Build Diagnostics & Smoother Updates (Jul 15, 2026)Export Any View to CSV & Regenerate AI Replies (Jul 13, 2026)Real Sign-In That Remembers Your Users (Jul 12, 2026)
Wiki
Taskade GenesisAI AgentsAutomation
ProjectsLiving DNAAutonomous Workspaces, Agents & AppsQuantum AI & Taskade Genesis QuantumPlatformIntegrationsProductivityMethodsProject ManagementAgileScrumAI ConceptsCommunityTerminologyFeatures
Prompts
Blog WritingBrandingPersonal Finance
Human ResourcesPublic RelationsTeam CollaborationProduct ManagementSupportAgencyReal EstateMarketingCodingResearchSalesAdvertisingSocial MediaCopywritingContentProject ManagementWebsite CreationDesignStrategyE-commerceEngineeringSEOEducationEmail MarketingUX/UIProductivityInfluencer MarketingAnalyticsEntrepreneurshipLegalVibe CodingCRMCustomer SupportRecruitingAll Categories
© 2026 Taskade
PrivacyTermsSecurity
Made withTaskade AIforBuilders
BlogAIOpen-Source LLM History:…

Open-Source LLM History: GPT-2 to Kimi K3 (2026)

The complete history of open-source LLMs, 2019 to 2026: from GPT-2's cautious release through BLOOM, LLaMA, Mistral, DeepSeek, Qwen, and the 2.8T open-weight Kimi K3. Open source vs open weight, the China surge, and the shrinking gap to the frontier. Updated July 2026.

A complete history and timeline of open-source and open-weight large language models, from GPT-2 in 2019 through BLOOM, LLaMA, Mistral, DeepSeek, Qwen, and the 2.8-trillion-parameter Kimi K3 in 2026
July 25, 202617 min readJohn XieAI·#open-source-ai#llm#ai-models
On this page (20)
What "Open-Source LLM" Actually Means (and Why "Open Weight" Is More Accurate)A Visual Timeline of Open-Source LLMs (2019 → 2026)2019: GPT-2 and the Cautious Birth of Open Models2021–2022: The Truly-Open Era — EleutherAI, OPT, and BLOOM2023: Llama Changes Everything (and the "Open" Debate Begins)Is Llama Really Open Source?2023–2024: Mistral and the Mixture-of-Experts Moment2024: The Field Scales Up and MaturesJanuary 2025: The DeepSeek ShockHow Much Did DeepSeek Cost to Train?Is DeepSeek Open Source?2025–2026: China's Open-Weight Surge — Qwen, Kimi, GLM, DeepSeek V4Why Are Chinese Open-Source Models Winning? The DebateOpen vs Closed: How Big Is the Gap in 2026?The Best Open-Source LLMs to Use Right NowOpen-Source LLM License Cheat-SheetHow to Use Open Models Without the Setup HeadacheWhat's Next for Open-Source LLMs🔗 Related Reading💬 Frequently Asked Questions About Open-Source LLMs

Open-source large language models went from "too dangerous to release" to larger and more capable than most people thought possible in the open — in about six years. In 2019, OpenAI held back the full weights of GPT-2, a 1.5-billion-parameter model, worried it was too risky to publish. In July 2026, a Chinese lab shipped Kimi K3, a 2.8-trillion-parameter open-weight model at frontier level — released in the open, its full weights set to follow within days for anyone to download.

How did we get from one to the other? Which models actually mattered, who built them, and what's the real difference between "open source" and "open weight" — a distinction almost every article gets wrong? This is the complete timeline. 🧬

TL;DR: Open-source LLMs evolved from GPT-2's cautious 2019 release → the truly-open era of EleutherAI and BLOOM (2021–22) → Meta's LLaMA (2023) detonating the ecosystem → Mistral bringing mixture-of-experts to open weights → DeepSeek's 2025 cost shock → a Chinese open-weight surge (Qwen, Kimi, GLM) that shrank the gap to the closed frontier to a matter of months. Most "open" models are technically open weight, not fully open source. See the 10 best open-source LLMs ranked →

What "Open-Source LLM" Actually Means (and Why "Open Weight" Is More Accurate)

Most models people call "open source" are actually open weight — a crucial difference. An open-weight model releases its trained parameters so you can download, run, and fine-tune it. A truly open-source model, by the Open Source Initiative's 2024 definition, releases the weights plus the training code, information about the data, and a license with no usage restrictions.

The Linux Foundation's Model Openness Framework — created by its Global CTO of AI, Matt White — grades this on three levels, and it's the cleanest way to think about the spectrum:

Level What's released Term people use Examples
Open Science (Class I) Weights + code + raw training data + checkpoints + paper Fully open source BLOOM, OLMo, LLM360
Open Tooling (Class II) Weights + full training/eval code, no raw data Open source (loosely) Some research releases
Open Model (Class III) Just the weights + light docs Open weight Llama, DeepSeek, Qwen, Kimi, GLM

Almost every "open" model you've heard of — Llama, DeepSeek, Qwen, Kimi — sits in that bottom row: open weight. That's still enormously valuable (you can run and adapt them), but it's not the same as fully reproducible. Matt White's own worry, after touring China's labs in 2026, was about disclosure shrinking:

"The research papers have shrunk and then become technical reports. And so there's not enough disclosure in a lot of these technical reports to be able to go and replicate the research."

— Matt White, Linux Foundation

Clone a live open-model tracker — track every open-weight release, license, and benchmark in Taskade Genesis

Don't just read the open-source timeline — clone a living dashboard that tracks every model, license, and benchmark in minutes. One click, no code, free.

A Visual Timeline of Open-Source LLMs (2019 → 2026)

Here's the whole arc at a glance. Each era changed what "open" could mean — from a cautious weights-withheld release to trillion-parameter models anyone can download.

2019GPT-2weights withheld 2021–22EleutherAI, OPTBLOOM (176B) 2023LLaMA + Llama 2ecosystem explodes 2023–24Mistral / MixtralMoE goes open Jan 2025DeepSeek R1the cost shock 2025–26Qwen, Kimi, GLMChina's open surge Jul 2026Kimi K3 (2.8T)largest open weights
2019GPT-2weights withheld 2021–22EleutherAI, OPTBLOOM (176B) 2023LLaMA + Llama 2ecosystem explodes 2023–24Mistral / MixtralMoE goes open Jan 2025DeepSeek R1the cost shock 2025–26Qwen, Kimi, GLMChina's open surge Jul 2026Kimi K3 (2.8T)largest open weights

Last updated: July 2026 — refreshed for Kimi K3 (July 16, 2026) and the July 27 open-weight drop, plus DeepSeek V4 and the latest Qwen and GLM releases. Model dates, parameter counts, and licenses move fast and vary by source; where figures are disputed we say so. Check Wikipedia and each vendor's model card before quoting exact specs.

2019: GPT-2 and the Cautious Birth of Open Models

The modern open-model story starts with a decision not to release. In February 2019, OpenAI announced GPT-2 — a 1.5-billion-parameter model — but held back the full weights, publishing a paper titled Release Strategies and the Social Impacts of Language Models and worrying openly that the model could be used to mass-produce fake text. It released the full 1.5B weights only in November 2019, after a staged rollout showed the feared harms didn't materialize at scale.

The myth-correction matters: GPT-2 was not "fully open" at launch. It was a careful, staged release — the first time a major lab treated a language model as something whose openness was a policy question, not a default. That framing has shaped every open-vs-closed debate since. (If you want the mechanics of how these models work at all, our explainer on how large language models work is the primer.)

2021–2022: The Truly-Open Era — EleutherAI, OPT, and BLOOM

The first genuinely open GPT-3-class models didn't come from a big lab — they came from a volunteer collective. EleutherAI released GPT-Neo (March 2021) and GPT-J-6B, then GPT-NeoX-20B (2022), reverse-engineering GPT-3-style capability and publishing everything. For the first time, anyone could download a large, capable model and its code.

Two 2022 releases pushed openness to its high-water mark:

  • Meta OPT-175B — Meta's attempt to open-source a GPT-3-scale model, released with training logs and a research license.
  • BLOOM — the crown jewel of the truly-open era. Built by BigScience, a collaboration of over 1,000 researchers, BLOOM was a 176-billion-parameter model trained on 46 natural languages and 13 programming languages, released with open weights, code, and full documentation. By the Model Openness Framework, BLOOM is one of the few models that reaches "Open Science" — the genuinely fully-open tier.

This era proved the community could build at scale. What it hadn't yet proved was that open models could be genuinely useful. That changed in 2023.

2023: Llama Changes Everything (and the "Open" Debate Begins)

Meta's LLaMA is the single most important release in this history, because it turned open weights into an ecosystem. LLaMA 1 (February 24, 2023) came in 7B to 65B sizes and — after the weights leaked online — detonated a wave of community fine-tunes (Alpaca, Vicuna, and hundreds more). Llama 2 (July 2023) made it official, shipping with a license that permitted commercial use.

But Llama also started the argument that still runs today.

Is Llama Really Open Source?

No, not by the strict definition. The Open Source Initiative classifies Llama as source-available or open weight, not open source, for two reasons: the Llama Community License restricts use above 700 million monthly active users, and it ships with an acceptable-use policy plus undisclosed training data. Critics call this "openwashing" — marketing a restricted release as open. Defenders point out that for 99.9% of developers, Llama was effectively open and did more to democratize LLMs than any purely-open model. Both are true, which is why the debate never fully resolves.

2023–2024: Mistral and the Mixture-of-Experts Moment

A French startup made the next leap. Mistral AI, founded by ex-Meta and DeepMind researchers Arthur Mensch, Guillaume Lample, and Timothée Lacroix, released Mistral 7B (September 2023) under the fully-permissive Apache 2.0 license — no MAU clause, no strings. Then came Mixtral 8x7B (December 2023), which brought mixture-of-experts (MoE) to open weights: instead of running the whole model for every token, MoE activates only a few specialist "experts," delivering big-model quality at small-model cost.

MoE is the architecture that made everything after it possible. Every trillion-parameter open model since — DeepSeek, Kimi, GLM — is an MoE. Mistral proved you could ship it in the open, and that Europe had a frontier lab of its own.

2024: The Field Scales Up and Matures

By 2024 the open ecosystem professionalized. Meta shipped Llama 3, 3.1 (a 405-billion-parameter flagship), 3.2 (multimodal), and 3.3. Google entered with Gemma. And Alibaba's Qwen family began its rise to become the most prolific open-weight lineage of all — dozens of models across sizes and specialties under Apache 2.0.

US / Europe pioneers China's open-weight surge GPT-2 (2019) EleutherAI (2021) BLOOM (2022) Llama 1–4 (2023–25) Mistral / Mixtral (2023–) Gemma (2024–) Qwen (2023–) DeepSeek V2→V4 (2024–26) Kimi K2→K3 (2025–26) GLM (Zhipu, 2024–26)
US / Europe pioneers China's open-weight surge GPT-2 (2019) EleutherAI (2021) BLOOM (2022) Llama 1–4 (2023–25) Mistral / Mixtral (2023–) Gemma (2024–) Qwen (2023–) DeepSeek V2→V4 (2024–26) Kimi K2→K3 (2025–26) GLM (Zhipu, 2024–26)

January 2025: The DeepSeek Shock

DeepSeek R1, released January 20, 2025, is the moment open weights went from "impressive" to "geopolitically significant." Built by a lab spun out of the Chinese hedge fund High-Flyer and led by Liang Wenfeng, R1 was a 671-billion-parameter MoE reasoning model, released under the permissive MIT license, that matched top proprietary reasoning models on math and code — and briefly rocketed DeepSeek's app to #1 on the US App Store.

How Much Did DeepSeek Cost to Train?

DeepSeek reported that the core training run cost under $6 million on roughly 2,000 Nvidia H800 GPUs over about two months. That figure covers the final run, not the full research and infrastructure bill, so it understates the true investment — Anthropic's CEO argued DeepSeek "were not substantially more resource-constrained than US AI companies." But even discounted, it reset expectations about what frontier-class training had to cost.

Is DeepSeek Open Source?

The newer generation, largely yes. DeepSeek's V4 models ship both code and weights under MIT, permitting commercial use — one of the most permissive frontier-class releases anywhere. Earlier V3 and R1 releases used MIT for code but a more restrictive model license, which drew the same "is it really open?" scrutiny Llama faced. The trajectory has been toward openness, and DeepSeek's low-cost, MIT-licensed models pushed the whole market's prices down.

2025–2026: China's Open-Weight Surge — Qwen, Kimi, GLM, DeepSeek V4

Through 2025 and into 2026, the center of gravity for open-weight releases shifted markedly toward Chinese labs, each shipping under permissive licenses:

Lab Family License Signature strength
Alibaba Qwen (Qwen3, Qwen3-Coder) Apache 2.0 Most prolific; coding + multilingual
DeepSeek V4-Pro / V4-Flash MIT Cheap reasoning + code
Moonshot AI Kimi K2 → K3 Modified MIT Agentic coding, long context
Zhipu GLM MIT Broad reasoning

The capstone landed in July 2026: Kimi K3, a 2.8-trillion-parameter open-weight model — the largest open-weight model to date — with native vision and a 1-million-token context window, released July 16 with full weights scheduled for July 27. A model that big, released in the open, would have been unthinkable when GPT-2's 1.5B weights were considered too risky to publish six years earlier. (The full Moonshot story is worth its own read: Moonshot AI & Kimi history.)

Why Are Chinese Open-Source Models Winning? The Debate

This is the most-discussed — and most-contested — question in the field, so let's be precise about what's argued versus what's settled.

The popular thesis: US GPU export controls (restricting the best chips to Chinese buyers) pushed Chinese labs toward software-efficiency breakthroughs — mixture-of-experts, latent attention, and cheaper training recipes — because "when you can't throw more compute at a problem, you find smarter solutions" (Matt White's framing). Layer on a culture of publishing and a "borrow-and-build" chain — one lab's technique (like DeepSeek's GRPO reinforcement-learning method) gets adopted and improved by the next — and you get rapid, compounding open progress.

The dissent is serious and named. Analysts including Gregory Allen (CSIS) argue that motivation to localize is not the same as acceleration, and that some Chinese chip-independence efforts predate the controls. Chris McGuire (CFR) argues the US chip lead is actually widening. And Anthropic's Dario Amodei argues DeepSeek's efficiency was "an expected point on an ongoing cost-reduction curve," not a resource-constraint miracle.

A careful reading of the download data supports humility on both sides. By one 2026 analysis, Chinese models edged ahead of US models in the share of downloads of newly released models — but that measures a rolling 12-month window, and on an all-time basis US models still dominate downloads. The honest summary: Chinese labs drove much of the open-weight momentum in 2025–26, the technical contributions are real, and the geopolitical cause is genuinely debated. (And a note on attribution: the widely used Muon optimizer originated in the international open-source community, not with any single lab or country — cite ideas by their lineage, not a flag.)

Open vs Closed: How Big Is the Gap in 2026?

The gap has never been smaller. On public benchmarks, the best open-weight models trail the top proprietary systems by a small and shrinking margin — often described in 2026 as a matter of a few months rather than the year-plus gap of 2023. Independent reviewers who tested Kimi K3 at launch placed it at or near frontier level for agentic coding, and one summed up the moment bluntly: "Open-source AI is no longer a few months behind the closed frontier models."

Two honest caveats keep this from becoming hype. First, the very hardest reasoning tasks still tend to favor the top proprietary models. Second, capability is not reliability — open frontier models can be slower, more token-hungry, and flakier on long runs than their polished closed rivals. The gap in raw capability is nearly closed; the gap in dependability is still real, and it's the reason AI-generated apps break more often than their demos suggest.

The Best Open-Source LLMs to Use Right Now

If you want a ranked answer to "which open model should I use in 2026," we keep a dedicated, continuously-updated guide: the 10 best open-source LLMs. The short version by job:

  • Coding: Qwen3-Coder and the Kimi K2/K3 line lead agentic coding.
  • Reasoning and math: DeepSeek V4 and the R1 lineage are the value picks.
  • Long context: Kimi (1M tokens) and Llama 4 Scout (very long windows).
  • Running locally: small Gemma, Qwen, and Llama variants fit a single GPU.

The larger point: you rarely want just one. The most effective setup keeps several models available and routes each task to the best one — a strategy we break down in AI agent cost optimization.

Open-Source LLM License Cheat-Sheet

Licenses are where "open" gets real. Here's the current lay of the land (always verify the specific version's model card — licenses change within families in both directions):

   MODEL FAMILY        LICENSE               COMMERCIAL USE
   ──────────────────────────────────────────────────────────
   Qwen                Apache 2.0            Yes, unrestricted
   Mistral / Mixtral   Apache 2.0            Yes, unrestricted
   DeepSeek (V4)       MIT                   Yes, unrestricted
   GLM (Zhipu)         MIT                   Yes, unrestricted
   Kimi (K2 family)    Modified MIT          Yes; attribution above
                                             100M MAU / $20M/mo rev
   Llama               Llama Community       Yes; restricted above
                                             700M MAU + use policy
   BLOOM               OpenRAIL              Yes; use restrictions
   ──────────────────────────────────────────────────────────
   Rule of thumb: Apache 2.0 and MIT = the freest. Everything
   else has a clause worth reading before you build on it.

How to Use Open Models Without the Setup Headache

Here's the practical payoff of this whole history. You don't have to pick one open model, and you certainly don't have to rent GPUs to use them. The lesson of the last six years — GPT-2's caution, Llama's ecosystem, DeepSeek's cost shock, Kimi K3 in the open — is that the frontier is becoming a commodity input. The durable advantage is the layer that lets you swap models freely.

That's how Taskade Genesis works. It gives you 15+ frontier models from OpenAI, Anthropic, Google, and open-weight providers — including DeepSeek, Qwen, and Moonshot's Kimi — selectable per agent, with Auto picking a sensible default. Route cheap high-volume work to an open-weight model and hard reasoning to a premium one, all in one workspace, no lock-in.

Pick your model per agent in Taskade — route open-weight and frontier models in one workspace

Every AI agent gets 34 built-in tools, persistent memory, and 100+ bidirectional integrations. Your projects become the memory, your agents the intelligence, your automations the execution:

MemoryProjects & knowledge IntelligenceAgents on any model ExecutionAutomations & apps
MemoryProjects & knowledge IntelligenceAgents on any model ExecutionAutomations & apps

Build your own AI app from a prompt → — describe it once, and Taskade Genesis builds it as living software you can run on whichever models you like.

What's Next for Open-Source LLMs

Three trends are already visible. Trillion-parameter open models are becoming normal — Kimi K3 (2.8T) and DeepSeek V4-Pro (~1.6T) put frontier-scale weights in the open. The gap keeps shrinking toward a few months, which changes the economics of building on open models. And attention is moving from raw chat quality to agentic and tool-use capability — models that can run long, multi-step tasks — plus on-device inference for privacy. The through-line: openness compounds, and the community keeps closing the distance.

▲ ■ ●  The labs build the models; you build the apps that run on them. Don't just read the open-source timeline — clone a living model tracker, point it at any lab, and own a research workspace in minutes. Memory feeds Intelligence, Intelligence triggers Execution. Clone the live tracker → or build your own from a prompt →.

🔗 Related Reading

Models & labs:

  • The 10 best open-source LLMs, ranked
  • Moonshot AI & Kimi history
  • Anthropic & Claude history
  • OpenAI & ChatGPT history
  • Google Gemini history
  • NVIDIA & Jensen Huang history

How it works:

  • How large language models work
  • What are reasoning models?
  • What are world models?
  • What is mechanistic interpretability?
  • The history of computing

🐑 Before you go — you don't have to choose one model. Inside Taskade Genesis you get:

  • 15+ frontier and open-weight models (DeepSeek, Qwen, Kimi, Claude, GPT, Gemini)
  • 34 built-in tools and persistent memory on every AI agent
  • 100+ bidirectional integrations to automate the busywork
  • Auto model routing — the right model per task, no lock-in

Start free → · Explore ready-made AI apps →

💬 Frequently Asked Questions About Open-Source LLMs

What was the first open-source LLM?

There's no clean answer. GPT-2 (2019) is often cited, but its full weights were withheld until November 2019, so it wasn't fully open at launch. EleutherAI's GPT-Neo and GPT-J (2021) were the first genuinely open GPT-3-class models, and BLOOM (2022) was the most fully open early release.

What's the difference between open source and open weight?

Open weight means you get the downloadable parameters. Open source (by the OSI's 2024 definition) means weights plus training code, data information, and an unrestricted license. Most popular "open" models — Llama, DeepSeek, Qwen, Kimi — are technically open weight.

Is Llama open source?

Not strictly. The OSI classifies it as source-available/open weight because of the 700-million-MAU clause, an acceptable-use policy, and undisclosed training data — though it's freely usable for almost everyone.

Is DeepSeek open source?

The V4 generation ships code and weights under the permissive MIT license, so largely yes. Earlier V3 and R1 releases had a more restrictive model license. The trend has been toward greater openness.

What's the largest open-source LLM?

Kimi K3 (2.8 trillion parameters, July 2026) is the largest open-weight model to date, followed by DeepSeek V4-Pro (~1.6 trillion). Both are mixture-of-experts, so only a fraction of the parameters run per token.

Are open models as good as GPT-5 or Claude?

On many tasks, yes, and the gap is down to months. The top proprietary models still lead on the hardest reasoning and on reliability. Most teams use both and route per task.

Can I run these locally?

Small open-weight models run on a single GPU or laptop; the biggest (like Kimi K3) need data-center hardware even with open weights. Managed platforms like Taskade Genesis let you use them without any setup.

Who makes the top open models?

Meta (Llama), Mistral, DeepSeek, Alibaba (Qwen), Moonshot (Kimi), and Zhipu (GLM), plus pioneers EleutherAI, BigScience (BLOOM), and Google (Gemma).

Taskade AI banner.

0%

On this page

What "Open-Source LLM" Actually Means (and Why "Open Weight" Is More Accurate)A Visual Timeline of Open-Source LLMs (2019 → 2026)2019: GPT-2 and the Cautious Birth of Open Models2021–2022: The Truly-Open Era — EleutherAI, OPT, and BLOOM2023: Llama Changes Everything (and the "Open" Debate Begins)Is Llama Really Open Source?2023–2024: Mistral and the Mixture-of-Experts Moment2024: The Field Scales Up and MaturesJanuary 2025: The DeepSeek ShockHow Much Did DeepSeek Cost to Train?Is DeepSeek Open Source?2025–2026: China's Open-Weight Surge — Qwen, Kimi, GLM, DeepSeek V4Why Are Chinese Open-Source Models Winning? The DebateOpen vs Closed: How Big Is the Gap in 2026?The Best Open-Source LLMs to Use Right NowOpen-Source LLM License Cheat-SheetHow to Use Open Models Without the Setup HeadacheWhat's Next for Open-Source LLMs🔗 Related Reading💬 Frequently Asked Questions About Open-Source LLMs

Related Articles

Fine-tuning vs RAG vs prompting: how to customize an LLM in 2026
June 20, 2026AI

Fine-Tuning vs RAG vs Prompting: How to Customize an LLM in 2026 (Cost, Effort, and a Decision Flowchart)

Fine-tuning, RAG, and prompting are the three ways to customize an LLM. Here is a decision flowchart, real cost math, an...

AI reasoning models explained: chain-of-thought and test-time compute in 2026
June 18, 2026AI

AI Reasoning Models Explained: Chain-of-Thought, Test-Time Compute, and When to Pay for Thinking (2026)

Reasoning models spend extra compute thinking before they answer. Here is how chain-of-thought, test-time compute, and R...

Vector databases and vector search explained: embeddings and similarity search in 2026
June 19, 2026AI

Vector Databases & Vector Search Explained: Embeddings, Similarity Search, and the Top Vector DBs in 2026

A vector database stores embeddings and finds the most similar ones fast. Here is how embeddings, ANN/HNSW search, and h...

What is LangChain? Complete history of LangChain, LangGraph, and the rise of AI agent frameworks 2022 to 2026
June 13, 2026AI

What Is LangChain? Complete History, LangGraph & the AI Agent Framework Era (2026)

The complete history of LangChain, from Harrison Chase's October 2022 side project to 100K+ GitHub stars, $35M in fundin...

A no-code custom AI agent builder showing role definition, document training, tool selection, and one-click deployment inside Taskade
July 16, 2026AI

How to Build Custom AI Agents Without Code (Step-by-Step, 2026)

Build a custom AI agent without code in under 30 minutes. Define its role, train it on your documents, give it 34 built-...

State of AI app building 2026 — an analytics dashboard tracking app usage, market stats, and adoption data
July 15, 2026AI

The State of AI App Building 2026: Market Map, Funding League Table & the Operation Turn

The state of AI app building in 2026, compiled as an industry report: the full funding league table (SpaceX's $60B Curso...

View All Articles