December 26, 2024

DeepSeek V3 released

DeepSeek released V3 on December 26, 2024 — a 671B-parameter MoE model with open weights, strong coding results, and training costs far below Western frontier labs.

What it was for

DeepSeekDeepSeek V3 preceded R1 and proved Chinese labs could ship competitive open-weight MoE models at scale. Its MITMassachusetts Institute of Technology — a research university central to many computing breakthroughs.-licensed weights and transparent training details made it a default choice for self-hosted inferenceRunning a trained model to produce predictions — as opposed to the training phase that learns weights. — and a wake-up call on AI compute economics before the R1 reasoningStep-by-step logical thinking in AI models — chain-of-thought before answering hard problems. shock in January.

Companies

  • DeepSeek

Why it's here

DeepSeek V3 showed open MoE models could rival closed APIs before the R1 reasoningStep-by-step logical thinking in AI models — chain-of-thought before answering hard problems. wave.

Why it mattered

It shifted assumptions about who could afford frontier-scale training and release open weights.

What it solved

TeamsMicrosoft Teams — chat, video calls, and Office integration for workplace collaboration. wanted strong coding and general models without proprietary APIApplication programming interface — a defined way for programs to talk to each other or to a service. lock-in or Western-only vendors.

Media

  • DeepSeek
    ImageDeepSeek

    DeepSeek, MIT, via Wikimedia Commons

Related