< Back to situations

Monitor this situation.

[SITUATION] · [ACTIVE] · [TECHNOLOGY]

25 clusters · 144 sources · 75 days · First seen · Last updated

DeepSeek expands agent frameworks and multimodal models

Overview

DeepSeek has expanded its ecosystem through the release of the DeepSeek Harness (DSH), an open-source, MIT-licensed agentic development environment built on the Cordis meta-framework. DSH utilizes a modular, plugin-based architecture that allows developers to swap models, tools, and file systems without modifying core code. It features Chain of Thought (CoT) tracking via append-only logs, enabling users to inspect or replay reasoning processes through a Trajectory view. The company has also transitioned its flagship DeepSeek-V4-Pro (build 0813) from preview to general availability. This 1.6-trillion-parameter Mixture-of-Experts model has shown improved performance in software engineering and cybersecurity benchmarks. To support local deployment, GEEKOM demonstrated a four-node AI cluster using A9 Mega Mini PCs to run DeepSeek V4 Flash locally via USB4. DeepSeek’s multimodal capabilities were extended by the release of DeepSeek-V4-Flash-Vision-Exp. Subsequently, FlashLabs released an experimental, uncensored GGUF quantized version, ‘DeepSeek-V4-Flash-Vision-Uncensored’, designed for security research.

In September 2026, DeepSeek released V4.1-Flash, a 552-billion-parameter multimodal Mixture-of-Experts model optimized for coding, agentic tasks, and cost efficiency. The model activates only 8 billion parameters for input and 16 billion for output, reportedly reducing KV-cache size to approximately one-quarter of the previous V4 Flash generation. V4.1-Flash achieved a score of 90.6 on Terminal-Bench 2.1, outperforming competitors like OpenAI’s GPT-5.6 Sol. The release triggered market volatility, contributing to a decline in shares for Samsung Electronics and SK Hynix as investors reassessed AI hardware requirements. Some developers have raised concerns regarding the rapid routing of V4 Pro requests to the new Flash model, citing potential risks to stability and research reproducibility.

Entities

DeepSeek · OpenAI · Anthropic · DeepSeek Harness · V4-Flash

Claims

What the coverage asserts, and how many sources carry each claim.

Coverage disagrees

Sources make claims that cannot both be true. CLSTR reports the disagreement; it does not decide who is right.

  • "The model has 552 billion total parameters, activating 8 billion for input and 16 billion for output." unwire.pro · thebridge.jp · time.news · gigazine.net

    vs

    "The V4 Pro model utilizes a Mixture-of-Experts (MoE) architecture with approximately 1.6 trillion total parameters."

    The claims provide different total parameter counts for the DeepSeek V4 Pro model (1.6 trillion vs 552 billion).

Timeline

  1. 5 days ago

    [TECHNOLOGY] 4 sources
    DeepSeek releases V4.1-Flash AI model for coding and agent tasks

    DeepSeek's new V4.1-Flash model targets high-speed coding and agentic tasks, offering lower inference costs that may shift competition toward model-independent AI development platforms.

  2. 8 days ago

    [TECHNOLOGY] 7 sources
    DeepSeek launches V4.1-Flash AI model to reduce inference costs

    DeepSeek launched V4.1-Flash, a multimodal AI model designed to slash inference costs and memory usage through a highly efficient Mixture-of-Experts architecture.

  3. 11 days ago

    [TECHNOLOGY] 12 sources
    DeepSeek launches V4. 1 Flash multimodal AI model

    DeepSeek has launched V4. 1 Flash, a multimodal MoE model featuring native vision support, improved inference speeds, and significantly lower API costs through a new Causal Encoder-Decoder architecture.

  4. 11 days ago

    [TECHNOLOGY] 5 sources
    DeepSeek announces hiring of 150 engineers to scale infrastructure

    DeepSeek is hiring approximately 150 engineers to overhaul its backend systems and infrastructure, focusing on scaling Agent-compute and server-side engineering rather than AI research.

  5. 17 days ago

    [TECHNOLOGY] 2 sources
    FlashLabs releases uncensored DeepSeek-V4 multimodal model for security research

    FlashLabs released a quantized, uncensored version of the DeepSeek-V4-Flash-Vision multimodal model for security research, while OpenAI’s GPT-5.6 Luna remains atop the OrcaRouter model leaderboard.

  6. 28 days ago

    [TECHNOLOGY] 11 sources
    DeepSeek launches V4 Flash Vision experimental multimodal AI model

    DeepSeek has launched DeepSeek-V4-Flash-Vision-Exp, an experimental multimodal AI model capable of interpreting screenshots and interfaces to automate complex tasks via API.

  7. about 1 month ago

    [TECHNOLOGY] 14 sources
    DeepSeek releases open-source Harness framework and updates API pricing

    DeepSeek has open-sourced its DeepSeek Harness (DSH) framework for AI agents and introduced peak-valley API pricing, while community controversy surrounds unverified performance claims from third-party tools.

  8. about 1 month ago

    [TECHNOLOGY] 3 sources
    DeepSeek implements new peak and off-peak API pricing

    DeepSeek has introduced a new peak and off-peak pricing model for its API services, implementing tiered rates based on time of day and cache hits to optimize server utilization.

  9. about 1 month ago

    [TECHNOLOGY] 6 sources
    DeepSeek launches V4 Pro flagship model alongside local AI cluster innovations

    DeepSeek has released its V4 Pro flagship model with major agentic performance upgrades, while GEEKOM demonstrated a local AI cluster using DeepSeek V4 Flash on Mini PCs for private enterprise use.

  10. about 1 month ago

    [TECHNOLOGY] 2 sources
    DeepSeek launches open-source agent framework DeepSeek Harness

    DeepSeek has launched DeepSeek Harness, an open-source agent framework featuring a modular plugin-based architecture and Chain of Thought tracking for AI development.

  11. about 1 month ago

    [TECHNOLOGY] 12 sources
    DeepSeek launches V4 Pro model and Harness agent framework

    DeepSeek has released its V4 Pro model and the open-source DeepSeek Harness framework for AI agents, while simultaneously implementing a new peak/off-peak API pricing structure.

  12. about 1 month ago

    [TECHNOLOGY] 38 sources
    DeepSeek launches V4 Pro flagship AI model with enhanced agent capabilities

    DeepSeek has launched its V4 Pro flagship AI model, featuring a 1-million-token context window and enhanced agentic capabilities. The release includes a new peak/off-peak API pricing structure.

  13. about 2 months ago

    [TECHNOLOGY] 16 sources
    Alibaba launches Qwen3.8-Max AI model and plans revenue share for large users

    Alibaba unveiled the 2.4‑trillion‑parameter Qwen3.8‑Max AI model, open‑weight and priced at $2/$6 per million tokens, and plans a revenue‑share deal for users earning over $20 M annually, spurring a 27.4 % July

  14. about 2 months ago

    [BUSINESS] 12 sources
    DeepSeek raises prices for V4-Flash AI model, shaking market

    DeepSeek will raise fees for its V4‑Flash AI model, ending its ultra‑low‑cost edge. The model, now the world’s most used AI service, currently costs $0.14/$0.28 per million tokens and may see higher rates, with

  15. about 2 months ago

    [TECHNOLOGY] 11 sources
    DeepSeek launches V4-Flash, the cheapest AI model

    DeepSeek’s V4-Flash, released 31 July, costs $0.14/$0.28 per M tokens (≈3¢ per test), over 100× cheaper than Anthropic’s Claude Fable 5, and scores 50/100, matching Google’s Gemini 3.6 Flash while trailing top‑

  16. about 2 months ago

    [TECHNOLOGY] 3 sources
    DeepSeek CEO Liang Wenfeng outlines AI strategy amid China-US competition

    DeepSeek CEO Liang Wenfeng said the firm values AGI and open‑source over profit, sees compute power as the main AI race bottleneck with the U.S., and plans its own large‑scale clusters after a $52 bn valuation.

  17. about 2 months ago

    [TECHNOLOGY] 10 sources
    DeepSeek CEO cites compute gap, pledges open‑source AI models

    DeepSeek CEO says the lab lacks compute power versus U.S. rivals, needs tens of thousands of GPUs, but will expand capacity and keep top models open‑source while prioritising AGI over profit.

  18. 2 months ago

    [TECHNOLOGY] 4 sources
    DeepSeek V4 and Alibaba Qwen 3.8 Launch Boost Chinese AI Race

    DeepSeek V4 and Alibaba Qwen 3.8 debut, offering low‑cost, high‑performance AI models that intensify China’s competition with US giants and the Kimi K3.

  19. 2 months ago

    [TECHNOLOGY] 2 sources
    Qwen3.5-9B NVFP4 Language Model Available for Windows Installation

    The Qwen3.5‑9B NVFP4 language model, a 9 B‑parameter AI system, is now available for Windows, offering fast, low‑memory inference with required specs of i5/Ryzen 5 CPU, 32 GB RAM, 80 GB NVMe SSD, and CUDA 8.0+.

  20. 2 months ago

    [TECHNOLOGY] 2 sources
    Qwen AI models receive new local deployment guide and Google TPU optimization playbook

    A guide details local deployment of Qwen3‑VL‑Reranker‑8B, while Google publishes a playbook boosting Alibaba’s Qwen 3.5‑397B on Ironwood TPUs with up to 4.7× speed gains.

  21. 2 months ago

    [TECHNOLOGY] 2 sources
    Local Deployment Guides for GLM-5-FP8 and Qwen3.5‑9B‑MLX‑8bit AI Models

    Guides detail offline deployment of GLM-5-FP8 (176 B, FP8) and Qwen3.5‑9B‑MLX‑8bit (9 B, 8‑bit) models on consumer hardware, outlining hardware needs and setup scripts.

  22. 2 months ago

    [TECHNOLOGY] 2 sources
    Gemma‑4‑26B and Qwen‑3.5‑2B Enable High‑Performance Local AI Deployment

    Gemma‑4‑26B (4‑bit AWQ) and Qwen‑3.5‑2B (2 B parameters) can be installed locally with scripts that auto‑configure hardware, enabling fast, low‑resource AI deployment.

  23. 2 months ago

    [TECHNOLOGY] 2 sources
    New Qwen3.5‑9B‑AWQ and Gemma‑4‑12B‑it models enable fast local AI deployment

    Qwen3.5‑9B‑AWQ and Gemma‑4‑12B‑it language models launch with offline deployment tools for fast, consumer‑grade AI inference and multilingual capabilities.

  24. 2 months ago

    [TECHNOLOGY] 4 sources
    Deepseek‑V4 and Qwen3.6‑27B AI models now runnable on consumer PCs

    Open‑source AI models Deepseek‑V4 (7 B) and Qwen3.6‑27B (27 B) can now be installed on typical PCs using GGUF packages, with step‑by‑step guides outlining required hardware and automatic setup.

  25. 3 months ago

    [TECHNOLOGY] 2 sources
    DeepSeek releases 284‑billion‑parameter V4 model for local laptop use

    DeepSeek launched a 284B‑parameter V4 model that runs on high‑end laptops using MoE and quantization, and issued a V3 API guide offering OpenAI‑compatible, low‑cost access for developers.

Sources

36kr.com · 36kr.jp · 4sysops.com · adslzone.net · agendadigitale.eu · androidgeek.pt · applaud.com · arquitectosdevalencia.es · arts-spectacles.com · ascii.jp · assodigitale.it · azeritimes.com · backendnews.net · begeek.fr · biznis.rs · blocktempo.com · blog-nouvelles-technologies.fr · bolnews.com · borncity.com · brandaktuell.at · brandspurng.com · brasil247.com · brasilemfolhas.com.br · bright.nl · cafef.vn · canaltech.com.br · chinesepress.com · cnmo.com · conteudos.cnnbrasil.com.br · cryptobriefing.com · deccanchronicle.com · dev.to · devx.com · dimsumdaily.hk · donanimhaber.com · dunya.com · enterprisesecuritytech.com · etailment.de · finance.technews.tw · fonearena.com · fortune.com · fourweekmba.com · frickingruvin.medium.com · gadgety.co.il · gamebusiness.jp · gamemag.it · games.yahoo.com.tw · geeky-gadgets.com

This summary has been updated 29 times: see revision history