惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

云风的 BLOG
云风的 BLOG
V
Visual Studio Blog
人人都是产品经理
人人都是产品经理
The GitHub Blog
The GitHub Blog
月光博客
月光博客
T
Tailwind CSS Blog
小众软件
小众软件
Y
Y Combinator Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
P
Proofpoint News Feed
B
Blog RSS Feed
博客园 - 司徒正美
A
About on SuperTechFans
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
博客园 - 聂微东
Microsoft Security Blog
Microsoft Security Blog
Recent Announcements
Recent Announcements
博客园 - Franky
U
Unit 42
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Microsoft Azure Blog
Microsoft Azure Blog
T
The Blog of Author Tim Ferriss
GbyAI
GbyAI
Apple Machine Learning Research
Apple Machine Learning Research

SiliconANGLE

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws - SiliconANGLE AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle - SiliconANGLE Team Cymru launches Total Insights Feed to replace legacy threat intelligence lists - SiliconANGLE AI Mode in Chrome adds split-screen view to enhance the web search experience - SiliconANGLE Resolve AI raises $40M at $1.5B valuation to optimize production environments - SiliconANGLE How Zscaler and OpenAI turn zero-trust security into an AI accelerator - SiliconANGLE OpenAI ratchets up Codex's agentic capabilities to rival Claude Code - SiliconANGLE Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements - SiliconANGLE Slash raises $100M at a $1.4B valuation to expand AI-powered banking platform for online businesses - SiliconANGLE Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work - SiliconANGLE Data center, consumer device chips boost TSMC’s revenue - SiliconANGLE Mission-critical security cannot be bolted on, says Oracle - SiliconANGLE Agentic infrastructure reshapes enterprise AI - SiliconANGLE Data quality, and data freedom, foundational for AI success - SiliconANGLE Data trust is a bedrock in successful, scalable AI outcomes - SiliconANGLE Google introduces new agentic AI-ready tools and resources for Android developers  - SiliconANGLE Agentic AI orchestration separates winners from laggards - SiliconANGLE Data-driven tools turning the tide against human trafficking - SiliconANGLE Achieving trusted AI development goes beyond 'vibes' - SiliconANGLE Impinj boosts edge computing power in updated R700 RAIN RFID reader - SiliconANGLE Certinia powers professional services with AI - SiliconANGLE Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M - SiliconANGLE Developer tooling startup Expo nabs $45M investment - SiliconANGLE Solidroad lands $25M to bring AI to customer support interactions - SiliconANGLE DuploCloud lands compliance and AI governance certifications as enterprise buyers tighten scrutiny - SiliconANGLE Lua lands $5.8M to help businesses build and manage AI agent workforces - SiliconANGLE Best of frenemies: Oracle's and AWS' clouds unite with dedicated, private connectivity - SiliconANGLE NIST shifts National Vulnerability Database to risk-based triage as CVE submissions hit record levels - SiliconANGLE Cisco goes to the races with new Churchill Downs multiyear partnership - SiliconANGLE Susecon 2026 will tackle the future of open-source platforms - SiliconANGLE
DeepSeek open-sources V4 large language model series - Si...
Maria Deutsc · 2026-04-25 · via SiliconANGLE

DeepSeek open-sources V4 large language model series

Chinese artificial intelligence developer DeepSeek today released a new series of open-source large language models.

V4, as the algorithm family is called, comprises 2 LLMs on launch. There’s the flagship V4-Pro and a smaller model called V4-Flash that trades off some output quality for lower hardware usage.

Both algorithms are based on a mixture of experts, or MoE, architecture. That means they comprise multiple neural networks rather than a single set of artificial neurons. V4-Pro has 1.6 trillion parameters and activates a subset of its neural networks with 49 billion parameters when answering user prompts. V4-Flash, in turn, contains 284 billion parameters and activates 13 billion at any given time.

One of the new architectural features in the LLM series is a so-called hybrid attention mechanism. An LLM’s attention mechanism ranks the data points in a user prompt based on their importance. The model takes the most relevant data points into consideration when generating responses and discards irrelevant details, which boosts output quality.

Attention mechanisms don’t process prompts in their original form, but rather use a mathematical representation called a KV cache. V4’s hybrid attention architecture uses two different compression methods to reduce the size of the KV cache, which lowers memory requirements. As a result, the model family’s KV cache uses 90% less memory during inference than the one in DeepSeek’s previous-generation LLMs.

Many of the other new features in the V4 lineup were added to optimize its training workflow.

A neural network comprises artificial neuron collections called layers that process data in a specific order. Prompts enter the first layer, which carries out a series of calculations and transmits the results to the second layer. The second layer then performs calculations of its own, sends the results to the third layer and so forth.

Data regularly moves between an LLM’s layers during training. V4 includes a feature called mHC that enables data to travel directly between distant layers without going through the intermediate neuron clusters between them. That approach reduces training errors, which in turn boosts AI output quality.

The neuron clusters between the first and last layers of an LLM are known as its hidden layers. According to DeepSeek, V4 uses a software module called Muon to optimize the hidden layers. It helps speed up training runs and reduce the associated infrastructure requirements.

DeepSeek carried out V4’s initial training using a dataset that compromised about 27 trillion tokens. It then applied a two-step post-training workflow. The first step separately optimized the neural networks that make up each V4 model, while the second improved their ability to coordinate their work. 

DeepSeek evaluated V4-Pro, the most capable LLM in the series, using about two dozen benchmarks. It then compared the model’s results against the scores of several other frontier models including Claude Opus 4.6. V4 bested all the competing LLMs across 3 of the benchmarks. Additionally, there were several cases where V4 completed a benchmark better than some of the other LLMs but not all of them.

V4-Pro and V4-Flash are available in preview on Hugging Face.

Image: Unsplash

A message from John Furrier, co-founder of SiliconANGLE:

Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities.

  • 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more
  • 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network.

About SiliconANGLE Media

SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement. As the parent company of SiliconANGLE, theCUBE Network, theCUBE Research, CUBE365, theCUBE AI and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.

Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.