惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

GbyAI
GbyAI
Martin Fowler
Martin Fowler
I
InfoQ
腾讯CDC
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
爱范儿
爱范儿
Microsoft Security Blog
Microsoft Security Blog
Google DeepMind News
Google DeepMind News
D
DataBreaches.Net
云风的 BLOG
云风的 BLOG
F
Fortinet All Blogs
N
Netflix TechBlog - Medium
博客园 - 聂微东
Microsoft Azure Blog
Microsoft Azure Blog
D
Docker
博客园 - 三生石上(FineUI控件)
Y
Y Combinator Blog
博客园 - Franky
Engineering at Meta
Engineering at Meta
B
Blog
罗磊的独立博客
Apple Machine Learning Research
Apple Machine Learning Research
Jina AI
Jina AI
V
Visual Studio Blog

Interesting Engineering

US firm to scale laser-based nuclear fusion ‘breakthrough’ with new partnership Military Archives - Interesting Engineering World’s first non-nuclear lead-cooled reactor to generate electricity begins installation US scientists devise new process to turn sewage sludge into 99% pure natural gas US firm unveils submarine-hunting drone with 9,200-mile-range, 35 mph top speed Military Archives - Interesting Engineering Supercomputer finds lithium-titanium tweak to boost sodium-ion batteries for grids Lockheed Martin demonstrates vertical launch missile system for mobile drone defense China’s 1116 MWe Taipingling Unit 1 reactor goes online, set to generate 9bn kWh yearly ChatGPT Images 2.0 update combines reasoning, research, and design with 2K output US Navy tests plug-and-play laser system on USS Bush carrier, downs drones at sea China’s CATL reveals 621-mile EV battery, under-7-minute charging to challenge BYD US uses world’s first exascale supercomputer to model supernovae, fusion reactors AI and Robotics Archives - Interesting Engineering First-in-human study confirms safety of graphene-based brain interface Tesla’s Optimus humanoid robot greets runners, poses for photos at Boston Marathon Interlocking materials offer high strength and flexibility for robotics, infrastructure US redeploys 100,000-ton nuclear-powered aircraft carrier in Red Sea after repairs US scientists unveil concept for ‘world’s first neutrino laser’ to unlock breakthroughs New military tech can maintain communication in contested electronic warfare environments Got a dark personality? Psychologists can help you choose your career wisely Humidity boosts performance of 3D-printed nanogenerator instead of degrading it China demonstrates microwave beam that recharges drones in flight, continues power delivery Scientists run compact free-electron laser for eight hours, cracks FEL stability problem China’s PLA considers to use minelaying underwater drones to enforce Taiwan blockade: Report 1-ton sharks may struggle for survival in waters exceeding 62.6°F, study suggests US firm’s thorium nuclear fuel bundles move to manufacturing for commercial reactors Tesla hits 0% charge in remote Chilean desert as YouTuber uses hood-mounted solar Humanoid robot surpasses human world record in Beijing half-marathon, clocking 50:26 mins New method extracts maximum work from unknown quantum states using symmetry tricks
China’s DeepSeek unveils V4 AI model with 1M context wind...
Aamir Kholla · 2026-04-25 · via Interesting Engineering

The artificial intelligence race is accelerating. OpenAI launched GPT-5.5 this week, while the White House accused China of copying US AI systems at scale. Now, DeepSeek has introduced preview versions of its V4 model, targeting direct competition with top US platforms.

DeepSeek released the V4 Flash and V4 Pro series, highlighting gains in coding, reasoning, and agent-driven tasks. The models incorporate architectural upgrades and optimization improvements, with a clear focus on efficiency as systems grow more expensive to run.

A key feature is what DeepSeek calls Hybrid Attention Architecture. The method improves how models retain context across long conversations and reduces memory loss in extended interactions.

The system also supports a 1 million-token context window, allowing users to input entire codebases or long documents in a single prompt. This could reshape workflows in software development and enterprise analysis.

DeepSeek reported strong benchmark performance against systems from Anthropic, Google, and OpenAI. However, it acknowledged that V4 still trails the most advanced models by three to six months, while emphasizing cost and deployment flexibility.

Cost and chip strategy

DeepSeek continues to focus on efficiency as a competitive edge. Its trillion-parameter system uses a Mixture-of-Experts approach, activating only a fraction of parameters per task. This reduces inference costs compared to traditional models, which typically activate all parameters for each request.

The models are also designed to run on domestic hardware. DeepSeek expects costs to drop further once clusters powered by Huawei Technologies Co.’s Ascend 950 chips come online later this year. This shift could reduce reliance on US chipmakers and strengthen China’s AI infrastructure.

Markets reacted quickly. Shares of Semiconductor Manufacturing International Corp. and Hua Hong Semiconductor rose, while rival AI firms declined. Investors appear to be betting on increased demand for Chinese-made chips.

DeepSeek said service capacity for the V4 Pro series remains constrained due to limited computing resources. It is also in talks with Tencent Holdings Ltd. and Alibaba Group Holding Ltd. for its first funding round, signaling plans to expand infrastructure.

Pressure on US rivals

The V4 launch follows the earlier R1 model, which shook the AI market and prompted a reassessment of spending on frontier systems. DeepSeek claimed R1 delivered competitive performance at a fraction of the cost of leading US models.

That debate has since shifted again. US tech firms are projected to invest around $650 billion in 2026 on AI infrastructure and data centers, balancing performance gains with long-term costs.

DeepSeek said V4 builds on that approach by improving both scale and efficiency. It continues to position open-source models as alternatives to closed systems, appealing to developers and enterprises seeking more control.

Still, the release comes under scrutiny. US officials have accused DeepSeek of using restricted chips, while Anthropic has alleged misuse of its Claude system. DeepSeek has not disclosed training costs or hardware details for V4.

The launch highlights intensifying global competition. By focusing on lower costs, scalable performance, and hardware flexibility, DeepSeek is challenging how AI systems are built and who leads the next phase of development.

The Blueprint

Get the latest in engineering, tech, space & science - delivered daily to your inbox.

Aamir is a seasoned tech journalist with experience at Exhibit Magazine, Republic World, and PR Newswire. With a deep love for all things tech and science, he has spent years decoding the latest innovations and exploring how they shape industries, lifestyles, and the future of humanity.