-
Emporio Armani names Dario Vitale new creative director
-
Stocks edge higher ahead of US Fed rate call
-
EU to ban social media for under 13s, curb access until 15
-
BASIS.pro Expands On-Chain Infrastructure with XDC Network Partnership and Zypher DAO as Auto Earn Goes Live
-
Wildfires push endangered Sumatran elephants to brink: Indonesian NGO
-
Pickle-flavoured tart? AI inspires Japan's convenience stores
-
Bank of Japan set to raise rates under pressure from inflation, US
-
EU chief hosts Canada's Carney in push to 'deepen' alliance
-
EU chief to unveil social media, gaming curbs for under-15s
-
Meta chief pushes back on AI slowdown calls
-
'Resident Evil' film brings video game's ethos to the big screen
-
US Fed to deliver rate decision with markets betting on hike
-
i-payout Expands U.S. Money Transmitter Licensing Footprint and Advances European EMI Strategy
-
Sanders, Bannon warn of AI dangers in rare US left-right alignment
-
US Senate crypto bill collapses amid partisan deadlock
-
Empty benches as schools reopen in Venezuela's quake-hit Guaira
-
OpenAI, Anthropic and Google are working to create an AI standards body
-
Tiny English village holds 'independence' vote over asylum seeker housing
-
Protesting shepherds, farmers clash with Romanian police
-
US Treasury chief says to meet Chinese counterpart at weekend
-
Foundation Lets Ledger Users Switch to Open Source Without Starting Over
-
BingX Evolves into a Multi-Asset Trading Platform, Connecting Users to Global Opportunities
-
EU to propose curbing social media, online games for under 15s
-
International Trading Institute Expands Professional Trader Development Beyond the Master’s in Trading
-
FX Junction Reaches 40,600 Members and 26 Million Trades
-
Lake Energy Secures $80 Million for U.S. Renewable Energy Expansion
-
'We unleashed the beast': the world's fears and hopes about AI
-
Stocks drop, oil climbs and Treasury yields hit 19-year high
-
Hospitality leaders ask Burnham to halve VAT amid rising costs
-
SkySail Strategies Outperforms Wall Street's 30-Year Risk Standard With Proprietary AI Inference Model
-
Worcestershire invited to nominate groups for 2027 King’s volunteer award
-
Just add fish: Scientists pursue more rice, less disease in Senegal paddies
-
Berlin's run-down public spaces at heart of election campaign
-
Belgium completes outer structure of offshore energy island
-
'Widow's Bay' and 'The Pitt' win big at Emmy Awards
-
China retail sales growth weakens further in August
-
Most markets drop as oil extends gains ahead of expected US rate hike
-
In Kashmir, lake weeds mix with traditional art
-
Sri Lanka's famed beach shack battles demolition
-
'Widow's Bay' and 'The Pitt' take key Emmy awards
-
Vserv targets ₹1,000 crore revenue by 2030 through AI expansion
-
Television's A-listers glitter on Emmys red carpet
-
India prohibits bank fees on UPI payments up to ₹2,000
-
Air India CEO summoned over alleged departure immigration lapse
-
Brazil's Amazon defender Raoni has cancer
-
HFCL expands planned fibre investment to ₹1,800 crore
-
Trump rejects AI slowdown concerns as UN urges coordinated controls
-
Kuwait schools restore classroom routines after remote learning
-
Higher wages draw migrant workers to Kashmir despite attacks
-
BIS flags debt and profitability risks in AI-driven market rally
China's DeepSeek releases long-awaited new AI model
Chinese startup DeepSeek released a new artificial intelligence model with "drastically reduced" costs Friday, more than a year after it stunned the world with a low-cost reasoning model that matched the capabilities of US rivals.
The AI race has intensified the rivalry between China and the United States, and the White House on Thursday accused Chinese entities of a massive effort to steal artificial intelligence technology.
Hangzhou-based DeepSeek burst onto the scene in January last year with a generative AI chatbot, powered by its R1 reasoning model, that upended assumptions of US dominance in the strategic sector.
DeepSeek-V4, "features an ultra-long context", the company said in a statement on social media platform WeChat, hailing it as "world-leading... with drastically reduced compute (and) memory costs" in a separate announcement on X.
V4 supports a context length of one million "tokens" -- small components of text including words or punctuation -- putting it on par with Google's Gemini.
Context length determines how much input a model is able to absorb to help it complete tasks.
The new V4 is released as two versions, DeepSeek-V4-Pro and DeepSeek-V4-Flash, with the latter being "a more efficient and economical choice" because it has smaller parameters.
In terms of "world knowledge", a benchmark for reasoning, V4-Pro trails only the latest Gemini model, DeepSeek said.
A "preview version" of the open source model is now available, the company said, without indicating when a final version would be released.
- 'Inflection point' -
Experts say V4's arrival marks an "inflection point" in terms of hardware and cost.
"This addresses the long-standing issues of slower performance and higher costs associated with long context lengths, marking a genuine inflection point for the industry," Zhang Yi, the founder of tech research firm iiMedia, told AFP.
"For end users, this will bring widespread, accessible benefits. For instance, if ultra-long context support becomes a standard feature, long-text processing is expected to move beyond high-end research labs and enter mainstream commercial applications," he said.
V4-Pro has 1.6 trillion parameters while the V4-Flash has 284 billion parameters, which refine models' decision-making ability.
The model has also been "optimised" for popular AI Agent products such as Claude Code, OpenClaw, OpenCode and CodeBuddy, the DeepSeek statement said.
DeepSeek's latest release is a "milestone" for Chinese firms, said veteran AI industry analyst Max Liu.
"It's a good thing for the entire domestic AI industry. It can provide better models for domestic users and we can now expect a lot more things -- more products (and a) more competitive market," he told AFP.
"This is no less shocking than when DeepSeek first came out" if its new model indeed matches the performance of leading models from Western labs, he added.
- 'Sputnik moment' -
Last year's so-called "DeepSeek shock" sparked a sell-off of AI-related shares and a reckoning on business strategy in what was also described as a "Sputnik moment" for the industry.
The chatbot performed at a similar level to ChatGPT and other top American offerings, but the company said it had taken significantly less computing power to develop.
However, its sudden popularity raised questions over data privacy and censorship, with the chatbot often refusing to answer questions on sensitive topics such as the 1989 Tiananmen crackdown.
At home, DeepSeek's AI tools have been widely adopted by Chinese municipalities and healthcare institutions as well as the financial sector and other businesses.
This has been partly driven by DeepSeek's decision to make its systems open source, with their inner workings public -- in contrast to the proprietary models sold by OpenAI and other Western rivals.
But the White House has accused Chinese firms of vying to "steal" American technology, ahead of an expected summit between Donald Trump and Xi Jinping in Beijing next month.
"The US has evidence that foreign entities, primarily in China, are running industrial-scale distillation campaigns to steal American AI," Trump's science and technology chief advisor Michael Kratsios said in a post on X.
Distillation is a common practice within AI development, often used by companies to create cheaper, smaller versions of their own models.
DeepSeek's Friday announcement also came as Meta said it planned to cut a tenth of its staff as it looks for productivity gains from the rest of the workforce while investing heavily in artificial intelligence. Reports said Microsoft was also looking to trim its ranks.
A.Agostinelli--CPN