<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	
	xmlns:georss="http://www.georss.org/georss"
	xmlns:geo="http://www.w3.org/2003/01/geo/wgs84_pos#"
	>

<channel>
	<title>RCA H100 Portable Audio Device &#8211; Digitex Solutions</title>
	<atom:link href="https://www.digiteex.com/tag/rca-h100-portable-audio-device/feed/" rel="self" type="application/rss+xml" />
	<link>https://www.digiteex.com</link>
	<description>Digitex Solutions</description>
	<lastBuildDate>Thu, 30 Jan 2025 10:34:32 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	
<site xmlns="com-wordpress:feed-additions:1">224764844</site>	<item>
		<title>India lauds Chinese AI lab DeepSeek, plans to host its models on local servers</title>
		<link>https://www.digiteex.com/india-lauds-chinese-ai-lab-deepseek-plans-to-host-its-models-on-local-servers/</link>
					<comments>https://www.digiteex.com/india-lauds-chinese-ai-lab-deepseek-plans-to-host-its-models-on-local-servers/#respond</comments>
		
		<dc:creator><![CDATA[digitex]]></dc:creator>
		<pubDate>Thu, 30 Jan 2025 10:34:28 +0000</pubDate>
				<category><![CDATA[AI]]></category>
		<category><![CDATA[AI Compute Facility]]></category>
		<category><![CDATA[China]]></category>
		<category><![CDATA[Chinese AI]]></category>
		<category><![CDATA[high-precision computing]]></category>
		<category><![CDATA[New Delhi]]></category>
		<category><![CDATA[Nvidia]]></category>
		<category><![CDATA[RCA H100 Portable Audio Device]]></category>
		<category><![CDATA[TikTok]]></category>
		<guid isPermaLink="false">https://www.digiteex.com/india-lauds-chinese-ai-lab-deepseek-plans-to-host-its-models-on-local-servers/</guid>

					<description><![CDATA[India’s IT minister on Thursday praised DeepSeek‘s progress and said the country will host the Chinese AI lab’s large language models on domestic servers, in a rare opening for Chinese technology in India. “You have seen what DeepSeek has done — $5.5 million and a very very powerful model,” IT Minister Ashwini Vaishnaw said on [&#8230;]]]></description>
										<content:encoded><![CDATA[
<br />
India’s IT minister on Thursday praised DeepSeek‘s progress and said the country will host the Chinese AI lab’s large language models on domestic servers, in a rare opening for Chinese technology in India.</p>
<p>“You have seen what DeepSeek has done — $5.5 million and a very very powerful model,” IT Minister Ashwini Vaishnaw said on Thursday, responding to criticism New Delhi has received for its own investment in AI, which has been much less than many other countries.</p>
<p>Since 2020, India has banned more than 300 apps and services linked to China, including TikTok and WeChat, citing national security concerns. The approval to allow DeepSeek to be hosted in India appears contingent on the platform storing and processing all Indian users’ data domestically, in line with India’s strict data localization requirements.</p>
<p>“Data privacy issues regarding DeepSeek can be addressed by hosting open-source models on Indian servers,” Vaishnaw said at an industry conference.</p>
<p>DeepSeek’s models will likely be hosted on India’s new AI Compute Facility. The facility is powered by 18,693 graphics processing units (GPUs), nearly double its initial target — almost 13,000 of those are Nvidia H100 GPUs, and about 1,500 are Nvidia H200 GPUs. Around 10,000 GPUs are ready to be deployed, and the facility is scheduled to begin operations “in the coming days,” according to the minister.</p>
<p>The facility will also offer computing services at steep discounts to firms in India. Vaishnaw said standard AI computing would be offered at a 42% discount to market rates, and high-precision computing would be discounted by 47%.</p>
<p>The minister’s remarks come a day after DeepSeek’s eponymous app was taken off Apple’s and Google’s app stores in Italy, after that country’s data protection regulator said it was asking how the Chinese firm was using and storing Italians’ personal data.</p>
<p>The release of DeepSeek’s R1 “reasoning” model, built on a purportedly modest budget, sent shockwaves through the tech industry this week, causing chip giant Nvidia’s market cap to decline by $600 billion. The model has quickly come under intense scrutiny, and has sparked heated debates around copyright issues, U.S. export controls, and how even more money needs to be poured into AI efforts.</p>
<p>Beyond hosting foreign AI models, India is also trying to drive development of AI models and related technology on its own turf. “Major chip designers are willing to work with India to develop indigenous GPUs,” Vaishnaw said.</p>
<p>Vaishnaw estimated that India would see investment of $30 billion in hyperscalers and data centres over the next two to three years. One of the country’s biggest conglomerates, Reliance, is planning to build what could become the world’s largest data center in the city of Jamnagar, with a capacity of 3 gigawatts, Bloomberg reported last week. </p>
<p>“We believe there are at least six major developers who can develop AI models in six to eight months on the outer limit, and four to six months on a more optimistic estimate. A common compute facility is the most important component for creating a robust AI ecosystem,” Vaishnaw said.</p>
<p>The computing facility will also support India’s broader AI initiatives. Vaishnaw said 18 AI-driven applications focusing on agriculture, climate change, and learning disabilities have been selected for initial funding.</p>
<p>To oversee development of these AI initiatives, India will establish a regulatory body using what Vaishnaw described as a “hub-and-spoke model,” allowing multiple institutions to collaborate on safety frameworks. “We will be keeping our models open and application-focused,” he said.</p>

<br /><a href="https://techcrunch.com/2025/01/30/india-to-host-china-deepseek-ai-model-locally-in-rare-tech-approval/" target="_blank" rel="noopener">Source link </a></p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.digiteex.com/india-lauds-chinese-ai-lab-deepseek-plans-to-host-its-models-on-local-servers/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">4527</post-id>	</item>
		<item>
		<title>DeepSeek: How a nerd with ‘terrible haircut’ shook up Silicon Valley&#8217;s AI race</title>
		<link>https://www.digiteex.com/deepseek-how-a-nerd-with-terrible-haircut-shook-up-silicon-valleys-ai-race/</link>
					<comments>https://www.digiteex.com/deepseek-how-a-nerd-with-terrible-haircut-shook-up-silicon-valleys-ai-race/#respond</comments>
		
		<dc:creator><![CDATA[digitex]]></dc:creator>
		<pubDate>Wed, 29 Jan 2025 15:52:55 +0000</pubDate>
				<category><![CDATA[AI]]></category>
		<category><![CDATA[Baidu]]></category>
		<category><![CDATA[Centre for Information Resilience]]></category>
		<category><![CDATA[CEO]]></category>
		<category><![CDATA[Chinese AI]]></category>
		<category><![CDATA[DeepSeek]]></category>
		<category><![CDATA[DeepSeek AI chatbot]]></category>
		<category><![CDATA[High-Flyer]]></category>
		<category><![CDATA[Liang Wenfeng]]></category>
		<category><![CDATA[librarian]]></category>
		<category><![CDATA[Nvidia chips]]></category>
		<category><![CDATA[RCA H100 Portable Audio Device]]></category>
		<category><![CDATA[Sam Altman]]></category>
		<guid isPermaLink="false">https://www.digiteex.com/deepseek-how-a-nerd-with-terrible-haircut-shook-up-silicon-valleys-ai-race/</guid>

					<description><![CDATA[The independent research group envisioned by High-Flyer took shape as DeepSeek. High-Flyer&#8217;s founder and controlling shareholder, Liang Wenfeng, also serves as DeepSeek&#8217;s discreet leader. China’s AI industry has delivered a seismic shock to the world. DeepSeek, a little-known startup, has unveiled a generative AI model that is reportedly as powerful as OpenAI’s ChatGPT—but developed at [&#8230;]]]></description>
										<content:encoded><![CDATA[
<br /> The independent research group envisioned by High-Flyer took shape as DeepSeek. High-Flyer&#8217;s founder and controlling shareholder, Liang Wenfeng, also serves as DeepSeek&#8217;s discreet leader. China’s AI industry has delivered a seismic shock to the world. DeepSeek, a little-known startup, has unveiled a generative AI model that is reportedly as powerful as OpenAI’s ChatGPT—but developed at a fraction of the cost. This breakthrough, described by Silicon Valley investor Marc Andreessen as a &#8220;Sputnik moment,&#8221; has rattled global tech markets.The fallout was immediate. US tech stocks suffered a bloodbath, losing nearly $1 trillion in value in a single day. Nvidia, the world’s most valuable chipmaker, saw $589 billion wiped off its market capitalization in what is now the largest one-day loss for any company in US history. The panic was fueled by fears that China’s AI could compete—and even surpass—US advancements while circumventing Washington’s chip restrictions.DeepSeek’s rise represents a stunning reversal of expectations. Just two years ago, China’s AI scene was seen as lagging behind the US. Beijing’s early attempts at ChatGPT-like models, such as Baidu’s Ernie and Tencent’s Hunyuan, were dismissed as inferior copies. Now, DeepSeek has positioned itself as a real contender, proving that China’s AI talent and research ecosystem are capable of groundbreaking innovation.Why it mattersDeepSeek’s AI is more than just another chatbot—it’s a symbol of China’s growing technological self-sufficiency. Its breakthrough challenges the long-standing assumption that the US holds an insurmountable lead in artificial intelligence.At the core of US confidence was the belief that China needed American technology—particularly advanced semiconductors from Nvidia—to train competitive AI models. To maintain its dominance, Washington banned the export of high-end AI chips to China in 2022 and further tightened restrictions in 2023. Many assumed these moves would kneecap China’s AI ambitions.Yet, DeepSeek has shattered that certainty. The company claims it trained its AI model using only 2,000 Nvidia chips, whereas OpenAI and Google typically require 16,000+ chips for comparable models. Even more astonishingly, OpenAI reportedly spent $1 billion on training ChatGPT, while DeepSeek did it with just about $6 million.If these claims hold true, it could fundamentally reshape the AI industry. The assumption that cutting-edge AI requires enormous computing power, billion-dollar investments, and thousands of advanced chips may no longer be valid. Instead, DeepSeek has demonstrated that a highly efficient approach to AI training could level the playing field—allowing smaller players, and even adversarial nations, to catch up quickly.Between the lines: How did DeepSeek pull this off?1. A different approach to AI training:DeepSeek’s success hinges on an unconventional training method. Traditional AI models process vast amounts of data, requiring massive computational resources. DeepSeek’s approach is different—it prioritizes knowing where to look for answers rather than memorizing everything.This is akin to a search engine librarian:Traditional AI models behave like a librarian who has read every book in the library and pulls from memory to answer questions. DeepSeek’s model doesn’t read every book in advance. Instead, it quickly finds the right book when asked a question—a more efficient and cost-effective approach. This technique, combined with a “mixture of experts” strategy—which assigns specialized AI models to different types of questions—dramatically reduces the computing power needed to train AI.2. A strategic stockpile of Nvidia chipsWhile the US tried to block China from accessing advanced AI chips, DeepSeek found a loophole. Before Washington closed the door in 2023, DeepSeek and other Chinese firms stockpiled tens of thousands of Nvidia’s A100 and H800 chips.These chips, though slightly less powerful than Nvidia’s cutting-edge H100, were still good enough for training DeepSeek’s AI. Some experts, including Elon Musk, have speculated that DeepSeek may have secretly acquired more high-end chips than disclosed.3. China’s ‘AI super geeks’ approachThe brains behind DeepSeek is a 39-year-old Chinese hedge funder named Liang Wenfeng. As per a CNN report, in interviews with the state-linked financial outlet Yicai, early business associates described the future DeepSeek founder as somewhat “nerdy” and recalled “a terrible haircut” he once had. Liang Wenfeng believes in hiring fresh graduates over experienced professionals. His rationale?Experienced engineers follow conventional approaches. Young engineers are more willing to experiment and think outside the box. This philosophy appears to have worked. Liang’s team of under 140 researchers—mostly graduates from China’s elite universities—designed a breakthrough AI system in record time. Their success has boosted morale in China’s tech industry, with Liang now celebrated as one of China’s &#8220;AI heroes.&#8221;A Trojan horse?The rise of DeepSeek raises a troubling question for Washington: Could this technology be a Trojan horse? Under Chinese law, all tech companies are required to “cooperate with national intelligence efforts.” This means that any data fed into DeepSeek’s chatbot could, in theory, be accessible to the Chinese state.Some analysts fear that DeepSeek’s chatbot could function as an AI-powered intelligence-gathering tool, subtly siphoning off user information under the guise of casual interactions. Others point to its built-in political censorship—when asked about topics such as the Tiananmen Square massacre or China’s treatment of Uyghurs, DeepSeek either refuses to answer or echoes the Chinese government’s official stance.With TikTok already under scrutiny in the US for alleged data harvesting, DeepSeek’s explosive rise is triggering similar concerns. The White House’s AI task force is reportedly investigating whether the platform poses a national security risk. “We should be alarmed,” said Ross Burley, co-founder of the Centre for Information Resilience. “Beijing has repeatedly weaponized its tech dominance for surveillance, control, and coercion,” Burley told Guardian.What they’re sayingDeepSeek’s sudden rise has sparked polarizing reactions across the tech world:President Donald Trump: Called DeepSeek a &#8220;wake-up call&#8221; for US industries, warning that China’s AI leap could threaten American economic and national security.Sam Altman (OpenAI CEO): Praised DeepSeek as “impressive” and welcomed the competition, despite previously framing AI development as a battle between democracy and authoritarianism.Elon Musk: Expressed doubts, suggesting that DeepSeek downplayed its use of banned US chips.US National Security Officials: Are now investigating potential data security risks, given that DeepSeek stores user data on servers in China—raising concerns about state surveillance.What’s next? DeepSeek’s rise throws a wrench into Washington’s AI strategy. The US has aggressively restricted China’s access to advanced AI chips, believing that these controls would slow Beijing’s progress. But DeepSeek’s breakthrough suggests that China has found ways to innovate despite these restrictions.Stronger US export controls: The Trump administrations will likely tighten AI-related export bans.Restrictions may expand beyond hardware, targeting software and cloud-based AI services to further limit China’s progress.Increased scrutiny of Chinese AI companies: US and European regulators may investigate DeepSeek’s funding sources to determine whether the Chinese government played a role.Concerns over data security could lead to bans on DeepSeek’s AI in Western markets, similar to TikTok’s ongoing legal battles.US push for AI supremacy: The DeepSeek shock may accelerate government funding for AI innovation.Trump’s $500 billion &#8220;Stargate&#8221; AI initiative could gain bipartisan support as Washington scrambles to ensure US leadership in AI.Bottom lineDeepSeek’s AI breakthrough signals a shift in the global AI landscape. While it’s unclear if China will ultimately overtake the US in AI, one thing is certain: Silicon Valley’s AI monopoly is no longer guaranteed.(With inputs from agencies)<br />
<br />
<br /><a href="https://timesofindia.indiatimes.com/business/international-business/deepseek-how-a-nerd-with-terrible-haircut-shook-up-silicon-valleys-ai-race/articleshow/117693138.cms" target="_blank" rel="noopener">Source link </a></p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.digiteex.com/deepseek-how-a-nerd-with-terrible-haircut-shook-up-silicon-valleys-ai-race/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">4475</post-id>	</item>
		<item>
		<title>DeepSeek: Everything you need to know about the AI chatbot app</title>
		<link>https://www.digiteex.com/deepseek-everything-you-need-to-know-about-the-ai-chatbot-app/</link>
					<comments>https://www.digiteex.com/deepseek-everything-you-need-to-know-about-the-ai-chatbot-app/#respond</comments>
		
		<dc:creator><![CDATA[digitex]]></dc:creator>
		<pubDate>Wed, 29 Jan 2025 00:29:13 +0000</pubDate>
				<category><![CDATA[AI]]></category>
		<category><![CDATA[AI algorithms]]></category>
		<category><![CDATA[Alibaba]]></category>
		<category><![CDATA[Chinese AI]]></category>
		<category><![CDATA[RCA H100 Portable Audio Device]]></category>
		<category><![CDATA[Sam Altman]]></category>
		<category><![CDATA[Taiwan]]></category>
		<category><![CDATA[the New York Times]]></category>
		<guid isPermaLink="false">https://www.digiteex.com/deepseek-everything-you-need-to-know-about-the-ai-chatbot-app/</guid>

					<description><![CDATA[DeepSeek has gone viral. Chinese AI lab DeepSeek broke into the mainstream consciousness this week after its chatbot app rose to the top of the Apple App Store charts. DeepSeek’s AI models, which were trained using compute-efficient techniques, have led Wall Street analysts — and technologists — to question whether the U.S. can maintain its lead in the AI race [&#8230;]]]></description>
										<content:encoded><![CDATA[
<br />
DeepSeek has gone viral.</p>
<p>Chinese AI lab DeepSeek broke into the mainstream consciousness this week after its chatbot app rose to the top of the Apple App Store charts. DeepSeek’s AI models, which were trained using compute-efficient techniques, have led Wall Street analysts — and technologists — to question whether the U.S. can maintain its lead in the AI race and whether the demand for AI chips will sustain.</p>
<p>But where did DeepSeek come from, and how did it rise to international fame so quickly?</p>
<p>DeepSeek’s trader origins</p>
<p>DeepSeek is backed by High-Flyer Capital Management, a Chinese quantitative hedge fund that uses AI to inform its trading decisions.</p>
<p>AI enthusiast Liang Wenfeng co-founded High-Flyer in 2015. Wenfeng, who reportedly began dabbling in trading while a student at Zhejiang University, launched High-Flyer Capital Management as a hedge fund in 2019 focused on developing and deploying AI algorithms.</p>
<p>In 2023, High-Flyer started DeepSeek as a lab dedicated to researching AI tools separate from its financial business. With High-Flyer as one of its investors, the lab spun off into its own company, also called DeepSeek.</p>
<p>From day one, DeepSeek built its own data center clusters for model training. But like other AI companies in China, DeepSeek has been affected by U.S. export bans on hardware. To train one of its more recent models, the company was forced to use Nvidia H800 chips, a less-powerful version of a chip, the H100, available to U.S. companies.</p>
<p>DeepSeek’s technical team is said to skew young. The company reportedly aggressively recruits doctorate AI researchers from top Chinese universities. DeepSeek also hires people without any computer science background to help its tech better understand a wide range of subjects, per The New York Times.</p>
<p>DeepSeek’s strong models</p>
<p>DeepSeek unveiled its first set of models — DeepSeek Coder, DeepSeek LLM, and DeepSeek Chat — in November 2023. But it wasn’t until last spring, when the startup released its next-gen DeepSeek-V2 family of models, that the AI industry started to take notice.</p>
<p>DeepSeek-V2, a general-purpose text- and image-analyzing system, performed well in various AI benchmarks — and was far cheaper to run than comparable models at the time. It forced DeepSeek’s domestic competition, including ByteDance and Alibaba, to cut the usage prices for some of their models, and make others completely free.</p>
<p>DeepSeek-V3, launched in December 2024, only added to DeepSeek’s notoriety.</p>
<p>According to DeepSeek’s internal benchmark testing, DeepSeek V3 outperforms both downloadable, openly available models like Meta’s Llama and “closed” models that can only be accessed through an API, like OpenAI’s GPT-4o. </p>
<p>Equally impressive is DeepSeek’s R1 “reasoning” model. Released in January, DeepSeek claims R1 performs as well as OpenAI’s o1 model on key benchmarks.</p>
<p>Being a reasoning model, R1 effectively fact-checks itself, which helps it to avoid some of the pitfalls that normally trip up models. Reasoning models take a little longer — usually seconds to minutes longer — to arrive at solutions compared to a typical non-reasoning model. The upside is that they tend to be more reliable in domains such as physics, science, and math.</p>
<p>There is a downside to R1, DeepSeek V3, and DeepSeek’s other models, however. Being Chinese-developed AI, they’re subject to benchmarking by China’s internet regulator to ensure that its responses “embody core socialist values.” In DeepSeek’s chatbot app, for example, R1 won’t answer questions about Tiananmen Square or Taiwan’s autonomy.</p>
<p>A disruptive approach</p>
<p>If DeepSeek has a business model, it’s not clear what that model is, exactly. The company prices its products and services well below market value — and gives others away for free.</p>
<p>The way DeepSeek tells it, efficiency breakthroughs have enabled it to maintain extreme cost competitiveness. Some experts dispute the figures the company has supplied, however.</p>
<p>Whatever the case may be, developers have taken to DeepSeek’s models, which aren’t open source as the phrase is commonly understood but are available under permissive licenses that allow for commercial use. According to Clem Delangue, the CEO of Hugging Face, one of the platforms hosting DeepSeek’s models, developers on Hugging Face have created over 500 “derivative” models of R1 that have racked up 2.5 million downloads combined.</p>
<p>DeepSeek’s success against larger and more established rivals has been described as “upending AI” and ushering in “a new era of AI brinkmanship.” The company’s success was at least in part responsible for causing Nvidia’s stock price to drop by 18% on Monday, and for eliciting a public response from OpenAI CEO Sam Altman.</p>
<p>As for what DeepSeek’s future might hold, it’s not clear. Improved models are a given. But the U.S. government appears to be growing wary of what it perceives as harmful foreign influence.</p>
<p>TechCrunch has an AI-focused newsletter! Sign up here to get it in your inbox every Wednesday.</p>

<br /><a href="https://techcrunch.com/2025/01/28/deepseek-everything-you-need-to-know-about-the-ai-chatbot-app/" target="_blank" rel="noopener">Source link </a></p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.digiteex.com/deepseek-everything-you-need-to-know-about-the-ai-chatbot-app/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">4441</post-id>	</item>
		<item>
		<title>Why DeepSeek Could Change What Silicon Valley Believe About A.I.</title>
		<link>https://www.digiteex.com/why-deepseek-could-change-what-silicon-valley-believe-about-a-i/</link>
					<comments>https://www.digiteex.com/why-deepseek-could-change-what-silicon-valley-believe-about-a-i/#respond</comments>
		
		<dc:creator><![CDATA[digitex]]></dc:creator>
		<pubDate>Tue, 28 Jan 2025 12:14:38 +0000</pubDate>
				<category><![CDATA[AI]]></category>
		<category><![CDATA[America]]></category>
		<category><![CDATA[artificial intelligence]]></category>
		<category><![CDATA[DeepSeek]]></category>
		<category><![CDATA[H100 chips]]></category>
		<category><![CDATA[OpenAI]]></category>
		<category><![CDATA[RCA H100 Portable Audio Device]]></category>
		<category><![CDATA[the New York Times]]></category>
		<guid isPermaLink="false">https://www.digiteex.com/why-deepseek-could-change-what-silicon-valley-believe-about-a-i/</guid>

					<description><![CDATA[The artificial intelligence breakthrough that is sending shock waves through stock markets, spooking Silicon Valley giants, and generating breathless takes about the end of America’s technological dominance arrived with an unassuming, wonky title: “Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.”The 22-page paper, released last week by a scrappy Chinese A.I. start-up called DeepSeek, didn’t [&#8230;]]]></description>
										<content:encoded><![CDATA[
<br />The artificial intelligence breakthrough that is sending shock waves through stock markets, spooking Silicon Valley giants, and generating breathless takes about the end of America’s technological dominance arrived with an unassuming, wonky title: “Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.”The 22-page paper, released last week by a scrappy Chinese A.I. start-up called DeepSeek, didn’t immediately set off alarm bells. It took a few days for researchers to digest the paper’s claims, and the implications of what it described. The company had created a new A.I. model called DeepSeek-R1, built by a team of researchers who claimed to have used a modest number of second-rate A.I. chips to match the performance of leading American A.I. models at a fraction of the cost.DeepSeek said it had done this by using clever engineering to substitute for raw computing horsepower. And it had done it in China, a country many experts thought was in a distant second place in the global A.I. race.Some industry watchers initially reacted to DeepSeek’s breakthrough with disbelief. Surely, they thought, DeepSeek had cheated to achieve R1’s results, or fudged their numbers to make their model look more impressive than it was. Maybe the Chinese government was promoting propaganda to undermine the narrative of American A.I. dominance. Maybe DeepSeek was hiding a stash of illicit Nvidia H100 chips, banned under U.S. export controls, and lying about it. Maybe R1 was actually just a clever re-skinning of American A.I. models that didn’t represent much in the way of real progress.Eventually, as more people dug into the details of DeepSeek-R1 — which, unlike most leading A.I. models, was released as open-source software, allowing outsiders to examine its inner workings more closely — their skepticism morphed into worry.And late last week, when lots of Americans started to use DeepSeek’s models for themselves, and the DeepSeek mobile app hit the number one spot on Apple’s App Store, it tipped into full-blown panic.I’m skeptical of the most dramatic takes I’ve seen over the past few days — such as the claim, made by one Silicon Valley investor, that DeepSeek is an elaborate plot by the Chinese government to destroy the American tech industry. I also think it’s plausible that the company’s shoestring budget has been badly exaggerated, or that it piggybacked on advancements made by American A.I. firms in ways it hasn’t disclosed.But I do think that DeepSeek’s R1 breakthrough was real. Based on conversations I’ve had with industry insiders, and a week’s worth of experts poking around and testing the paper’s findings for themselves, it appears to be throwing into question several major assumptions the American tech industry has been making.The first is the assumption that in order to build cutting-edge A.I. models, you need to spend huge amounts of money on powerful chips and data centers.It’s hard to overstate how foundational this dogma has become. Companies like Microsoft, Meta and Google have already spent tens of billions of dollars building out the infrastructure they thought was needed to build and run next-generation A.I. models. They plan to spend tens of billions more — or, in the case of OpenAI, as much as $500 billion through a joint venture with Oracle and SoftBank that was announced last week.DeepSeek appears to have spent a small fraction of that building R1. We don’t know the exact cost, and there are plenty of caveats to make about the figures they’ve released so far. It’s almost certainly higher than $5.5 million, the number the company claims it spent training a previous model.But even if R1 cost 10 times more to train than DeepSeek claims, and even if you factor in other costs they may have excluded, like engineer salaries or the costs of doing basic research, it would still be orders of magnitude less than what American A.I. companies are spending to develop their most capable models.The obvious conclusion to draw is not that American tech giants are wasting their money. It’s still expensive to run powerful A.I. models once they’re trained, and there are reasons to think that spending hundreds of billions of dollars will still make sense for companies like OpenAI and Google, which can afford to pay dearly to stay at the head of the pack.But DeepSeek’s breakthrough on cost challenges the “bigger is better” narrative that has driven the A.I. arms race in recent years by showing that relatively small models, when trained properly, can match or exceed the performance of much bigger models.That, in turn, means that A.I. companies may be able to achieve very powerful capabilities with far less investment than previously thought. And it suggests that we may soon see a flood of investment into smaller A.I. start-ups, and much more competition for the giants of Silicon Valley. (Which, because of the enormous costs of training their models, have mostly been competing with each other until now.)There are other, more technical reasons that everyone in Silicon Valley is paying attention to DeepSeek. In the research paper, the company reveals some details about how R1 was actually built, which include some cutting-edge techniques in model distillation. (Basically, that means compressing big A.I. models down into smaller ones, making them cheaper to run without losing much in the way of performance.)DeepSeek also included details that suggested that it had not been as hard as previously thought to convert a “vanilla” A.I. language model into a more sophisticated reasoning model, by applying a technique known as reinforcement learning on top of it. (Don’t worry if these terms go over your head — what matters is that methods for improving A.I. systems that were previously closely guarded by American tech companies are now out there on the web, free for anyone to take and replicate.)Even if the stock prices of American tech giants recover in the coming days, the success of DeepSeek raises important questions about their long-term A.I. strategies. If a Chinese company is able to build cheap, open-source models that match the performance of expensive American models, why would anyone pay for ours? And if you’re Meta — the only U.S. tech giant that releases its models as free open-source software — what prevents DeepSeek or another start-up from simply taking your models, which you spent billions of dollars on, and distilling them into smaller, cheaper models that they can offer for pennies?DeepSeek’s breakthrough also undercuts some of the geopolitical assumptions many American experts had been making about China’s position in the A.I. race.First, it challenges the narrative that China is meaningfully behind the frontier, when it comes to building powerful A.I. models. For years, many A.I. experts (and the policymakers who listen to them) have assumed that the United States had a lead of at least several years, and that copying the advancements made by American tech firms was prohibitively hard for Chinese companies to do quickly.But DeepSeek’s results show that China has advanced A.I. capabilities that can match or exceed models from OpenAI and other American A.I. companies, and that breakthroughs made by U.S. firms may be trivially easy for Chinese firms — or, at least, one Chinese firm — to replicate in a matter of weeks.(The New York Times has sued OpenAI and its partner, Microsoft, accusing them of copyright infringement of news content related to A.I. systems. OpenAI and Microsoft have denied those claims.)The results also raise questions about whether the steps the U.S. government has been taking to limit the spread of powerful A.I. systems to our adversaries — namely, the export controls used to prevent powerful A.I. chips from falling into China&#8217;s hands — are working as designed, or whether those regulations need to adapt to take into account new, more efficient ways of training models.And, of course, there are concerns about what it would mean for privacy and censorship if China took the lead in building powerful A.I. systems used by millions of Americans. Users of DeepSeek’s models have noticed that they routinely refuse to respond to questions about sensitive topics inside China, such as the Tiananmen Square massacre and Uyghur detention camps. If other developers build on top of DeepSeek’s models, as is common with open-source software, those censorship measures may get embedded across the industry.Privacy experts have also raised concerns about the fact that data shared with DeepSeek models may be accessible by the Chinese government. If you were worried about TikTok being used as an instrument of surveillance and propaganda, the rise of DeepSeek should worry you, too.I’m still not sure what the full impact of DeepSeek’s breakthrough will be, or whether we will consider the release of R1 a “Sputnik moment” for the A.I. industry, as some have claimed.But it seems wise to take seriously the possibility that we are in a new era of A.I. brinkmanship now — that the biggest and richest American tech companies may no longer win by default, and that containing the spread of increasingly powerful A.I. systems may be harder than we thought.At the very least, DeepSeek has shown that the A.I. arms race is truly on, and that after several years of dizzying progress, there are still more surprises left in store.<br />
<br />
<br /><a href="https://www.nytimes.com/2025/01/28/technology/why-deepseek-could-change-what-silicon-valley-believes-about-ai.html" target="_blank" rel="noopener">Source link </a></p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.digiteex.com/why-deepseek-could-change-what-silicon-valley-believe-about-a-i/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">4365</post-id>	</item>
		<item>
		<title>DeepSeek just blew up the AI industry’s narrative that it needs more money and power</title>
		<link>https://www.digiteex.com/deepseek-just-blew-up-the-ai-industrys-narrative-that-it-needs-more-money-and-power/</link>
					<comments>https://www.digiteex.com/deepseek-just-blew-up-the-ai-industrys-narrative-that-it-needs-more-money-and-power/#respond</comments>
		
		<dc:creator><![CDATA[digitex]]></dc:creator>
		<pubDate>Tue, 28 Jan 2025 11:04:55 +0000</pubDate>
				<category><![CDATA[AI]]></category>
		<category><![CDATA[America]]></category>
		<category><![CDATA[artificial intelligence]]></category>
		<category><![CDATA[far fewer chips]]></category>
		<category><![CDATA[H100 chips]]></category>
		<category><![CDATA[Mile Island]]></category>
		<category><![CDATA[RCA H100 Portable Audio Device]]></category>
		<category><![CDATA[Sam Altman]]></category>
		<category><![CDATA[Trump administration]]></category>
		<guid isPermaLink="false">https://www.digiteex.com/deepseek-just-blew-up-the-ai-industrys-narrative-that-it-needs-more-money-and-power/</guid>

					<description><![CDATA[A version of this story appeared in CNN Business’ Nightcap newsletter. To get it in your inbox, sign up for free here. New York CNN  —  The story of AI in the 2020s has gone something like this: Sam Altman: Look, a toy that can write your book report. VCs: This will fix everything! Doomers: [&#8230;]]]></description>
										<content:encoded><![CDATA[
</p>
<p> A version of this story appeared in CNN Business’ Nightcap newsletter. To get it in your inbox, sign up for free here.</p>
<p>  New York<br />
  CNN<br />
   — </p>
<p>   The story of AI in the 2020s has gone something like this:</p>
<p>  Sam Altman: Look, a toy that can write your book report.<br />
  VCs: This will fix everything!<br />
  Doomers: This will ruin everything.<br />
  Tech: We need money!<br />
  Everyone else: Could we maybe not destroy the environment over this?<br />
  Tech: Let’s restart Three Mile Island.<br />
  Tech: We need money!<br />
  Wall Street: Where’s our return?<br />
  Tech: (Chants) More power! More power! More power!</p>
<p>   And finally, in the year 2025, here comes DeepSeek to blow up the industry’s whole narrative about AI’s bottomless appetite for power, and potentially break the spell that had kept Wall Street funneling money to anyone with the words “harnessing artificial intelligence” in their pitch deck.</p>
<p>   ICYMI: DeepSeek dropped a bomb known as R1 that’s got all of Silicon Valley and much of Wall Street in a tizzy.</p>
<p>   The Chinese company’s large language model is basically a cheaper, more efficient ChatGPT, built on a fraction of OpenAI’s budget and using far fewer chips than any other leading chatbot.</p>
<p>   “That is a massive earthquake in the AI sector,” Gil Luria, head of tech research at investment group D.A. Davidson, told me. “Everybody is looking at it and saying, ‘We didn’t think this is possible. And since it is possible, we have to rethink everything that we have been planning.’”</p>
<p>   Suddenly, all that money and computing power that the Sam Altmans, Mark Zuckerbergs and Elon Musks have been saying are crucial to their AI projects — and thus America’s continued leadership in the industry — may end up being wildly overblown.</p>
<p>   DeepSeek, which on Monday climbed to No. 1 on the Apple app store, claims to have built its base model for less than $6 million (versus the more than $100 million Altman has said it cost to build GPT-4).</p>
<p>   It also claims to have used just 2,000 Nvidia chips that it obtained before US export restrictions were put in place. (OpenAI says it used 25,000 of the more powerful Nvidia H100 chips to build GPT-4.)</p>
<p>   It’s awkward timing for the Trump administration, which last week announced a half-trillion-dollar private-sector investment to build more data centers and keep the United States ahead of China in the AI race. (Oops!)</p>
<p>   And it’s incredibly bad news for Nvidia, the American chip maker powering the AI gold rush. Nvidia shares sank 17% Monday, shedding $600 billion in market cap in a single session — the biggest one-day loss for a single stock in history. Alphabet, Microsoft, Oracle, TSMC and plenty of others sank, and because tech stocks are so dominant, that dragged the broader stock market down, too.</p>
<p>   The tech-heavy Nasdaq plunged by 3% and the broader S&amp;P 500 fell 1.5%. (The Dow, buoyed by health care and consumer companies, ended the day up less than 1%.)</p>
<p>   Of course, one bad day on Wall Street does not an apocalypse make. (That’s for later, when one of these AI labs creates superintelligent murder bots. Kidding! Kind of.)</p>
<p>   But DeepSeek is forcing investors to take a beat and question tech companies’ assumptions. By its own reasoning, the AI industry needed to keep increasing “compute” (or computational power), which meant buying tens of thousands of Nvidia’s state-of-the-art chips and building giant data centers.</p>
<p>   “DeepSeek makes it very clear that that the current trajectory of scaling up of data centers is highly unlikely to be economic to Nvidia’s customers,” Luria said.</p>
<p>   The AI industry, and OpenAI in particular, has been going down two paths at once.</p>
<p>   There’s the business of designing AI models with better algorithms and sounder reasoning — the kind of stuff that requires “finesse, as opposed to brute force,” Luria says. And then there’s the Stargate path of giant energy investments.</p>
<p>   The first task is still “valid and important,” while the second path looks “ridiculous,” Luria said. “DeepSeek makes it clear that that scale and that spend would be, at the very least, wasteful.”</p>
<p>   In other words, AI isn’t dead. But the landscape is shifting faster than anyone, perhaps most of all Nvidia, expected.</p>
<p>  Picks and shovels</p>
<p>   Nvidia has become the ultimate “picks and shovels” play on Wall Street, transforming it into a $3 trillion company in the span of a couple of years. Up until now, the demand for Nvidia chips appeared boundless — tech companies were going to keep gobbling them up faster than Nvidia could produce them.</p>
<p>   But if DeepSeek really did manage to build a ChatGPT competitor using a handful of old processors, then maybe Nvidia’s tech customers soon won’t need as many as they’d thought. Many on Wall Street seemed to think so Monday as the stock went into a tailspin. (For its part, Nvidia seemed to shrug at the selloff in a statement to Bloomberg, calling DeepSeek’s model an “excellent AI advancement” that “illustrates how new models can be created.”)</p>
<p>   It’s also good to keep in mind that Wall Street is prone to tantrums, which is how some tech investors chalked up Monday’s selloff.</p>
<p>   “At the end of the day, there is only one chip company in the world launching autonomous, robotics, and broader AI use cases, and that is Nvidia,” Wedbush analysts wrote in a letter to clients. “Launching a competitive LLM model for consumer use cases is one thing… launching broader AI infrastructure is a whole other ballgame, and nothing with DeepSeek makes us believe anything different.”</p>

<br /><a href="https://www.cnn.com/2025/01/28/business/deepseek-ai-nvidia-nightcap/index.html" target="_blank" rel="noopener">Source link </a></p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.digiteex.com/deepseek-just-blew-up-the-ai-industrys-narrative-that-it-needs-more-money-and-power/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">4300</post-id>	</item>
	</channel>
</rss>
