Semiconductor News & Analysis Feed

20 articles
2026-07-27
news.google.com 2026-07-27
2026-07-26
news.google.com 2026-07-26
2026-07-21
blogs.nvidia.com 2026-07-21
2026-07-21
www.theregister.com 2026-07-21
2026-07-21
tomshardware.com 2026-07-21
According to a report by The Information, Google is reportedly developing a server chip codenamed 'Frozen v2,' which would integrate parts of the Gemini model's architecture directly into the silicon. This design aims to significantly boost the number of tokens processed per unit of energy, potentia
2026-07-19
en.sedaily.com 2026-07-19
2026-07-16
news.google.com 2026-07-16
2026-07-16
seekingalpha.com 2026-07-16
2026-07-09
www.benzinga.com 2026-07-09
According to Futurum Research analyst Shay Bolo, the semiconductor industry's core issue lies in the deep dependence of major global chip manufacturers on TSMC's manufacturing capabilities. This phenomenon reflects the industry's trend toward centralization and highlights TSMC's irreplaceable positi
2026-07-08
www.morningstar.com 2026-07-08
Parasail has announced a collaboration with d-Matrix to combine NVIDIA AI infrastructure with d-Matrix's Corsair accelerators, achieving up to 10x faster token generation. This marks one of the first large-scale commercial deployments of heterogeneous disaggregated inference in production. By pairin
2026-07-08
uk.finance.yahoo.com 2026-07-08
Parasail has partnered with d-Matrix to combine NVIDIA's AI infrastructure with d-Matrix's Corsair accelerators, achieving up to 10x faster inference performance. This integration targets token generation and aims to enhance both speed and cost-efficiency in AI inference services. By pairing NVIDIA
2026-07-08
www.prnewswire.com 2026-07-08
On July 8, 2026, Parasail, an AI inference service provider, and d-Matrix, a pioneer in low-latency AI inference platforms, announced a collaboration to combine d-Matrix's Corsair inference accelerators with NVIDIA Hopper and Blackwell GPUs, achieving up to a 10x increase in token generation speed.
2026-07-03
tomshardware.com 2026-07-03
Palantir CEO Alex Karp delivered a blistering critique of leading AI companies like OpenAI and Anthropic during a CNBC interview, accusing them of stealing customer data while charging for unproductive tokens. He argued that these firms are exploiting customer data to improve their models while simu
2026-07-02
www.tomshardware.com 2026-07-02
NVIDIA has introduced a new business model enabling it to generate revenue twice from the same hardware: once through initial hardware sales and again via a percentage of the ongoing revenue generated by its partners' cloud services. This 'revenue-sharing and credit-support model' aims to provide co
2026-07-02
www.coindesk.com 2026-07-02
Ondo Finance's launch of blockchain-based tokenized versions of BlackRock's iShares Core S&P 500 ETF and Micron Technology shares represents a significant convergence of traditional finance and blockchain technology. This innovation operates within the existing U.S. securities framework, adhering to
2026-07-02
news.google.com 2026-07-02
2026-07-01
diginomica.com 2026-07-01
Qualcomm's $3.9 billion acquisition of Modular marks a pivotal strategic shift in its AI data center ambitions. This move underscores Qualcomm's vision of a future where compute is no longer defined by a single chip but by a distributed, heterogeneous system. Modular's platform aims to unify the AI
2026-07-01
wccftech.com 2026-07-01
NVIDIA has achieved a groundbreaking 5x reduction in token cost for DeepSeek v4 AI models just one month after its launch, through full-stack inference software optimizations on its Blackwell GPU platform. This development underscores the importance of 'cost per token' as a key metric for AI total c
2026-07-01
wccftech.com 2026-07-01
NVIDIA has achieved a groundbreaking 5x reduction in token cost for DeepSeek v4 AI models just one month after its launch, thanks to full-stack inference software optimizations on its Blackwell GPU platform. This development underscores the importance of 'cost per token' as a key metric for AI total
2026-06-30
blogs.nvidia.com 2026-06-30
As organizations transition from AI pilot projects to full-scale AI factories, infrastructure decisions have shifted from peak chip specifications to cost per token—measuring how many useful tokens can be delivered per dollar, watt, and within required latency targets. NVIDIA’s inference software st