- 资讯公开站Digitimes
DIGITIMES Insight: AI servers will carry nearly twice as many CPUs per accelerator by 2027
Demand for agentic AI in commercial and consumer markets was rising rapidly, and the market appeal of consumer services such as Meta Muse was expanding. In agentic AI tasks, large language model (LLM) inference was mainl…
- 资讯公开站Digitimes
AMD to expand Taiwan supply chain investment as AI demand drives capacity needs
AMD CEO Lisa Su said the company plans to further increase investment in Taiwan's semiconductor supply chain as demand for CPUs, GPUs, and AI computing continues to outpace available capacity.
- 资讯公开站gnews_nvidia_supply
Calif. CEO Charged in $300M Nvidia GPU Smuggling Case [2026] - tech-insider.org
- 资讯公开站rss_arxiv_cs_ai
FlashSinkhorn 2: Block-Sparse Entropic Optimal Transport
arXiv:2610.02395v1 Announce Type: new Abstract: Streaming GPU solvers for entropic optimal transport (EOT), such as FlashSinkhorn, avoid storing the dense kernel but still evaluate all $n\times m$ point pairs in every S…
- 资讯公开站Digitimes
AM Intelligence expands India AI infrastructure push with 20,000 more Nvidia Rubin GPUs
AM Intelligence (AMI) is widening its AI infrastructure push in India and Malaysia, adding to a regional buildout that could matter well beyond Asia. The latest orders expand access to frontier computing for cloud provid…
- 资讯公开站Digitimes
Biren readies BR20X, sharpening China's AI GPU challenge to Nvidia
Biren Technology is preparing its next-generation GPU for mass production, advancing a domestic AI accelerator roadmap built around broader low-precision support and locally available manufacturing.
- 资讯公开站Digitimes
GMI Cloud secures NT$14.05 billion loan for AI factory
CTBC Bank and GMI Cloud on October 1 announced the closing of an NT$14.05 billion (about US$445 million) syndicated credit facility to fund Taiwan's first AI factory. The deal is the first in Taiwan to fully back GPU fin…
- 资讯公开站rss_arxiv_cs_ai
Decode-Latency Feedback Prefill: A Model-Free Controller and Its Generalization Limits
arXiv:2609.38386v1 Announce Type: new Abstract: Concurrent autoregressive inference creates a fundamental interference problem: prefilling a newly arrived long prompt can delay tokens for requests that are already decod…
- 资讯公开站Digitimes
Neweb Information targets late-October listing on Taipei Exchange
Neweb Information, a Taiwan-based IT systems integrator, is targeting a late-October listing on the Taipei Exchange. Long focused on enterprise core systems, the company has expanded in recent years into AI GPU computing…
- 资讯公开站Digitimes
2027年ASIC出货量超越GPU,ABF载板瓶颈转移
ASIC shipments top GPUs in 2027 as ABF substrate bottlenecks shift
2027年高端云AI加速器市场将发生结构性转变,ASIC总出货量首次超过GPU,受谷歌、亚马逊和华为产量上升推动,市场进入英伟达与ASIC双雄竞争格局。
- 资讯公开站rss_arxiv_cs_ai
PowerZooJax:面向强化学习的基于JAX的电力系统基准
PowerZooJax: A JAX-based Power System Benchmark for Reinforcement Learning
arXiv:2609.36052v1 公告类型:新 摘要:电力系统运行是一个安全关键的序贯决策问题,因此是强化学习(RL)的天然试验平台。然而,现有的电力系统RL环境往往范围狭窄,且受限于基于CPU的仿真工作流,难以进行大规模评估。我们提出PowerZooJax,一个基于JAX的电力系统运行RL基准套件。它提供五个约束马尔可夫决策过程任务,涵盖发电、输电、配电、分布式能源和数据中心微电网。通过将潮流计算、经济调度、市场出清和设备动态重写…
- 资讯公开站rss_arxiv_cs_ai
有效的大语言模型微调是否必须依赖人类可读文本?
Is Human-Readable Text Necessary for Effective LLM Fine-Tuning?
arXiv:2609.35868v1 公告类型:新 摘要:有效微调大语言模型是否必须依赖人类可读性?我们研究模型条件化的训练表示能否在无需人类可读文本形式的情况下保持或提升适配效用。我们提出 Desired-Update-Aligned Synthetic Data(DASA),利用冻结参考模型的激活梯度反馈来指导连续合成输入嵌入的优化。受激活梯度在局部风险降低中作用的启发,DASA 针对有用的适配更新,而非源文本重建或语言流畅性。所得…
- 资讯公开站rss_arxiv_cs_ai
面向资源受限边缘设备可靠推理的神经符号路由
Neurosymbolic Routing for Reliable Reasoning on Resource-Constrained Edge Devices
在边缘硬件上运行语言模型可在无网络连接的情况下提供私密、低延迟的推理,但适合此类设备的小模型在计算机本应擅长的任务(如算术、代数和形式逻辑问题)上并不可靠。我们认为这种不可靠性在很大程度上是可以避免的。许多看似需要推理的查询实际上是结构确定性的,可以用快速且精确的符号方法求解。因此,强迫概率模型去近似它们只会牺牲准确性和能耗而收益甚微。我们提出一种神经符号路由器,对每个传入查询进行分类,并将其分派给成本最低的正确求解器:将结构化任务发送…
- 资讯公开站Digitimes
大同集团目标180天部署AI数据中心
Tatung targets 180-day AI data center deployment
为展示AI基础设施竞争力,大同集团9月30日举办“Power Ready for AI”论坛,发布基于“CUBE AI Data Center”概念的模块化AI数据中心,集成电力、冷却、监控与计算。
- 资讯公开站Semiconductor Engineering
E系列GPU IP:迈向融合加速的第一步
E-Series GPU IP: The First Step Towards Converged Acceleration
在单一灵活架构和可编程软件栈上融合图形、计算与AI,每核在1GHz下提供高达32 TOPS Int8算力,并可同时运行图形和AI工作负载。
- 资讯公开站Digitimes
专栏:AI改变平衡,内存与逻辑碰撞
Column: Memory and logic collide as AI shifts the balance
随着HBM成本接近GPU-HBM CoWoS封装的一半,内存还能被视为AI计算的被动组件吗?答案要回到“内存墙”,这一挑战塑造了近半个世纪的半导体发展。
- 资讯公开站Digitimes
Meta与Firmus达成协议,在东南亚获取AI算力
Meta reaches agreements with Firmus on securing AI compute in Southeast Asia
Meta与澳大利亚新型云服务商Firmus达成多项协议,后者将在其东南亚AI工厂租赁GPU算力,支持Meta在亚太地区的AI扩张。
- 资讯公开站Semiconductor Engineering
芯片产业技术论文汇总:9月29日
Chip Industry Technical Paper Roundup: Sept. 29
涵盖A7 CFET与A10 NSFET对比、晶圆级亚5nm MoS₂晶体管、3D HI多千瓦供电、铜微结构与TSV残余应力、GPU Rowhammer攻击、门级RTL木马定位、LLM推理中HBM与主机内存并发访问等。
- 资讯公开站Digitimes
韩国AI算力紧张冲击研究人员与初创企业
South Korea AI compute squeeze hits researchers and startups
韩国AI行业面临日益严重的GPU短缺,研究机构和小企业难以获得足够的算力用于模型训练和部署。随着政府支持结束,研究团队发现训练更难持续,初创企业也受到挤压。
- 资讯公开站gnews_nvidia_supply
英伟达可能减少每颗GPU的HBM用量,但总体采购量增加
Nvidia Could Use Less HBM Per GPU—and Buy More Overall - drrobertcastellano.substack.com
- 资讯公开站gnews_nvidia_supply
英伟达 B200 价格达每小时 8.01 美元,GPU 成本走势分化
Nvidia B200 Price Hits $8.01/Hr as GPU Costs Diverge - shattered.io
- 资讯公开站rss_arxiv_cs_ai
BioEVAL:面向生物工程的全球多机构大型语言与多模态模型基准
BioEVAL: A global, multi-institutional benchmark of large language and multimodal models for bioengineering
arXiv:2609.30489v1 公告类型:新 摘要:大型语言模型(LLM)在通用推理方面已取得历史性突破,并在生物医学科学中初获成功。然而,现有 LLM 基准测试侧重事实回忆,对模型在前沿和多模态任务上的表现洞察有限。我们构建了 BioEVAL(AI 与 LLM 的生物工程验证),这是一项全球多机构计划,旨在评估生物工程(BE)各子领域的实验推理能力。BioEVAL 涵盖 11 个主要 BE 子领域及一组未分类项目,汇集 22 个…
- 资讯公开站iccircle_rss
【VISION GUIDE - 36】片上网络 NoC 大科普
【关键词】#前端 #NoC。【摘要】基于NoC的多处理器系统是一种使用网络互连的架构,将多核CPU、GPU、FPGA等处理器和加速器通过高带宽、低延迟的通信通道连接起来,实现高性能、可扩展的并行计算。它提供了灵活性、节能性和可靠性,适用于高性能计算、嵌入式系统等领域,加速图形处理、人工智能和机器学习等任务。
- 资讯公开站CNX Software
Critical Link MitySOM-AM62x/A/P:面向边缘AI与工业自动化的TI Sitara AM62 SO-DIMM模块
Critical Link MitySOM-AM62x/A/P – TI Sitara AM62 SO-DIMM SoMs for edge AI and industrial automation
摘要显示,Critical Link 发布 MitySOM-AM62x/A/P 系列工业级 SO-DIMM 系统模块,基于 TI Sitara AM62 处理器,面向嵌入式 Linux、工业自动化与 HMI。AM62 通用型搭载最高四核 1.4GHz Cortex-A53、400MHz Cortex-M4F MCU 及双 PRU;AM62A 增加 C7x DSP 与矩阵乘法加速器,算力最高 2 TOPS,并含 ISP;AM62P 配备 …
意义:为边缘 AI 与工业自动化提供可插拔 SO-DIMM 模块方案,AM62A 的 2 TOPS 与 ISP 适合视觉推理,AM62P 的 GPU 与 MIPI DSI 便于多媒体人机界面。
- 资讯公开站EE Times
Delos Data 以数据接口瞄准异构 AI
Delos Data Targets Heterogeneous AI with Data Interface
摘要显示,Delos Data 的 Apollo 芯片粒旨在桥接不同的端点语义与互连,构建一个跨越 GPU、加速器、CPU 和内存的低延迟域,从而应对异构 AI 计算场景。该方案试图解决多类型计算单元协同时的数据接口与延迟问题,但原文未披露具体性能指标或客户信息。
意义:异构计算中跨芯片数据接口与延迟是瓶颈,该方案若落地可简化多加速器协同,对 AI 基础设施开发者有参考价值。