- 资讯公开站Entrepreneur
The Biggest ChatGPT Update Yet Gives Entrepreneurs 7 New Ways to Grow Their Business Fast
OpenAI's always-on AI agents keep working after you close the chat, so you need to decide what to delegate and what to keep.
- 资讯公开站Ars Technica
Apple changes full-disk access permissions to curb abuse from AI agents
- 资讯公开站Semiconductor Engineering
HBF for High-Throughput LLM Serving (UC Berkeley, FuriosaAI)
Researchers at the UC Berkeley and FuriosaAI published a technical paper titled “Characterizing High Bandwidth Flash for LLM Serving.” Abstract: “Large language model (LLM) serving requires substantial memory to store mo…
- 资讯公开站rss_arxiv_cs_ai
Build2SPARQL: A Large-Scale Text-to-SPARQL Benchmark Dataset for Building Knowledge Graph Querying
arXiv:2610.00224v1 Announce Type: new Abstract: Building automation systems are increasingly represented as semantic knowledge graphs (KGs) using ontologies such as Brick and ASHRAE 223P, creating a machine-readable sub…
- 资讯公开站rss_arxiv_cs_ai
EviGraph: Proof-Carrying Selective Recommendation over Temporal Public-Service Knowledge Graphs
arXiv:2610.00212v1 Announce Type: new Abstract: Public-service recommendations require evidence that matches the requested service, scope, and date. Yet treating every missing detail as decisive can withhold useful reco…
- 资讯公开站rss_arxiv_cs_ai
Scientific Agents: Evaluating Profession-Specific System Prompts on Scientific Tasks
arXiv:2610.00084v1 Announce Type: new Abstract: Detailed profession-specific system prompts raise token use and estimated cost per response without a consistent accuracy gain. We evaluate Scientific Agents, an open-sour…
- 资讯公开站rss_arxiv_cs_ai
K-Dense BYOK: An Open-Source AI Research Assistant That Runs Locally and Keeps a Hash-Chained Lab Notebook
arXiv:2610.00074v1 Announce Type: new Abstract: K-Dense BYOK (bring your own keys) is a free, open-source AI research assistant for scientists in any field that runs on the researcher's own computer. The researcher supp…
- 资讯公开站Digitimes
Global annual AI server shipments, 2025-2026
Continued advances in top-tier LLM capabilities, coupled with the potential for business automation enabled by agentic AI, are prompting major North American cloud providers, Neo Clouds, and leading AI labs to accelerate…
- 资讯公开站Digitimes
Life after Meta: Manus regains independence, debuting Manus 2.0 and 'Cue' AI agent
Following its return to independent operations, AI startup Manus unveiled two new products in late September 2026: Manus 2.0 and "Cue," a personal agentic AI application. The product rollouts mark the company's first maj…
- 资讯公开站Digitimes
Meta hires MongoDB CEO to lead AI enterprise platform
Meta has launched a new business unit, "Meta Enterprise Platform," to package its artificial intelligence (AI) models, agent tools, and infrastructure for enterprise customers, and has tapped MongoDB chief executive CJ D…
- 论文公开站arXiv
InterEvolve: Test-Time Evolution of Reward Programs for Humanoid Loco-Manipulation
We study test-time evolution for humanoid loco-manipulation: solving tasks that a controller was never trained for by repurposing its existing skills, improving from its own attempts, and retaining what it learns, withou…
- 论文公开站arXiv
VISTA: A Visual Harness for Reasoning in an Interactive World
We show that multimodal models possess strong reasoning abilities and that an appropriate harness can unlock their potential to solve tasks across diverse interactive environments. We introduce VISTA, a visual harness th…
- 论文公开站arXiv
ScholarCatalyst: A Benchmark for Retrieving Papers That Inspire New Research
What makes great scientists great? Even as AI systems start to make progress on open problems, scientists remain far ahead of them at sensing which prior idea, buried in an ever-growing archive of research, a new problem…
- 论文公开站arXiv
Reconstruct, Practice, Go Real: Guided Self-Improvement for Embodied Agents
Building reliable robot capabilities across diverse tasks requires substantial human effort to develop and maintain skills, design rewards, and integrate perception with control. We present Reconstruct, Practice, Go Real…
- 论文公开站arXiv
KaliBench: A Fine-Grained Benchmark for Cybersecurity Tool Use on Kali Linux with Runtime-Free Verifiable Rewards
LLMs are increasingly applied to cybersecurity workflows, where they are expected to translate analysts' intent into tool invocations. However, existing evaluations focus on knowledge-based assessments or end-to-end agen…
- 资讯公开站EE Times
Qualcomm Doubles Down on Agentic AI at Snapdragon Summit 2026
- 资讯公开站rss_arxiv_cs_ai
SimTrace: Grounded Multimodal User Trajectories Generation for Online User Modeling
arXiv:2609.38397v1 Announce Type: new Abstract: Virtual clients offer a cost-effective approach to support applications such as A/B testing, recommender system development, and interface evaluation. However, building th…
- 资讯公开站rss_arxiv_cs_ai
Self-Evolving Harness on Multiple Tasks with the Agent as Its Own Optimizer
arXiv:2609.38372v1 Announce Type: new Abstract: A harness is the code around a language-model agent that organizes prompts, calls tools, manages context, and controls execution. As models grow stronger, recent work has …
- 资讯公开站Digitimes
Meta's Muse lifts server CPU demand, with AMD EPYC orders extending into 2028
Meta's rapid rollout of its Muse AI agent is emerging as an early test of how agentic AI could reshape server infrastructure demand, with CPU suppliers already seeing longer lead times and AMD's latest EPYC processors re…
- 资讯公开站Digitimes
Focus: The hidden price of Nvidia's AI agent security play
Nvidia's new AI agent security platform puts hardware at the centre of agent control, but it also forces enterprises to weigh stronger protection against higher infrastructure costs and deeper dependence on Nvidia's stac…
- 资讯公开站Digitimes
Samsung AI Forum 2026 puts agentic AI at center of transformation across devices, robotics, and chipmaking
Samsung Electronics brought global AI experts and researchers together in Seoul on September 30 for the Samsung AI Forum 2026, exploring how agentic AI and physical AI can deliver measurable business value across enterpr…
- 资讯公开站Digitimes
Nvidia ties AI agent security to BlueField-4, Vera CPU
Over the past year, Nvidia has repeatedly observed AI agents drifting beyond approved workflows. In some cases, agents quietly escalated read permissions to write access, sent requests to unauthorized endpoints, or click…
- 论文公开站arXiv
Cogentic:面向自动证明发现的多智能体编排
Cogentic: Multi-Agent Orchestration for Automated Proof Discovery
提出 Cogentic,一个用于开放研究问题自动证明发现的多智能体框架。前沿语言模型虽能单次生成优质数学想法,但面对需多路径探索的开放问题,单次生成往往不足。
- 论文公开站arXiv
WorldAuditBench:多模态智能体的交互式3D世界审计
WorldAuditBench: Interactive 3D World Auditing with Multimodal Agents
随着交互式3D世界被广泛用于研究智能行为,识别其中异常(如悬浮物体、可穿墙、与环境不一致的物体)的高效流程变得重要,本文提出相应基准。
- 论文公开站arXiv
Turbo Harness:实例自适应框架优化
Turbo Harness: Instance-Adaptive Harness Optimization
自动化搜索有效框架是实现智能体递归自我改进的重要一步。现有优化通常产出统一应用于所有任务的单一全局框架,但适用于某类任务的框架未必通用。