Grafana 正式发布 gcx 和 MCP 服务器,助力基于遥测的智能代理开发
Grafana Labs 宣布正式上线两款工具 gcx CLI 和 Grafana MCP 服务器,使 AI 编码代理能够在开发过程中实时查询可观测性数据。

Grafana Labs 宣布正式上线两款工具 gcx CLI 和 Grafana MCP 服务器,使 AI 编码代理能够在开发过程中实时查询可观测性数据。

Cloudflare 发布了一个 CI SDK,允许开发者使用 TypeScript 而不是 YAML 定义持续集成(CI)管道,并将每个步骤作为持久化的 Cloudflare Workflow 来运行。

IT 工程项目领导者最常见的假设之一,是报告的事故数量不断增加就意味着系统可靠性正在下降。然而,Great Circle 最近的一篇文章提出,实际情况往往恰恰相反:事故数量增加,反而可能意味着组织的事故管理文化正在改善。随着团队不断投入资源完善流程、工具、培训和运维规范,他们会更愿意正式地将那些过去大概只会被悄悄处理、甚至被隐瞒的事故报告上去。这样带来的结果就是组织对运营问题的可见性提高了,但却不一定意味着系统的健康状况恶化。

Cloudflare 近日推出缓存响应规则(Cache Response Rules),这是一套新的规则引擎,运行在源站返回响应之后、内容写入 Cloudflare 缓存之前。此前,缓存规则(Cache Rules)只能根据请求属性进行判断;缓存响应规则则新增了一个响应处理阶段,可以在响应进入缓存之前检查源站返回的内容。

(翻译)借助 NVIDIA DSX MaxLPS 最大化 AI 工厂每瓦性能
AI factories are power-constrained industrial systems. The question is no longer how many GPUs fit in a data center, but how much AI output each available…

(翻译)NVIDIA BlueField-4 驱动面向 Agentic AI 工厂的新型 Scale-In 网络基础设施
Traditional cloud infrastructure was designed for predictable, general-purpose workloads and standard interfaces. Agentic AI factories connect diverse users…

(翻译)同一集群,利用率提升33个百分点:改变的是顺序
A Blog post by Dharma-AI on Hugging Face

手头这台闲置的SurfacePro7在桌角放了挺长一段时间。当年买它看重的是二合一的便携形态,但放在Windows11下只要长时间亮屏就一定会过热、偶尔跳屏,基本处于不可用的状态。直到前阵子在网上看到 ...

(翻译)GPU 管理:为何闲置 GPU 是新的停场飞机
A Blog post by Dharma-AI on Hugging Face

(翻译)如何为NVIDIA AI工厂选择全栈可观测性
AI infrastructure spans multiple layers, from compute and networking to storage, orchestration, and applications. When performance degrades…

(翻译)前沿实验室智能体入侵剖析:2026年7月事件技术时间线
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
(翻译)使用 Amazon OpenSearch Service MCP Apps 实现智能体可观测性
Amazon OpenSearch Service now supports MCP Apps, which return interactive visualizations alongside your AI agent's text responses. Learn how a single, locally run MCP server lets your agent move from alert to trace to logs to root cause in one conversation, and how you can verify every step inline w