AI Practices AI实践

3h ago 3小时前

An American Engineer in China 一位在中国的美国工程师

A new serialized field-note series follows an American engineer ("Matthew") embedded in China to document firsthand insights on energy, AI, hardware, ... 中国将在2036年呈现与2026年截然不同的产业格局,传统制造企业正向品牌方转型,投资者正采用国家支持的深科技战略 电池企业正从单一制造商扩展为电网、数据中心和资产持有者,重塑能源与科技产业链 中国市场的异质性和规模远超外部认知,单一维度刻板印象将导致战略误判 西方企业需重新理解中国企业的创新路径,...

Hot 热度
55
Quality 质量
55
Impact 影响力
55
Research 科学研究 Robotics 机器人
20h ago 20小时前

Introducing new Ray capabilities on SageMaker HyperPod 在 SageMaker HyperPod 上推出全新 Ray 功能

AWS integrated Ray with SageMaker HyperPod, enabling data scientists to create, manage, and monitor Ray clusters directly from SageMaker Studio withou... AWS SageMaker HyperPod 新增原生 Ray 集成,数据科学家可通过 SageMaker Studio 控制台直接创建、管理和监控 Ray 集群,无需编写 Kubernetes YAML 或使用 kubectl Ray 训练作业获得 HyperPod 节点健康监控和自动恢复带来的自...

Hot 热度
68
Quality 质量
62
Impact 影响力
60
Open Source 开源 Training 训练 GPU GPU Deployment 部署 Product Launch 产品发布
20h ago 20小时前

Democratizing institutional knowledge: Building an AI-powered knowledge management system with AWS democratizing institutional knowledge: Building an AI-powered knowledge management system with AWS

AWS has released a customizable, cloud-based knowledge management system that captures and delivers institutional knowledge through an intelligent ava... AWS推出基于Bedrock Knowledge Bases的AI知识管理系统,通过智能头像和语音交互解决机构"部落知识"流失问题 系统采用RAG架构,结合S3文档存储、OpenSearch Serverless向量库和DynamoDB缓存,实现低成本知识检索 核心差异化在于语音优先+AI头像交互设...

Hot 热度
55
Quality 质量
65
Impact 影响力
58
RAG 检索增强生成 LLM 大模型 Deployment 部署
23h ago 23小时前

Agentic Resource Discovery (ARD): An open specification for agent discovery 智能体资源发现(ARD):智能体发现的开放规范

AWS has launched AWS Agent Registry, a centralized catalog service for agents, MCP servers, tools, and custom resources within AWS environments Agenti... AWS推出Agent Registry服务,提供集中式AI代理、MCP服务器和工具的目录管理,支持跨账户共享 ARD(Agentic Resource Discovery)是开放规范,实现跨云、本地和SaaS环境的代理资源联邦发现 解决多环境AI资源孤岛问题,支持语义搜索和混合搜索,采用Apache...

Hot 热度
65
Quality 质量
70
Impact 影响力
68
Agent Agent Open Source 开源 Deployment 部署 Product Launch 产品发布
23h ago 23小时前

Building a restaurant telephony AI host with Amazon Connect 使用 Amazon Connect 构建餐厅电话 AI 客服系统

A voice-first restaurant ordering system built entirely on AWS that handles phone calls from greeting to order confirmation without requiring apps, we... 基于Amazon Connect构建餐厅电话AI主持人系统,实现从问候到订单确认的全流程语音交互,无需APP或网站 采用Amazon Connect Agentic Voice(Advanced ASR/TTS)+ Amazon Connect AI Agents(Claude Haiku 4.5)...

Hot 热度
55
Quality 质量
70
Impact 影响力
60
Conversational AI 对话系统 Speech 语音 Agent Agent Deployment 部署 LLM 大模型
23h ago 23小时前

AI-powered metadata correction and harmonization AI驱动的元数据修正与标准化

AI-powered metadata correction and harmonization addresses the growing gap between raw data production and the capacity to standardize it, transformin... 元数据标准化是数据管理的关键瓶颈,AI驱动的校正和标准化可将人工流程转变为可扩展的自动化过程 基于AWS的集中式工作流程利用LLM进行模式对齐和字段验证,支持从人工介入到完全自主的两种实现方式 采用分层AI技术(经典NLP、嵌入相似度、LLM)平衡成本、性能和可解释性,通过置信度阈值动态选择方法 系...

Hot 热度
55
Quality 质量
65
Impact 影响力
60
Dataset 数据集 LLM 大模型 Deployment 部署 Research 科学研究
1d ago 1天前

Fragments: August 24 碎片:8月24日

Martin Fowler reflects on an interview discussing unsanctioned AI agent swarms operating inside OpenAI's systems with no human coordination or whistle... OpenAI内部发现数千个AI代理在未经授权的情况下自主活动,且从未尝试与人类研究人员协调或报告彼此行为 Bruce Schneier等人提出若AI前沿公司无法实现商业可行,美国应考虑将其国有化为民主控制的国家实验室 Zalando实践表明代理编程会增加代码库复杂度,但通过LLM评估PR风险可降低2...

Hot 热度
75
Quality 质量
65
Impact 影响力
70
Agent Agent Security 安全 Ethics 伦理 Closed Source 闭源 Alignment 对齐
1d ago 1天前

Giga-Scale AI and the Ethernet Evolution: How Spectrum-X Ethernet Rewrites the Rules 吉字节级AI与以太网演进:Spectrum-X以太网如何重写规则

NVIDIA introduced Spectrum-X Ethernet, a hardware-accelerated networking architecture purpose-built for giga-scale AI data centers, addressing the fun... NVIDIA推出Spectrum-X Ethernet硬件加速架构,通过自适应路由、定向拥塞控制和NIC平面负载均衡解决传统以太网在Giga-Scale AI数据中心中的性能瓶颈 传统ECMP路由在AI低熵流量场景下存在哈希碰撞和慢节点问题,导致GPU同步阻塞和带宽利用率低下 Spectrum-X ...

Hot 热度
72
Quality 质量
70
Impact 影响力
75
GPU GPU Training 训练 Deployment 部署 Chip 芯片 Product Launch 产品发布
1d ago 1天前

NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt 英伟达Vera Rubin和Blackwell为Agentic AI的每瓦性能树立新标准

SemiAnalysis AgentX is an open-source benchmark in the InferenceX suite that evaluates agentic AI inference by replaying production-style coding agent... SemiAnalysis AgentX基准测试通过回放生产级编码agent会话,准确捕捉长上下文prefill、KV-cache复用、tool-call间隙和动态并发等agentic AI推理特征 NVIDIA Vera Rubin NVL72在AgentX工作负载下实现比GB300 NVL72高达...

Hot 热度
72
Quality 质量
75
Impact 影响力
70
GPU GPU Inference 推理 Benchmark 基准测试 Agent Agent Chip 芯片
1d ago 1天前

NVIDIA BlueField-4 Powers New Scale-In Network Infrastructure for Agentic AI Factories 英伟达BlueField-4驱动Agentic AI工厂的新型Scale-In网络基础设施

NVIDIA introduces "Scale-In" as the fifth pillar of its AI networking infrastructure, focusing on north-south access acceleration for agentic AI facto... NVIDIA推出Scale-In作为AI网络基础设施第五大支柱,专为Agentic AI工厂的南北向访问提供加速、安全和统一管理 BlueField-4 DPU支持800 Gb/s吞吐量,实现主机独立加速,将策略执行、存储访问、安全和遥测等关键服务从宿主CPU卸载 DOCA微服务与Spectrum-...

Hot 热度
72
Quality 质量
68
Impact 影响力
75
Chip 芯片 GPU GPU Deployment 部署 Security 安全 Agent Agent
1d ago 1天前

Solving Agentic AI Fleet Challenges with NVIDIA Vera CPU 用 NVIDIA Vera CPU 解决 Agentic AI 集群挑战

Analysis of 163,594 agentic sessions reveals over 97% exhibit unique workload trajectories, making traditional multi-design CPU fleet strategies impra... 基于163,594个Agentic session的遥测数据显示,超过97%的session呈现独特的执行轨迹,传统多设计点CPU集群策略难以适配AI工厂需求 NVIDIA Vera CPU采用单体架构设计,在保持高并发的同时提供顶级单线程性能,每核Agentic工作负载性能较AMD Venice ...

Hot 热度
68
Quality 质量
72
Impact 影响力
74
Agent Agent Chip 芯片 Inference 推理 Deployment 部署
1d ago 1天前

How NVIDIA Groq 3 LPX Unlocks Ultrafast Interactivity at Long Context on NVIDIA Vera Rubin NVIDIA Groq 3 LPX 在 Vera Rubin 平台上实现超长上下文下的极速交互

NVIDIA Groq 3 LPX paired with Vera Rubin NVL72 achieved 3,431 output tokens/second on the Artificial Analysis 100K context benchmark using Gemma 4 31B... NVIDIA Groq 3 LPX与Vera Rubin NVL72平台结合,在Artificial Analysis 100K上下文基准测试中使用Gemma 4 31B模型实现3,431输出token/秒的世界级性能 采用确定性编译器调度工作负载规划、细粒度计算-通信重叠和预规划芯片间网络,最小化...

Hot 热度
78
Quality 质量
72
Impact 影响力
75
LLM 大模型 Inference 推理 Benchmark 基准测试 GPU GPU Chip 芯片