AI Practices AI实践

21h ago 21小时前

Fragments: September 8 碎片:9月8日

AI is drastically reducing the cost of content generation while verification costs remain unchanged, creating a dangerous imbalance in capability The ... AI生成成本大幅降低但验证成本未同步下降,自动化边界从"常规vs非常规工作"转向"可衡量vs不可衡量工作" "虚假效用"(counterfeit utility)现象:过度依赖AI自动化而使用不完整的效能度量,导致短期仪表板数据上升但长期风险累积 "空洞经济"(Hollow Economy)风险:在...

Hot 热度
62
Quality 质量
72
Impact 影响力
68
Conversational AI 对话系统 Image Generation 图像生成 Code Generation 代码生成 LLM 大模型 Evaluation 评测
23h ago 23小时前

Do you even need a presentation? 你真的需要演示文稿吗?

Most corporate communication should use linear documents rather than slide decks, as documents support coherent, scannable, and self-contained informa... 大多数情况下线性文档足以传递信息,幻灯片格式会导致信息碎片化且丢失上下文 Infodecks是Martin Fowler提出的轻量级阅读材料概念,用幻灯片软件制作图文结合的简洁文档 录制音视频可替代现场演示,让观众自主控制节奏;现场演示仅适用于需要实时互动的场景 演示是叙事艺术,幻灯片只是工具而非演...

Hot 热度
52
Quality 质量
65
Impact 影响力
50
Programming 编程 Research 科学研究
23h ago 23小时前

Everyone is Reading Lonesome Dove 大家都在读《孤独的大篷车》

Larry McMurtry's 1985 Pulitzer Prize-winning novel "Lonesome Dove" has experienced a significant cultural resurgence, with the author noting that five... 拉里·麦克默特里(Larry McMurtry)1985年荣获普利策奖的小说《孤独之鸽》(Lonesome Dove)经历了显著的文化复兴。作者指出,在一个Slack群组中,九位朋友里有五位正在阅读或计划阅读此书。 这一复兴似乎由多种因素共同推动:麦克默特里于2021年3月逝世、斯蒂芬·金在科伦秀(...

Hot 热度
52
Quality 质量
55
Impact 影响力
50
Research 科学研究
1d ago 1天前

Introducing CUDA Rust: Two Tracks for Writing GPU Kernels 介绍 CUDA Rust:编写 GPU 内核的两种路径

NVIDIA announced two Rust-based GPU kernel programming tracks: cuda-oxide (SIMT-style) and cutile-rs (Tile-based), closing the gap that previously for... NVIDIA推出CUDA Rust双轨方案:cuda-oxide(SIMT风格)和cutile-rs(Tile风格),允许用Rust原生编写GPU内核并编译为PTX 两个项目均在编译时强制内存安全:cuda-oxide使用DisjointSlice和launch contracts防止别名,cuti...

Hot 热度
65
Quality 质量
70
Impact 影响力
65
GPU GPU Programming 编程 Open Source 开源 Chip 芯片
4d ago 4天前

Deploy a multimodal WhatsApp ordering assistant with Amazon Bedrock AgentCore 使用 Amazon Bedrock AgentCore 部署多模态 WhatsApp 点餐助手

Amazon introduces a reference architecture for a multimodal WhatsApp ordering assistant using Bedrock AgentCore and Amazon Nova 2 models (Lite for tex... 基于 Amazon Bedrock AgentCore 与 Nova 2 模型构建多模态 WhatsApp 点餐助手,支持文本、语音笔记和实时语音通话三种交互渠道 通过统一后端与跨渠道记忆(AgentCore Memory)解决多渠道订单系统碎片化问题,实现同一客户在不同渠道的身份识别与历史连续性 ...

Hot 热度
62
Quality 质量
68
Impact 影响力
65
Agent Agent Multimodal 多模态 Deployment 部署 Conversational AI 对话系统 LLM 大模型
4d ago 4天前

Building a Memory-Driven Agent with NVIDIA NemoClaw 构建基于 NVIDIA NemoClaw 的记忆驱动智能体

NVIDIA NemoClaw enables a memory-driven Chief of Staff agent that maintains a structured self model of people, projects, priorities, and working patte... NVIDIA NemoClaw构建了记忆驱动的Chief of Staff Agent,通过"self model"维护人员、项目、优先级和工作模式的结构化知识 采用Markdown存储知识、SQLite记录义务和审计事件的设计,实现证据与判断的分离,提升Agent推理能力 通过意图门控机制优先处理...

Hot 热度
65
Quality 质量
72
Impact 影响力
68
Agent Agent LLM 大模型 GPU GPU
4d ago 4天前

Designing lifecycle policies for AgentCore memory 设计 AgentCore 记忆的生存周期策略

Amazon Bedrock AgentCore introduces memory lifecycle management to prevent long-running AI agents from accumulating outdated context that degrades res... AgentCore内存生命周期管理通过系统化评分、合并与剪枝,解决长期运行Agent的上下文累积问题 提出情景/语义/程序三类内存分类框架,对应差异化保留策略(30-60天/6-12个月/长期保留) 实现三阶段夜间工作流:TTL硬过期→相关性衰减评分→低分记忆合并/删除 提供完整AWS CDK部署方...

Hot 热度
58
Quality 质量
65
Impact 影响力
60
Agent Agent Deployment 部署 Security 安全
4d ago 4天前

Frontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson 前沿推理抵达边缘:如何在 NVIDIA Jetson 上部署和优化模型

Compact open models released in 2026 now deliver reasoning and agentic capabilities previously requiring data center-scale systems, making edge deploy... 2026年发布的紧凑型开源模型(Nemotron 3.5 Lightning、Qwen3.8-27B)已具备此前需大型数据中心才能实现的推理与智能体能力,可在NVIDIA Jetson边缘设备本地运行 Nemotron 3.5 Lightning采用MoE架构(300亿总参数/每token激活30亿...

Hot 热度
68
Quality 质量
65
Impact 影响力
62
Open Source 开源 LLM 大模型 Deployment 部署 Inference 推理 GPU GPU
4d ago 4天前

Build a Physical AI model factory with NVIDIA Cosmos 3 on SageMaker HyperPod 在 SageMaker HyperPod 上使用 NVIDIA Cosmos 3 构建物理 AI 模型工厂

NVIDIA Cosmos 3 is an omnimodal world foundation model using a Mixture-of-Transformers (MoT) design with per-layer joint attention between a reasoner ... NVIDIA Cosmos 3 采用 Mixture-of-Transformers (MoT) 架构,通过单层联合注意力机制实现推理器与生成器的深度融合,支持视频、图像、动作和声音的统一多模态处理 模型采用训练-推理非对称设计:训练时运行完整去噪流程并解码视频,推理时仅执行少量去噪步骤并跳过视频解...

Hot 热度
70
Quality 质量
72
Impact 影响力
68
Robotics 机器人 Autonomous Driving 自动驾驶 Training 训练 GPU GPU Deployment 部署
4d ago 4天前

Run agent-driven Amazon SageMaker HyperPod operations with InstantStart 使用 InstantStart 运行基于代理的 Amazon SageMaker HyperPod 操作

Amazon SageMaker HyperPod InstantStart is an open-source control plane that automates the multi-stage provisioning and management of foundation model ... HyperPod InstantStart 是开源控制平面,将 SageMaker HyperPod 集群的部署、容量管理、训练/推理工作负载和存储集成封装为可编排的 API。 提供 Web UI 与终端两种入口,AI 代理通过 MCP 工具调用同一后端,自动规划多阶段工作流并轮询异步 AWS 操作...

Hot 热度
62
Quality 质量
68
Impact 影响力
58
Agent Agent Training 训练 Deployment 部署 GPU GPU
4d ago 4天前

Customizing your knowledge base on Amazon Bedrock for large and complex documents using Amazon Textract 使用 Amazon Textract 自定义 Amazon Bedrock 知识库以处理大型复杂文档

Amazon Textract integrated with Amazon Bedrock enables accurate extraction of structured and unstructured content from complex, multi-page utility bil... Amazon Bedrock与Textract集成方案解决复杂多页文档(如公用事业账单)的信息提取难题 直接对原始文档使用RAG会导致关键信息遗漏、模型幻觉和跨格式性能不一致三大问题 Textract提供高精度文本提取、数据清洗去噪和上下文标注能力,显著提升RAG准确性 支持PDF、DOCX、TXT...

Hot 热度
60
Quality 质量
70
Impact 影响力
65
RAG 检索增强生成 LLM 大模型 Deployment 部署
4d ago 4天前

How Intuit built an agentic disaster recovery assistant with Amazon Bedrock Intuit 如何利用 Amazon Bedrock 构建智能体灾难恢复助手

Intuit built "EWOK Agent," an AI-powered disaster recovery assistant using Amazon Bedrock, to automate decision-making in failover scenarios that prev... Intuit基于Amazon Bedrock构建了EWOK Agent,将AI智能体能力集成到灾难恢复系统中,解决传统DR流程中依赖工程师"部落知识"进行决策的痛点 核心设计原则是"模型决定做什么,EWOK Agent确定性执行怎么做",通过分层架构实现AI推理与基础设施执行的清晰边界 EWOK系统...

Hot 热度
62
Quality 质量
70
Impact 影响力
65
Agent Agent Closed Source 闭源 LLM 大模型 Deployment 部署