InfoQ 中文发布于 08/23 01:17

Cloudflare 推出 Cache Response Rules,在源站响应后进一步控制缓存

Cloudflare 近日推出缓存响应规则(Cache Response Rules),这是一套新的规则引擎,运行在源站返回响应之后、内容写入 Cloudflare 缓存之前。此前,缓存规则(Cache Rules)只能根据请求属性进行判断;缓存响应规则则新增了一个响应处理阶段,可以在响应进入缓存之前检查源站返回的内容。

查看原文
AWS Machine Learning Blog发布于 08/29 00:20

Spreading the load: How Salesforce met Multi-AZ HA with SageMaker Inference Components

(翻译)分摊负载:Salesforce 如何借助 SageMaker Inference Components 满足多可用区高可用

Learn how Salesforce used Amazon SageMaker AI Inference Component placement (the SchedulingConfig parameter) to distribute model copies across multiple Availability Zones, meeting their Multi-AZ high availability compliance requirements without sacrificing the cost efficiency of multi-model co-hosti

查看原文
Solidot发布于 08/31 23:07

Linux 7.3-rc1 释出

Solidot是至顶网的科技资讯网站,主要面对开源自由软件和关心科技资讯读者群,包括众多中国开源软件的开发者,爱好者和布道者。口号是“奇客的知识,重要的东西”。

查看原文
AWS Machine Learning Blog发布于 09/01 03:08

Build observable enterprise agentic retrieval using Managed Amazon Bedrock Knowledge Base with AWS CloudFormation

(翻译)使用 Amazon Bedrock 托管知识库与 AWS CloudFormation 构建可观测的企业级智能体检索

This post builds an enterprise agentic retrieval solution on the Amazon Bedrock Managed Knowledge Base and Amazon Bedrock AgentCore. An agent reasons, routes across multiple knowledge bases, and returns cited answers, with seven layers of observability and both on-demand and continuous evaluation, a

查看原文
Solidot发布于 09/01 15:30

Paint.NET 实验性支持 Wine/Linux

Solidot是至顶网的科技资讯网站,主要面对开源自由软件和关心科技资讯读者群,包括众多中国开源软件的开发者,爱好者和布道者。口号是“奇客的知识,重要的东西”。

查看原文
Solidot发布于 09/10 19:42

微软九月例行更新修复近千个 Bug

Solidot是至顶网的科技资讯网站,主要面对开源自由软件和关心科技资讯读者群,包括众多中国开源软件的开发者,爱好者和布道者。口号是“奇客的知识,重要的东西”。

查看原文
AWS Machine Learning Blog发布于 09/11 05:37

Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

(翻译)通过模型缓存减少 Amazon SageMaker HyperPod 上的推理冷启动

Amazon SageMaker HyperPod now supports model caching for inference, which pre-loads model weights and container images onto cluster nodes so pods read from local NVMe storage instead of downloading over the network. Learn how model caching cuts cold starts from tens of minutes to seconds, how it wor

查看原文