AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-07-07 0 浏览 会员

Hugging Face模型在Foundry托管计算上可用

微软在Build 2026上宣布Foundry托管计算支持一键部署Hugging Face开源模型,结合企业级安全与智能体服务。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 会员 POST / 2026-07-07 23:20:06

原贴

查看原文
作者:Hugging Face Blog 来源站点:huggingface.co 原贴时间:

原文

At Microsoft Build 2026, we announced Foundry Managed Compute and Hugging Face models on Foundry — a curated catalog of open-weight models from the Hugging Face ecosystem, refreshed weekly, deployable in one click onto Foundry Managed Compute. Weights are pre-staged in Azure, runtimes are built and scanned by Microsoft, and every model in the Collection ships with the same enterprise security, governance, observability, and billing that applies to every other model on Foundry. Microsoft Foundry is a platform for building and operating agentic AI applications. Foundry starts with the widest model selection on any cloud — models from Microsoft, OpenAI, Anthropic, Meta, Mistral, DeepSeek, Hugging Face, and others, spanning frontier, open-source, and custom weights — all accessible through a single endpoint and a single set of SDKs in Python, C#, JavaScript, and Java. On top of those models sits the Foundry Agent Service : multi-agent orchestration with built-in memory, knowledge grounding through Foundry IQ, and a catalog of connectable tools via agentic protocols, so agents can work with enterprise data. Once agents are running, Foundry provides end-to-end tracing, real-time monitoring, continuous evaluations, and a prompt optimizer that improves agent behavior based on eval results — observability and quality loops that are part of the platform. Alongside that, developers get access to: An AI Red Teaming Agent for adversarial testing Azure Policy integration directly within the platform Alongside pay-per-token (lowest-friction path to get started) and provisioned throughput (predictable, high-performance production workloads on frontier models), Foundry Managed Compute is the third deployment option in Foundry: a managed GPU platform-as-a-service for open-source and custom models. You deploy a model instance described by the things that matter to your workload — parameter count, context length, and whether you want to optimize for latency or throughput — and Foundry handles the GPU topology underneath, whether the instance lands on one accelerator or several, so you think and plan in model terms. Microsoft takes care of the machine: container updates, runtime upgrades, and security patches happen automatically on the supported runtimes — vLLM, SGLang, TensorRT-LLM, NIM, TEI, llama.cpp — without redeploying your model, while model configuration, deployment behavior, and routing stay with you. That consistency carries through the developer surface — pay-per-token, provisioned throughput, and Managed Compute share: Open-source models integrate with Foundry Agents the same way frontier models do, so you can mix model types in a single agent without a separate integration path. Global deployments — broadest capacity and best pricing

中文翻译

在Microsoft Build 2026上,我们宣布了Foundry托管计算和Foundry上的Hugging Face模型——来自Hugging Face生态系统的策划的开源权重模型目录,每周刷新,可一键部署到Foundry托管计算上。权重预置于Azure,运行时由Microsoft构建和扫描,集合中的每个模型都享有与企业安全、治理、可观测性和计费相同的特性,适用于Foundry上的所有其他模型。Microsoft Foundry是一个用于构建和运营智能体AI应用的平台。Foundry从任何云上最广泛的模型选择开始——来自Microsoft、OpenAI、Anthropic、Meta、Mistral、DeepSeek、Hugging Face等的模型,涵盖前沿、开源和自定义权重——全部通过单一端点和单一SDK集(Python、C#、JavaScript和Java)访问。在这些模型之上是Foundry智能体服务:多智能体编排,带有内置记忆、通过Foundry IQ的知识基础,以及通过智能体协议的可连接工具目录,使智能体能够与企业数据协作。一旦智能体运行,Foundry提供端到端追踪、实时监控、持续评估,以及基于评估结果改进智能体行为的提示优化器——作为平台一部分的可观测性和质量循环。除此之外,开发者可以访问:用于对抗性测试的AI红队智能体、平台内直接集成的Azure Policy。在按token付费(启动的最低摩擦路径)和预置吞吐量(前沿模型上可预测的高性能生产工作负载)的基础上,Foundry托管计算是Foundry中的第三个部署选项:一个用于开源和自定义模型的托管GPU平台即服务。您部署一个由对工作负载重要的因素(参数数量、上下文长度,以及是否优化延迟或吞吐量)描述的模型实例,Foundry处理底层的GPU拓扑,无论实例落在单个加速器还是多个加速器上,因此您可以按模型术语思考和规划。Microsoft负责机器:容器更新、运行时升级和安全补丁在支持的运行时(vLLM、SGLang、TensorRT-LLM、NIM、TEI、llama.cpp)上自动进行,无需重新部署模型,而模型配置、部署行为和路由仍由您保留。这种一致性贯穿开发者表面——按token付费、预置吞吐量和托管计算共享:开源模型与Foundry智能体集成的方式与前沿模型相同,因此您可以在单个智能体中混合模型类型,无需单独的集成路径。全球部署——最广泛的能力和最佳定价。

核心信息

微软在Build 2026上宣布Foundry托管计算支持一键部署Hugging Face开源模型,结合企业级安全与智能体服务。

  • 微软在Build 2026上宣布Foundry托管计算支持一键部署Hugging Face开源模型,结合企业级安全与智能体服务。
  • 原贴提到:At Microsoft Build 2026, we announced Foundry Managed Compute and Huggin
  • 来源:huggingface.co
试看内容

成为会员查看完整内容

你已经看到了这篇内容的前置整理,剩余深度部分仅对会员开放。

详细解读 信息差价值 参考来源
成为会员查看完整内容
上一篇 中国考虑对其顶尖AI模型实施出口限制,欧洲夹在中间 下一篇 腾讯Hy3模型