T
traeai
Sign in

模型

GPT-OSS 20B

210亿参数的开源稀疏计算基准模型。

已跟踪 5 条高相关材料

TraeAI 观察

相关材料

已收录 5 条与 GPT-OSS 20B 相关的内容,按评分排序。

Comprehensive observability for Amazon SageMaker AI LLM inference: From GPU utilization to LLM quality

AWS proposes a full-stack observability solution for SageMaker LLM inference, collecting infrastructure metrics (GPU utilization, latency) and custom quality metrics (response accuracy, compliance) via CloudWatch, visualized in Managed Grafana—enabling dual-dimension monitoring to address cases where systems appear healthy but produce poor outputs, or deliver high-quality responses inefficiently.

入选理由:SageMaker AI Inference 支持单 endpoint 多 inference components 部署(如 gpt-oss-20b + Qwen2.5-7B-Instruct),实现模型隔离与共享资源协同。

FeaturedArticle#LLM#Observability#Amazon SageMaker#CloudWatch#Grafana英文
MLCommons Releases MLPerf Training v6.0 Results

MLCommons Releases MLPerf Training v6.0 Results

MLCommons1195 字 (约 5 分钟)
85

MLPerf v6.0新增稀疏计算基准,推动AI系统多样性,DeepSeek V3参数达6710亿。

入选理由:DeepSeek V3使用6710亿参数,激活370亿参数/令牌,验证大规模稀疏训练系统。

FeaturedArticle#MLPerf#AI#MoE#基准测试#稀疏计算英文
Simon Willison's Weblog 图标

Prompt Injection as Role Confusion

Simon Willison's Weblog529 字 (约 3 分钟)
85

模型无法有效区分特权文本与用户输入,导致提示注入攻击风险显著增加。

入选理由:模型更关注文本风格而非内容,导致角色混淆。

FeaturedArticle#AI#LLM#安全#Prompt Injection英文
https://t.co/NrFGBIq9fx

https://t.co/NrFGBIq9fx

OpenRouter(@OpenRouterAI)34 字 (约 1 分钟)
60

OpenRouter 提供了 gpt-oss-20b 模型的免费 API 接口,并展示了其性能基准。

入选理由:gpt-oss-20b 模型可通过 OpenRouter 免费使用。

FeaturedTweet#AI模型#API#OpenRouter#gpt-oss-20b英文

跨材料问答 · GPT-OSS 20B

回答基于:GPT-OSS 20B 相关 5 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.