T
traeai
Sign in

人物

Philipp Schmid

别名:_philschmid

推文作者,技术博主

已跟踪 30 条高相关材料

TraeAI 观察

相关材料

已收录 30 条与 Philipp Schmid 相关的内容,按评分排序。

Managed Agents in the Gemini API. One API call gives you a sandboxed Linux with code execution, web ...

The Gemini API provides a sandboxed Linux environment through a single API call, supporting code execution, web access, and file I/O. The article offers a complete example to build a data science assistant.

入选理由:Gemini API 通过单个 API 调用提供沙盒化 Linux 环境

FeaturedTweet#Gemini API#Data Science中文
Gemini Managed Agents Dev Guide: 1 API call = Gemini 3.5 Flash + Antigravity Harness + remote Linux ...

Philipp Schmid published an article about the Gemini Managed Agents Dev Guide, introducing a feature that can be achieved through a single API call, including Gemini 3.5 Flash, Antigravity Harness, and a remote Linux sandbox, without any infrastructure or orchestration.

入选理由:通过单个 API 调用实现功能,包括 Gemini 3.5 Flash、Antigravity Harness 和远程 Linux 沙盒。

FeaturedTweet#Gemini#API#DevGuide中文
Why (Senior) Engineers Struggle to Build AI Agents — Philipp Schmid, Google DeepMind

Senior engineers struggle with AI Agent development because the paradigm shifted from deterministic programming to iterative prompt-feedback loops; text is now the new state representation, and engineers must transition from 'traffic controllers' to 'dispatchers'.

入选理由:AI Agent开发采用“定义目标→运行→观察→调整提示/工具→再运行”的迭代闭环,而非传统“需求→编码→测试→部署”线性流程。

FeaturedVideo#AI Agent#LLM Engineering#Software Paradigm Shift#Prompt Engineering英文
We released Gemma 4 12B yesterday. Here is a visual guide that explains the full architecture.

→ Ho...

Gemma 4 12B Released: Visual Guide to Native Multimodal Architecture

Philipp Schmid(@_philschmid)169 字 (约 1 分钟)
75

Gemma 4 12B achieves native multimodal processing for text, images, and audio by removing separate vision and audio encoders. This architecture replaces traditional encoder-patching approaches with joint representation learning, reducing inference latency and improving edge deployment efficiency.

入选理由:Gemma 4 12B移除独立视觉/音频编码器,采用原生多模态统一架构

FeaturedTweet#Gemma 4#Multimodal LLM#Native Multimodality#Edge AI英文
TIL: You can optimize any agent (cli) with GEPA to automatically optimize your prompts. 

GEPA accep...

TIL: Optimize Any Agent (CLI) with GEPA to Automatically Optimize Prompts

Philipp Schmid(@_philschmid)170 字 (约 1 分钟)
75

GEPA is a general framework for automatically optimizing prompts of CLI tools, supporting custom CLI, local models, and API agents.

入选理由:GEPA 可以优化任何 CLI 工具的提示,只需提供一个 `(str) -> str` 类型的可调用函数。

FeaturedTweet#CLI#Automation#Prompt Optimization#Python#Toolchain英文
Together with @OfficialLoganK and @alihcevik we sat down and discussed the launch of Managed Agents ...

Google launched Managed Agents in the Gemini API, enabling developers to spin up AI agents that reason, write and run code with a single API call, all within a hosted Linux sandbox, simplifying AI agent development.

入选理由:通过单个 API 调用即可启动一个具备推理、编码和文件管理能力的 AI 代理。

FeaturedTweet#Gemini API#AI Agents#Managed Agents#Google#API英文
What if you can build an Agent with it own computer in a single api call?

At my @Google I/O talk, I...

What if you can build an Agent with it own computer in a single api call?

Philipp Schmid(@_philschmid)198 字 (约 1 分钟)
75

Gemini Managed Agents combined with the Interactions API enables AI agents to have secure Linux sandbox environments for executing code and managing memory autonomously.

入选理由:通过单个 API 调用即可构建具备独立计算资源的 AI Agent。

FeaturedTweet#AI Agent#Gemini#API#Sandboxing#Google I/O英文
We made a collection @GoogleDeepMind scientific agent skils for research tasks, genomics, structural...

Google DeepMind Open-Sources Science Skills for Research Agents

Philipp Schmid(@_philschmid)81 字 (约 1 分钟)
72

Google DeepMind released Science Skills, an open-source toolkit providing standardized agent interfaces for genomics, structural biology, and literature search to accelerate scientific AI workflows.

入选理由:Google DeepMind发布science-skills开源库,专为科研AI Agent设计。

FeaturedTweet#Google DeepMind#AI Agent#Scientific Computing#Open Source英文
We just launched a Gemma 4 12B! Our first mid-sized model with native audio inputs. Gemma 4 12 B is ...

Gemma 4 12B Launch: First Mid-Sized Model with Native Audio Inputs

Philipp Schmid(@_philschmid)112 字 (约 1 分钟)
72

Gemma 4 12B is the first mid-sized multimodal model with native audio input, featuring a unified encoder-free architecture that runs on 16GB VRAM, matches 26B benchmark performance, and uses Apache 2.0 license.

入选理由:Gemma 4 12B采用无编码器统一架构,直接将视觉与音频信号输入LLM,降低推理延迟。

FeaturedTweet#Gemma 4#Multimodal Model#Audio Understanding#Apache 2.0英文
Philipp Schmid(@_philschmid) 图标

Blog: https://t.co/d0IjCP8nd4 Code: https://t.co/QSGciiZA6Y

Philipp Schmid(@_philschmid)48 字 (约 1 分钟)
70

本文介绍如何使用 Gemini Live API、LiveKit 和 Google Cloud Run 构建实时翻译应用,适合前端工程师参考。

入选理由:Gemini Live API 支持实时语音翻译,适用于多语言场景。

FeaturedTweet#Gemini#LiveKit#Google Cloud Run#实时翻译#前端英文
https://t.co/v54oLXJXnF

https://t.co/v54oLXJXnF

Philipp Schmid(@_philschmid)39 字 (约 1 分钟)
50

推文介绍了一种通过中间代理层实现GitHub令牌安全管理的方案,但未提供技术细节和验证数据。

入选理由:使用OAuth 2.0和令牌管理技术可降低GitHub令牌泄露风险

FeaturedTweet#GitHub#OAuth#安全#DevOps英文
Guide (correct link): https://t.co/mTokUFY359

Guide (correct link): https://t.co/mTokUFY359

Philipp Schmid(@_philschmid)45 字 (约 1 分钟)
50

文章内容信息密度低,缺乏技术深度和实用价值,主要为社交媒体上的链接分享。

入选理由:文章未提供具体技术细节或实用建议。

FeaturedTweet#社交媒体#链接分享中英混合
https://t.co/9dFYdBs5W2

https://t.co/9dFYdBs5W2

Philipp Schmid(@_philschmid)46 字 (约 1 分钟)
50

文章内容为社交媒体帖子,信息密度低,缺乏技术深度和实用价值。

入选理由:文章为社交媒体帖子,未提供具体技术内容。

FeaturedTweet#社交媒体#技术讨论中英混合
CLI + SKILL: https://t.co/PnH7cOQkZO

CLI + SKILL: https://t.co/PnH7cOQkZO

Philipp Schmid(@_philschmid)44 字 (约 1 分钟)
50

文章内容信息密度低,缺乏具体技术细节和深度分析,仅提供了一个 CLI 工具的链接。

入选理由:文章未提供具体技术细节或深度分析。

FeaturedTweet#CLI#工具英文
Reminder: Your skills are way too prescriptive and over-structured.

Reminder: Your skills are way too prescriptive and over-structured.

Philipp Schmid(@_philschmid)48 字 (约 1 分钟)
50

This tweet highlights that current AI models have overly structured and prescriptive skill descriptions, limiting flexibility and creativity, and suggests rethinking skill definition to support more natural interactions.

入选理由:当前AI模型的技能描述普遍采用高度结构化的格式,限制了其表达多样性。

FeaturedTweet#AI#Skill Modeling#Natural Language Processing#Model Architecture英文
We made a skill for and using Gemma. 

```
npx skills add google-gemma/gemma-skills --skill gemma-de...

We made a skill for and using Gemma

Philipp Schmid(@_philschmid)64 字 (约 1 分钟)
40

This article introduces a skill named gemma-dev for using Google's Gemma model in development, installed via npx command, but lacks technical details.

入选理由:使用命令 `npx skills add google-gemma/gemma-skills --skill gemma-dev` 安装 Gemma 开发技能。

FeaturedTweet#Gemma#AI#Development Tool#X Platform英文
https://t.co/r780truprJ

Philipp Schmid Shares Gemma-4-12B-it Model Link

Philipp Schmid(@_philschmid)53 字 (约 1 分钟)
30

This tweet only shares a Hugging Face link to Google's Gemma-4-12B-it model without technical analysis, benchmarks, or engineering guidance, offering minimal informational value for practitioners.

入选理由:Gemma-4-12B-it是Google发布的120亿参数指令微调模型,托管于Hugging Face平台。

FeaturedTweet#Gemma-4#Hugging Face#LLM英文

跨材料问答 · Philipp Schmid

回答基于:Philipp Schmid 相关 30 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.