Gemini API
PulseAugur coverage of Gemini API — every cluster mentioning Gemini API across labs, papers, and developer communities, ranked by signal.
- developed by Gemini Enterprise Agent Platform 95%
- used by Gemini 3.5 Transcribe 95%
- developed by Gemini 3.6 Flash 95%
- instance of Gemini 3.5 Flash 95%
- instance of Gemini 3.6 Flash 95%
- developed by Gemini 3.5 Transcribe 95%
- instance of Gemini-3.* 95%
- developed Gemini 3.8 Live Extended Thinking 95%
- instance of Google AI Studio 90%
- developed Google AI Studio 90%
- used by Gemini Enterprise Agent Platform 90%
- instance of Gemini app 90%
- 2026-09-06 product_launch Google launched Managed Agents for its Gemini API, featuring the Gemini 3.6 Flash model. 来源
- 2026-08-27 product_launch The Gemini API was released with version v2.1.1, addressing cookie cache overwriting issues. 来源
- 2026-08-15 product_launch The Gemini API was released with version 2.1.0, introducing new features like guest mode and account quotas. 来源
- 2026-07-28 product_launch Gemini API managed agents now include execution hooks and default to the 3.6 Flash model. 来源
- 2026-07-27 product_launch Google DeepMind expanded Managed Agents in the Gemini API with four new capabilities for production use. 来源
- 2026-07-18 product_launch Google updated its Gemini API with enhanced Managed Agents, introducing Background Execution and direct MCP server connectivity. 来源
- 2026-07-08 product_launch Google DeepMind added four new features to Managed Agents in the Gemini API, including background execution and MCP support. 来源
- 2026-07-08 product_launch Google DeepMind added new features to Gemini API Managed Agents, including background execution and MCP support. 来源
- 2026-05-20 product_launch Google launched Managed Agents within the Gemini API, simplifying agent development and deployment. 来源
- 2026-05-20 product_launch Google launched Managed Agents as a new feature for its Gemini API.
- 2026-05-20 product_launch Google launched a preview of managed agents within its Gemini API. 来源
- 2026-05-07 product_launch Google expanded the Gemini API File Search tool to support multimodal RAG systems with text and image data, and added page citations. 来源
12 天有情绪数据
-
AI 艺术展展示 Flux-Pro-Ultra、Dreamshaper、SDXL 模型
一位用户分享了在 2024 年至 2025 年间创作的一系列 AI 生成图像,其中包含 Flux-Pro-Ultra、Dreamshaper 和 SDXL 等模型。该帖子还呼吁安装一个利用 Gemini API 的 AI 检测工具。
-
Google 在 Gemini API 上发布 Nano Banana 2.1 图像模型
Google 已通过其 Gemini API 发布了新的图像生成和编辑模型 Nano Banana 2.1。该模型是目前可用的四款 Nano Banana 模型套件的一部分,其中 Nano Banana 2.1 是最新推出的。这些模型在 Gemini API 上的可用性为开发人员提供了先进的图像创建和处理工具。
-
Google发布Nano Banana 2.1图像AI,提升一致性和编辑功能
Google发布了更新的图像生成AI模型Nano Banana 2.1。据报道,新版本在生成一致的人物形象和增强编辑能力方面有所改进。该模型已集成到Google搜索的AI模式中,并通过Gemini API提供,Google销售四款Nano Banana图像模型。
-
Ollama 发布 mistral-large-4 和 embeddinggemma-2 模型
Ollama 发布了两个新模型:mistral-large-4 和 embeddinggemma-2。这些模型可通过 Ollama 库获取,并提供了详细链接。此次发布带有相关的 AI 和 LLM 标签,并提到了 Gemini 和 Gemini API。
-
Google Cloud 和 AWS 为 AI 服务推出支出上限
Google Cloud 于 7 月推出了预览版支出上限,其中包括 Gemini API。Amazon Web Services (AWS) 于 9 月跟进,实施了支出限制,最初仅针对新账户。这些措施旨在让用户更好地控制其云支出。
-
通过使用廉价模型处理常规任务来优化 LLM 工作流
开发人员可以通过策略性地使用更便宜、更快的模型来处理常规任务,并将 Claude Opus 等昂贵、强大的模型用于复杂推理,从而优化 LLM 工作流。这种常被忽视的方法包括使用基本模型进行分类、提取和摘要,同时将高级模型保留用于模糊或高风险的决策。这种管道设计降低了成本,最大限度地减少了上下文窗口的使用,并提高了整体工作流效率,这一策略得到了 Anthropic 和 Google 等提供商分级定价的支持。
-
AI视频生成:API、UI还是代理优先工具?
文章比较了不同的人工智能驱动的视频生成方法,将工具分为API优先、UI优先和代理优先。对于能够管理流程的高质量或大批量输出,推荐使用Runway、Gemini API和Kling等API优先工具。UI优先工具适用于人类监督决定视觉风格的工作流程。CoAnimator等代理优先工具专为需要通过提供可编辑项目状态和清晰反馈机制来修改现有视频的人工智能代理而设计。
-
AI代理引发数据泄露;谷歌发布Gemini 3.8 Live;中美提议AI安全保障
一起重大的数据泄露事件被归咎于一个AI代理,这是首次有记录显示此类事件在实际中发生。谷歌发布了Gemini 3.8 Live,据报道该模型在语音到语音的质量方面树立了新标杆。此外,报告还强调了日益加剧的关于数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心电力的数据中心…
-
Google推出Gemini 3.8 Live音频模型,实现流畅的AI对话
Google推出了两款新的音频模型Gemini 3.8 Live和Gemini 3.8 Live Extended Thinking,旨在增强对话式AI代理。这些模型支持实时交互,使代理能够在继续流式传输音频响应的同时执行API调用和工具执行等任务。Extended Thinking变体特别支持可配置的“思考”过程,允许用户打开或关闭AI的内部推理,并且可以在不中断用户对话的情况下在后台执行任务。这一进步旨在降低交互延迟,提高语音代理…
-
Google发布Gemini 3.8 Live和Extended Thinking音频模型
Google发布了Gemini 3.8 Live和Gemini 3.8 Live Extended Thinking,这两款新的先进音频模型旨在增强语音代理功能并创造更自然的AI对话。Gemini 3.8 Live专注于规模、速度和成本效益,能够处理句子中途打断和多语言转换,而Gemini 3.8 Live Extended Thinking通过并行推理和语音处理,为复杂的、多步骤任务提供更深层次的智能。这些模型正通过API向消费者、…
-
Google DeepMind 发布 Gemini 3.8 Live 对话式 AI · 已追踪 2 个来源
Google DeepMind 推出了 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,将其定位为最先进的对话式 AI 模型。这些新模型旨在通过在后台执行任务和进行通信而不中断用户当前交互来增强用户体验。主要功能包括升级的推理能力、近乎实时的视觉理解、自动检测 97 种语言以及无缝执行后台工具调用的能力。
-
GitHub Action 使用 Gemini API 自动发布 DEV.to 文章
一位开发者创建了一个GitHub Actions工作流,该工作流每天早上自动生成并发布文章到DEV Community。该系统利用Gemini API提出一个工程角度并撰写内容,整个过程从主题选择到发布都是自动化的。这种方法利用现有的CI/CD基础设施,无需手动撰写和发布,从而提高效率并节省成本。
-
Google Gemini Interactions API 需要更新 SDK 以进行图像编辑
开发人员正在更新 Python MCP 服务器以适应 Google Gemini Interactions API 的更改以及相关的 google-genai SDK。Interactions API 已放弃其旧版模式,需要将 google-genai Python SDK 升级到 2.0.0 或更高版本。此更新可确保服务器能够正确处理图像生成和编辑请求,这些请求以交互 ID 的形式存储在服务器端以便继续编辑。进行这些更改是必要的,因…
-
GPT Image 2、Nano Banana 2和FLUX.2引领图像生成模型
在图像生成方面,没有单一的最佳模型,而是根据具体需求进行选择。OpenAI的GPT Image 2提供最高的质量和最佳的提示遵循度,尽管它是最昂贵且最慢的。Google的Nano Banana 2 (Gemini 3.1 Flash Image)以显著较低的成本提供了接近顶级的质量和编辑能力的强大平衡,使其成为一个不错的默认选择。对于那些需要开源模型进行自托管的用户,FLUX.2是领先的选择,并提供各种许可选项。
-
Ollama 发布新的 DeepSeek-V4.1-Flash 模型
Ollama 发布了新的模型 DeepSeek-V4.1-Flash,可通过其库获取。此次发布是 Ollama 持续提供各种大型语言模型访问的一部分。
-
使用 Google Sheets 和 Gemini API 构建免费 AI 新闻管道
本文详细介绍了如何使用 Google Sheets、Gemini API 和 Telegram 构建一个免费的 AI 新闻处理管道。通过演示使用 UrlFetchApp.fetchAll() 并行化 Gemini 的独立请求的方法,解决了 Google Apps Script 执行时间的限制。这种方法显著加快了处理数百篇外语行业新闻文章的速度,将处理 50 篇公告的时间从两分钟缩短到 20 秒,同时保持在免费 API 配额内。
-
Claude Code 重定向到其他模型会破坏协议并增加成本
开发人员尝试将 Claude Code 重定向到 OpenAI 或 Gemini 等其他 AI 模型时,由于传输协议和负载格式不兼容,会面临重大挑战。Claude Code 被设计为仅与 Anthropic Messages API 通信,其专门的系统提示并非所有模型都能普遍理解。这种不兼容可能导致代码生成错误、因提示缓存丢失而导致的代币成本增加,以及执行循环中断。像 Bifröst 这样的 AI 网关可以通过翻译请求并为编码代理启用…
-
Google Gemini API 推出托管代理及 3.6 Flash 模型
Google 为其 Gemini API 推出了托管代理(Managed Agents),搭载 Gemini 3.6 Flash 模型。此次更新包括了 Hooks 和 Modals 等新功能,旨在增强 AI 代理在应用程序中的功能和集成性。此次发布专注于为开发者提供更强大的工具,以构建复杂的 AI 驱动体验。
-
Google AI 发布 Gemini 3.8 Flash、Lyria 3.5 音乐模型和 WeatherNext 3
Google AI 宣布了一系列新的模型发布和更新。这包括 Gemini 3.8 Flash,一个在编码和代理工作流方面有所改进的增强型主力模型,以及 Gemini 3.8 Flash Cyber,专门用于漏洞检测等网络安全任务。此外,新的音乐生成模型 Lyria 3.5 现已通过各种 Google 平台提供。该公司还推出了先进的天气 AI 模型 WeatherNext 3,以及用于视频分析并降低代币使用量和成本的 Agentic V…
-
LLM API故障不可避免;构建多模型备用系统
本文讨论了LLM API在生产环境中故障的不可避免性,例如速率限制、区域性中断和配额耗尽。它提出了一个多模型备用系统作为超越简单重试逻辑的解决方案。作者概述了实现此备用的Python模式,并强调了协议不兼容和需要统一的API网关来管理各种LLM提供商等潜在陷阱。