<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
<title>AI Index News</title>
<link>https://ai.gotry.io/news</link>
<atom:link href="https://ai.gotry.io/news.xml" rel="self" type="application/rss+xml"/>
<description>AI model and API news from providers’ official pages: releases, price changes, retirements and new capabilities.</description>
<language>en</language>
<lastBuildDate>Sat, 03 Oct 2026 00:00:00 GMT</lastBuildDate>
<item>
<title>xAI retires grok-voice-transcribe-1.0</title>
<link>https://ai.gotry.io/news/2026-10-03-xai-retires-grok-voice-transcribe-1-0</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-10-03-xai-retires-grok-voice-transcribe-1-0</guid>
<pubDate>Sat, 03 Oct 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>xAI</category>
<description>grok-voice-transcribe-1.0 reached end of life on October 2, 2026. Requests to its slug are routed to grok-voice-transcribe-2.0 at the same price. Source: xAI API release notes (https://docs.x.ai/developers/release-notes#grok-voice-transcribe-10-end-of-life)</description>
</item>
<item>
<title>Google retires Gemini 2.5 Flash Image</title>
<link>https://ai.gotry.io/news/2026-10-02-google-retired-gemini-2-5-flash-image</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-10-02-google-retired-gemini-2-5-flash-image</guid>
<pubDate>Fri, 02 Oct 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>Google</category>
<description>Gemini 2.5 Flash Image is now retired. Source: Release notes | Gemini API | Google AI for Developers (https://ai.google.dev/gemini-api/docs/changelog?hl=en)</description>
</item>
<item>
<title>OpenAI deprecates gpt-4o-mini-tts (2025-12-15) and gpt-4o-mini-tts (2025-03-20)</title>
<link>https://ai.gotry.io/news/2026-10-01-openai-deprecated-gpt-4o-mini-tts-2025-12-15-and-1-more</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-10-01-openai-deprecated-gpt-4o-mini-tts-2025-12-15-and-1-more</guid>
<pubDate>Thu, 01 Oct 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>OpenAI</category>
<description>gpt-4o-mini-tts (2025-12-15) and gpt-4o-mini-tts (2025-03-20) are now deprecated and shut down on 2027-01-06. Source: OpenAI API deprecations (https://developers.openai.com/api/docs/deprecations)</description>
</item>
<item>
<title>OpenAI deprecates tts-1, tts-1-hd, GPT-5.4 nano, GPT-5.3 Codex and GPT-5.1</title>
<link>https://ai.gotry.io/news/2026-10-01-openai-deprecated-tts-1-and-4-more</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-10-01-openai-deprecated-tts-1-and-4-more</guid>
<pubDate>Thu, 01 Oct 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>OpenAI</category>
<description>tts-1, tts-1-hd, GPT-5.4 nano, GPT-5.3 Codex and GPT-5.1 are now deprecated and shut down on 2027-01-06. Source: OpenAI API deprecations (https://developers.openai.com/api/docs/deprecations)</description>
</item>
<item>
<title>OpenAI retires gpt-5.4-cyber</title>
<link>https://ai.gotry.io/news/2026-10-01-openai-retired-gpt-5-4-cyber</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-10-01-openai-retired-gpt-5-4-cyber</guid>
<pubDate>Thu, 01 Oct 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>OpenAI</category>
<description>gpt-5.4-cyber is now retired. Source: OpenAI API deprecations (https://developers.openai.com/api/docs/deprecations)</description>
</item>
<item>
<title>Alibaba Qwen retires deepseek-r1-distill-llama-8b</title>
<link>https://ai.gotry.io/news/2026-09-30-alibaba-qwen-retired-deepseek-r1-distill-llama-8b</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-alibaba-qwen-retired-deepseek-r1-distill-llama-8b</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>Alibaba Qwen</category>
<description>deepseek-r1-distill-llama-8b is now retired. Source: Alibaba Cloud Model Studio: text generation models (https://www.alibabacloud.com/help/en/model-studio/text-generation-model)</description>
</item>
<item>
<title>Anthropic schedules Claude Sonnet 4.5 retirement</title>
<link>https://ai.gotry.io/news/2026-09-30-anthropic-schedules-claude-sonnet-4-5-retirement</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-anthropic-schedules-claude-sonnet-4-5-retirement</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>Anthropic</category>
<description>Anthropic plans to retire Claude Sonnet 4.5 (claude-sonnet-4-5-20250929) from the Claude API on November 30, 2026. It recommends migrating to Claude Sonnet 5.5. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-30-2026)</description>
</item>
<item>
<title>Black Forest Labs releases FLUX 3 Image</title>
<link>https://ai.gotry.io/news/2026-09-30-black-forest-labs-releases-flux-3-image</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-black-forest-labs-releases-flux-3-image</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Black Forest Labs</category>
<description>FLUX 3 Image is available through one endpoint for image generation and editing, with up to ten reference images and output up to 4K. Per-image prices range from $0.041 at 768sq to $0.607 at 4K. Source: BFL API release notes (https://docs.bfl.ai/release-notes#october-1-2026)</description>
</item>
<item>
<title>Google announces Gemini 4 Argon</title>
<link>https://ai.gotry.io/news/2026-09-30-google-announces-gemini-4-argon</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-google-announces-gemini-4-argon</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Google announced Gemini 4 Argon, with an output token limit of 1 million. It is rolling out to trusted cyber defenders and is not yet available to developers; Google says developer access will come later. Source: Google DeepMind blog (https://deepmind.google/blog/gemini-4-argon-our-next-era-of-frontier-intelligence/)</description>
</item>
<item>
<title>Google retires gemini-omni-flash-preview</title>
<link>https://ai.gotry.io/news/2026-09-30-google-retired-gemini-omni-flash-preview</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-google-retired-gemini-omni-flash-preview</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>Google</category>
<description>gemini-omni-flash-preview is now retired. Source: Gemini API deprecations (https://ai.google.dev/gemini-api/docs/deprecations?hl=en)</description>
</item>
<item>
<title>Runway adds Eleven v4 to Runway Dev</title>
<link>https://ai.gotry.io/news/2026-09-30-runway-adds-eleven-v4-to-runway-dev</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-30-runway-adds-eleven-v4-to-runway-dev</guid>
<pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Runway</category>
<description>Eleven v4 is available on Runway Dev for speech generation, with scripts up to 2,500 characters. It costs 2.2 credits per 1,000 characters through October 12, 2026 PT, then 5 credits, with a 1-credit minimum. Source: Runway API changelog (https://docs.dev.runwayml.com/api-details/api_changelog/#eleven-v4-on-runway-dev)</description>
</item>
<item>
<title>Inception makes Mercury Voice generally available</title>
<link>https://ai.gotry.io/news/2026-09-29-inception-makes-mercury-voice-generally-available</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-inception-makes-mercury-voice-generally-available</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Inception</category>
<description>Inception made Mercury Voice generally available to enterprise customers. It supports 128K-token contexts, up to 50K output tokens and three reasoning settings. It costs $0.40 per 1M input tokens and $1.50 per 1M output tokens, with 50% off at launch ($0.20 and $0.75). Source: Inception blog (https://www.inceptionlabs.ai/blog/introducing-mercury-voice)</description>
</item>
<item>
<title>MiniMax releases minimax-m3.1-flash-preview</title>
<link>https://ai.gotry.io/news/2026-09-29-minimax-m3-1-flash-preview-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-minimax-m3-1-flash-preview-release</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>MiniMax</category>
<description>1M-token context window. Source: Subscription Plan Upgrade Announcement - MiniMax API Docs (https://platform.minimax.io/docs/token-plan/announcements)</description>
</item>
<item>
<title>Mistral AI retires Leanstral 1.5 and GLM 5.2</title>
<link>https://ai.gotry.io/news/2026-09-29-mistral-retires-leanstral-1-5-and-glm-5-2</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-mistral-retires-leanstral-1-5-and-glm-5-2</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>Mistral AI</category>
<description>Leanstral 1.5 retires on September 30, 2026; Z.ai GLM 5.2 retires on October 31, 2026. Mistral AI recommends Z.ai GLM 5.3 as its replacement at the same price. Source: Mistral AI changelog (https://docs.mistral.ai/resources/changelogs#date-2026-09-29)</description>
</item>
<item>
<title>OpenAI adds computer use to the Agents API</title>
<link>https://ai.gotry.io/news/2026-09-29-openai-adds-computer-use-to-the-agents-api</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-openai-adds-computer-use-to-the-agents-api</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>OpenAI</category>
<description>OpenAI added computer use to the Agents API, allowing agents to complete tasks in an OpenAI-hosted browser. Applications handle website access approvals and sign-in. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>OpenAI adds Ultrafast mode for GPT-6 Astra</title>
<link>https://ai.gotry.io/news/2026-09-29-openai-adds-ultrafast-mode-for-gpt-6-astra</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-openai-adds-ultrafast-mode-for-gpt-6-astra</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>OpenAI</category>
<description>OpenAI added Ultrafast mode for GPT-6 Astra in the Responses API. API customers can use the ultrafast service tier to reduce the time between generated output tokens, subject to rate limits. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>OpenAI releases GPT-6.1 Sol</title>
<link>https://ai.gotry.io/news/2026-09-29-openai-releases-gpt-6-1-sol</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-29-openai-releases-gpt-6-1-sol</guid>
<pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>OpenAI released GPT-6.1 Sol (gpt-6.1-sol) for complex coding and professional work. Standard prices for prompts of up to 272K input tokens are $2 per 1M input tokens, $0.10 cached input and $10 output. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>Anthropic releases Claude Sonnet 5.5</title>
<link>https://ai.gotry.io/news/2026-09-28-anthropic-releases-claude-sonnet-5-5</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-28-anthropic-releases-claude-sonnet-5-5</guid>
<pubDate>Mon, 28 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Anthropic</category>
<description>Anthropic launched Claude Sonnet 5.5 (claude-sonnet-5-5) on the Claude API, Amazon Bedrock, Claude Platform on AWS, Google Cloud, and Microsoft Foundry. Context window, output limits, and prices are on its model page. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-28-2026)</description>
</item>
<item>
<title>ElevenLabs releases Eleven v4 and Eleven v4 Turbo</title>
<link>https://ai.gotry.io/news/2026-09-28-elevenlabs-releases-eleven-v4-and-eleven-v4-turbo</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-28-elevenlabs-releases-eleven-v4-and-eleven-v4-turbo</guid>
<pubDate>Mon, 28 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>ElevenLabs</category>
<description>Eleven v4 and Eleven v4 Turbo are now available. Eleven v4 supports voice cloning in more than 90 languages; Turbo is intended for real-time use and has approximately 100 ms median inference latency. Source: ElevenLabs changelog (https://elevenlabs.io/docs/changelog/2026/9/28)</description>
</item>
<item>
<title>Meituan releases LongCat-2.5-Preview</title>
<link>https://ai.gotry.io/news/2026-09-25-longcat-2-5-preview-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-25-longcat-2-5-preview-release</guid>
<pubDate>Fri, 25 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Meituan</category>
<description>1M-token context window. On LongCat API: $0.30 input, $1.20 output per 1M tokens. Source: LongCat API: changelog (https://longcat.chat/platform/docs/ChangeLog.html)</description>
</item>
<item>
<title>Perplexity cuts fast preset search price</title>
<link>https://ai.gotry.io/news/2026-09-25-perplexity-cuts-fast-preset-search-price</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-25-perplexity-cuts-fast-preset-search-price</guid>
<pubDate>Fri, 25 Sep 2026 00:00:00 GMT</pubDate>
<category>Pricing</category>
<category>Perplexity</category>
<description>The Agent API fast preset now uses Fast Search. web_search falls from $2.50 to $1.00 per 1,000 invocations and is about 800 ms faster. Source: Perplexity API changelog (https://docs.perplexity.ai/docs/resources/changelog#september-2026)</description>
</item>
<item>
<title>Perplexity adds Fast Search option</title>
<link>https://ai.gotry.io/news/2026-09-24-perplexity-adds-fast-search-option</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-24-perplexity-adds-fast-search-option</guid>
<pubDate>Thu, 24 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Perplexity</category>
<description>Search API and Agent API web_search can set search_type to fast for a lower-latency path at $1.00 per 1,000 requests or invocations. Agent API model tokens are billed separately; search_type web is standard search. Source: Perplexity API changelog (https://docs.perplexity.ai/docs/resources/changelog#september-2026)</description>
</item>
<item>
<title>Black Forest Labs releases FLUX 3 Action</title>
<link>https://ai.gotry.io/news/2026-09-23-black-forest-labs-releases-flux-3-action</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-23-black-forest-labs-releases-flux-3-action</guid>
<pubDate>Wed, 23 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Black Forest Labs</category>
<description>Black Forest Labs released FLUX 3 Action, an open-weight 7B world-action model from its FLUX 3 backbone, with weights available. On RoboLab-120 it uses under half the parameters of the prior best open model and runs up to 3.95x faster. Source: Black Forest Labs blog (https://bfl.ai/blog/flux-3-action)</description>
</item>
<item>
<title>Meta introduces Muse Realtime Avatar</title>
<link>https://ai.gotry.io/news/2026-09-23-meta-introduces-muse-realtime-avatar</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-23-meta-introduces-muse-realtime-avatar</guid>
<pubDate>Wed, 23 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Meta</category>
<description>On September 23, 2026, Meta introduced Muse Realtime Avatar. It uses Muse Realtime Voice speech tokens and reference media to generate live, synchronized facial, hand, and full-body video. Source: Meta Superintelligence Labs research (https://research.meta.ai/blog/bringing-your-muse-to-life)</description>
</item>
<item>
<title>Anthropic adds mid-conversation tool definitions</title>
<link>https://ai.gotry.io/news/2026-09-22-anthropic-adds-mid-conversation-tool-definitions</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-anthropic-adds-mid-conversation-tool-definitions</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Anthropic</category>
<description>On the Claude API, the inline-tools-2026-09-15 beta header lets a system message added mid-conversation carry a full tool definition or, with the MCP connector beta header, an MCP toolset, without editing the tools list. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-22-2026)</description>
</item>
<item>
<title>Anthropic previews fast mode for Claude Opus 5.5</title>
<link>https://ai.gotry.io/news/2026-09-22-anthropic-previews-fast-mode-for-claude-opus-5-5</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-anthropic-previews-fast-mode-for-claude-opus-5-5</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Anthropic</category>
<description>Fast mode is available as a research preview for Claude Opus 5.5 on the Claude API. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-22-2026)</description>
</item>
<item>
<title>Anthropic releases Claude Opus 5.5</title>
<link>https://ai.gotry.io/news/2026-09-22-anthropic-releases-claude-opus-5-5</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-anthropic-releases-claude-opus-5-5</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Anthropic</category>
<description>Anthropic released Claude Opus 5.5 (claude-opus-5-5): 1M context, 128k max output, always-on adaptive thinking, at $4/$20 per MTok (Claude Opus 5 is $5/$25). It is on the Claude API, Amazon Bedrock, AWS, Google Cloud, and Microsoft Foundry. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-22-2026)</description>
</item>
<item>
<title>Google releases Gemini 3.8 Flash TTS models</title>
<link>https://ai.gotry.io/news/2026-09-22-google-releases-gemini-3-8-flash-tts-models</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-google-releases-gemini-3-8-flash-tts-models</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) and Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) are generally available, with a Gemini API Voices endpoint. Flash-Lite TTS is meant to replace gemini-3.1-flash-tts-preview. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-22-2026)</description>
</item>
<item>
<title>Xiaomi releases MiMo-V2.6-Flash</title>
<link>https://ai.gotry.io/news/2026-09-22-mimo-v2-6-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-mimo-v2-6-flash-release</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Xiaomi</category>
<description>1M-token context window. On Xiaomi MiMo API: $0.14 input, $0.28 output per 1M tokens. Source: Xiaomi MiMo: model updates (https://mimo.mi.com/docs/en-US/updates/model)</description>
</item>
<item>
<title>Xiaomi releases MiMo-V2.6-Pro</title>
<link>https://ai.gotry.io/news/2026-09-22-mimo-v2-6-pro-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-mimo-v2-6-pro-release</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Xiaomi</category>
<description>1M-token context window. On Xiaomi MiMo API: $0.435 input, $0.87 output per 1M tokens. Source: Xiaomi MiMo: model updates (https://mimo.mi.com/docs/en-US/updates/model)</description>
</item>
<item>
<title>Xiaomi releases MiMo-V2.6-Pro-UltraSpeed</title>
<link>https://ai.gotry.io/news/2026-09-22-mimo-v2-6-pro-ultraspeed-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-mimo-v2-6-pro-ultraspeed-release</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Xiaomi</category>
<description>1M-token context window. On Xiaomi MiMo API: $4.35 input, $8.70 output per 1M tokens. Source: Xiaomi MiMo: model updates (https://mimo.mi.com/docs/en-US/updates/model)</description>
</item>
<item>
<title>OpenAI improves prompt caching for GPT-6</title>
<link>https://ai.gotry.io/news/2026-09-22-openai-improves-prompt-caching-for-gpt-6</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-openai-improves-prompt-caching-for-gpt-6</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>OpenAI</category>
<description>OpenAI says GPT-6 models hit the prompt cache more often by default, and discounts now apply to shared prefixes reused within 30 minutes. New tools include a caching dashboard, a diagnostics tool that explains cache misses, explicit cache breakpoints, and changing reasoning effort mid-conversation without losing the cache. Source: OpenAI news (https://openai.com/index/better-prompt-caching-for-gpt-6)</description>
</item>
<item>
<title>OpenAI releases GPT-6 Sol and GPT-6 Luna</title>
<link>https://ai.gotry.io/news/2026-09-22-openai-releases-gpt-6-sol-and-gpt-6-luna</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-22-openai-releases-gpt-6-sol-and-gpt-6-luna</guid>
<pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>OpenAI released GPT-6 Sol (gpt-6-sol) and GPT-6 Luna (gpt-6-luna), reasoning models that take text and images and return text through the Responses and Chat Completions APIs. Standard prices per 1M tokens for prompts of up to 272K input tokens: Sol $2 input, $0.20 cached input, $10 output; Luna $0.10, $0.01 and $0.50. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>xAI releases Grok 4.7</title>
<link>https://ai.gotry.io/news/2026-09-21-grok-4-7-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-21-grok-4-7-release</guid>
<pubDate>Mon, 21 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>xAI</category>
<description>500K-token context window. On xAI API: $2.00 input, $6.00 output per 1M tokens. Source: xAI release notes (https://docs.x.ai/developers/release-notes)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen3.8-omni-flash-realtime</title>
<link>https://ai.gotry.io/news/2026-09-21-qwen3-8-omni-flash-realtime-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-21-qwen3-8-omni-flash-realtime-release</guid>
<pubDate>Mon, 21 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>197K-token context window. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Qwen releases Qwen-Image-2.1</title>
<link>https://ai.gotry.io/news/2026-09-20-alibaba-releases-qwen-image-2-1</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-20-alibaba-releases-qwen-image-2-1</guid>
<pubDate>Sun, 20 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>Qwen open-sourced Qwen-Image-2.1 on September 20, 2026. The 7B-parameter image model combines text-to-image generation and image editing, with native support for transparent images. Source: Qwen research (https://qwen.ai/blog?id=qwen-image-2.1)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen-audio-3.1-realtime-plus</title>
<link>https://ai.gotry.io/news/2026-09-20-qwen-audio-3-1-realtime-plus-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-20-qwen-audio-3-1-realtime-plus-release</guid>
<pubDate>Sun, 20 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>262K-token context window. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Qwen releases Qwen3.8-LiveTranslate</title>
<link>https://ai.gotry.io/news/2026-09-18-alibaba-releases-qwen3-8-livetranslate</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-18-alibaba-releases-qwen3-8-livetranslate</guid>
<pubDate>Fri, 18 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>Qwen announced Qwen3.8-LiveTranslate on September 18, 2026, describing an Interleave architecture for real-time simultaneous interpretation. Its average lagging metric drops from 2.8 seconds to 2.3 seconds. Source: Qwen research (https://qwen.ai/blog?id=qwen3.8-livetranslate)</description>
</item>
<item>
<title>Google limits access to Gemini 2.5 models</title>
<link>https://ai.gotry.io/news/2026-09-18-google-limits-access-to-gemini-2-5-models</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-18-google-limits-access-to-gemini-2-5-models</guid>
<pubDate>Fri, 18 Sep 2026 00:00:00 GMT</pubDate>
<category>Other</category>
<category>Google</category>
<description>Google is limiting Gemini 2.5 API access to users who have actively used those models. They are not deprecated and remain available until further notice. New projects should use 3.5 Flash-Lite or 3.8 Flash. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-18-2026)</description>
</item>
<item>
<title>Google releases Antigravity Agent 09-2026</title>
<link>https://ai.gotry.io/news/2026-09-17-google-releases-antigravity-agent-09-2026</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-17-google-releases-antigravity-agent-09-2026</guid>
<pubDate>Thu, 17 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Google released antigravity-preview-09-2026, replacing antigravity-preview-05-2026. Local tool parameters and file edits changed; remote output-only use needs only the new agent string. The May preview shuts down on October 5, 2026. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-17-2026)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen3.8-omni-flash</title>
<link>https://ai.gotry.io/news/2026-09-17-qwen3-8-omni-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-17-qwen3-8-omni-flash-release</guid>
<pubDate>Thu, 17 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>1M-token context window. On Alibaba Cloud Model Studio: $0.15 input, $0.47 output per 1M tokens. Source: 模型上下架与更新-大模型服务平台百炼-阿里云_大模型服务平台百炼(Model Studio)-大模型服务平台百炼(Model Studio)-阿里云帮助中心 (https://help.aliyun.com/zh/model-studio/newly-released-models)</description>
</item>
<item>
<title>Runway adds Enhance Frame Rate on Dev</title>
<link>https://ai.gotry.io/news/2026-09-17-runway-adds-enhance-frame-rate-on-dev</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-17-runway-adds-enhance-frame-rate-on-dev</guid>
<pubDate>Thu, 17 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Runway</category>
<description>Runway Dev can now convert video to 24, 25, 30, 48, 50, 60, 120, 23.98, 29.97, or 59.94 fps. Inputs are at most 300 seconds, billed at 1 credit per 2 seconds via the video upscale endpoint with model enhance_frame_rate. Source: Runway API changelog (https://docs.dev.runwayml.com/api-details/api_changelog/#enhance-frame-rate-on-runway-dev)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen-mt-uni</title>
<link>https://ai.gotry.io/news/2026-09-16-qwen-mt-uni-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-16-qwen-mt-uni-release</guid>
<pubDate>Wed, 16 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>qwen-mt-uni is now available. Source: 模型上下架与更新-大模型服务平台百炼-阿里云_大模型服务平台百炼(Model Studio)-大模型服务平台百炼(Model Studio)-阿里云帮助中心 (https://help.aliyun.com/zh/model-studio/newly-released-models)</description>
</item>
<item>
<title>Google releases Gemini 3.8 Live models</title>
<link>https://ai.gotry.io/news/2026-09-15-google-releases-gemini-3-8-live-models</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-15-google-releases-gemini-3-8-live-models</guid>
<pubDate>Tue, 15 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Gemini 3.8 Live (gemini-3.8-live) and Gemini 3.8 Live Extended Thinking (gemini-3.8-live-extended-thinking) are generally available as audio-to-audio models on the Live API. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-15-2026)</description>
</item>
<item>
<title>TypeSafe releases Jev (preview)</title>
<link>https://ai.gotry.io/news/2026-09-15-jev-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-15-jev-release</guid>
<pubDate>Tue, 15 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>TypeSafe</category>
<description>64K-token context window. On TypeSafe API: $0.042 input, $0.00 output per 1M tokens. Source: Introducing System One models &amp; Jev (https://typesafe.ai/blog/introducing-system-one-models-and-jev)</description>
</item>
<item>
<title>Kling AI releases Virtual Try-On 3.0</title>
<link>https://ai.gotry.io/news/2026-09-15-kuaishou-releases-virtual-try-on-3-0</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-15-kuaishou-releases-virtual-try-on-3-0</guid>
<pubDate>Tue, 15 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Kuaishou Kling</category>
<description>Virtual Try-On 3.0 improves face consistency and image quality. It accepts flat-lay, mannequin, and on-model clothing images, and can lock pose, keep the face, and choose whether to keep the original background. V1 and V1.5 parameters map to the 3.0 pipeline. Source: Kling AI API updates (https://kling.ai/document-api/updates/api)</description>
</item>
<item>
<title>Anthropic adds on-demand compaction to Messages API</title>
<link>https://ai.gotry.io/news/2026-09-14-anthropic-adds-on-demand-compaction-to-messages-api</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-14-anthropic-adds-on-demand-compaction-to-messages-api</guid>
<pubDate>Mon, 14 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Anthropic</category>
<description>Anthropic's Messages API now supports on-demand conversation compaction in beta via the compact-2026-09-04 header. A compaction parameter returns a signed summary block you can send later instead of the original messages, while keeping recent turns verbatim. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-14-2026)</description>
</item>
<item>
<title>Perplexity adds custom MCP connectors</title>
<link>https://ai.gotry.io/news/2026-09-13-perplexity-adds-custom-mcp-connectors</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-13-perplexity-adds-custom-mcp-connectors</guid>
<pubDate>Sun, 13 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Perplexity</category>
<description>A Project can register a remote MCP server once and Perplexity stores its credential. Agent API calls use the connector ID with type connector. API-key or no auth, and Streamable HTTP or SSE, are supported. Source: Perplexity API changelog (https://docs.perplexity.ai/docs/resources/changelog#september-2026)</description>
</item>
<item>
<title>ElevenLabs makes Scribe v2 Medical generally available</title>
<link>https://ai.gotry.io/news/2026-09-11-elevenlabs-makes-scribe-v2-medical-generally-available</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-11-elevenlabs-makes-scribe-v2-medical-generally-available</guid>
<pubDate>Fri, 11 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>ElevenLabs</category>
<description>Scribe v2 Medical is now generally available for batch medical and clinical speech recognition. It is billed at the same rate as Scribe v2. Source: ElevenLabs changelog (https://elevenlabs.io/docs/changelog/2026/9/11)</description>
</item>
<item>
<title>Black Forest Labs adds 2K and 4K FLUX 3 Video</title>
<link>https://ai.gotry.io/news/2026-09-10-black-forest-labs-adds-2k-and-4k-flux-3-video</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-10-black-forest-labs-adds-2k-and-4k-flux-3-video</guid>
<pubDate>Thu, 10 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Black Forest Labs</category>
<description>The FLUX 3 Video endpoint now returns qhd (2560×1440) and uhd (3840×2176) clips in one request, and draft_enhance can commit a draft at either size. Added per-second prices are $0.40 and $0.80 for text or image input, and $0.65 and $0.95 for continuation. Source: BFL API release notes (https://docs.bfl.ai/release-notes#september-10-2026)</description>
</item>
<item>
<title>DeepSeek releases DeepSeek-V4.1-Flash</title>
<link>https://ai.gotry.io/news/2026-09-10-deepseek-releases-deepseek-v4-1-flash</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-10-deepseek-releases-deepseek-v4-1-flash</guid>
<pubDate>Thu, 10 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>DeepSeek</category>
<description>DeepSeek released DeepSeek-V4.1-Flash, its smallest new-architecture model with native multimodal vision, callable as deepseek-flash. The retired names deepseek-v4-flash and deepseek-v4-flash-vision-exp route to it for now. V4 Pro continues after 14 September 2026 with unchanged billing, and API prices were reduced with the release. Source: DeepSeek API change log (https://api-docs.deepseek.com/updates)</description>
</item>
<item>
<title>OpenAI makes GPT-Live 1 generally available</title>
<link>https://ai.gotry.io/news/2026-09-10-openai-makes-gpt-live-1-generally-available</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-10-openai-makes-gpt-live-1-generally-available</guid>
<pubDate>Thu, 10 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>GPT-Live 1 is generally available for full-duplex voice sessions that continue while a backend model or agent handles reasoning and tools. Voice costs $0.05 per minute, billed per second; model and tool use is charged separately. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>OpenAI releases Agents API in public beta</title>
<link>https://ai.gotry.io/news/2026-09-10-openai-releases-agents-api-in-public-beta</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-10-openai-releases-agents-api-in-public-beta</guid>
<pubDate>Thu, 10 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>OpenAI</category>
<description>OpenAI released the Agents API in public beta, with a managed Codex harness for session orchestration, context compaction, and recovery. Sessions can stream progress, use custom tools and MCP servers, and run in OpenAI or customer sandboxes. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>Inception releases Mercury 2.5</title>
<link>https://ai.gotry.io/news/2026-09-08-inception-releases-mercury-2-5</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-08-inception-releases-mercury-2-5</guid>
<pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Inception</category>
<description>Inception released Mercury 2.5, a quality step up from Mercury 2, with a 260K-token context and 1,107 tokens per second. List price is $0.20 per million input tokens and $0.75 per million output tokens, 80% off at launch. Source: Inception blog (https://www.inceptionlabs.ai/blog/introducing-mercury-2-5)</description>
</item>
<item>
<title>OpenAI makes GPT-Rosalind generally available</title>
<link>https://ai.gotry.io/news/2026-09-08-openai-makes-gpt-rosalind-generally-available</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-08-openai-makes-gpt-rosalind-generally-available</guid>
<pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>OpenAI made GPT-Rosalind (gpt-rosalind-research) generally available for approved internal life sciences research under its trusted-access program. Prices are $5, $0.50 cached, and $25 per 1M input, cached input, and output tokens. Billing starts October 5, 2026. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>OpenAI releases GPT Image 2.5 Sunburst and Flare</title>
<link>https://ai.gotry.io/news/2026-09-08-openai-releases-gpt-image-2-5-sunburst-and-flare</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-08-openai-releases-gpt-image-2-5-sunburst-and-flare</guid>
<pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>OpenAI released GPT Image 2.5 Sunburst and GPT Image 2.5 Flare for generation and editing via the Image API and Responses image tool. Both add xhigh and max quality and use GPT Image 2 token rates. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>Kling AI launches Video Commerce API</title>
<link>https://ai.gotry.io/news/2026-09-07-kuaishou-launches-video-commerce-api</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-07-kuaishou-launches-video-commerce-api</guid>
<pubDate>Mon, 07 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Kuaishou Kling</category>
<description>Kling AI launched Video Commerce via API and Agent. A character image plus script makes a talking-head video; adding product images produces promotion shots. Options cover audio speed, aspect ratio, resolution, voice, speech rate, and background music. Source: Kling AI API updates (https://kling.ai/document-api/updates/api)</description>
</item>
<item>
<title>Ant Group releases Ling-3.0-flash-VL</title>
<link>https://ai.gotry.io/news/2026-09-04-ant-group-releases-ling-3-0-flash-vl</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-04-ant-group-releases-ling-3-0-flash-vl</guid>
<pubDate>Fri, 04 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Ant Group</category>
<description>On September 4, 2026, Ant Group launched multimodal Ling-3.0-flash-VL. It can be tried in the chat UI and called through OpenAI-compatible and Anthropic-compatible APIs. Source: Ant Ling changelog (https://developer.ant-ling.com/en/docs/getting-started/changelog/)</description>
</item>
<item>
<title>Google releases Lyria 3.5 music model</title>
<link>https://ai.gotry.io/news/2026-09-03-google-releases-lyria-3-5-music-model</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-03-google-releases-lyria-3-5-music-model</guid>
<pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Lyria 3.5 (lyria-3.5) is generally available for full-length song generation. It accepts text and image inputs and outputs 44.1 kHz stereo audio, with duration and structure controls. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-03-2026)</description>
</item>
<item>
<title>OpenAI adds long-running controls for GPT-6 Astra</title>
<link>https://ai.gotry.io/news/2026-09-03-openai-adds-long-running-controls-for-gpt-6-astra</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-03-openai-adds-long-running-controls-for-gpt-6-astra</guid>
<pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>OpenAI</category>
<description>The Responses API now offers async tool calling, mid-turn steering over WebSockets, and mid-conversation reasoning-effort changes for GPT-6 Astra, while preserving the cached prompt prefix. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>OpenAI releases GPT-6 Astra</title>
<link>https://ai.gotry.io/news/2026-09-03-openai-releases-gpt-6-astra</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-03-openai-releases-gpt-6-astra</guid>
<pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>OpenAI</category>
<description>OpenAI released GPT-6 Astra on the Responses and Chat Completions APIs for reasoning, coding, computer use, research, and documents. It rejects none reasoning effort, custom temperature, top_p, and logprobs. Tool calling requires the Responses API. Source: OpenAI API changelog (https://developers.openai.com/api/docs/changelog)</description>
</item>
<item>
<title>Google releases Gemini 3.8 Flash</title>
<link>https://ai.gotry.io/news/2026-09-02-google-releases-gemini-3-8-flash</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-02-google-releases-gemini-3-8-flash</guid>
<pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>gemini-3.8-flash is generally available. Google describes it as its most intelligent Flash model, aimed at long-horizon software engineering, agents, and enterprise workflows. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-02-2026)</description>
</item>
<item>
<title>Meta releases Muse Spark 1.3</title>
<link>https://ai.gotry.io/news/2026-09-02-meta-releases-muse-spark-1-3</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-02-meta-releases-muse-spark-1-3</guid>
<pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Meta</category>
<description>On September 2, 2026, Meta released Muse Spark 1.3, including a max-reasoning option, on Muse Code and Meta Model API. It targets stronger agentic and coding work and more reliable long instructions. Source: Meta Superintelligence Labs research (https://research.meta.ai/blog/introducing-muse-spark-1-3)</description>
</item>
<item>
<title>Alibaba Qwen releases Qwen3.8 Max (2026-09-02)</title>
<link>https://ai.gotry.io/news/2026-09-02-qwen3-8-max-0902-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-02-qwen3-8-max-0902-release</guid>
<pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>1M-token context window. On Alibaba Cloud Model Studio: ¥12.00 input, ¥36.00 output per 1M tokens. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Anthropic prices Claude Fable 5.1 cache reads</title>
<link>https://ai.gotry.io/news/2026-09-01-anthropic-prices-claude-fable-5-1-cache-reads</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-anthropic-prices-claude-fable-5-1-cache-reads</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Pricing</category>
<category>Anthropic</category>
<description>Cache-read price for Claude Fable 5.1 and Claude Mythos 5.1 is $0.25 per million tokens, 0.025 times base input, versus 0.1 times on other models. Cache write prices are unchanged. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-1-2026)</description>
</item>
<item>
<title>Anthropic releases Claude Fable 5.1</title>
<link>https://ai.gotry.io/news/2026-09-01-anthropic-releases-claude-fable-5-1</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-anthropic-releases-claude-fable-5-1</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Anthropic</category>
<description>Anthropic launched Claude Fable 5.1 (claude-fable-5-1), successor to Claude Fable 5. Context is 1M tokens, max output 128k, price $10/$50 per MTok, cache reads $0.25 per MTok, on Claude API, Bedrock, AWS, Google Cloud, and Microsoft Foundry. Source: Claude Developer Platform release notes (https://platform.claude.com/docs/en/release-notes/overview#september-1-2026)</description>
</item>
<item>
<title>Anthropic releases Claude Mythos 5.1</title>
<link>https://ai.gotry.io/news/2026-09-01-claude-mythos-5-1-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-claude-mythos-5-1-release</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Anthropic</category>
<description>1M-token context window. Source: Claude Platform release notes (https://platform.claude.com/docs/en/release-notes/overview)</description>
</item>
<item>
<title>Google adds agentic video understanding to Gemini</title>
<link>https://ai.gotry.io/news/2026-09-01-google-adds-agentic-video-understanding-to-gemini</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-google-adds-agentic-video-understanding-to-gemini</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Capability</category>
<category>Google</category>
<description>Agentic video understanding is available for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite on the Interactions and GenerateContent APIs. Google says it can use up to 88% fewer tokens on long-form video than static processing. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog#09-01-2026)</description>
</item>
<item>
<title>Meta releases Muse Voice Transcribe</title>
<link>https://ai.gotry.io/news/2026-09-01-meta-releases-muse-voice-transcribe</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-meta-releases-muse-voice-transcribe</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Meta</category>
<description>On September 1, 2026, Meta released Muse Voice Transcribe, a real-time audio model with streaming speech recognition, diarization for more than 20 speakers, endpointing, and multilingual code-switching. Source: Meta Superintelligence Labs research (https://research.meta.ai/blog/introducing-muse-voice-transcribe)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen3.7-text-embedding-flash</title>
<link>https://ai.gotry.io/news/2026-09-01-qwen3-7-text-embedding-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-qwen3-7-text-embedding-flash-release</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>qwen3.7-text-embedding-flash is now available. Source: 模型上下架与更新-大模型服务平台百炼-阿里云_大模型服务平台百炼(Model Studio)-大模型服务平台百炼(Model Studio)-阿里云帮助中心 (https://help.aliyun.com/zh/model-studio/newly-released-models)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen3.7-text-rerank</title>
<link>https://ai.gotry.io/news/2026-09-01-qwen3-7-text-rerank-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-qwen3-7-text-rerank-release</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>qwen3.7-text-rerank is now available. Source: 模型上下架与更新-大模型服务平台百炼-阿里云_大模型服务平台百炼(Model Studio)-大模型服务平台百炼(Model Studio)-阿里云帮助中心 (https://help.aliyun.com/zh/model-studio/newly-released-models)</description>
</item>
<item>
<title>MongoDB releases rerank-3-lite</title>
<link>https://ai.gotry.io/news/2026-09-01-rerank-3-lite-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-rerank-3-lite-release</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>MongoDB</category>
<description>rerank-3-lite is now available. Source: Model Deprecations, Lifecycle States, and Support - Voyage AI by MongoDB - MongoDB Docs (https://www.mongodb.com/docs/voyageai/models/lifecycle/)</description>
</item>
<item>
<title>MongoDB releases rerank-3</title>
<link>https://ai.gotry.io/news/2026-09-01-rerank-3-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-09-01-rerank-3-release</guid>
<pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>MongoDB</category>
<description>rerank-3 is now available. Source: Model Deprecations, Lifecycle States, and Support - Voyage AI by MongoDB - MongoDB Docs (https://www.mongodb.com/docs/voyageai/models/lifecycle/)</description>
</item>
<item>
<title>Mistral AI makes OCR 4.1 generally available</title>
<link>https://ai.gotry.io/news/2026-08-31-mistral-makes-ocr-4-1-generally-available</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-31-mistral-makes-ocr-4-1-generally-available</guid>
<pubDate>Mon, 31 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Mistral AI</category>
<description>On August 31, 2026, Mistral AI made OCR 4.1 (mistral-ocr-4-1) generally available. Source: Mistral AI changelog (https://docs.mistral.ai/resources/changelogs#date-2026-08-31)</description>
</item>
<item>
<title>Tencent releases Hy4 Preview</title>
<link>https://ai.gotry.io/news/2026-08-28-hy4-preview-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-28-hy4-preview-release</guid>
<pubDate>Fri, 28 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Tencent</category>
<description>1.02M-token context window. On Tencent Cloud TokenHub: ¥6.00 input, ¥18.00 output per 1M tokens. Source: Tencent: Tencent releases and open-sources Tencent Hy4 preview (https://www.tencent.com/tencent-releases-and-open-sources-tencent-hy4-preview/)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen-mt-image-2.0</title>
<link>https://ai.gotry.io/news/2026-08-28-qwen-mt-image-2-0-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-28-qwen-mt-image-2-0-release</guid>
<pubDate>Fri, 28 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>qwen-mt-image-2.0 is now available. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Google releases Gemini Omni Flash</title>
<link>https://ai.gotry.io/news/2026-08-27-gemini-omni-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-27-gemini-omni-flash-release</guid>
<pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Gemini Omni Flash is now available. Source: Gemini API deprecations (https://ai.google.dev/gemini-api/docs/deprecations?hl=en)</description>
</item>
<item>
<title>Google releases Gemini 3.5 Transcribe Live</title>
<link>https://ai.gotry.io/news/2026-08-26-gemini-3-5-transcribe-live-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-26-gemini-3-5-transcribe-live-release</guid>
<pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Gemini 3.5 Transcribe Live is now available. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog?hl=en)</description>
</item>
<item>
<title>Google releases Gemini 3.5 Transcribe</title>
<link>https://ai.gotry.io/news/2026-08-26-gemini-3-5-transcribe-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-26-gemini-3-5-transcribe-release</guid>
<pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>Gemini 3.5 Transcribe is now available. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog?hl=en)</description>
</item>
<item>
<title>Zhipu AI releases GLM-5.3-Flash</title>
<link>https://ai.gotry.io/news/2026-08-26-glm-5-3-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-26-glm-5-3-flash-release</guid>
<pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Zhipu AI</category>
<description>1M-token context window. On Z.ai API: $0.15 input, $0.50 output per 1M tokens. Source: Z.ai new releases (https://docs.z.ai/release-notes/new-released)</description>
</item>
<item>
<title>OpenAI deprecates whisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribe and gpt-4o-transcribe-diarize</title>
<link>https://ai.gotry.io/news/2026-08-26-openai-deprecated-whisper-1-and-3-more</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-26-openai-deprecated-whisper-1-and-3-more</guid>
<pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
<category>Retirement</category>
<category>OpenAI</category>
<description>whisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribe and gpt-4o-transcribe-diarize are now deprecated and shut down on 2027-02-26. Source: OpenAI API deprecations (https://developers.openai.com/api/docs/deprecations)</description>
</item>
<item>
<title>Alibaba Qwen releases Qwen3.8 Flash</title>
<link>https://ai.gotry.io/news/2026-08-26-qwen3-8-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-26-qwen3-8-flash-release</guid>
<pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>1M-token context window. On Alibaba Cloud Model Studio: $0.15 input, $0.47 output per 1M tokens. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Alibaba Qwen releases Wan3.0 Video Prime</title>
<link>https://ai.gotry.io/news/2026-08-20-wan3-0-video-prime-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-20-wan3-0-video-prime-release</guid>
<pubDate>Thu, 20 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>Wan3.0 Video Prime is now available. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Zhipu AI releases GLM-5.3</title>
<link>https://ai.gotry.io/news/2026-08-18-glm-5-3-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-18-glm-5-3-release</guid>
<pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Zhipu AI</category>
<description>1M-token context window. On Z.ai API: $1.40 input, $4.40 output per 1M tokens. Source: Z.ai new releases (https://docs.z.ai/release-notes/new-released)</description>
</item>
<item>
<title>Alibaba Qwen releases qwen3.8-27b</title>
<link>https://ai.gotry.io/news/2026-08-17-qwen3-8-27b-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-17-qwen3-8-27b-release</guid>
<pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Alibaba Qwen</category>
<description>1M-token context window. On Alibaba Cloud Model Studio: $0.50 input, $3.00 output per 1M tokens. Source: Alibaba Cloud Model Studio: newly released models (https://www.alibabacloud.com/help/en/model-studio/newly-released-models)</description>
</item>
<item>
<title>Google releases Gemini 3.7 Flash</title>
<link>https://ai.gotry.io/news/2026-08-13-gemini-3-7-flash-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-13-gemini-3-7-flash-release</guid>
<pubDate>Thu, 13 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>Google</category>
<description>1.05M-token context window. On Gemini API: $0.75 input, $3.75 output per 1M tokens. Source: Gemini API changelog (https://ai.google.dev/gemini-api/docs/changelog?hl=en)</description>
</item>
<item>
<title>xAI releases Grok 4.6</title>
<link>https://ai.gotry.io/news/2026-08-12-grok-4-6-release</link>
<guid isPermaLink="false">ai.gotry.io/news/2026-08-12-grok-4-6-release</guid>
<pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate>
<category>Release</category>
<category>xAI</category>
<description>500K-token context window. On xAI API: $2.00 input, $6.00 output per 1M tokens. Source: xAI release notes (https://docs.x.ai/developers/release-notes)</description>
</item>
</channel>
</rss>
