← AP News
NO. 02 · Weekly News

Weekly AI News #1 — Google's Gemini 4 Arrives, and Why the Strongest Models Aren't Given to Everyone

Translated from the Korean original by AI; native-speaker review pending.

On September 30 Google unveiled Gemini 4 Argon, and in early September three companies all opened their security-focused models only to "verified defenders." As models get faster, the way they are released is changing.

This article was drafted together with AI, and an editor verified the sources. All figures are based on the sources below, and nothing here is investment advice.

This week in one line

The strongest models are now opened first to "authorized parties" rather than as "an app anyone can use." Meanwhile, beneath that, the pace of model releases and price competition keeps accelerating.

1. Google unveils Gemini 4 Argon (September 30)

Fact. According to Axios, Google announced Gemini 4 Argon on September 30. The model targets long, complex work such as software engineering, finance, law and cybersecurity, and Google said it beats OpenAI's GPT-6 Astra on coding and knowledge-work benchmarks. It will first be released in limited form to a small number of cybersecurity partners, with plans to widen access to paying subscribers after further testing. The same report, citing Bloomberg, said some employees had pointed out shortcomings in internal testing, which Google denied.

Interpretation. A benchmark lead is a claim measured on criteria chosen by the company announcing it. Until independent evaluations appear, it is safer to read it as a "claim."

What we would do. We do not switch models. We run a single comparison, feeding the same inputs into the work we actually do (writing, image descriptions, data cleanup), and switch only if there is a noticeable difference.

2. Three companies' security-focused models, all "verified defenders first"

Fact. According to The Hacker News, in early September Google (Gemini 3.8 Flash Cyber), Anthropic (Claude Mythos 5.1 and Fable 5.1) and OpenAI (Astra) each launched an access program for cybersecurity. Google said it is offering its model to trusted defenders in government, healthcare, telecommunications and other sectors, together with more than 650 partners. Anthropic said it opens its least restricted model only to a trusted access program, while also introducing defenses against malicious requests and prompt injection and detection of sandbox-escape attempts. OpenAI provides its model to a group of testers with an added abuse-prevention classifier, and cited a 91.5% refusal rate for jailbreak requests.

Interpretation. The ability to find vulnerabilities serves both defense and attack. That all three companies made the same choice, "give it to the defenders first and widen access later," reads as a signal that as capabilities grow, the way a model is released becomes part of product design itself.

What we would do. The direct impact is small, but we apply two things right away. For tasks where AI reads external documents or web pages, we set a rule that it must not follow instructions found inside them, and for any irreversible action AI takes (deleting, sending, paying), we require human confirmation.

3. A flood of models, a race to the bottom on price

Fact. An independent tracking blog compiled that more than 20 models launched over the two weeks from September 12 to 25, and that the token-price gap between the most expensive and the cheapest model was about 119x. On September 6, CNBC covered the industry mood of weariness at the rapid release pace under the phrase "model fatigue."

Interpretation. The tracked figures were compiled by a personal blog, so they should be checked against each vendor's official price list. The direction, though, is clear: the range of work that can be done with a much cheaper model keeps widening.

What we would do. Rather than using one "best model" for everything, we split work into hard tasks and repetitive tasks, and run the repetitive ones on cheaper models. We cover this standard in more detail in the next Insight article.

What to watch next week

  • Independent evaluations of Gemini 4 and the timing of its general release
  • Whether the security-model access programs expand
  • Changes to major API pricing

Sources

← Welcome to AP News — We Write …When AI Reads the Ingredient L… →