最新記事

Fable-5-Grade Intelligence at Last Gen's Price: How Claude Opus 5 Rebuilt "the Opus for Long-Running Agents"

Fable-5-Grade Intelligence at Last Gen's Price: How Claude Opus 5 Rebuilt "the Opus for Long-Running Agents"

Anthropic released Claude Opus 5 on July 24. At the unchanged $5/$25 pricing, it delivers performance approaching the publicly top-tier Fable 5 and replaces the default Opus in Claude Code. Here's a rundown of the shifts in pricing, benchmarks, and agent operations—along with the caveats to keep in

By FF
Fable 5級の知能を、据え置きの値段で回す──Claude Opus 5が「長時間エージェントのOpus」を作り直した

Fable 5級の知能を、据え置きの値段で回す──Claude Opus 5が「長時間エージェントのOpus」を作り直した

Anthropicが7月24日にClaude Opus 5を公開。据え置きの$5/$25で公開最上位Fable 5に迫る性能を打ち出し、Claude Codeの既定Opusも置き換え。料金・ベンチ・エージェント運用の変化と留保点を整理します。

FF
Handing the "Which Model Should I Write With" Decision to AI — How Cursor Router Tackles the Model Bill of Running Agents Nonstop

Handing the "Which Model Should I Write With" Decision to AI — How Cursor Router Tackles the Model Bill of Running Agents Nonstop

On July 22, Cursor announced Cursor Router, which automatically selects the AI model to use for each request. It's said to be up to 50% cheaper than pinning to Opus, but concerns remain — such as it being hard to see which model wrote your code. We break down the cost, mechanics, and caveats.

By FF
「どのモデルで書くか」を選ぶ手をAIに預ける──Cursor Routerが突く、エージェントを回し続けるモデル代

「どのモデルで書くか」を選ぶ手をAIに預ける──Cursor Routerが突く、エージェントを回し続けるモデル代

Cursorが7月22日、リクエストごとに使うAIモデルを自動で選ぶCursor Routerを発表。Opus固定より最大50%安いとされる一方、どのモデルで書いたか見えにくい等の懸念も。コスト・仕組み・注意点を整理する。

FF
Reviews Move Out of the Conversation, Deep Dives Wait to Be Called — How Claude Code v2.1.218 Draws the Line at "No Running Off on Its Own"

Reviews Move Out of the Conversation, Deep Dives Wait to Be Called — How Claude Code v2.1.218 Draws the Line at "No Running Off on Its Own"

Claude Code v2.1.218 looks like a bundle of minor fixes, but a single thread runs through it. /code-review now runs in the background, /deep-research launches only when invoked manually, and fork skills default to running behind the scenes—pushing heavy automated work out of the conversation so noth

By FF
レビューは会話の外へ、深掘りは呼ぶまで動かない──Claude Code v2.1.218が引いた「勝手に走らせない」の線

レビューは会話の外へ、深掘りは呼ぶまで動かない──Claude Code v2.1.218が引いた「勝手に走らせない」の線

Claude Code v2.1.218は細かな修正の束に見えて、一本の筋が通っている。/code-reviewをバックグラウンド実行にし、/deep-researchは手動起動のみに、forkスキルも既定で裏へ——重い自動作業を会話の外へ出し、勝手に走らせない。無人運用で効く「既定の締め直し」を開発者視点で整理する。

FF
Leaving the 'Pro' Seat Empty, Three Cheap-and-Fast Flash Models Instead — How Gemini's Refresh Targets the Token Bill of Keeping Agents Running

Leaving the 'Pro' Seat Empty, Three Cheap-and-Fast Flash Models Instead — How Gemini's Refresh Targets the Token Bill of Keeping Agents Running

On July 21, Google refreshed three Gemini Flash models. It again held back the flagship 3.5 Pro, instead strengthening the mid-tier — including a 3.6 Flash that cuts output tokens by 17% — to target "the cost of keeping agents running." A developer-focused rundown covering price tags, benchmarks, an

By FF
『Pro』の席は空けたまま、安く速いFlashを3枚──Gemini刷新が突く、エージェントを回し続けるトークン代

『Pro』の席は空けたまま、安く速いFlashを3枚──Gemini刷新が突く、エージェントを回し続けるトークン代

Googleが7月21日にGemini Flashを3モデル更新。旗艦3.5 Proは今回も見送り、出力トークン17%減の3.6 Flashなど「エージェントを回し続けるコスト」に効く中位帯を強化した。値札・ベンチ・セキュリティ特化Cyberの限定配布まで、開発者視点で整理する。

FF
Agents' Output Now Shows Up on the Management Dashboard — CloudWatch Starts Measuring Claude Code, Codex, and Copilot Side by Side

Agents' Output Now Shows Up on the Management Dashboard — CloudWatch Starts Measuring Claude Code, Codex, and Copilot Side by Side

Announced by AWS on July 20, CloudWatch Coding Agent Insights lines up the consumption, cost, and output of Claude Code, Codex, and Copilot on a single screen. This piece weighs both sides — the benefit of measuring adoption gains in hard numbers, and the danger of repurposing individuals' line coun

By FF
エージェントの働きぶりが、経営の画面に載る──CloudWatchがClaude Code/Codex/Copilotを横並びで測り始めた

エージェントの働きぶりが、経営の画面に載る──CloudWatchがClaude Code/Codex/Copilotを横並びで測り始めた

AWSが7月20日に発表したCloudWatch Coding Agent Insightsは、Claude Code・Codex・Copilotの消費・コスト・成果を一枚の画面に並べる。導入効果を数字で測れる利点と、個人の行数を監視に転用する危うさの両面を整理する。

FF