Ashley Goolam

Ashley Goolam

Claude Sonnet 5.5 Pricing: The Full Cost Breakdown (API, Caching, Batch, and Plans)Tutorials

Claude Sonnet 5.5 Pricing: The Full Cost Breakdown (API, Caching, Batch, and Plans)

Claude Sonnet 5.5 pricing: $2/$10 per million tokens, $0.20 cache reads, batch at half price. Worked cost examples, effort costs, and plan prices.

Ashley Goolam

September 29, 2026

What Is Claude Sonnet 5.5?Viewpoint

What Is Claude Sonnet 5.5?

What is Claude Sonnet 5.5? Anthropic's Sept 28, 2026 model: $2/$10 pricing, 1M context, benchmarks, what changed from Sonnet 5, and where to use it.

Ashley Goolam

September 29, 2026

GPT-6 prompt caching: 90% off cached reads, and how to actually hit the cacheTutorials

GPT-6 prompt caching: 90% off cached reads, and how to actually hit the cache

GPT-6 prompt caching: a 90% discount on cached input reads, higher hit rates by default, cache-safe reasoning effort and tool changes, explicit breakpoints, and how to test your hit rate in CI.

Ashley Goolam

September 23, 2026

What Is GPT-6 Luna? Model ID, $0.10/$0.50 Pricing, and a 1M Token Context WindowTutorials

What Is GPT-6 Luna? Model ID, $0.10/$0.50 Pricing, and a 1M Token Context Window

GPT-6 Luna explained: model ID gpt-6-luna, $0.10/$0.50 pricing, a 1M token context window that is larger than Sol's, DeepSWE 66.6% at 93% less per task than Claude Opus 5, and where you can call it.

Ashley Goolam

September 23, 2026

GPT-6 Sol Pricing: What the 50% Cut Is Actually Measured AgainstTutorials

GPT-6 Sol Pricing: What the 50% Cut Is Actually Measured Against

GPT-6 Sol costs $2/$10 per million tokens with 872k context and 90% off cached reads. The advertised 50% cut is against GPT-5.6 promotional pricing; against list rates it is 60% and 67%. Plus the cost-per-task numbers that moved further.

Ashley Goolam

September 23, 2026

Migrating from Claude Opus 5 to Opus 5.5: What Changes for API CallersTutorials

Migrating from Claude Opus 5 to Opus 5.5: What Changes for API Callers

Claude Opus 5 to Opus 5.5 for API callers: the new claude-opus-5-5 id, 20% off input and output, $5 cache writes and $0.20 cache reads, fast mode at $8/$40, and a 128k output ceiling.

Ashley Goolam

September 23, 2026

GPT-6 Astra crossed OpenAI's Critical cyber line. What it means for the APIs you run.Viewpoint

GPT-6 Astra crossed OpenAI's Critical cyber line. What it means for the APIs you run.

GPT-6 Astra is the first OpenAI model rated Critical for cyber capability. What the rating means, what ships by default, what Daybreak unlocks, and six API checks to run this week.

Ashley Goolam

September 5, 2026

Gemini 3.7 Flash to 3.8 Flash: API migration guideTutorials

Gemini 3.7 Flash to 3.8 Flash: API migration guide

Migrate from Gemini 3.7 Flash to 3.8 Flash: 9 API changes with before/after JSON, the minimal thinking-level error, call_id rules, token budgets, and rollback.

Ashley Goolam

September 3, 2026

Gemini 3.8 Flash pricing: intro rates, thinking tokens, and the real cost per taskViewpoint

Gemini 3.8 Flash pricing: intro rates, thinking tokens, and the real cost per task

Gemini 3.8 Flash pricing: $0.75/$3.75 intro rates doubling Jan 1 2027, thinking tokens billed as output, caching, batch, and why cost per task rose to $0.58.

Ashley Goolam

September 3, 2026

Gemini 3.8 Flash thinking levels: low vs medium vs high (and why minimal is gone)Tutorials

Gemini 3.8 Flash thinking levels: low vs medium vs high (and why minimal is gone)

Gemini 3.8 Flash thinking levels explained: what low, medium, and high do, why minimal now errors, per-level cost and latency numbers, and how to set each one.

Ashley Goolam

September 3, 2026

Function calling with Gemini 3.8 Flash: call_id, iterative tool loops, and how to test themTutorials

Function calling with Gemini 3.8 Flash: call_id, iterative tool loops, and how to test them

Gemini 3.8 Flash function calling step by step: declare a tool, read the function_call step, return function_result with call_id and name, cap loops, test.

Ashley Goolam

September 3, 2026

Claude Fable 5.1 Preserved Thinking: Fixing "The Block Is Bound to a Different Conversation"Tutorials

Claude Fable 5.1 Preserved Thinking: Fixing "The Block Is Bound to a Different Conversation"

Fix the Claude Fable 5.1 400 "bound to a different conversation": what preserved thinking checks, who is enforced, drop_block, the audit, and append-only patterns.

Ashley Goolam

September 2, 2026

Migrating to Claude Fable 5.1 from Fable 5 or Opus 5: Every Breaking ChangeTutorials

Migrating to Claude Fable 5.1 from Fable 5 or Opus 5: Every Breaking Change

Migrate to Claude Fable 5.1 from Fable 5 or Opus 5: forced tool_choice 400, one-way thinking blocks, the history-editing check, every fix, and a full checklist.

Ashley Goolam

September 2, 2026

How to Upgrade Your Coding Agent With 5 Open Source Tools in 2026Tutorials

How to Upgrade Your Coding Agent With 5 Open Source Tools in 2026

Five open source repos upgrading Claude Code, Cursor and Codex in 2026, what each really does, and the two gaps none of them close.

Ashley Goolam

September 1, 2026

How to Turn Claude Code or Cursor Into a Video Production StudioTutorials

How to Turn Claude Code or Cursor Into a Video Production Studio

Give your coding agent 11 video pipelines and 700+ skill files, plus the AGPL catch and the API cost problem under the creative work.

Ashley Goolam

September 1, 2026