Skip to main content

Blog

Page 16 of 59

All Articles

Insights on AI, machine learning, and technology strategy

A cool-paper run-control strip connects one workflow file to four large controls: job-scoped permission, cost center and credit cap, usage evidence with a named reviewer, and a preselected stop action with rollback ready.
Industry Insights·

The secret is gone. The Copilot Actions run still needs a card.

GitHub just let Copilot CLI run in Actions without a personal access token. That closes one risk and opens a quieter one: a workflow that spends organization AI credits with no owner, no cap, and no reviewer.

10 min read
Textless layered glass workbook with light paths moving through source wells, rule channels, and a reviewer ring.
Industry Insights·

Before AI Edits the Forecast, Give the Workbook a Rulebook

Copilot in Excel is moving from formula helper to workflow runner. Microsoft's real answer to 'can we trust it' is a worksheet that travels with the file. Here's the smaller packet that makes that answer hold.

7 min read
A model facts register showing model rows with freshness dates, source links, capability boundaries, latency and cost notes, approval owners, and recheck triggers.
AI Development·

The model decision is stale before the meeting ends

AWS just admitted, in its own release notes, that the facts a model-picking meeting needs are scattered across console pages, documentation, and regional API calls. Its fix is a catalog. Yours still needs an owner.

8 min read
An analytics source-of-truth register showing a canonical metric with owner, definition, freshness check, dashboard and query source, exception note, and update cadence fields.
Industry Insights·

Before Claude answers the dashboard question, make someone own the metric

Anthropic says Claude automates 95% of its internal business analytics queries at 95% accuracy. The accuracy came from a maintained metric layer, not from pointing an agent at a warehouse.

8 min read
An AI vendor handoff dossier showing exit criteria, owned system, runbook, unresolved risk, internal owner, rollback path, proof packet, and transfer route fields.
Industry Insights·

AWS will send the AI pod. Ask what stays when it leaves.

AWS is putting $1 billion behind Forward Deployed Engineering teams that embed with customers to build agentic AI fast. The durable question for buyers is not whether the demo works. It is what evidence, ownership, and operating muscle remain after the outside team goes home.

8 min read
A machine visitor terms sheet showing content surfaces, visitor classes, allow, charge, and block decisions, enforcement paths, breakage tests, and revisit dates.
Industry Insights·

Before September 15, write your site's machine visitor terms

Starting September 15, 2026, new sites on Cloudflare will block AI training and agent crawlers by default on any page that shows ads, while search crawlers stay open. Existing sites can opt out before the deadline, but the harder problem isn't the checkbox. It's that "crawler" was never one category to begin with.

8 min read
A fleet migration contract packet showing a pilot repository, stop conditions, code-owner routing, rollback owner, and a status board for many repository changes.
Industry Insights·

Before an AI agent opens 80 migration PRs, write the fleet contract

An engineer at Mercari went looking for one deprecated call and found roughly 80 repositories that needed the same fix. That number is the real story in Sourcegraph's new agentic migration tool: not whether an agent can write the change, but whether your team has a plan for repo two before repo one finishes.

8 min read
A failure population ledger with many logged AI workflow failures grouped into recurrence clusters by category, tool path, environment, reviewer label, and proof of disappearance.
AI Development·

One bug or two? What OpenAI's 18-year-old crash teaches AI teams about counting failures

OpenAI spent years chasing a crash that looked like one bug and turned out to be two, a bad server and an 18-year-old race condition, both wearing the same symptom. The breakthrough wasn't a clever fix. It was refusing to explain any single crash until they'd counted every crash. AI workflows fail the same way, and most teams still debug them one weird case at a time.

8 min read
An agent effort budget card for a Claude Sonnet 5 migration showing effort level, tool limit, latency target, review gate, fallback model, tokenizer adjustment, and stop rule fields.
AI Development·

Claude Sonnet 5 didn't just get cheaper. Your agents now need an effort policy.

The upgrade note said Sonnet 5 was the most agentic version yet, and everyone read it as a price cut. The operator question buried in the release is different: how hard should this workflow be allowed to try?

8 min read
Three transparent glass test chambers connected by blue, violet, and green light paths, representing compile, deploy, and behavior gates for agent-led Java migration.
Industry Insights·

Before an agent migrates Java, make it earn three receipts

ScarfBench shows AI coding agents can compile migrated Java code and still fail deploy or behavior. Use a migration acceptance bench before giving agents modernization work.

8 min read
A smartphone with a blank glowing control strip sends cyan and violet light threads into translucent panels, suggesting text fields becoming action surfaces.
Small Business AI·

Your phone keyboard is becoming an automation surface

Acti's new agentic keyboard puts AI actions directly under your thumbs, inside the text field you were already typing in, with no chat window and no dashboard to sign off on. That makes it a different kind of rollout, and it means every business with a phone in an employee's hand needs an answer to one question before someone else answers it for you.

7 min read
A cool-toned terrain tile with glowing geospatial cells connected to blank glass verification blocks and beads.
Industry Insights·

The map is not the receipt

An AI can sound certain about a supplier plot, field site, or flood claim. That does not make the answer replayable. emem shows what real-world agents need next: a field-fact receipt that pins down place, source, time, signature, and the decision the fact is allowed to support.

7 min read