Skip to main content

Blog

Page 15 of 60

All Articles

Insights on AI, machine learning, and technology strategy

Cyan, amber, and violet glass beads loop through a clear ring as cyan and amber light trails enter from opposite sides.
Industry Insights·

Bot detection now reads the journey between clicks. Observe before you enforce.

Cloudflare Precursor can evaluate behavior across a session. Before tightening a login or checkout rule, test how legitimate input journeys appear and where people leave.

7 min read
Constructed diagram
AI Development·

Before you switch AI models, put the bake-off on trial

Ploy's GPT-5.6 migration looked worse until its team repaired the evaluation harness, tool schema, cache design, and reasoning replay.

9 min read
Three illustrative case records from the same workflow diverge into successful, pending-resolution, and failed states while each retains its proof, owner, safe next transition, and retry authority.
Small Business AI·

Three transactions enter the same AI workflow. Only one finishes cleanly.

A clean completion, a human pause, and a technical failure show what managers need from an AI workflow after the demo ends.

9 min read
An illustrative six-clip timeline shows a human save creating revision 26 while an agent full write based on revision 25 stops at a stale-state conflict; reversal and export remain separate decisions.
AI Development·

When the editor and the agent grab the same cut

FableCut exposes one shared video timeline to humans and agents. Its most revealing behavior appears when both try to change the same cut.

8 min read
An illustrative comparison packet tests a claimed AI output against an accessible public alternative under comparable conditions before a locked capability-gain step allows broader severity factors to be considered.
Technical Tutorials·

Start AI jailbreak triage with capability gain

Anthropic's early Cyber Jailbreak Severity proposal gives security teams a useful first question: what attacker capability did the AI output add beyond public tools and information?

10 min read
An illustrative route-audition field strip follows one robot route through a person crossing, reflective wall, lighting change, moved shelf, and endpoint-orientation check before holding the sensor decision.
Machine Learning·

One camera may be enough for the robot. The route still has to prove it.

Mistral's Robostral Navigate follows language instructions with one RGB camera. Use one local route to evaluate camera-only behavior under ordinary disruption, changing light, moved objects, and endpoint tolerances.

10 min read
An illustrative contact sheet follows one creative asset through maker, freelancer, approver, platform declaration, and customer disclosure while an AI-edited origin record stays attached to each version.
Small Business AI·

Your ad may disclose AI use. Your creative handoff still has to remember it.

Google is adding AI-use details to My Ad Center. The disclosure will only be as reliable as the origin fact that survives the creative handoff.

7 min read
A review sheet holds the question “Are onboarding delays getting worse?” above three chart intents: median activation time falls, overdue account count rises, and overdue cohort rate falls.
Technical Tutorials·

Your AI agent can compile a beautiful chart. It can still answer the wrong question.

Microsoft Flint gives AI agents a compact chart language. Use a chart-intent diff and one-question/three-intents test to inspect fields, denominators, cohorts, and viewer inference before approval.

9 min read
A middle-market AI fit map connects six pre-built vendor lanes to local prerequisites like ERP custom code, shared-drive customer data, data owners, integration gaps, and rollout proof.
Industry Insights·

The middle market is becoming an AI product category

Accenture and Google Cloud packaged enterprise AI into six pre-built lanes for companies between $300 million and $3 billion in revenue. The technology is standardized. The hard part is still local.

7 min read
A public-output quarantine test receipt checks source request, private context touched, leak terms, reply scope, and reviewer sign-off before an agent posts a public reply.
Industry Insights·

Before a repo agent comments in public, quarantine the output

A crafted public GitHub issue tricked an agentic workflow into posting private repo contents as a public comment. Narrower read access wouldn't have stopped it alone — the write path needed its own check.

8 min read
A conversation baton log records caller intent, active voice model, delegated worker, handoff time, returned context, and owner of record for a live voice AI call.
Industry Insights·

Voice AI can delegate mid-call now. Log who's holding the baton.

The voice model keeps listening while it hands the hard part to another model in the background, then picks the conversation back up like nothing happened. That's the feature. It's also the reason nobody can reconstruct what occurred during the handoff.

8 min read
A docs witness record turns a merged implementation PR into a draft documentation PR with captured facts, review owner, and landing gate.
Industry Insights·

The docs PR should land before the memory fades

Microsoft's Aspire team turned merged product PRs into draft documentation PRs automatically. The numbers are good. The reason it works is that almost none of the judgment calls were left to the agent.

9 min read