Skip to main content

Blog

Page 10 of 57

All Articles

Insights on AI, machine learning, and technology strategy

A hand holds a light tan, creased defective coffee bean above a white cupping bowl with surrounding walnut bean trays and plain coffee-quality tools.
AI Development·

Cisco Antares can find likely vulnerable files. It cannot decide whether to patch.

Cisco trained compact models to locate files related to a known vulnerability class. The benchmark shows where that helps and where analyst verification still begins.

8 min read
A constructed approval diagram shows a named data request entering one API endpoint, crossing an intentionally opaque model route, and returning an answer with token usage and cost before a reviewer chooses allow, restrict, or stop.
AI Development·

Fugu-Cyber hides its model route. Can your security policy allow that?

Sakana AI exposes Fugu-Cyber through one API and reports cost per request, but it does not disclose the model route behind each query. Approval depends on whether policy can govern that opaque boundary.

10 min read
A constructed repository review diagram separates GitHub Code Quality costs into active committers, AI credits, and scan compute, then shows keep, narrow, or disable as owner decisions. No usage, score, savings, or outcome is filled in.
AI Development·

GitHub Code Quality is billing now. Decide which repositories justify all three meters.

GitHub Code Quality now has active-committer, AI-credit, and scan-compute costs. Review enabled repositories before preview scope becomes unexamined spend.

8 min read
A constructed audience request splits into purchase-gap, churn-risk, sports-interest, and store-proximity clauses. Each clause has an illustrative source, freshness check, and supported or review status.
Industry Insights·

The AI can write the segment. Your customer record still has to prove every clause.

Dotdigital’s Segment Agent turns a plain-language audience request into a segment. Its example reveals the data, definitions, and fallback decisions marketers still need.

8 min read
Pilot handoff folder organized into the workflow change, evidence and limits, and the next decision.
Small Business AI·

What a first AI pilot should leave behind

A useful first AI pilot leaves a workflow map, source boundary, acceptance evidence, known exclusions, a prototype or decision memo, and a clear next decision.

8 min read
A constructed permission-rule test compares Edit(src/**) at a root src file and a nested packages/api/src file. Allow and hook-if match only the root path, while ask and deny match both depths; a precedence rail places deny before ask before allow.
AI Development·

Claude Code 2.1.214 changed permission behavior across shells and rules

Claude Code 2.1.214 changed how path rules, shell commands, remote confirmations, and Docker or Podman daemon flags reach allow, prompt, and block decisions.

8 min read
A constructed six-row independence ledger separates Apache-2.0 source rights from the build, model endpoint, authentication, update, network, and operator-ownership evidence required for vendor independence.
AI Development·

Grok Build is open source. Vendor independence is a separate question.

Grok Build's Apache-2.0 source release grants real rights to inspect, use, modify, fork, and redistribute the coding client. Those rights do not by themselves replace its hosted models, authentication, updates, or other runtime services.

8 min read
Constructed diagram
AI Development·

Copilot code review now reads instructions from the pull request branch

GitHub moved Copilot code review instructions to the pull request head branch and separated review setup, runner, and firewall controls. Teams should protect the files that shape automated review.

7 min read
A constructed comparison bench sends the same long coding task through Codex 0.144.5 and 0.144.6. Each client lane shows a different bundled-metadata and compaction checkpoint, then both expose retained constraints, cited files, tests and build evidence, and unresolved failures for review.
AI Development·

Codex 0.144.6 corrected its context-window metadata. Test long sessions before standardizing it.

Codex CLI 0.144.6 changed bundled context-window metadata for three GPT-5.6 models from 372,000 to 272,000 tokens. Test one known long-running workflow under both client versions before making the patch standard.

7 min read
A constructed decision frame requires separate task-acceptance and operations-ownership reviews before a team chooses an open, hybrid, or managed deployment.
AI Development·

Open models are easy to access. Production still needs an owner.

A 2026 survey commissioned by Mozilla and fielded by SlashData found that 51% of open-model adopters reported reaching production, compared with 63% of closed-model adopters. Before switching models, separate task quality from the operating work your team must own.

8 min read
A constructed multilingual guardrail replay matrix compares complete threads and benign or harmful reviewer labels across four required language slices, leaving false-positive, false-negative, p90, and p99 evidence open and holding release until every critical slice passes.
AI Development·

How to test AI guardrails on multilingual, multi-turn traffic

A strong guardrail result on one benchmark does not settle a production decision. Test representative languages, full conversations, false positives, and tail latency on the traffic your assistant will actually handle.

7 min read
A constructed comparison bench sends the same task through Google Search grounding and Parallel Web Search grounding. The Parallel lane crosses a separate-provider boundary with a rewritten query, and both lanes return evidence and citation packets for review.
Industry Insights·

How to evaluate Parallel Web Search grounding in Gemini

Gemini can now use Parallel Web Search for grounding. The option also adds a separate provider, rewritten-query transfer, Preview terms, and search charges.

8 min read