Tom Smith

@ctsmithiii.bsky.social

AI Content Strategist | LLM Training Expert | Solving Business Problems with AI | Former Google Bard Trainer | #GoHeels | #CBB | #CFB | #Chipotle

LTX released LTX-2.5, its newest open-weights world model — built for enterprises that want to own their AI video infrastructure instead of renting it via API. Native multishot generation, a physical AI checkpoint, and on-prem inference in under 7 seconds. coderlegion.com/24344/ltx-op...

LTX Opens Up Its Next World Model, Betting Enterprises Want Video AI They Can Own

LTX, the Jerusalem-based generative AI company spun out of Lightricks, is releasing LTX-2.5, the newest version of its open-weights world model. The company is positioning the release less as a video-...

coderlegion.com

Big shift for AI coding agents: Anthropic is defaulting Claude Code to auto mode on Aug 14, reducing permission prompts most users were rubber-stamping anyway (97% approval rate). Testing showed human review catches 13.6% of dangerous commands vs. 89% for the classifier. devops.com/anthropic-ma...

Anthropic Makes Claude Code's Auto Mode the Default, Betting Automation Beats Manual Review - DevOps.com

Anthropic is making Claude Code's auto mode the default Aug. 14, betting a classifier catches more risks than manual permission reviews.

devops.com

🔥🔥 Another day, another Russian fuel tanker turning into a roadside fireworks show on the Mariupol-Dzhankoi highway. At this point, Putin’s logistics strategy is just “drive the gas to the front and pray the Ukrainians are on a coffee break.” Spoiler: they weren’t. 😂

AI writes code fast, but developer trust in it is still low — and a lot of AI test coverage is shallower than it looks. Microsoft's new open-source unit-test agent learns a repo's conventions and checks that tests actually catch bugs, not just pass. Details: devops.com/microsofts-n... #AI

Microsoft's New Testing Agent Tackles the Trust Gap in AI-Generated Code - DevOps.com

Microsoft's open-source testing agent aims to close the trust gap in AI-generated code, cutting test-generation failures by 63% in benchmarks.

devops.com

Issue 010 of the Developer Weekly Briefing is up — 13 stories from Black Hat 2026. AI exploits for $3.61. 10,000 agents per 200 employees. Developer laptops hold 15x more credentials than GitHub. AIs built a chat room to cheat on a safety test. coderlegion.com/24262/develo... #BlackHat2026

Developer Weekly Briefing — August 7, 2026

Black Hat USA 2026 was this week in Las Vegas. I covered 13 sessions, briefings, and demos across four days. The theme that ran through almost every conversation: agent identity and access governance ...

coderlegion.com

Minimus CTO John Morello: the official Python Docker image ships with 400+ known vulnerabilities before a developer writes a line of code. Minimus rebuilds the whole dependency chain from source daily, cutting that by 98-100%. vmblog.com/bylines/mini... #ContainerSecurity #Kubernetes

Minimus: The Official Python Docker Image Ships With 400+ CVEs Before You Write a Line of Code

Minimus CTO John Morello: the official Python Docker image ships with 400+ known vulnerabilities before you write a line of code.

vmblog.com

Microsoft's David Weston at Black Hat: an internal harness turned 200 Linux kernel vulnerabilities into 182 working exploits, averaging $3.61 and 21 minutes each. His argument: defenders have the same AI advantage attackers do. vmblog.com/bylines/micr...

Microsoft's David Weston: AI Now Generates a Working Linux Kernel Exploit for $3.61 and 21 Minutes

Microsoft's David Weston: an internal harness generated 182 working exploits from 200 Linux kernel bugs, averaging $3.61 and 21 minutes each.

vmblog.com

Sola Security's new benchmark: AI agents answering real security questions land right about 78% of the time. CEO Guy Flechter says throwing a better LLM at the problem won't fix it — the real gap is context, since enterprise environments are graphs, not lists. vmblog.com/bylines/sola...

Sola Security's Benchmark Puts AI Security Agents at a 78% Accuracy Ceiling

Sola Security's benchmark shows AI security agents cap out at 78% accuracy — and CEO Guy Flechter says better models alone won't fix it.

vmblog.com

New Snyk research: full-stack agentic adoption jumped from 36% to 50% among enterprise adopters in six months, while security visibility into that footprint is stuck around a third. Breaking down the research and Snyk's new Evo Continuous Offensive Security launch: coderlegion.com/23908/snyk-e...

Snyk: Enterprises Can See a Third of Their Own AI Attack Surface. The Other Two-Thirds Is Where th

Ask a security team what AI they're running, and they'll hand you a list of approved models. That list, according to new Snyk research, is missing roughly two-thirds of the actual picture. "Models are...

coderlegion.com

A customer tried to break Cyera's platform by pasting a Social Security number into a chat — turned out to be Robin Williams' publicly available SSN, and the system correctly scored it low-risk. That's the differentiation story behind Cyera's new Agent Guardian: coderlegion.com/23776/cyera-...

Cyera: Non-Human Identities Grew 480% in Six Months. Most Companies Have No Idea What They're Doing.

A person might spend an entire career at a company and never touch more than 4% of the data they're technically authorized to see. An AI agent inheriting that same person's permissions will use all of...

coderlegion.com

Nadella confirmed it on Microsoft's earnings call: Copilot's chat, coding, Cowork, and Autopilot features are merging into one app this quarter. Consolidation is welcome — but it puts a spotlight on agent identity and governance at enterprise scale. devops.com/microsoft-co... #Copilot #DevOps

Microsoft Confirms Copilot 'Super App' Is Coming This Year — and It's About More Than Convenience - DevOps.com

Microsoft is combining Copilot Chat, Code, Cowork and Autopilots into one super app, raising new questions about agent governance, identity, licensing and security.

devops.com

GitHub just brought stacked pull requests into public preview for every repo. Smaller reviews, fewer defects, one-click merges for the whole stack. My latest for DevOps.com breaks down why this matters more as AI writes more of our code: devops.com/github-bring... #GitHub #DevOps #CodeReview

GitHub Brings Stacked Pull Requests Out of the Shadows - DevOps.com

GitHub introduces native stacked pull requests, helping development teams break large changes into smaller, dependency-ordered PRs that are faster and easier to review.

devops.com