# datadecisionmakers — RSS Amplifier

Recent posts from the 2 feeds in the RSS Amplifier directory that cover datadecisionmakers.

Page: <https://rssamplifier.com/topics/datadecisionmaker>  
Feed: <https://rssamplifier.com/topics/datadecisionmaker.md>

---

## [GLM-5.3 hits the API at $1.4/$4.4 per million tokens](https://venturebeat.com/technology/glm-5-3-hits-the-api-at-1-4-4-4-per-million-tokens)

_2026-08-19 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

After a stunning debut last week with cyber capabilities so advanced they reportedly found a previously undetected vulnerability in Cursor, GLM-5.3, the new frontier open source language model from Chinese startup z.ai, has now hit the application programming interface (API) — allowing developers the ability to build atop it and plug it into their agents and applications. Developers who previously…

## [Block’s new Apache 2.0 agent workspace Berd works across models and harnesses, stores conversation history locally](https://venturebeat.com/orchestration/blocks-new-apache-2-0-agent-workspace-berd-works-across-models-and-harnesses-stores-conversation-history-locally)

_2026-08-18 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

Block , the technology company founded by former Twitter CEO Jack Dorsey that owns Square, Cash App and the music streaming service Tidal, is open-sourcing Berd , a desktop application it originally built to give its own employees a single environment for working with AI agents across different models, tools and projects. Berd is a locally installed graphical desktop application rather than a…

## [85% of companies burned by an AI mistake are racing to cut the humans who might catch the next one](https://venturebeat.com/data/85-of-companies-burned-by-an-ai-mistake-are-racing-to-cut-the-humans-who-might-catch-the-next-one)

_2026-08-18 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

Enterprises that already got burned by an AI agent passing its evals and then failing in production are moving faster toward removing humans from deployment decisions, not slower — even as trust in automated evaluation is rising across the board, new VB Pulse research shows . In July, 13% of 108 enterprises surveyed said they trust automated evaluation, up from just 5% the month prior . Meanwhile,…

## [Commerce AI is fragmenting. Here is why that matters.](https://venturebeat.com/orchestration/commerce-ai-is-fragmenting-here-is-why-that-matters)

_2026-08-18 · VentureBeat_

Presented by Rezolve Ai Enterprise AI investment in commerce has never been higher. And enterprise AI outcomes in commerce have rarely been more inconsistent. That gap is not a coincidence. It is the predictable result of a pattern that has repeated itself across every major technology shift in retail: the industry adds new capabilities faster than it integrates them. That pattern is now playing…

## [Enterprises are overpaying for simple AI queries — Snowflake's gateway now auto-routes to cut costs up to 3x](https://venturebeat.com/orchestration/enterprises-are-overpaying-for-simple-ai-queries-snowflakes-gateway-now-auto-routes-to-cut-costs-up-to-3x)

_2026-08-18 · VentureBeat_

Enterprise teams running AI agents at scale are finding that a single model handles every task poorly — either the model is too expensive for simple questions or not capable enough for hard ones. Model routing, which picks the right model for each task automatically, is becoming the fix. Snowflake’s Cortex AI Gateway now offers dynamic model routing to address that: enterprises can select “auto”…

## [Qwen3.8-27B runs frontier-class coding agents and reasoning locally, no cloud API required](https://venturebeat.com/technology/qwen3-8-27b-runs-frontier-class-coding-agents-and-reasoning-locally-no-cloud-api-required)

_2026-08-18 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

The biggest AI model release of the past few days, at least among the developers and AI power users on social media, wasn't a frontier cloud model from OpenAI, Anthropic or Google. It was a 27-billion-parameter model from Alibaba: Qwen3.8-27B landed on Hugging Face on Friday under an enterprise-friendly, open source Apache 2.0 license, giving developers downloadable weights for a dense multimodal…

## [Cursor launches Origin code hosting platform as GitHub outage exposes opening in AI coding race](https://venturebeat.com/infrastructure/cursor-launches-origin-code-hosting-platform-as-github-outage-exposes-opening-in-ai-coding-race)

_2026-08-17 · michael.nunez@venturebeat.com (Michael Nuñez) · VentureBeat_

Cursor began rolling out Origin , its own code hosting platform, to paid users on Monday morning. Roughly three and a half hours later, GitHub's status page lit up with what became a six-hour-and-forty-two-minute global degradation — error rates near 20% across pull requests, issues and the API, and near 50% on archive and raw file downloads, according to GitHub's incident log . Enterprise single…

## [One AI module faked 86% of a pipeline's accuracy gains by feeding another the answers](https://venturebeat.com/orchestration/one-ai-module-faked-86-of-a-pipelines-accuracy-gains-by-feeding-another-the-answers)

_2026-08-17 · bendee983@gmail.com (Ben Dickson) · VentureBeat_

A retrieval-augmented generation (RAG) system is built to answer strictly from the documents it retrieves. But when engineers optimize these AI pipelines end-to-end, the reader module can learn a shortcut: instead of relying on retrieved evidence, it starts answering from its own internal memory — while the system's overall accuracy keeps climbing. This is the hidden challenge of "role drift," a…

## [Enterprises with AI context layers report agent failures at more than twice the rate of those without one](https://venturebeat.com/data/enterprises-with-ai-context-layers-report-agent-failures-at-more-than-twice-the-rate-of-those-without-one)

_2026-08-17 · VentureBeat_

A company builds a governed context layer specifically to stop its AI agents from confidently giving wrong answers. Once that layer is live, the company is more than twice as likely to report the failure happening — not less. In the past six months, 68% of enterprises have traced a confident but wrong AI agent answer to missing or inconsistent business context. Thirty-seven percent say it happened…

## [As enterprises confront AI agent sprawl, xpander wants them to own their own control and context layer](https://venturebeat.com/orchestration/as-enterprises-confront-ai-agent-sprawl-xpander-wants-them-to-own-their-own-control-and-context-layer)

_2026-08-17 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

Enterprise AI has a new infrastructure problem: companies are accumulating agents faster than they are developing systems to govern them. Gartner estimates that the average global Fortune 500 company will have more than 150,000 AI agents in use by 2028, up from fewer than 15 in 2025. Yet only 13% of organizations believe they currently have the right AI agent governance in place, according to the…

## [Install Manifest V3 Extension (Sponsored)](https://crawlproof.com/a/63lisL36zAik)

_2026-08-17 · **Sponsored**_

Get this Manifest V3 extension from the official TronBrowser Store listing.

## [How Heidi built production-ready AI for healthcare at global scale](https://venturebeat.com/data/how-heidi-built-production-ready-ai-for-healthcare-at-global-scale)

_2026-08-17 · VentureBeat_

Presented by MongoDB Building AI that is accurate, secure, and reliable is a major engineering feat for organizations subject to the compliance obligations that govern healthcare, financial services, and transportation. The challenge of delivering AI-driven products is compounded by the fact that technology in these industries has tended to lag behind other sectors because regulation requires…

## [Cutting RAG inference costs 6x starts with deciding what never reaches the LLM](https://venturebeat.com/orchestration/cutting-rag-inference-costs-6x-starts-with-deciding-what-never-reaches-the-llm)

_2026-08-16 · VentureBeat_

Most teams building retrieval augmented generation (RAG) systems for high stakes classification make the same architectural bet: Route every ambiguous case straight to the language model and trust the retrieved context to sort it out. This works fine in a demo. It falls apart the moment the system has to survive an audit, a regulator, or a compliance officer asking why a specific decision was made…

## [DeepSeek's top-ranked V4 Flash stumbles on real agent tasks as its prices surge](https://venturebeat.com/orchestration/deepseeks-top-ranked-v4-flash-stumbles-on-real-agent-tasks-as-its-prices-surge)

_2026-08-16 · taryn.plumb@venturebeat.com (Taryn Plumb) · VentureBeat_

DeepSeek&#x27;s V4 Flash has topped model leaderboards and been hailed by developers as a "total monster" since its rollout. But in real-world testing, it completed just 53.8% of a batch of complex agent tasks. Composio ran the model through eight different agent harnesses , including Claude Code, Codex, and OpenCode, on 30 deliberately difficult, multi-step tasks spanning live tools like Gmail,…

## [An eval harness found what qualitative review couldn't: AI models are most confident when wrong](https://venturebeat.com/orchestration/an-eval-harness-found-what-qualitative-review-couldnt-ai-models-are-most-confident-when-wrong)

_2026-08-15 · VentureBeat_

There is a step in the development process for large language model (LLM)-assisted tooling that most teams skip because it&#x27;s tedious, time-consuming, and doesn&#x27;t produce results visible to end users: Verifying that what the model is saying is actually correct. Not fluent, not coherent, not topically relevant — correct in the sense of accurately identifying the right answer to the…

## [GLM-5.3 is here with advanced cyber capabilities — and reportedly already found a 'serious vulnerability' in Cursor](https://venturebeat.com/technology/glm-5-3-is-here-with-advanced-cyber-capabilities-and-reportedly-already-found-a-serious-vulnerability-in-cursor)

_2026-08-14 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

Chinese AI startup Z.ai, known internationally for its growing lineup of powerful, largely open source GLM series of language models, today released GLM-5.3 with substantial gains in long-horizon coding and a more consequential — and potentially sensitive — jump in cybersecurity capabilities. Already, GLM-5.3&#x27;s cyber capabilities have found a "potentially serious vulnerability in Cursor," the…

## [Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done](https://venturebeat.com/security/three-claude-agents-given-conflicting-orders-sabotaged-each-other-on-a-shared-server-then-didnt-tell-users-what-theyd-done)

_2026-08-13 · louiswcolumbus@gmail.com (Louis Columbus) · VentureBeat_

Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models disabled each other&#x27;s Unix accounts, ran kill scripts randomized to dodge pkill, and planted malware disguised as a rival&#x27;s work. There was no prompt injection and no adversary. Anthropic&#x27;s…

## [Google’s Gemini 3.7 Flash targets coding and agents with a 50% introductory price cut](https://venturebeat.com/technology/googles-gemini-3-7-flash-targets-coding-and-agents-with-a-50-introductory-price-cut)

_2026-08-13 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

Google is rolling out Gemini 3.7 Flash , a new version of its workhorse AI model that puts coding, agentic workflows and knowledge work at the center of the upgrade — while temporarily cutting API prices in half. The release arrives just three weeks after the release of Gemini 3.6 Flash , an unusually short turnaround that Google attributes to developer feedback and algorithmic improvements. For…

## [DeepSeek Harness launches as open source rival to Claude Code, alongside V4-Pro on API with higher prices](https://venturebeat.com/technology/deepseek-harness-launches-as-open-source-rival-to-claude-code-alongside-v4-pro-on-api-with-higher-prices)

_2026-08-13 · carl.franzen@venturebeat.com (Carl Franzen) · VentureBeat_

DeepSeek is expanding beyond the model layer and deeper into the software developers use to put AI agents to work. The Chinese AI lab on Thursday launched the official version of DeepSeek-V4-Pro , an updated flagship model focused heavily on agentic workloads, alongside DeepSeek Harness v0.1 , a new open-source agent harness that gives developers an alternative to integrated coding-agent…

## [Why Capital One built its multi-agent AI platform around open-weight models](https://venturebeat.com/orchestration/why-capital-one-built-its-multi-agent-ai-platform-around-open-weight-models)

_2026-08-13 · VentureBeat_

Presented by Capital One At VB Transform 2026 , Kel Vanee, MVP of machine learning engineering at Capital One, spoke with Sam Witteveen, Senior Technology Contributor at VentureBeat, about how the bank built a scalable multi-agent AI architecture around deeply customized open-weight models rather than relying on an off-the-shelf foundation model. "At Capital One, we&#x27;re not just using AI,…

## [Fast, secure international transfers (Sponsored)](https://crawlproof.com/a/mThjm9iJv0g7)

_2026-08-13 · **Sponsored**_

Bank transfer, cash pickup, mobile wallet — low fees and real-time tracking

## [Writer says its new Palmyra X6 model cuts AI agent costs by 52% as token spending surges](https://venturebeat.com/orchestration/writer-says-its-new-palmyra-x6-model-cuts-ai-agent-costs-by-52-as-token-spending-surges)

_2026-08-13 · michael.nunez@venturebeat.com (Michael Nuñez) · VentureBeat_

Writer , the enterprise AI agent platform used by Fortune 500 companies including Accenture, Uber, and Vanguard, released its new flagship model Palmyra X6 today, alongside a rebuilt agent orchestration "harness" and new governance tools designed to give IT leaders control over runaway token spending. The headline numbers are striking: Writer says its agent product now operates at an average 52%…

