Same ChatGPT Prompt Can Yield Different Answers, Study Finds
Anthropic CEO Rejects Banning Chinese Open-Weight AI Models
July 28, 2026
D.A.D. today covers 5 stories — about a 2-minute read. What's New, What's Innovative, What's Controversial, What's in the Lab, and What's in Academe.
The Daily AI Digest is a daily AI briefing automated by Alexander Panetta — a veteran political journalist tracking the field during a Master's in AI Management at Georgetown University.
D.A.D. Joke of the Day: I trained an AI on all my emails so it could write like me. Now it opens every message with "Sorry for the delay."
What's New
AI developments from the last 24 hours
Anthropic CEO Rejects Calls to Ban Chinese Open-Weight AI Models
Anthropic CEO Dario Amodei published a post distancing his company from proposals to ban Chinese open-weights AI models, a debate that's split the industry (D.A.D., July 21-25). Amodei said Anthropic has never advocated banning open-weights models outright. His actual concerns, he wrote, are authoritarian governments building militarily superior AI and bad actors misusing open models for cyberattacks or bioweapons—problems he says are better addressed through chip export controls and cracking down on smuggling than blanket bans.
Why it matters: As Washington weighs restrictions on Chinese open-weights models like Kimi and DeepSeek, Anthropic's public reframe signals it wants to shape that policy debate without being cast as anti-open-source.
Discuss on Hacker News · Source: anthropic.com
What's in the Lab
New announcements from major AI labs
Cohere Lets Businesses Chain Multiple AI Agents Into One Workflow
Cohere launched North Automations, a feature for its North enterprise AI platform that lets companies chain together multiple AI agents into coordinated workflows using plain-language instructions, with built-in spending controls and oversight. Cohere points to Gartner research suggesting many companies' AI agent projects underperform because agents work in isolation rather than together, and cites an internal example: its own marketing team using the tool to pull campaign data from BigQuery automatically. The company also published a hypothetical case study aimed at regulated fields, showing how a wealth manager could automate client research, market monitoring, and compliance checks. No independent performance data was provided.
Why it matters: Most companies today have scattered, single-purpose AI agents that don't talk to each other—orchestration tools like this aim to turn them into an actual workflow, and Cohere is clearly targeting regulated industries like wealth management where compliance work is time-consuming and high-stakes.
What's in Academe
New papers on AI and its effects from researchers
Dementia Therapists Get an AI Assistant for Memory Sessions
Researchers unveiled RemiAssist, an AI tool built to help therapists—not patients—run photo-based reminiscence therapy for people with dementia. The system organizes life-event photos into a structured map for session planning and offers real-time prompts to guide conversations and navigate emotionally sensitive moments. In a small field study of eight therapist-patient pairs, RemiAssist was linked to a 44% gain in planning efficiency and 54% longer conversations, along with better-handled sensitive exchanges.
Why it matters: It's an early example of AI aimed at augmenting a clinician's judgment in a delicate, human-centered therapy rather than automating the interaction itself.
Same ChatGPT Prompt Can Yield Different Answers, Study Finds
A new paper delivers an inconvenient finding for researchers using ChatGPT and similar tools to classify or score data: the same prompt sent twice can yield different answers, even with settings meant to eliminate randomness. The culprits go beyond deliberate randomization—silent model updates, rounding errors, and internal routing all inject variability, and none of this disappears just by turning "temperature" (a randomness dial) down to zero. Even open-weight models run locally aren't fully reproducible unless the entire hardware setup is identical. The authors propose reporting standards so studies using LLMs can be properly replicated.
Why it matters: As more academic and business research leans on AI to classify documents, score sentiment, or extract data, this suggests those outputs should be treated as statistical estimates with margins of error—not fixed facts—which has real implications for anyone citing AI-generated analysis as evidence.
A Way to Prove AI Can't Misuse Your Mental Health Data
Digital phenotyping tools infer mental health conditions from behavioral data like spending patterns and phone use, raising obvious privacy and consent concerns. Researchers built a formal system that encodes ethical rules—like when data can be used—as logical constraints, then used automated verification software to prove those rules can't be violated within the model. Applied to a case combining financial and mental health data, the approach checks compliance continuously rather than through after-the-fact audits or paperwork. There is no real-world performance data yet—this is a proof-of-concept.
Why it matters: As AI systems increasingly infer sensitive personal states from mundane data like transactions, this kind of built-in, verifiable guardrail could become a compliance requirement rather than a research curiosity.
What's Happening on Capitol Hill
Upcoming AI-related committee hearings
Wednesday, July 29 — Hearings to examine the impact of AI on the workplace. Senate · Senate Health, Education, Labor, and Pensions Subcommittee on Employment and Workplace Safety (Open Hearing) 430, Dirksen Senate Office Building
Wednesday, July 29 — Hearings to examine the AI deception machine, focusing on deepfakes, chatbots, and the new frontier of senior fraud. Senate · Senate Aging (Special) (Open Hearing) 562, Dirksen Senate Office Building
Thursday, July 30 — Hearings to examine intelligent networks, focusing on powering artificial intelligence and transforming communications. Senate · Senate Commerce, Science, and Transportation Subcommittee on Telecommunications and Media (Open Hearing) 253, Russell Senate Office Building
What's On The Pod
Some new podcast episodes
The Cognitive Revolution — Nathan Goes to China – Part 1: Tech & Agent Setup, Chinese AI UX, WAIC, and Attitudes on AI
How I AI — From zero coding background to hardware hacker: How Cursor + a Raspberry Pi makes AI fun
AI in Business — Building the Infrastructure Behind AI-Enabled Field Service - with Deniz Mullis of Cytiva