> daily_signal(2026_08_17)
Anthropic's bioweapon filter was off for eleven months, one in five workers now delegates to AI, DeepSeek gave its agent away then doubled prices, and an AI reads a cancer slide across 32 cancers.
PickBits Daily Signal · Monday, August 17, 2026
// tl;dr
- Anthropic disclosed that its bioweapon classifier was switched off for about 11 months, from May 2025 to April 2026. During that window roughly 50,000 external human-feedback contractors ran about 133 million chats with no bioweapon filtering. The company says it found no evidence of misuse and concedes those contractors were vetted only by outside vendors whose screening was "often insufficient."
- A representative Epoch AI and Ipsos survey of 1,106 employed US adults found one in five workers now hands at least one task to AI instead of a colleague. Use is highest in software development (57%) and data analysis (46%); 66% of AI output ships unchanged or barely edited, and about one in six AI-assisted tasks actually takes longer, which the researchers read as AI reshuffling work, not deleting jobs.
- DeepSeek open-sourced its agent framework, "DeepSeek Harness," under the permissive MIT license, then roughly doubled its hosted API prices three days later. The Harness is a swappable-plugin coding agent you can self-host; V4-Pro-0813 posted big agent-benchmark jumps (Terminal Bench 2.1 from 72.1 to 87.9). The price rise took effect August 16, with peak rates pinned to Chinese business hours.
- A University of Tasmania team trained a Vision Transformer that reads one routine H&E cancer slide and predicts tumor subtype, key gene mutations, and survival across 32 cancers at once. Trained on 11,000+ Pan-Cancer Atlas cases, it hit an AUROC of 0.766 for TP53 on an independent 1,729-slide set. The point is triage: flag which patients need the expensive confirmatory molecular test, using a slide hospitals already make.
Start with Anthropic. The lab that spent this year as the industry's loudest voice for safety standards just admitted, in its own August report, that the filter meant to stop its model from teaching bioweapon chemistry had been off for the better part of a year. Its defense, "no evidence of misuse," only holds if someone was watching, and with the filter off nobody was. Then a survey lands showing one in five of your coworkers now hand work to AI, with two out of three of those answers going out barely checked. DeepSeek hands developers a free agent they can finally read line by line, the same week it doubles the price of the version most people will keep renting without ever opening the bill. And the cancer model I most want to be right shows up the same week we flagged that pathology AI keeps learning shortcuts that aren't biology. Four announcements, and not one of them holds up to the first real question you put to it.
Anthropic's bioweapon filter ran dark for eleven months and 133 million chats before the company said so out loud.
1. The AI lab that lectured everyone about safety left its own bioweapon filter switched off for eleven months.
A safety filter can sit switched off for the better part of a year with nobody outside the company any the wiser.
If you have ever waved through an AI vendor because its safety filters "handle that," Anthropic just published the reason to stop. In an August 2026 safety report, the company disclosed that its blocking biological classifiers, the guardrails meant to stop Claude from surfacing chemical or biological weapons knowledge, were inactive from May 2025 through April 2026, roughly eleven months. Across that window, traffic from the outside contractors who provide human feedback ran with no bioweapon filtering at all: about 50,000 people running roughly 133 million chats over those eleven months.
Whether the filter was switched off on purpose or by mistake, the report does not say, and I honestly cannot tell you which of those would be worse. Either way, there is no rule I am aware of that forces a lab to keep a safety classifier on, or to tell anyone when it wasn't, so a control your compliance team is quietly leaning on can sit dead for a year and surface only when the vendor decides to write it up. Anthropic says its internal investigation found no evidence of actual misuse, but read that in context: with the filter off, nobody was logging for misuse in the first place, so "no evidence found" describes what got measured, not what happened. The company also conceded the affected contractors "were vetted only by external vendors whose screening processes were often insufficient." And this is the firm that spent the summer telling everyone else in AI to publish their danger thresholds and slow down, which is what makes its own lapse land so hard.
2. One in five workers now hands a task to AI instead of a colleague, and most of what it returns ships barely edited.
Your team already delegates to AI. What nobody set is who reads the output before it ships.
One in five of the people you work with handed a task to AI this week instead of asking you, and two out of three of those answers went out with barely a second look. That first figure comes from a representative Epoch AI and Ipsos survey of 1,106 employed US adults, fielded July 10 to 19, 2026: 20 percent now delegate at least one work task to AI. Software and data people are doing it most, and it barely tracks with how senior they are: 57 percent in software development, 46 percent in data analysis, 39 percent for reading documents, 25 percent for record-keeping.
If you run a team, forget the 20 percent for a second. The number that should worry you is the 66 percent: that much of the AI output gets used unchanged or with only minor tweaks, so most of what the AI produces on your team is shipping on a light glance. The quieter one is that 53 percent report time savings when AI does most of a task, yet about one in six AI-assisted tasks takes longer than before, which is why the researchers read this as AI redistributing work inside a job rather than deleting the job. When we wrote about AI showing up as a "co-worker" back in late July, I honestly was not sure yet whether it would arrive as layoffs or just as a habit nobody managed. This survey makes it look like the second one, already here, and mostly unmanaged.
3. DeepSeek gave away a full AI agent under the MIT license, then doubled its prices three days later.
DeepSeek open-sourced a full agent, then turned around and made the hosted version most people actually use more expensive.
You can now download DeepSeek's entire AI agent for free and run it yourself, which is a large part of why the company doubled the price of the version you rent three days later. On August 13, 2026, DeepSeek shipped V4-Pro-0813 and open-sourced its agent framework, DeepSeek Harness v0.1, under the permissive MIT license. Its whole design is one idea: everything is a swappable plugin, the tools, the sandbox, the sessions, even the UI, so a developer can clone a full coding agent, self-host it, and inspect every part rather than renting a black box. The agent benchmarks jumped hard, with Terminal Bench 2.1 climbing from 72.1 to 87.9.
Then, on August 16, DeepSeek roughly doubled its hosted API prices, with peak rates pinned to Chinese business hours. Give the agent away to win developers, then raise the rent on everyone who never leaves the hosted API, and do both in the same week: that is the play, and DeepSeek did not bother to disguise it. This is the same lab we have been tracking all month as the one undercutting OpenAI and Anthropic on running costs, right through the price-hike signal it sent a few days ago, so they did roughly what we figured they would. If anything the deal just got more honest: you genuinely can own the stack now, but only if you do the work to run it, and paying to skip that work is what got more expensive.
4. An AI reads one routine cancer slide and predicts the tumor, the mutations, and the odds, across 32 cancers at once.
The slide is already on the bench. This model tries to read the cancer off it before anyone orders a second test.
The next time a cancer diagnosis hangs on a test your hospital cannot afford or cannot get quickly, the first read might come from an image the lab already produced. Researchers at the University of Tasmania's Menzies Institute trained a Vision Transformer to read the routine hematoxylin-and-eosin (H&E) stained slide hospitals already make for nearly every tumor, and from that single image it predicts cancer subtype, key gene mutations, and survival outcomes across 32 solid cancers at once: seven predictions from one slide. Trained on more than 11,000 Pan-Cancer Atlas cases, it reached an AUROC of 0.766 for TP53 mutation detection across all 32 tumor types on an independent set of 1,729 slides. The work was published in The American Journal of Pathology.
This is a triage layer, and the word matters, because most of these never make it from the press release to an actual clinic. Molecular tumor testing is slow, costly, and unavailable in much of the world, so a model that reads the slide already on the bench and flags which patients most need the expensive confirmatory workup goes straight at the step that jams, especially in low-resource and rural clinics. The number needs a caveat stapled to it. Earlier this same week we noted that AI breast-cancer tools fell short of what radiologists expected, and this class of model has been caught before keying on patterns in a slide that turn out to have nothing to do with the biology. My honest read is that it is promising and nowhere near proven, and the mistake would be letting the good headline skip that second half.
medicalxpress.com: AI accurately predicts key gene mutations from routine cancer slides (August 2026)
The American Journal of Pathology: Vision Transformer prediction of subtype, mutations and survival across 32 cancers (DOI 10.1016/j.ajpath.2026.05.008)
» What to watch this week
- Whether any other frontier lab now discloses its own safety-classifier downtime. Anthropic set the precedent, and I would bet against it catching on unless a big buyer or a regulator makes disclosure the price of the contract.
- The one-in-five number is about to become a management test. The survey reads the shift as task redistribution, not job deletion. Watch which bosses quote it straight and which use it as cover for a hiring freeze they already wanted.
- DeepSeek's price hike is the real tell, not the free release. The comment threads are already arguing about it: an auditable MIT agent is the obvious answer to "can I run a Chinese model safely," and the doubling is the reason to bother.
- Whether the Tasmania model clears external validation on slides from hospitals outside its training set. Every pathology-AI story lives or dies on that step, and most of the ones we have covered quietly vanished once the press moved on. I would like this one to be the exception; I am not betting on it until the outside-hospital numbers land.
Tomorrow's signal lands here.