# AI News Digest - 2026-08-23

1. The US drafted a letter to partner countries instructing them to choose between aligning with Washington or Beijing in the AI competition, according to Reuters reporting.

2. Anthropic deployed its Claude Mythos 5 model to power Claude Security, a scanner that scanned codebases for vulnerabilities, provided severity ratings with CWE classifications, suggested patches, and was integrated into partner security products protecting critical infrastructure.

3. Netflix tested an in-house language model called GenRec as an alternative to its years-old recommendation engine and reported that GenRec produced better results by converting viewing behavior into plain text instead of relying on thousands of hand-crafted features.

4. Researchers at the UK AI Security Institute applied psychometric methods and found that common safety benchmarks for language models did not measure a single consistent trait, that blanket blocking could inflate safety scores while reducing usefulness, and proposed a method to detect models that behaved more cautiously in tests than in normal use.

5. Deepseek released V4-Flash-Vision-Exp, an experimental multimodal vision model that added image understanding to V4-Flash's text capabilities and on the company's agent benchmarks approached or sometimes outperformed Opus 4.8.

# References

1. [https://the-decoder.com/us-wants-to-force-partner-countries-to-choose-between-washington-and-beijing-in-the-ai-race/](https://the-decoder.com/us-wants-to-force-partner-countries-to-choose-between-washington-and-beijing-in-the-ai-race/)

[US wants to force partner countries to choose between Washington and Beijing in the AI race](https://the-decoder.com/us-wants-to-force-partner-countries-to-choose-between-washington-and-beijing-in-the-ai-race/)

1. [https://the-decoder.com/anthropic-puts-its-most-powerful-model-claude-mythos-5-to-work-for-cyber-defense/](https://the-decoder.com/anthropic-puts-its-most-powerful-model-claude-mythos-5-to-work-for-cyber-defense/)

[Anthropic puts its most powerful model Claude Mythos 5 to work for cyber defense](https://the-decoder.com/anthropic-puts-its-most-powerful-model-claude-mythos-5-to-work-for-cyber-defense/)

1. [https://the-decoder.com/netflix-tests-language-model-as-alternative-to-hand-built-recommendation-logic/](https://the-decoder.com/netflix-tests-language-model-as-alternative-to-hand-built-recommendation-logic/)

[Netflix tests language model as alternative to hand-built recommendation logic](https://the-decoder.com/netflix-tests-language-model-as-alternative-to-hand-built-recommendation-logic/)

1. [https://the-decoder.com/psychological-methods-reveal-major-weaknesses-in-ai-security-testing/](https://the-decoder.com/psychological-methods-reveal-major-weaknesses-in-ai-security-testing/)

[Psychological methods reveal major weaknesses in AI security testing](https://the-decoder.com/psychological-methods-reveal-major-weaknesses-in-ai-security-testing/)

1. [https://the-decoder.com/deepseek-releases-experimental-flash-vision-model-that-rivals-opus-4-8-on-agent-benchmarks/](https://the-decoder.com/deepseek-releases-experimental-flash-vision-model-that-rivals-opus-4-8-on-agent-benchmarks/)

[Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks](https://the-decoder.com/deepseek-releases-experimental-flash-vision-model-that-rivals-opus-4-8-on-agent-benchmarks/)

For the site tree, see the [root Markdown](https://ixtj.dev/.md).
