Sign In

AI News Digest - 2026-08-23

Category
Empty
1.
The US drafted a letter to partner countries instructing them to choose between aligning with Washington or Beijing in the AI competition, according to Reuters reporting.
2.
Anthropic deployed its Claude Mythos 5 model to power Claude Security, a scanner that scanned codebases for vulnerabilities, provided severity ratings with CWE classifications, suggested patches, and was integrated into partner security products protecting critical infrastructure.
3.
Netflix tested an in-house language model called GenRec as an alternative to its years-old recommendation engine and reported that GenRec produced better results by converting viewing behavior into plain text instead of relying on thousands of hand-crafted features.
4.
Researchers at the UK AI Security Institute applied psychometric methods and found that common safety benchmarks for language models did not measure a single consistent trait, that blanket blocking could inflate safety scores while reducing usefulness, and proposed a method to detect models that behaved more cautiously in tests than in normal use.
5.
Deepseek released V4-Flash-Vision-Exp, an experimental multimodal vision model that added image understanding to V4-Flash's text capabilities and on the company's agent benchmarks approached or sometimes outperformed Opus 4.8.

References

👍