Sign In

AI News Digest - 2026-07-20

Category
Empty
1.
Alibaba unveiled Qwen 3.8, a multimodal AI model with 2.4 trillion parameters, made an open-weight preview available, and stated the model rivaled leading systems while trailing only Fable 5.
2.
Moonshot's Kimi K3 topped the Code Arena: Frontend rankings, outperforming Claude Fable 5 and GPT-5.6 Sol, but scored about 39% on FrontierMath Tier 4 compared with roughly 90% for models from OpenAI and Anthropic.
3.
Google DeepMind repurposed a video generator in GenCeption to perform classic computer vision tasks such as depth estimation and segmentation, matching state-of-the-art systems while training on far less data and using predominantly synthetic videos.
4.
RadLE 2.0 benchmark showed many AI models for radiology produced incorrect findings with full confidence, and human radiologists remained substantially more accurate.
5.
Epoch AI tested three leading AI text detectors—Pangram, GPTZero, and Originality.ai—and found up to 18% of AI-generated passages went undetected overall and up to 48% for scientific writing.

References

👍