Multimodal models
Models that read or produce images, audio or video, and attacks that hide inside those inputs.
- All items
- 19
- Last 90 days
- 10
- Change
- +100%vs 5 before
Items per month
| Month | Items |
|---|---|
| May 2025 | 0 |
| Jun 2025 | 0 |
| Jul 2025 | 0 |
| Aug 2025 | 0 |
| Sep 2025 | 0 |
| Oct 2025 | 0 |
| Nov 2025 | 0 |
| Dec 2025 | 0 |
| Jan 2026 | 1 |
| Feb 2026 | 1 |
| Mar 2026 | 2 |
| Apr 2026 | 2 |
| May 2026 | 1 |
| Jun 2026 | 2 |
| Jul 2026 | 1 |
| Aug 2026 | 0 |
| Sep 2026 | 4 |
| Oct 2026 | 5 |
2 items
ThreatsDay: Android Spyware, PLC Attacks, AI Image Prompt Injection + 12 More Stories
Jul 23, 2026MediumNewsSecurityResearchGitHub will begin rejecting command-line support bundle uploads from older GHES appliances starting August 18, 2026, unless they are patched. A separate npm package, @copilot-mcp/apex, acts as a postinstall dropper that installs a macOS infostealer, and a fake VS Code extension, "Markdown All Pro", impersonates Markdown All in One to beacon machine details and fetch remote payloads.
Fix: To avoid disruption when submitting support bundles, update your GHES instance to the latest patch release available for your current version line. At minimum, the required patch versions are: 3.21.3, 3.20.5, 3.19.9, 3.18.12, and 3.17.18.
The Hacker NewsIntroducing Gemma 4 12B: a unified, encoder-free multimodal model
Jun 9, 2026InfoNewsIndustryResearchGoogle DeepMind introduced Gemma 4 12B, a mid-sized model with native audio inputs designed to run locally on laptops with 16GB of VRAM or unified memory. Its encoder-free architecture feeds vision and audio inputs directly into the LLM backbone, and the model is released under an Apache 2.0 license with Multi-Token Prediction drafters to reduce latency.
DeepMind Safety Research
Topic added 2026-10-09. An item belongs to this topic when its title matches one of the topic's patterns or its summary mentions the topic at least twice. Report a wrong match with the feedback button on the item.