Google Launches Efficient Multimodal AI Model for Laptops
In brief
- Google has unveiled Gemma 4 12B, a new multimodal AI model designed to run efficiently on laptops.
- This model eliminates the need for separate encoders for vision and audio, allowing these inputs to directly interact with the language processing core.
- This streamlined approach reduces memory usage while maintaining performance comparable to larger models.
- The model's key features include advanced reasoning capabilities, laptop-friendly size requirements (16GB of VRAM), and native support for audio inputs.
- It also introduces Multi-Token Prediction drafters, reducing latency for smoother user experience.
- Gemma 4 12B is open-source under Apache 2.0, making it accessible to developers worldwide.
- With over 150 million downloads across its models, Google expects Gemma 4 12B to unlock new possibilities in AI development, from enterprise security tools to assistive wearables.
- The model's efficiency and versatility aim to bring cutting-edge AI capabilities to everyday devices, setting the stage for broader AI adoption in personal computing.
Terms in this brief
- Gemma 4 12B
- A new multimodal AI model by Google designed to efficiently run on laptops. It processes vision and audio directly with its language core, using less memory while still performing well like larger models. This makes it ideal for devices with limited resources, enabling smarter applications in everyday tech.
Read full story at Hugging Face Blog →, DeepMind Safety →
More briefs
SK Hynix Sees 1242 Percent Net Profit Boost
SK hynix said its second-quarter net profit soared 1242 percent year-on-year. This was driven by demand for its memory chips from the artificial intelligence industry. The company's quarterly net profit was 94 trillion won, an all-time high. Operating profit jumped 557 percent from last year to 60 trillion won. Revenue stood at 79 trillion won. SK hynix will invest around 40 trillion won this year. The company expects demand for its memory chips to persist as tech companies increase their AI infrastructure investments. SK hynix will continue to grow with the evolving AI technology.
UK Introduces AI-Enabled Smart Lamp-Posts
The UK has introduced AI-enabled smart lamp-posts that can recognize number plates and faces. These lamp-posts can power themselves through solar panels and have the potential to fight crime and track down missing persons. They can also incorporate new AI capability such as gait analysis and suspicious behaviors. The use of these lamp-posts has raised concerns about surveillance and data ownership, with 50,000 of them set to be installed in Nigeria. The technology will continue to develop and expand in the future.
AI Companies Destroying Millions of Books for Training Data
AI companies are destroying millions of physical books to use for training data. They scan the books and then discard them. This matters because the books are a valuable source of human-authored text. The AI companies need this text to improve their models. One company, Anthropic, was sued for copyright infringement and paid a $1.5 billion settlement. The AI companies are now hiring middlemen to buy the books for them. They want to keep their involvement a secret because they know it is not popular. The demand for human-authored text will continue to drive the destruction of physical books.
AI Sees Through Leaves to Help Farmers
Scientists created a new AI technology that can see through leaves to identify and measure hidden fruit. This technology can help farmers by creating a complete 3D model of each plant, including what is behind the leaves. It can open the door to automating tasks like crop forecasting. Current methods can produce inaccuracies of up to 23%, but the new technology has achieved fruit counts within 2% to 3% of the correct figure. The technology can save large growers millions of dollars and reduce waste. It will help farmers understand how much fruit they will produce and the size of the fruit. Farmers will be able to plan better with this new technology.
Microsoft's New Cybersecurity Model Delivers Big Efficiency Gains
Microsoft has unveiled MAI-Cyber-1-Flash, a compact cybersecurity model that achieves an impressive 96 percent score on the CyberGym benchmark when integrated into its MDASH multi-agent system. This new tool significantly reduces costs by cutting down reliance on expensive pure frontier models-costs are expected to drop by 50 percent since only the most challenging cases will be handled by GPT-5.4. While MAI-Cyber-1-Flash excels in efficiency, Microsoft still turns to OpenAI for complex reasoning tasks. This hybrid approach allows Microsoft to leverage its own model where it shines brightest while relying on OpenAI's expertise for tougher problems. Looking ahead, this cost-effective solution could expand cybersecurity capabilities for businesses, making advanced protection more accessible. The integration with MDASH suggests a broader push toward multi-agent systems in security, potentially leading to even smarter and more coordinated defenses in the future.