DeepMind

DeepMind on Beyond Market Intelligence: a running collection of 8 stories we have gathered and hand-picked because they are worth your time. Every post here touches on deepmind in some way — the news, the analysis, the deep dives, and the occasional surprise find. Acme AI is the next-generation, AI-powered spreadsheet platform built to replace Excel and redefine how analysts, data scientists, and enterprise teams work with data. New stories are added to this page as we find them, so check back if you want to keep up with what is happening around deepmind, or subscribe to the RSS feed to get them as soon as they are published. Browse the collection below, or head back to the homepage to see everything Beyond Market Intelligence is covering right now.

I built an open-source roguelike specifically for training game-playing agents [P]
Machine Learning

I built an open-source roguelike specifically for training game-playing agents [P]

For researchers and AI practitioners seeking a streamlined environment for reinforcement learning agent training, meet DelveRL: an open-source roguelike built specifically for that purpose. Inspired by DeepMind and OpenAI’s work, DelveRL offers a human-playable game with a structured API, deterministic simulation, and procedural generation—addressing a common integration hurdle. The included baseline agent achieves a median floor of 18, showcasing its potential.

Inherent, founded by DeepMind alumni, says its AI ‘teammate’ just outperformed Anthropic and OpenAI at replicating research
TechCrunch

Inherent, founded by DeepMind alumni, says its AI ‘teammate’ just outperformed Anthropic and OpenAI at replicating research

Inherent, a British AI lab founded by DeepMind alumni, has unveiled Faraday, an AI agent demonstrating remarkable capabilities in replicating scientific research. Initial tests show Faraday outperforming both Anthropic and OpenAI in this crucial area, suggesting a significant step forward in AI-driven scientific exploration. This breakthrough could accelerate innovation by automating literature review and hypothesis generation. For those interested in the broader challenges of AI agent development, our recent article, "Building a Proper Backend for My LangGraph AI Agent," explores practical considerations for real-world applications.

Google’s Gemini 3.7 Flash targets coding and agents with a 50% introductory price cut
VentureBeat

Google’s Gemini 3.7 Flash targets coding and agents with a 50% introductory price cut

Google is accelerating AI innovation with the release of Gemini 3.7 Flash, its "most intelligent workhorse model yet" for coding and agentic workflows. This upgrade prioritizes diligent planning and disciplined execution, showing significant gains in debugging, web development, and enterprise automation—potentially reducing human intervention. Notably, Google is offering a 50% introductory price cut through the end of 2026, making it a compelling option for high-volume applications.

Europe got its own TBPN-style live show, and everyone’s angling for a guest spot
TechCrunch

Europe got its own TBPN-style live show, and everyone’s angling for a guest spot

Europe's media landscape is evolving rapidly, and a new live show is taking center stage. Inspired by the TBPN model, this network recently secured a $1.6 million seed round from influential investors including Powerhouse Capital, Axel Springer SE, and LadBible, alongside angel investors from OpenAI and DeepMind. This substantial funding fuels the network’s largest expansion yet, signaling a future-focused approach to content creation. For deeper exploration into AI-powered tools, see our recent article, "How to Give an LLM Agent a Browser."

Machine Learning

Did blatant AI Slop just win a 25K USD Deepmind / Kaggle Grand Prize? [D]

A recent DeepMind/Kaggle competition, "Measuring Progress Toward AGI," has sparked considerable debate following the announcement of its results. The 25,000 USD grand prize was awarded to a submission critiqued as presenting “nonsensical number generation” and questionable methodology. The work, intended to assess LLM reasoning through viewpoint comparison, appears to have been overlooked for critical review. Explore a deeper investigation of this outcome, detailing the methodology and data—a journey that may challenge conventional understanding.

Google's AlphaEvolve Reaches General Availability with Evolutionary Code Optimization as a Service
InfoQ

Google's AlphaEvolve Reaches General Availability with Evolutionary Code Optimization as a Service

Google’s AlphaEvolve is now generally available on the Gemini Enterprise Agent Platform, marking a significant shift in code optimization. This service, born from DeepMind research, leverages evolutionary algorithms to enhance code performance—with evaluators running client-side, ensuring data remains within your infrastructure. Early adopters, like Klarna, have already seen substantial gains, doubling ML training throughput where a measurable evaluation function is present.

How a former DeepMind researcher raised at a $300M pre-seed valuation before launching a product
TechCrunch

How a former DeepMind researcher raised at a $300M pre-seed valuation before launching a product

Andrew Dai, a former DeepMind researcher with over a decade of experience shaping influential AI systems—including work that informed ChatGPT—is pioneering a new frontier: visual AI. He recently secured a remarkable $300 million pre-seed valuation before even launching his product, signaling immense confidence in this emerging field. Dai articulates a clear vision for how visual AI will transform data management. For further insights into the evolving landscape of AI, explore our recent article, "Google continues its renaming streak by turning NotebookLM to Gemini Notebook."

DeepMind CEO calls for an independent standards body to regulate frontier AI
TechCrunch

DeepMind CEO calls for an independent standards body to regulate frontier AI

Frontier AI demands responsible development, and DeepMind CEO Demis Hassabis is advocating for a crucial step: an independent standards body. Modeled after FINRA, this organization would rigorously test advanced AI models and establish best practices prior to release, ensuring safety and alignment. This proposal underscores the growing need for robust oversight as AI capabilities rapidly advance. Explore the nuances of prompt engineering, a foundational element of effective AI interaction—as detailed in our article, "What is Meta Prompting and How does it work?".