AskVantage

AI and intelligence

DeepMind’s newest AI learns by itself and creates its own knowledge

An AI that can learn by itself, with no human interaction, and that can "create its own knowledge" could give us a new perspective on issues affecting the world, and transform industry, innovation and society.

Key takeaways

  • During each turn of the game it played the network looked at the positions of the pieces on the Go board and calculated which of the millions of possible moves, based on probability, would be the most likely to give it a win.
  • While it was far better than previous versions, AlphaGo Zero is actually a simpler program and it mastered Go faster despite training on less data and running on a smaller computer.
Cite or link to this article

Griffin, M. (2017) 'DeepMind’s newest AI learns by itself and creates its own knowledge', 311 Institute, 12 December. Available at: https://www.311institute.com/deepminds-newest-ai-learns-by-itself-and-creates-its-own-knowledge/ (Accessed: 1 October 2026).

A couple of month’s ago Google’s Artificial Intelligence (AI) group, DeepMind, unveiled the latest incarnation of its Go playing program, AlphaGo Zero, an AI so powerful that it managed to cram thousands of years of human knowledge of playing the game, before inventing better moves of its own, into just three days.

Hailed as a major breakthrough in AI learning because, unlike previous versions of AlphaGo, which went on to beat the world Go champion as well as take the Go online player community to the cleaners, AlphaGo Zero mastered the ancient Chinese board game from nothing more than a clean slate, with no more help from humans than being told the rules of the game. However, and as if that wasn’t already impressive enough, it took its predecessor, AlphaGo, the AI that famously beat Lee Sedol, the South Korean grandmaster, to the cleaners as well, hammering it 100 games to nil.

AlphaGo Zero’s ability to learn for itself, and without human input, is a milestone on the road to one day realising Artificial General Intelligence (AGI), something that the same company, DeepMind, published an architecture for last year, and it will undoubtedly help us create the next generation of more “general” AI’s that can do a lot more than just thrash humans at board games.

AlphaGo Zero amassed its impressive skills using a technique called Reinforcement Learning, and at the heart of the program are a group of software “neurons” that are connected together to form a digital neural network. During each turn of the game it played the network looked at the positions of the pieces on the Go board and calculated which of the millions of possible moves, based on probability, would be the most likely to give it a win. Then, after each game it updated the network, making it stronger player for the next game, and so on and so on.

While it was far better than previous versions, AlphaGo Zero is actually a simpler program and it mastered Go faster despite training on less data and running on a smaller computer.

“Given more time, it could have learned the rules for itself too,” said Demis Hassabis, CEO of DeepMind and a researcher on the team.

Now DeepMind, which is based in London, has set it the task of working out how proteins fold – a massive scientific challenge that could give drug discovery a big shot in the arm.

“For us, AlphaGo wasn’t just about winning the game of Go,” said Habbis,”it was also a big step for us towards building these general-purpose algorithms.”

Most AIs are described as “narrow” because they perform only a single task, such as translating languages or recognising faces, but, asides from putting us firmly on the path to AGI, these general-purpose AIs could potentially outperform humans at many different tasks, and in the next decade, Hassabis believes that AlphaGo’s descendants will work alongside humans, for example, as scientific and medical experts.

“Using the new technique is more powerful than previous approaches because by not using human data, or human expertise in any fashion, we’ve removed the constraints of human knowledge and therefore it’s able to create knowledge itself,” said David Silver, AlphaGo’s lead researcher.

Let’s back up there for a moment because there’s another milestone right there lurking in that sentence – “able to create knowledge for itself.”

However, that AlphaGo Zero can only work on problems that can be simulated in a computer, making tasks such as driving out of the question, but no doubt it will master driving too one day, something that still, arguably, many humans still haven’t mastered… another milestone in the making?

FAQ

Why does this matter?

An AI that can learn by itself, with no human interaction, and that can "create its own knowledge" could give us a new perspective on issues affecting the world, and transform industry, innovation and society.

Matthew Griffin

About the author

Matthew Griffin Founder, 311 Institute

Matthew Griffin is a multi-award winning Futurist and expert in Disruption and Innovation, Geopolitics, Leadership, and Technology, who NASA have described as a "walking encyclopaedia of the future" and a "futurist Polymath."

Read full bio

Matthew Griffin is a multi-award winning Futurist and expert in Disruption and Innovation, Geopolitics, Leadership, and Technology, who NASA have described as a "walking encyclopaedia of the future" and a "futurist Polymath." 15-time best selling author of the "Codex of the Future" series, Matthew is the Founder and Futurist in Chief of the 311 Institute, a global Futures and Deep Futures advisory firm working with royal households, world leaders, G7, G20, and G77 governments, NGOs, and multi-national mid and mega cap firms to help them explore, shape, and lead the next 50 years of business and society.

An award-winning YouTube creator with over a million followers, with an unrivalled global reach and impact, Matthew is a highly sought-after international keynote speaker, lecturer, and mentor who collaborates with global leaders through the United Nations Alliance of Civilizations (UNAOC) and United Nations General Assembly (UNGA) to shape pivotal initiatives such as the UN’s AI for Humanity program, the United Nations Conference of the Parties (UN COP), and the World Economic Forum in Davos.

As the former Global Head of Cloud, National Security, and Enterprise Sales for companies including Atos, Dell-EMC, and IBM, Matthew has a proven track record of building multi-billion dollar business units and turning failing divisions into market leaders. His ability to identify, analyse, and communicate the implications of hundreds of emerging technologies and trends is unparalleled, and his insights are trusted by many of the world’s most respected organisations, including ABB, Accenture, Adidas, AON, ARM, BCG, Centrica, Citi, Coca-Cola, Dentons, Deloitte, Dow Jones, EY, Google, KPMG, Lego, Legal & General, LinkedIn, Microsoft, PepsiCo, Qualcomm, RWE, Samsung, Siemens AG and Siemens Energy, T-Mobile, UBS, VISA, Walmart, Workday, Worldpay and many others.

Regularly featured in the global media including the AP, BBC, Bloomberg, CNBC, Discovery, Forbes, Khaleej Times, Telegraph, TIME, ViacomCBS, WIRED, and the WSJ, Matthews mission is to help organisations create a fair and sustainable future whose benefits are shared by everyone irrespective of their ability, background, or circumstances.

What future do you need to see?

Choose one to get started on AI and intelligence and the future of your organisation.

Where should Matthew reply?

Takes 30 seconds. No obligation. Matthew replies quickly. Privacy

Tag Cloud

Starburst opens that technology on the interactive 311 Starburst.

Sources and further reading

  1. Reinforcement learning en.wikipedia.org

Source: first published by the 311 Institute on 12 December 2017. Cite as: Griffin, M. (2017). DeepMind’s newest AI learns by itself and creates its own knowledge. 311 Institute. https://www.311institute.com/deepminds-newest-ai-learns-by-itself-and-creates-its-own-knowledge/

You are welcome to quote this article with credit and a link to the original.

Book a Keynote