If it wasnt bad enough that Moores Law improvements in the density and cost of transistors is slowing. At the same time, the cost of designing chips and of the factories that are used to etch them is also on the rise. Any savings on any of these fronts will be most welcome to keep IT innovation leaping ahead.
One of the promising frontiers of research right now in chip design is using machine learning techniques to actually help with some of the tasks in the design process. We will be discussing this at our upcoming The Next AI Platform event in San Jose on March 10 with Elias Fallon, engineering director at Cadence Design Systems. (You can see the full agenda and register to attend at this link; we hope to see you there.) The use of machine learning in chip design was also one of the topics that Jeff Dean, a senior fellow in the Research Group at Google who has helped invent many of the hyperscalers key technologies, talked about in his keynote address at this weeks 2020 International Solid State Circuits Conference in San Francisco.
Google, as it turns out, has more than a passing interest in compute engines, being one of the large consumers of CPUs and GPUs in the world and also the designer of TPUs spanning from the edge to the datacenter for doing both machine learning inference and training. So this is not just an academic exercise for the search engine giant and public cloud contender particularly if it intends to keep advancing its TPU roadmap and if it decides, like rival Amazon Web Services, to start designing its own custom Arm server chips or decides to do custom Arm chips for its phones and other consumer devices.
With a certain amount of serendipity, some of the work that Google has been doing to run machine learning models across large numbers of different types of compute engines is feeding back into the work that it is doing to automate some of the placement and routing of IP blocks on an ASIC. (It is wonderful when an idea is fractal like that. . . .)
While the pod of TPUv3 systems that Google showed off back in May 2018 can mesh together 1,024 of the tensor processors (which had twice as many cores and about a 15 percent clock speed boost as far as we can tell) to deliver 106 petaflops of aggregate 16-bit half precision multiplication performance (with 32-bit accumulation) using Googles own and very clever bfloat16 data format. Those TPUv3 chips are all cross-coupled using a 3232 toroidal mesh so they can share data, and each TPUv3 core has its own bank of HBM2 memory. This TPUv3 pod is a huge aggregation of compute, which can do either machine learning training or inference, but it is not necessarily as large as Google needs to build. (We will be talking about Deans comments on the future of AI hardware and models in a separate story.)
Suffice it to say, Google is hedging with hybrid architectures that mix CPUs and GPUs and perhaps someday other accelerators for reinforcement learning workloads, and hence the research that Dean and his peers at Google have been involved in that are also being brought to bear on ASIC design.
One of the trends is that models are getting bigger, explains Dean. So the entire model doesnt necessarily fit on a single chip. If you have essentially large models, then model parallelism dividing the model up across multiple chips is important, and getting good performance by giving it a bunch of compute devices is non-trivial and it is not obvious how to do that effectively.
It is not as simple as taking the Message Passing Interface (MPI) that is used to dispatch work on massively parallel supercomputers and hacking it onto a machine learning framework like TensorFlow because of the heterogeneous nature of AI iron. But that might have been an interesting way to spread machine learning training workloads over a lot of compute elements, and some have done this. Google, like other hyperscalers, tends to build its own frameworks and protocols and datastores, informed by other technologies, of course.
Device placement meaning, putting the right neural network (or portion of the code that embodies it) on the right device at the right time for maximum throughput in the overall application is particularly important as neural network models get bigger than the memory space and the compute oomph of a single CPU, GPU, or TPU. And the problem is getting worse faster than the frameworks and hardware can keep up. Take a look:
The number of parameters just keeps growing and the number of devices being used in parallel also keeps growing. In fact, getting 128 GPUs or 128 TPUv3 processors (which is how you get the 512 cores in the chart above) to work in concert is quite an accomplishment, and is on par with the best that supercomputers could do back in the era before loosely coupled, massively parallel supercomputers using MPI took over and federated NUMA servers with actual shared memory were the norm in HPC more than two decades ago. As more and more devices are going to be lashed together in some fashion to handle these models, Google has been experimenting with using reinforcement learning (RL), a special subset of machine learning, to figure out where to best run neural network models at any given time as model ensembles are running on a collection of CPUs and GPUs. In this case, an initial policy is set for dispatching neural network models for processing, and the results are then fed back into the model for further adaptation, moving it toward more and more efficient running of those models.
In 2017, Google trained an RL model to do this work (you can see the paper here) and here is what the resulting placement looked like for the encoder and decoder, and the RL model to place the work on the two CPUs and four GPUs in the system under test ended up with 19.3 percent lower runtime for the training runs compared to the manually placed neural networks done by a human expert. Dean added that this RL-based placement of neural network work on the compute engines does kind of non-intuitive things to achieve that result, which is what seems to be the case with a lot of machine learning applications that, nonetheless, work as well or better than humans doing the same tasks. The issue is that it cant take a lot of RL compute oomph to place the work on the devices to run the neural networks that are being trained themselves. In 2018, Google did research to show how to scale computational graphs to over 80,000 operations (nodes), and last year, Google created what it calls a generalized device placement scheme for dataflow graphs with over 50,000 operations (nodes).
Then we start to think about using this instead of using it to place software computation on different computational devices, we started to think about it for could we use this to do placement and routing in ASIC chip design because the problems, if you squint at them, sort of look similar, says Dean. Reinforcement learning works really well for hard problems with clear rules like Chess or Go, and essentially we started asking ourselves: Can we get a reinforcement learning model to successfully play the game of ASIC chip layout?
There are a couple of challenges to doing this, according to Dean. For one thing, chess and Go both have a single objective, which is to win the game and not lose the game. (They are two sides of the same coin.) With the placement of IP blocks on an ASIC and the routing between them, there is not a simple win or lose and there are many objectives that you care about, such as area, timing, congestion, design rules, and so on. Even more daunting is the fact that the number of potential states that have to be managed by the neural network model for IP block placement is enormous, as this chart below shows:
Finally, the true reward function that drives the placement of IP blocks, which runs in EDA tools, takes many hours to run.
And so we have an architecture Im not going to get a lot of detail but essentially it tries to take a bunch of things that make up a chip design and then try to place them on the wafer, explains Dean, and he showed off some results of placing IP blocks on a low-powered machine learning accelerator chip (we presume this is the edge TPU that Google has created for its smartphones), with some areas intentionally blurred to keep us from learning the details of that chip. We have had a team of human experts places this IP block and they had a couple of proxy reward functions that are very cheap for us to evaluate; we evaluated them in two seconds instead of hours, which is really important because reinforcement learning is one where you iterate many times. So we have a machine learning-based placement system, and what you can see is that it sort of spreads out the logic a bit more rather than having it in quite such a rectangular area, and that has enabled it to get improvements in both congestion and wire length. And we have got comparable or superhuman results on all the different IP blocks that we have tried so far.
Note: I am not sure we want to call AI algorithms superhuman. At least if you dont want to have it banned.
Anyway, here is how that low-powered machine learning accelerator for the RL network versus people doing the IP block placement:
And here is a table that shows the difference between doing the placing and routing by hand and automating it with machine learning:
And finally, here is how the IP block on the TPU chip was handled by the RL network compared to the humans:
Look at how organic these AI-created IP blocks look compared to the Cartesian ones designed by humans. Fascinating.
Now having done this, Google then asked this question: Can we train a general agent that is quickly effective at placing a new design that it has never seen before? Which is precisely the point when you are making a new chip. So Google tested this generalized model against four different IP blocks from the TPU architecture and then also on the Ariane RISC-V processor architecture. This data pits people working with commercial tools and various levels tuning on the model:
And here is some more data on the placement and routing done on the Ariane RISC-V chips:
You can see that experience on other designs actually improves the results significantly, so essentially in twelve hours you can get the darkest blue bar, Dean says, referring to the first chart above, and then continues with the second chart above. And this graph showing the wireline costs where we see if you train from scratch, it actually takes the system a little while before it sort of makes some breakthrough insight and was able to significantly drop the wiring cost, where the pretrained policy has some general intuitions about chip design from seeing other designs and people that get to that level very quickly.
Just like we do ensembles of simulations to do better weather forecasting, Dean says that this kind of AI-juiced placement and routing of IP block sin chip design could be used to quickly generate many different layouts, with different tradeoffs. And in the event that some feature needs to be added, the AI-juiced chip design game could re-do a layout quickly, not taking months to do it.
And most importantly, this automated design assistance could radically drop the cost of creating new chips. These costs are going up exponentially, and data we have seen (thanks to IT industry luminary and Arista Networks chairman and chief technology officer Andy Bechtolsheim), an advanced chip design using 16 nanometer processes cost an average of $106.3 million, shifting to 10 nanometers pushed that up to $174.4 million, and the move to 7 nanometers costs $297.8 million, with projections for 5 nanometer chips to be on the order of $542.2 million. Nearly half of that cost has been and continues to be for software. So we know where to target some of those costs, and machine learning can help.
The question is will the chip design software makers embed AI and foster an explosion in chip designs that can be truly called Cambrian, and then make it up in volume like the rest of us have to do in our work? It will be interesting to see what happens here, and how research like that being done by Google will help.
See the original post here:
Google Teaches AI To Play The Game Of Chip Design - The Next Platform
- Chess - Wikipedia [Last Updated On: May 3rd, 2017] [Originally Added On: May 3rd, 2017]
- Chess Engines list @wiki - Computer Chess Wiki [Last Updated On: May 3rd, 2017] [Originally Added On: May 3rd, 2017]
- Top Chess Engine Championship - Wikipedia [Last Updated On: May 3rd, 2017] [Originally Added On: May 3rd, 2017]
- Complete mastery: Gaylord Perry's durable legacy - Kitsap Sun [Last Updated On: May 8th, 2017] [Originally Added On: May 8th, 2017]
- Chess notes - The Boston Globe [Last Updated On: May 8th, 2017] [Originally Added On: May 8th, 2017]
- Russia's richest billionaire Alexei Mordashov's incredible 40million Lady M 'super yacht' dwarfs fishing boats as ... - The Sun [Last Updated On: May 11th, 2017] [Originally Added On: May 11th, 2017]
- Garry Kasparov's next move: teaming up with machines - Toronto Star [Last Updated On: May 11th, 2017] [Originally Added On: May 11th, 2017]
- Final Frontier Friday: 'Q Who' - Science Fiction [Last Updated On: May 13th, 2017] [Originally Added On: May 13th, 2017]
- chess set - Hackaday [Last Updated On: May 30th, 2017] [Originally Added On: May 30th, 2017]
- Download free chess engines - Komodo 10, Houdini [Last Updated On: May 30th, 2017] [Originally Added On: May 30th, 2017]
- New Star Trek VR Game Really Is Like Manning Your Own Starfleet Vessel - Kotaku Australia [Last Updated On: June 1st, 2017] [Originally Added On: June 1st, 2017]
- Detonation; Enthusiastic Racing - TruckTrend Network [Last Updated On: June 8th, 2017] [Originally Added On: June 8th, 2017]
- Carlsen-Nakamura Norway Clash Ends In Draw - Chess.com [Last Updated On: June 8th, 2017] [Originally Added On: June 8th, 2017]
- Rouhani should play chess where Trump is playing the fool - Trend News Agency [Last Updated On: June 8th, 2017] [Originally Added On: June 8th, 2017]
- Landry: 5 takeaways from the first week of pre-season - CFL.ca [Last Updated On: June 12th, 2017] [Originally Added On: June 12th, 2017]
- Literature, Films on Chess Captivates Enthusiasts - High on Sports (blog) [Last Updated On: June 14th, 2017] [Originally Added On: June 14th, 2017]
- Ditmas Park's City Council Candidates Debate Major Issues - BKLYNER [Last Updated On: June 16th, 2017] [Originally Added On: June 16th, 2017]
- The Fourth Industrial Revolution Is About Empowering People, Not The Rise Of The Machines - Forbes [Last Updated On: June 16th, 2017] [Originally Added On: June 16th, 2017]
- Worry about people, not jobs: Garry Kasparov - Economic Times [Last Updated On: June 17th, 2017] [Originally Added On: June 17th, 2017]
- ET Recommendations: Get Google Daydream View for Rs 6499 - Economic Times [Last Updated On: June 18th, 2017] [Originally Added On: June 18th, 2017]
- Free Chess Engine recommendation? - Chess Forums - Chess.com [Last Updated On: June 22nd, 2017] [Originally Added On: June 22nd, 2017]
- Calendar of events for June 29 and beyond - Ocala [Last Updated On: June 29th, 2017] [Originally Added On: June 29th, 2017]
- Ford Daytona Notes and Quotes - 13abc Action News [Last Updated On: July 4th, 2017] [Originally Added On: July 4th, 2017]
- How logic games have advanced AI thinking - ComputerWeekly.com [Last Updated On: August 6th, 2017] [Originally Added On: August 6th, 2017]
- Carlsen Falters In Winning Position, Loses To MVL - Chess.com [Last Updated On: August 6th, 2017] [Originally Added On: August 6th, 2017]
- What Can You Do with Continuous Intelligence? - RTInsights [Last Updated On: October 16th, 2019] [Originally Added On: October 16th, 2019]
- Fifty years ago, it was Boris Spassky's turn to shine at the chessboard - Washington Times [Last Updated On: October 16th, 2019] [Originally Added On: October 16th, 2019]
- Lennart Ootes: "Chess is a sport and sport is emotion" - Chessbase News [Last Updated On: October 16th, 2019] [Originally Added On: October 16th, 2019]
- How To Win With The Halloween Gambit - Chess.com [Last Updated On: November 3rd, 2019] [Originally Added On: November 3rd, 2019]
- GM Larry Kaufman Interview: 'New Repertoire For Black And White' - Chess.com [Last Updated On: November 3rd, 2019] [Originally Added On: November 3rd, 2019]
- Geek of the Week: If theres roadwork ahead, Kurt Stiles uses 3D modeling and more to drive project - GeekWire [Last Updated On: November 17th, 2019] [Originally Added On: November 17th, 2019]
- Introducing Fritz 17 with Fat Fritz and other goodies - Chessbase News [Last Updated On: November 17th, 2019] [Originally Added On: November 17th, 2019]
- Hamburg Grand Prix Final Goes To Tiebreak - Chess.com [Last Updated On: November 17th, 2019] [Originally Added On: November 17th, 2019]
- 100 Years Ago | 22 November 2019 - The Statesman [Last Updated On: November 23rd, 2019] [Originally Added On: November 23rd, 2019]
- Magnus Carlsen takes on the Vishy Anand best games quiz - Chessbase News [Last Updated On: November 23rd, 2019] [Originally Added On: November 23rd, 2019]
- Garry Kasparov on chess, tech, Trump and Putin - Chessbase News [Last Updated On: November 23rd, 2019] [Originally Added On: November 23rd, 2019]
- Tata Steel 2: Wesley So beats Anand as five lead - chess24 [Last Updated On: January 17th, 2020] [Originally Added On: January 17th, 2020]
- Ju vs Goryachkina all tied at the half - Chessbase News [Last Updated On: January 17th, 2020] [Originally Added On: January 17th, 2020]
- Xavier Litt: Chess shows that humans and AI work better together - Irish Examiner [Last Updated On: January 17th, 2020] [Originally Added On: January 17th, 2020]
- The Clipper Race Leg 5 - Race 6, Day 3: Le Mans Race start and finding the wind - Sail World [Last Updated On: January 27th, 2020] [Originally Added On: January 27th, 2020]
- Top 10 Richest Tech Company CEO's Ranked By Net Worth | TheTalko - TheTalko [Last Updated On: March 13th, 2020] [Originally Added On: March 13th, 2020]
- Beating the Philidor - BusinessWorld Online [Last Updated On: March 13th, 2020] [Originally Added On: March 13th, 2020]
- Out-preparing the Candidates with Fat Fritz (Part 1) - Chessbase News [Last Updated On: March 24th, 2020] [Originally Added On: March 24th, 2020]
- 8 Reasons Vanderpump Rules Needs to Be Rebooted - Variety [Last Updated On: April 11th, 2020] [Originally Added On: April 11th, 2020]
- Chess greats face off online, webcams, arbiters to watch moves - The Indian Express [Last Updated On: April 24th, 2020] [Originally Added On: April 24th, 2020]
- Chess: Breaking the Code - TheArticle [Last Updated On: April 24th, 2020] [Originally Added On: April 24th, 2020]
- "Chess makes me happy": An interview with Boris Gelfand - Chessbase News [Last Updated On: April 24th, 2020] [Originally Added On: April 24th, 2020]
- With new rules and a new normal, NASCAR set to return this weekend - ESPN [Last Updated On: May 15th, 2020] [Originally Added On: May 15th, 2020]
- Who Are The 8 Best U.S. Chess Players Ever? - Chess.com [Last Updated On: July 6th, 2020] [Originally Added On: July 6th, 2020]
- Welcome to the Status Quo of the Streaming Wars - The Ringer [Last Updated On: July 25th, 2020] [Originally Added On: July 25th, 2020]
- The Cockroach's Carapace (and other opening disasters) - Chessbase News [Last Updated On: July 25th, 2020] [Originally Added On: July 25th, 2020]
- These are the best Chess games you can play on Android phone - The Indian Express [Last Updated On: July 25th, 2020] [Originally Added On: July 25th, 2020]
- Early Fire Season Puts Weary Northern California Firefighters On Front Lines For Months - CBS San Francisco [Last Updated On: September 15th, 2020] [Originally Added On: September 15th, 2020]
- AI Ruined Chess. Now, It's Making the Recreation Lovely Once more - editorials360.com [Last Updated On: September 15th, 2020] [Originally Added On: September 15th, 2020]
- The 10 Best Chess Moves Of All Time - Chess.com [Last Updated On: September 15th, 2020] [Originally Added On: September 15th, 2020]
- Is Creativity Dying in Sports? - NYU Washington Square News [Last Updated On: September 15th, 2020] [Originally Added On: September 15th, 2020]
- Norway Chess: Caruana and Firouzja get off to a good start - Chessbase News [Last Updated On: October 7th, 2020] [Originally Added On: October 7th, 2020]
- Chess Online: How to Play and Win Chess | Chess Tips & Strategies - Popular Mechanics [Last Updated On: October 7th, 2020] [Originally Added On: October 7th, 2020]
- How to Experience the Best Games of the Star Wars Universe - Fantha Tracks [Last Updated On: October 27th, 2020] [Originally Added On: October 27th, 2020]
- Plumbing the Depths of Ethanol Ignorance - The Auto Channel [Last Updated On: October 27th, 2020] [Originally Added On: October 27th, 2020]
- Best Free Chess Engines Every Chess Player Should Download ... [Last Updated On: October 27th, 2020] [Originally Added On: October 27th, 2020]
- Netflix's 'The Queen's Gambit' is the best sports show on TV right now - Business Insider - Business Insider [Last Updated On: November 6th, 2020] [Originally Added On: November 6th, 2020]
- The Queen's Gambit: That ending explained and all your questions answered - CNET [Last Updated On: November 6th, 2020] [Originally Added On: November 6th, 2020]
- The joy of hacking - Chessbase News [Last Updated On: November 6th, 2020] [Originally Added On: November 6th, 2020]
- Cognitive Abilities Of Humans Peak At The Age Of 35: Chess Study - Analytics India Magazine [Last Updated On: November 6th, 2020] [Originally Added On: November 6th, 2020]
- Ed Miliband: 'If the Conservatives want a climate election in the next election, I say bring it on' - PoliticsHome.com [Last Updated On: November 29th, 2020] [Originally Added On: November 29th, 2020]
- Online Chess and Working from Home - Chessbase News [Last Updated On: December 4th, 2020] [Originally Added On: December 4th, 2020]
- Superfinals: Nepomniachtchi and Karjakin still tied on top - Chessbase News [Last Updated On: December 19th, 2020] [Originally Added On: December 19th, 2020]
- Adults, children, cheating, and online chess - Chessbase News [Last Updated On: December 19th, 2020] [Originally Added On: December 19th, 2020]
- Technology - AI and yachting - Superyacht News - The Superyacht Report [Last Updated On: December 29th, 2020] [Originally Added On: December 29th, 2020]
- DeepMind's MuZero AI masters games without knowing the rules - The Burn-In [Last Updated On: December 29th, 2020] [Originally Added On: December 29th, 2020]
- 2020: The year of a pandemic of cheating in online chess - Livemint [Last Updated On: December 29th, 2020] [Originally Added On: December 29th, 2020]
- How Tech Has Changed Traditional Indian Games - United News of India [Last Updated On: December 29th, 2020] [Originally Added On: December 29th, 2020]
- Tata Steel R12: Almost there - Chessbase News [Last Updated On: January 31st, 2021] [Originally Added On: January 31st, 2021]
- Komodo - Chess Engines - Chess.com [Last Updated On: January 31st, 2021] [Originally Added On: January 31st, 2021]
- Computer Chess Engines: A Quick Guide - Chess.com [Last Updated On: January 31st, 2021] [Originally Added On: January 31st, 2021]
- Gravwell 2nd Edition Will Be Coming Out Later This Year - Bleeding Cool News [Last Updated On: February 14th, 2021] [Originally Added On: February 14th, 2021]
- Fat Fritz 2: The Best of Both Worlds - Chessbase News [Last Updated On: February 14th, 2021] [Originally Added On: February 14th, 2021]
- Fat Fritz 2.0 - The new number 1 - Chessbase News [Last Updated On: February 14th, 2021] [Originally Added On: February 14th, 2021]
- The 25th anniversary of Deep Blue beating Garry Kasparov in a chess game. - Slate [Last Updated On: February 14th, 2021] [Originally Added On: February 14th, 2021]