Google DeepMind’s new AI uses large language models to crack real-world problems

Stay Ahead, Stay ONMINE

Google DeepMind’s new AI uses large language models to crack real-world problems

Google DeepMind has once again used large language models to discover new solutions to long-standing problems in math and computer science. This time the firm has shown that its approach can not only tackle unsolved theoretical puzzles, but improve a range of important real-world processes as well. Google DeepMind’s new tool, called AlphaEvolve, uses the Gemini 2.0 family of large language models (LLMs) to produce code for a wide range of different tasks. LLMs are known to be hit and miss at coding. The twist here is that AlphaEvolve scores each of Gemini’s suggestions, throwing out the bad and tweaking the good, in an iterative process, until it has produced the best algorithm it can. In many cases, the results are more efficient or more accurate than the best existing (human-written) solutions. “You can see it as a sort of super coding agent,” says Pushmeet Kohli, a vice president at Google DeepMind who leads its AI for Science teams. “It doesn’t just propose a piece of code or an edit, it actually produces a result that maybe nobody was aware of.” In particular, AlphaEvolve came up with a way to improve the software Google uses to allocate jobs to its many millions of servers around the world. Google DeepMind claims the company has been using this new software across all of its data centers for more than a year, freeing up 0.7% of Google’s total computing resources. That might not sound like much, but at Google’s scale it’s huge. Jakob Moosbauer, a mathematician at the University of Warwick in the UK, is impressed. He says the way AlphaEvolve searches for algorithms that produce specific solutions—rather than searching for the solutions themselves—makes it especially powerful. “It makes the approach applicable to such a wide range of problems,” he says. “AI is becoming a tool that will be essential in mathematics and computer science.” AlphaEvolve continues a line of work that Google DeepMind has been pursuing for years. Its vision is that AI can help to advance human knowledge across math and science. In 2022, it developed AlphaTensor, a model that found a faster way to solve matrix multiplications—a fundamental problem in computer science—beating a record that had stood for more than 50 years. In 2023, it revealed AlphaDev, which discovered faster ways to perform a number of basic calculations performed by computers trillions of times a day. AlphaTensor and AlphaDev both turn math problems into a kind of game, then search for a winning series of moves. FunSearch, which arrived in late 2023, swapped out game-playing AI and replaced it with LLMs that can generate code. Because LLMs can carry out a range of tasks, FunSearch can take on a wider variety of problems than its predecessors, which were trained to play just one type of game. The tool was used to crack a famous unsolved problem in pure mathematics. AlphaEvolve is the next generation of FunSearch. Instead of coming up with short snippets of code to solve a specific problem, as FunSearch did, it can produce programs that are hundreds of lines long. This makes it applicable to a much wider variety of problems. In theory, AlphaEvolve could be applied to any problem that can be described in code and that has solutions that can be evaluated by a computer. “Algorithms run the world around us, so the impact of that is huge,” says Matej Balog, a researcher at Google DeepMind who leads the algorithm discovery team. Survival of the fittest Here’s how it works: AlphaEvolve can be prompted like any LLM. Give it a description of the problem and any extra hints you want, such as previous solutions, and AlphaEvolve will get Gemini 2.0 Flash (the smallest, fastest version of Google DeepMind’s flagship LLM) to generate multiple blocks of code to solve the problem. It then takes these candidate solutions, runs them to see how accurate or efficient they are, and scores them according to a range of relevant metrics. Does this code produce the correct result? Does it run faster than previous solutions? And so on. AlphaEvolve then takes the best of the current batch of solutions and asks Gemini to improve them. Sometimes AlphaEvolve will throw a previous solution back into the mix to prevent Gemini from hitting a dead end. When it gets stuck, AlphaEvolve can also call on Gemini 2.0 Pro, the most powerful of Google DeepMind’s LLMs. The idea is to generate many solutions with the faster Flash but add solutions from the slower Pro when needed. These rounds of generation, scoring, and regeneration continue until Gemini fails to come up with anything better than what it already has. Number games The team tested AlphaEvolve on a range of different problems. For example, they looked at matrix multiplication again to see how a general-purpose tool like AlphaEvolve compared to the specialized AlphaTensor. Matrices are grids of numbers. Matrix multiplication is a basic computation that underpins many applications, from AI to computer graphics, yet nobody knows the fastest way to do it. “It’s kind of unbelievable that it’s still an open question,” says Balog. The team gave AlphaEvolve a description of the problem and an example of a standard algorithm for solving it. The tool not only produced new algorithms that could calculate 14 different sizes of matrix faster than any existing approach, it also improved on AlphaTensor’s record-beating result for multipying two four-by-four matrices. AlphaEvolve scored 16,000 candidates suggested by Gemini to find the winning solution, but that’s still more efficient than AlphaTensor, says Balog. AlphaTensor’s solution also only worked when a matrix was filled with 0s and 1s. AlphaEvolve solves the problem with other numbers too. “The result on matrix multiplication is very impressive,” says Moosbauer. “This new algorithm has the potential to speed up computations in practice.” Manuel Kauers, a mathematician at Johannes Kepler University in Linz, Austria, agrees: “The improvement for matrices is likely to have practical relevance.” By coincidence, Kauers and a colleague have just used a different computational technique to find some of the speedups AlphaEvolve came up with. The pair posted a paper online reporting their results last week. “It is great to see that we are moving forward with the understanding of matrix multiplication,” says Kauers. “Every technique that helps is a welcome contribution to this effort.” Real-world problems Matrix multiplication was just one breakthrough. In total, Google DeepMind tested AlphaEvolve on more than 50 different types of well-known math puzzles, including problems in Fourier analysis (the math behind data compression, essential to applications such as video streaming), the minimum overlap problem (an open problem in number theory proposed by mathematician Paul Erdős in 1955), and kissing numbers (a problem introduced by Isaac Newton that has applications in materials science, chemistry, and cryptography). AlphaEvolve matched the best existing solutions in 75% of cases and found better solutions in 20% of cases. Google DeepMind then applied AlphaEvolve to a handful of real-world problems. As well as coming up with a more efficient algorithm for managing computational resources across data centers, the tool found a way to reduce the power consumption of Google’s specialized tensor processing unit chips. AlphaEvolve even found a way to speed up the training of Gemini itself, by producing a more efficient algorithm for managing a certain type of computation used in the training process. Google DeepMind plans to continue exploring potential applications of its tool. One limitation is that AlphaEvolve can’t be used for problems with solutions that need to be scored by a person, such as lab experiments that are subject to interpretation. Moosbauer also points out that while AlphaEvolve may produce impressive new results across a wide range of problems, it gives little theoretical insight into how it arrived at those solutions. That’s a drawback when it comes to advancing human understanding. Even so, tools like AlphaEvolve are set to change the way researchers work. “I don’t think we are finished,” says Kohli. “There is much further that we can go in terms of how powerful this type of approach is.”

Google DeepMind’s new tool, called AlphaEvolve, uses the Gemini 2.0 family of large language models (LLMs) to produce code for a wide range of different tasks. LLMs are known to be hit and miss at coding. The twist here is that AlphaEvolve scores each of Gemini’s suggestions, throwing out the bad and tweaking the good, in an iterative process, until it has produced the best algorithm it can. In many cases, the results are more efficient or more accurate than the best existing (human-written) solutions.

“You can see it as a sort of super coding agent,” says Pushmeet Kohli, a vice president at Google DeepMind who leads its AI for Science teams. “It doesn’t just propose a piece of code or an edit, it actually produces a result that maybe nobody was aware of.”

In particular, AlphaEvolve came up with a way to improve the software Google uses to allocate jobs to its many millions of servers around the world. Google DeepMind claims the company has been using this new software across all of its data centers for more than a year, freeing up 0.7% of Google’s total computing resources. That might not sound like much, but at Google’s scale it’s huge.

Jakob Moosbauer, a mathematician at the University of Warwick in the UK, is impressed. He says the way AlphaEvolve searches for algorithms that produce specific solutions—rather than searching for the solutions themselves—makes it especially powerful. “It makes the approach applicable to such a wide range of problems,” he says. “AI is becoming a tool that will be essential in mathematics and computer science.”

AlphaEvolve continues a line of work that Google DeepMind has been pursuing for years. Its vision is that AI can help to advance human knowledge across math and science. In 2022, it developed AlphaTensor, a model that found a faster way to solve matrix multiplications—a fundamental problem in computer science—beating a record that had stood for more than 50 years. In 2023, it revealed AlphaDev, which discovered faster ways to perform a number of basic calculations performed by computers trillions of times a day. AlphaTensor and AlphaDev both turn math problems into a kind of game, then search for a winning series of moves.

FunSearch, which arrived in late 2023, swapped out game-playing AI and replaced it with LLMs that can generate code. Because LLMs can carry out a range of tasks, FunSearch can take on a wider variety of problems than its predecessors, which were trained to play just one type of game. The tool was used to crack a famous unsolved problem in pure mathematics.

AlphaEvolve is the next generation of FunSearch. Instead of coming up with short snippets of code to solve a specific problem, as FunSearch did, it can produce programs that are hundreds of lines long. This makes it applicable to a much wider variety of problems.

In theory, AlphaEvolve could be applied to any problem that can be described in code and that has solutions that can be evaluated by a computer. “Algorithms run the world around us, so the impact of that is huge,” says Matej Balog, a researcher at Google DeepMind who leads the algorithm discovery team.

Survival of the fittest

Here’s how it works: AlphaEvolve can be prompted like any LLM. Give it a description of the problem and any extra hints you want, such as previous solutions, and AlphaEvolve will get Gemini 2.0 Flash (the smallest, fastest version of Google DeepMind’s flagship LLM) to generate multiple blocks of code to solve the problem.

It then takes these candidate solutions, runs them to see how accurate or efficient they are, and scores them according to a range of relevant metrics. Does this code produce the correct result? Does it run faster than previous solutions? And so on.

AlphaEvolve then takes the best of the current batch of solutions and asks Gemini to improve them. Sometimes AlphaEvolve will throw a previous solution back into the mix to prevent Gemini from hitting a dead end.

When it gets stuck, AlphaEvolve can also call on Gemini 2.0 Pro, the most powerful of Google DeepMind’s LLMs. The idea is to generate many solutions with the faster Flash but add solutions from the slower Pro when needed.

These rounds of generation, scoring, and regeneration continue until Gemini fails to come up with anything better than what it already has.

Number games

The team tested AlphaEvolve on a range of different problems. For example, they looked at matrix multiplication again to see how a general-purpose tool like AlphaEvolve compared to the specialized AlphaTensor. Matrices are grids of numbers. Matrix multiplication is a basic computation that underpins many applications, from AI to computer graphics, yet nobody knows the fastest way to do it. “It’s kind of unbelievable that it’s still an open question,” says Balog.

The team gave AlphaEvolve a description of the problem and an example of a standard algorithm for solving it. The tool not only produced new algorithms that could calculate 14 different sizes of matrix faster than any existing approach, it also improved on AlphaTensor’s record-beating result for multipying two four-by-four matrices.

AlphaEvolve scored 16,000 candidates suggested by Gemini to find the winning solution, but that’s still more efficient than AlphaTensor, says Balog. AlphaTensor’s solution also only worked when a matrix was filled with 0s and 1s. AlphaEvolve solves the problem with other numbers too.

“The result on matrix multiplication is very impressive,” says Moosbauer. “This new algorithm has the potential to speed up computations in practice.”

Manuel Kauers, a mathematician at Johannes Kepler University in Linz, Austria, agrees: “The improvement for matrices is likely to have practical relevance.”

By coincidence, Kauers and a colleague have just used a different computational technique to find some of the speedups AlphaEvolve came up with. The pair posted a paper online reporting their results last week.

“It is great to see that we are moving forward with the understanding of matrix multiplication,” says Kauers. “Every technique that helps is a welcome contribution to this effort.”

Real-world problems

Matrix multiplication was just one breakthrough. In total, Google DeepMind tested AlphaEvolve on more than 50 different types of well-known math puzzles, including problems in Fourier analysis (the math behind data compression, essential to applications such as video streaming), the minimum overlap problem (an open problem in number theory proposed by mathematician Paul Erdős in 1955), and kissing numbers (a problem introduced by Isaac Newton that has applications in materials science, chemistry, and cryptography). AlphaEvolve matched the best existing solutions in 75% of cases and found better solutions in 20% of cases.

Google DeepMind then applied AlphaEvolve to a handful of real-world problems. As well as coming up with a more efficient algorithm for managing computational resources across data centers, the tool found a way to reduce the power consumption of Google’s specialized tensor processing unit chips.

AlphaEvolve even found a way to speed up the training of Gemini itself, by producing a more efficient algorithm for managing a certain type of computation used in the training process.

Google DeepMind plans to continue exploring potential applications of its tool. One limitation is that AlphaEvolve can’t be used for problems with solutions that need to be scored by a person, such as lab experiments that are subject to interpretation.

Moosbauer also points out that while AlphaEvolve may produce impressive new results across a wide range of problems, it gives little theoretical insight into how it arrived at those solutions. That’s a drawback when it comes to advancing human understanding.

Even so, tools like AlphaEvolve are set to change the way researchers work. “I don’t think we are finished,” says Kohli. “There is much further that we can go in terms of how powerful this type of approach is.”

Stay Ahead

Explore More Insights

Stay ahead with more perspectives on cutting-edge power, infrastructure, energy, bitcoin and AI solutions. Explore these articles to uncover strategies and insights shaping the future of industries.

Presidential order addresses quantum computing gaps

By comparison, in AI, there are a number of benchmarks comparing AI models on everything from how well they do at math, to their ability to generate images, to whether they can find bugs in computer code. Benchmarks are important so that people can do fair comparisons of quantum systems,

Researchers cast new doubt on Microsoft’s quantum computing advance

In 2018, Microsoft said its researchers had detected evidence of their existence, an apparently major breakthrough it was forced to retract when the data was successfully challenged. Nature’s editors subsequently backed this up with the blunt note: “The results in this manuscript do not represent evidence for the presence of

IBM unveils sub-1 nanometer chip with nearly 100 billion transistors

It’s the world’s first sub-1 nm chip technology, IBM claims. Researcher holds IBM’s sub-1 nm node wafer.

IBM, Red Hat, Palo Alto team to secure open-source software

“The clearinghouse will serve as a security coordination layer, using advanced AI capabilities to validate and test fixes across an unprecedented volume of open source code,” IBM stated in May. “These capabilities will be offered through commercial subscriptions, allowing enterprises to integrate secure patches directly into their existing software supply

Trump Administration Keeps Coal-Fired Power Generation Alive in Colorado

WASHINGTON—U.S. Secretary of Energy Chris Wright today issued an emergency order to keep a Colorado coal plant operational to ensure Americans maintain access to affordable, reliable, and secure electricity. The order directs Tri-State Generation and Transmission Association (Tri-State), Platte River Power Authority, Salt River Project, PacifiCorp, and Public Service Company of Colorado (a subsidiary of Xcel Energy), to take all measures necessary to ensure that Craig Unit 1 is available to operate at the direction of the Southwest Power Pool (SPP). For the duration of this Order, SPP is directed to take every step to employ economic dispatch of Craig Unit 1 to minimize costs to ratepayers. Unit 1 of the coal plant was originally scheduled to shut down at the end of 2025, but in December 2025 and again in March 2026, Secretary Wright issued emergency orders directing Tri-State and the co-owners to ensure that Unit 1 at the Craig Station remains available to operate. “Taking reliable generation off the grid compromises energy reliability and needlessly raises energy costs for Americans,” said Energy Secretary Wright. “During peak summer demand, Coloradans deserve continued access to affordable, reliable, and secure energy to power and cool their homes.” Thanks to President Trump’s leadership, coal plants across the country are being saved from premature retirement and reversing plans to shut down. In 2025, more than 17 gigawatts of coal-power electricity generation were saved. According to DOE’s Resource Adequacy Report, blackouts were on track to potentially increase 100 times by 2030 if the U.S. continued to take reliable power offline as it did during the Biden administration. The North American Electric Reliability Corporation (NERC) 2025 Long-Term Reliability Assessment warns that the WECC-Rocky Mountain assessment area faces challenges from an aging thermal resource fleet, which can lead to unplanned outages, exacerbated by supply chain issues, and vendor availability. This order

Energy Department Analysis Finds Proposed International Building Codes Would Cost Americans $9.2 Billion Annually

WASHINGTON—The U.S. Department of Energy (DOE) today released a new analysis finding that nationwide adoption of the 2024 International Energy Conservation Code (IECC) would significantly increase housing construction costs and burden American families with costly Green New Scam mandates. DOE’s analysis found that the 2024 IECC would increase residential construction costs by more than $9.2 billion annually compared to the 2006 code levels, adding more than $127 billion in cumulative costs nationwide. If states choose to update their energy codes to the 2024 IECC, construction costs for a typical single-family home could increase by as much as $14,000. These costly mandates force American families to pay thousands of dollars more upfront for a new home, while projected energy savings may take decades to materialize. In most states, estimated payback periods exceed 10 years, with some exceeding 20 years—locking American families into decades-long repayment timeframes and restricting consumer choice. “American families should not be forced to pay more for a home because of nonsensical energy-related mandates,” said U.S. Energy Secretary Chris Wright. “For too long, climate activists have pushed regulations that increase housing costs, reduce consumer choice, and make it harder for Americans to build and own a home. Thankfully, President Trump will continue fighting for the American people so they can enjoy affordable energy access and the ability to buy the home they desire with the features they choose.” “This analysis shows how unnecessary regulations and ineffective building codes have drastically increased housing costs with little to no benefit for homeowners or communities,” said Assistant Secretary of Energy (EERE) Audrey Robertson. “An average payback period of 11 years—as long as 22 years in some cases—for new residential building codes is unacceptable. Standard-setting bodies should take note: we prioritize the American homeowner and will not allow erroneous building requirements to push

U.S., Qatar, Nigeria, and Algeria Warn Proposed E.U. Methane Regulations Could Disrupt Europe’s Oil and Gas Supply

WASHINGTON—U.S. Secretary of Energy Chris Wright, Qatari Minister of State for Energy Affairs Saad Sherida Al-Kaabi, Nigerian Minister of State for Petroleum Resources Ekperikpe Ekpo, and Algerian Minister of State, Minister of Hydrocarbons Mohamed Arkab yesterday sent a letter to the Leaders of the European Commission, European Council, and European Union (EU) Member States, regarding the European Union’s proposed EU Methane Regulations (EUMR). Click here to read the letter or see the full text below. Open Letter to Leaders of the European Commission, European Council, and European Union (EU) Member States on the EU Methane Regulation Dear President von der Leyen, President Costa, and EU Member State Leaders: As your largest energy suppliers, we are committed to strengthening our economic and strategic partnerships and ensuring Europe’s energy security. We fully support your objectives of increasing EU economic competitiveness, prosperity, sustainability, and energy security through provision of reliable energy supplies for the European Union and its citizens. It is with these shared goals in mind that we write to urge the EU to take swift, necessary actions to clarify and to adopt targeted amendments to the EU Methane Regulation (EUMR), some of which have already been requested by several EU Member States, industry, and members of European Parliament. These amendments should also be preceded by the: (i) adoption of a stop the clock mechanism, to provide time to develop necessary methodologies and compliance pathways that work for all; (ii)grandfathering of new contracts signed while these additional legislative adjustments are underway; and (iii) removal of penalties for noncompliance during this transitional period. As a large and diverse importing region, the EU purchases oil and natural gas from a wide variety of exporters, the majority of which cannot meet the EUMR methane emissions measuring, reporting, and verification (MRV) requirements on the prescribed timeline.

Department of Energy Announces American Nuclear Supply Chain Loans

WASHINGTON—The U.S. Department of Energy’s (DOE) Office of Energy Dominance Financing (EDF) issued a conditional loan commitment to finance the purchase of long-lead time items needed to rebuild America’s commercial nuclear supply chain. The $17.5 billion American Nuclear Supply Chain Loans will help finance five eligible projects sponsored by utilities and energy companies nationwide to accelerate the deployment of 10 large-scale commercial nuclear reactors across the United States by up to three years. The project marks a major step toward advancing President Trump’s Executive Order, Reinvigorating the Nuclear Industrial Base, by supporting the objective of having 10 new large nuclear reactors with complete designs under construction by 2030. “Just over one year ago, President Trump directed the Energy Department and its agency partners to unleash the next American nuclear renaissance,” U.S. Energy Secretary Chris Wright said. “To accomplish that mission, these conditional loans will play an important role in reviving the supply chain needed for America to once again build large-scale commercial reactors. They will also help accelerate the timeline of building those large-scale reactors by up to three years, lowering construction costs and ensuring the United States is able to deliver on President Trump’s bold and ambitious energy addition agenda.” Westinghouse’s AP1000® units are the only licensed large-scale advanced commercial reactors operating in the United States today. Long-lead items are complex components of a nuclear power plant that require the longest time for manufacturing and delivery. EDF financing will support up to five loans, each loan supporting two reactors at a project site. Westinghouse will partner with up to five eligible utilities and energy companies nationwide to procure the long-lead items at a fixed price. Each project will be jointly owned by Westinghouse and a utility or energy company partner. Both Westinghouse and the partner are required to

FPSO ready for Santos-led Barossa LNG project

BW Offshore completed the Interim Performance Test (IPT) for the BW Opal floating production, storage, and offloading vessel (FPSO) as part of the commissioning program for the Santos Ltd.-operated Barossa LNG project about 285 km offshore from Darwin in the Northern Territory of Australia. The milestone is part of early-stage technical testing and adjustments following first gas from the FPSO in September and the beginning of flow from subsea wells. BW Offshore confirmed that key production, processing, and utility systems on the FPSO were operating in an integrated manner and capable of delivering stable performance under production conditions. Following the restart of production in early May, BW Opal has continued gas production and export. Production is being managed in close coordination with Santos during this phase of the ramp-up and commissioning program. BW Opal contains a 358-m hull and accommodation for up to 140 personnel. It has gas handling capacity of 850 MMscfd and condensate handling capacity of 11,000 b/d. The FPSO will feed the Darwin LNG plant for the next two decades. The Barossa LNG project consists of the FPSO, a subsea production system, supporting in-field subsea infrastructure, a gas export pipeline, and a Darwin pipeline duplication. Up to eight subsea wells are planned (six wells from three drill centers) with contingency plans for an additional two wells. Gas and condensate is gathered from the wells through the subsea production system and then brought to the FPSO via a network of subsea infrastructure. Santos operates the Barossa LNG project (50%) with joint venture partners PRISM Energy International Australia Pty Ltd. (37.5%) and JERA Australia (12.5%).

Equinor mulls additional Johan Sverdrup development phase

Equinor Energy AS is considering further development of the Johan Sverdrup area resources in the North Sea. Production from discoveries in Tonjer west and east and Geitungen would form the basis for the maturation of a potential phase 4 development in the northern part of the field. The volumes would be developed via subsea tieback to existing Johan Sverdrup infrastructure. Tonjer lies in the northernmost part of the Geitungen terrace in the Johan Sverdrup area. Oil was discovered in the area, but volumes and potential have been uncertain. The drilling of two appraisal wells and a sidetrack have provided a more precise assessment of the resource base. Preliminary estimates for Tonjer and Geitungen combined are 20-30 MMboe. Further analyses of subsurface data will form the basis for more precise resource estimates. Phase 4 is now being matured towards an investment decision with a possible production start-up in 2029. Johan Sverdrup Johan Sverdrup, which accounts for about one third of Norwegian oil production, lies on the Utsira High (Utsirahøyden) in the central part of the North Sea, 65 km northeast of Sleipner field in water depths of 115 m. The main reservoir contains oil in Upper Jurassic intra-Draupne sandstone. The reservoir depth is 1,900 m. The quality of the main reservoir is excellent with very high permeability. The remaining oil resources are in sandstone in the Upper Triassic Statfjord Group and Middle to Upper Jurassic Vestland Group, as well as in spiculites in the Upper Jurassic Viking Group. Oil was also proven in Permian Zechstein carbonates. Equinor is operator of Johan Sverdrup (42.62%) with partners Aker BP (31.57%), Petoro (17.36%), and TotalEnergies (8.44%).

You can’t build sovereign infrastructure with Broadcom, says CISPE

CISPE has cited several reasons why VCF doesn’t fit the bill, in particular highlighting its lack of portability. This means that it doesn’t qualify as resilient under CISPE’s Sovereign and Resilient Cloud Framework. Earlier this month, the EU unveiled proposals for its Cloud and AI Development Act (CADA) to strengthen Europe’s digital economy. CADA will encourage investment in European research, lay down conditions for European data centers, and provide a single EU-wide assessment framework for cloud and AI sovereignty. CISPE said that Broadcom is a long way short of fulfilling the conditions proposed for CADA. Broadcom would fail to meet anything but a Level 1 certification under the CADA sovereignty framework, CISPE said, adding that Broadcom’s terms and conditions offer limited maintenance commitments, no source-code escrow, no substitution plan and no Data Act certification, all likely to fall foul of CADA’s recommendations.

Break legacy lock-in: Strategic options for enterprises facing the vSphere 8 deadline

The acquisition of VMware by Broadcom has caused many enterprise IT leaders to reexamine their infrastructure strategies. For organizations running vSphere 8, the October 2027 end-of-support deadline is rapidly becoming a planning priority. What may appear to be a routine upgrade is driving bigger discussions about cost, flexibility, cloud strategy, and long-term infrastructure direction. Many organizations have not only begun evaluating alternatives but also are leaving VMware. “VMware has been a great, innovative company,” says Harsha Kotikela, senior director of product and solutions marketing at Nutanix. “But since the acquisition, their business model has fundamentally changed, and that is what is forcing IT leaders to adapt.” Sticker shock, vendor lock-in, and the need for flexibility One of the biggest catalysts has been licensing costs. Organizations that had grown accustomed to predictable contracts have encountered significant pricing increases, creating what Kotikela describes as “sticker shock.” At the same time, some enterprises are reevaluating their vendor relationships due to concerns about support availability and changes in partner engagement models. Beyond immediate operational concerns, IT leaders are also focused on future requirements. Hybrid cloud environments have become the norm, with applications and data distributed across data centers, public clouds, and edge locations. AI initiatives are adding another layer of complexity, requiring infrastructure that can support workloads wherever they need to run. “The future is about flexibility,” Kotikela says. “If enterprises want to implement AI at the edge, in the data center, or in the cloud, they need the capability to manage that environment without creating silos.” That flexibility is becoming a critical factor in infrastructure decisions. Organizations increasingly want platforms that support multiple deployment models, open APIs, and cloud-native technologies to minimize the risk of vendor lock-in. How a future-ready platform addresses IT and business requirements Nutanix positions its architecture around openness and choice, according to

Qualcomm’s $3.9 billion purchase of Modular aims to change the data center dynamic

“Nvidia has something like 85% of the AI accelerator chip market,” he pointed out. “Sure, they have nowhere to go but down, but that’s still going to take them a while. More importantly, they have literally spent decades working with practitioners in AI and ML and compute-intensive fields, indoctrinating them into their CUDA software ecosystem. Rewriting that tool chain will take institutional change at most organizations, which means years, if not decades, to uncouple.” “Organizations that think they’ve achieved agnosticism because they’re using high-level abstractions like PyTorch, well, they have come closest,” he observed. “But just cutting and pasting the same code into AMD Instinct can lead to memory and dependency errors. It’s like VM lift and shifts to the public cloud 10 years ago. Easier, but still possible to screw up.” Nonetheless, Annand said that the deal, if it goes through, is still good news for enterprises.

KKR Bets Big on AI Infrastructure With Helix Launch, Tapping Former AWS CEO Adam Selipsky to Build a New Hyperscale Model

To close industry watchers, it’s really no secret that the AI infrastructure race has entered another phase; one where capital formation itself may become as strategically important as GPUs, power procurement, or liquid cooling. And in launching Helix Digital Infrastructure, investment giant KKR is making a calculated wager that hyperscalers no longer simply need developers or financiers. They need a partner capable of orchestrating capital, energy, connectivity, and data center execution as a unified platform. The significance of that strategy is underscored by the executive chosen to lead it. Adam Selipsky, the former CEO of Amazon Web Services and one of the industry’s most experienced cloud operators, will serve as Co-Founder and CEO of Helix, bringing firsthand experience from the very class of customers the new venture intends to serve. A New Model for AI Infrastructure Helix launches with more than $10 billion in long-duration committed capital from founding investors including KKR, the Kuwait Investment Authority (KIA), NVIDIA, and Vistra. But the headline number tells only part of the story. The company has been structured around an increasingly important thesis: that AI infrastructure can no longer be assembled piecemeal. Rather than treating data centers, electrical supply, transmission capacity, and fiber connectivity as separate procurement exercises, Helix proposes a vertically coordinated approach in which a single organization manages and finances the entire infrastructure stack. According to KKR, the objective is to reduce execution risk and accelerate deployment for hyperscale customers facing unprecedented AI demand. As AI factories grow from hundreds of megawatts toward gigawatt-scale campuses, synchronization among land acquisition, utility planning, financing, construction, and technology deployment has emerged as one of the industry’s defining challenges. Helix is effectively positioning itself as an operating platform designed to simplify that complexity. Why Selipsky Matters The appointment of Adam Selipsky may be the announcement’s

Beyond Hyperscale: Why Enterprise Data Centers Still Matter in the AI Era

“The enterprise data centers, even the new ones, tend to be far, far smaller than new hyperscale deployments,” Killian said. “Not uncommon to see enterprises deploy a quarter meg or one meg or two, maybe up to 10 megs. Whereas the hyperscale guys are deploying 40 up to 300 meg facilities.” But scale alone does not tell the story. For every one of the roughly 20 hyperscale users that dominate headlines, Killian noted, there may be 50 to 100 times as many large and mid-sized enterprise users. Those companies run critical business systems, purchase hardware, software, telecom and services, employ large data center teams, and often operate multiple facilities across domestic, edge, EMEA and Asia-Pacific footprints. In other words, enterprise demand may be smaller in unit size, but it remains massive in aggregate. And as AI shifts from training to inference, the enterprise data center could become newly strategic. Enterprise AI Is Not Hyperscale AI Killian’s central point is that enterprise infrastructure requirements differ materially from hyperscale requirements. Hyperscalers are primarily optimizing for massive scale and speed to market. Enterprises, by contrast, tend to prioritize reliability, flexibility, integration into broader IT systems, and audit and compliance. That difference has major implications for developers and colocation providers. “The real industry opportunity is to take some of the innovation and the economies of scale that we’re seeing from the hyperscale builds to deliver smaller chunks of data center capacity,” Killian said. That might mean adapting lessons from 40 MW or 100 MW campuses into enterprise-ready deployments of 2 MW, 4 MW or 8 MW. Killian pointed to providers such as DataBank and Flexential as examples of companies working to deliver hyperscale-derived efficiencies in smaller enterprise increments. He also noted that QTS and other large campus developers may reserve portions of multi-building campuses

Revolutionizing Data Center Cooling: Innovations for AI and HPC Growth

This is a crucial point for AI infrastructure. In some markets, water can be as politically and operationally difficult as power. Evaporative cooling and cooling towers can consume large volumes of water, while discharge permits can slow projects or limit operations. Gradiant claims HyperSolved can expand access to alternative sources such as municipal reuse and impaired supplies, reduce reliance on freshwater, protect cooling performance through integrated treatment and AI-enabled operations, and minimize discharge through high-recovery concentration and reuse. The platform uses containerized systems for immediate or temporary capacity while also supporting permanent infrastructure and lifecycle operations from commissioning onward. That fits the AI data center buildout, where developers may need bridge capacity during construction, phased water infrastructure, or interim systems while permanent treatment plants are completed. This can address the speed of deployment issue that plagues many data center solutions. Water is becoming a siting and scaling variable that has to be addressed. A site may have land and power prospects, but if water sourcing, reuse, or discharge cannot be solved, the project will face higher costs, delays, and local opposition. Gradiant is positioning itself as the managed water layer for hyperscale AI, similar to how power providers, cooling vendors, and network suppliers each own critical infrastructure domains. The Pattern: Hybridization, Standardization, and Industrial Scale The announcements included here make it clear that cooling is seeing significant attention from technology vendors, and not just state-of-the-art new technologies such as direct-to-chip, but also traditional data center air cooling. T-Global and SiPearl are working on high-conductivity materials and two-phase modules for HPC chips. Castrol is providing fluids for direct-to-chip and immersion environments. These are technologies aimed at the heat source itself, where higher chip power and rack density are overwhelming conventional approaches. The reference design offerings from Johnson Controls acknowledges the importance

Microsoft will invest $80B in AI data centers in fiscal 2025

And Microsoft isn’t the only one that is ramping up its investments into AI-enabled data centers. Rival cloud service providers are all investing in either upgrading or opening new data centers to capture a larger chunk of business from developers and users of large language models (LLMs). In a report published in October 2024, Bloomberg Intelligence estimated that demand for generative AI would push Microsoft, AWS, Google, Oracle, Meta, and Apple would between them devote $200 billion to capex in 2025, up from $110 billion in 2023. Microsoft is one of the biggest spenders, followed closely by Google and AWS, Bloomberg Intelligence said. Its estimate of Microsoft’s capital spending on AI, at $62.4 billion for calendar 2025, is lower than Smith’s claim that the company will invest $80 billion in the fiscal year to June 30, 2025. Both figures, though, are way higher than Microsoft’s 2020 capital expenditure of “just” $17.6 billion. The majority of the increased spending is tied to cloud services and the expansion of AI infrastructure needed to provide compute capacity for OpenAI workloads. Separately, last October Amazon CEO Andy Jassy said his company planned total capex spend of $75 billion in 2024 and even more in 2025, with much of it going to AWS, its cloud computing division.

John Deere unveils more autonomous farm machines to address skill labor shortage

Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More Self-driving tractors might be the path to self-driving cars. John Deere has revealed a new line of autonomous machines and tech across agriculture, construction and commercial landscaping. The Moline, Illinois-based John Deere has been in business for 187 years, yet it’s been a regular as a non-tech company showing off technology at the big tech trade show in Las Vegas and is back at CES 2025 with more autonomous tractors and other vehicles. This is not something we usually cover, but John Deere has a lot of data that is interesting in the big picture of tech. The message from the company is that there aren’t enough skilled farm laborers to do the work that its customers need. It’s been a challenge for most of the last two decades, said Jahmy Hindman, CTO at John Deere, in a briefing. Much of the tech will come this fall and after that. He noted that the average farmer in the U.S. is over 58 and works 12 to 18 hours a day to grow food for us. And he said the American Farm Bureau Federation estimates there are roughly 2.4 million farm jobs that need to be filled annually; and the agricultural work force continues to shrink. (This is my hint to the anti-immigration crowd). John Deere’s autonomous 9RX Tractor. Farmers can oversee it using an app. While each of these industries experiences their own set of challenges, a commonality across all is skilled labor availability. In construction, about 80% percent of contractors struggle to find skilled labor. And in commercial landscaping, 86% of landscaping business owners can’t find labor to fill open positions, he said. “They have to figure out how to do

2025 playbook for enterprise AI success, from agents to evals

Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More 2025 is poised to be a pivotal year for enterprise AI. The past year has seen rapid innovation, and this year will see the same. This has made it more critical than ever to revisit your AI strategy to stay competitive and create value for your customers. From scaling AI agents to optimizing costs, here are the five critical areas enterprises should prioritize for their AI strategy this year. 1. Agents: the next generation of automation AI agents are no longer theoretical. In 2025, they’re indispensable tools for enterprises looking to streamline operations and enhance customer interactions. Unlike traditional software, agents powered by large language models (LLMs) can make nuanced decisions, navigate complex multi-step tasks, and integrate seamlessly with tools and APIs. At the start of 2024, agents were not ready for prime time, making frustrating mistakes like hallucinating URLs. They started getting better as frontier large language models themselves improved. “Let me put it this way,” said Sam Witteveen, cofounder of Red Dragon, a company that develops agents for companies, and that recently reviewed the 48 agents it built last year. “Interestingly, the ones that we built at the start of the year, a lot of those worked way better at the end of the year just because the models got better.” Witteveen shared this in the video podcast we filmed to discuss these five big trends in detail. Models are getting better and hallucinating less, and they’re also being trained to do agentic tasks. Another feature that the model providers are researching is a way to use the LLM as a judge, and as models get cheaper (something we’ll cover below), companies can use three or more models to

OpenAI’s red teaming innovations define new essentials for security leaders in the AI era

Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More OpenAI has taken a more aggressive approach to red teaming than its AI competitors, demonstrating its security teams’ advanced capabilities in two areas: multi-step reinforcement and external red teaming. OpenAI recently released two papers that set a new competitive standard for improving the quality, reliability and safety of AI models in these two techniques and more. The first paper, “OpenAI’s Approach to External Red Teaming for AI Models and Systems,” reports that specialized teams outside the company have proven effective in uncovering vulnerabilities that might otherwise have made it into a released model because in-house testing techniques may have missed them. In the second paper, “Diverse and Effective Red Teaming with Auto-Generated Rewards and Multi-Step Reinforcement Learning,” OpenAI introduces an automated framework that relies on iterative reinforcement learning to generate a broad spectrum of novel, wide-ranging attacks. Going all-in on red teaming pays practical, competitive dividends It’s encouraging to see competitive intensity in red teaming growing among AI companies. When Anthropic released its AI red team guidelines in June of last year, it joined AI providers including Google, Microsoft, Nvidia, OpenAI, and even the U.S.’s National Institute of Standards and Technology (NIST), which all had released red teaming frameworks. Investing heavily in red teaming yields tangible benefits for security leaders in any organization. OpenAI’s paper on external red teaming provides a detailed analysis of how the company strives to create specialized external teams that include cybersecurity and subject matter experts. The goal is to see if knowledgeable external teams can defeat models’ security perimeters and find gaps in their security, biases and controls that prompt-based testing couldn’t find. What makes OpenAI’s recent papers noteworthy is how well they define using human-in-the-middle

Stay Ahead, Stay ONMINE