The Power Problem in AI Drug Discovery: Lessons from BMS
If energy is the new bottleneck, then efficiency is the only metric that actually matters for production.

Automation needs a narrow first win
The best first AI workflow is usually a repeated task with a clear input, clear output, and a human approval step.
Bristol Myers Squibb (BMS) is moving into Nvidia’s Vera Rubin architecture with a new DGX SuperPOD to support its drug discovery and development operations. This move marks BMS as the first life sciences group to acquire infrastructure based on this specific architecture, signaling a shift toward more specialized, high-performance computing in the pharmaceutical space. But this isn't just a hardware refresh; it’s a strategic effort to centralize proprietary model training and large-scale scientific predictions into a unified environment.
Breaking Down Research Silos
The new infrastructure consists of eight DGX Vera Rubin NVL72 systems, combining Nvidia Vera CPUs and Rubin GPUs. While the raw specs are impressive, the real value lies in the integration. By combining these with an older SuperPOD into a shared computing environment, BMS is effectively institutionalizing its learnings. Instead of data and insights remaining trapped in individual research pockets across different sites, researchers worldwide can now access a shared pool of resources. This includes Nvidia’s BioNeMo Agent Toolkit for protein-structure prediction, molecular generation, and genomics. The goal is clear: move away from fragmented "lab-scale" compute toward a shared corporate infrastructure that ensures every scientist is working from the same foundational models.

Phugialy Picks

The Agentic AI Bible: The Complete and Up-to-Date Guide to Design, Develop, and Scale Goal-Driven, LLM-Powered Agents that Think, Execute...
Some Phugialy Picks use affiliate links. If you buy through one, Phugialy may earn a commission. It doesn't change what we recommend. Full disclosure →
Scaling From Feasibility to Throughput
The practical goal of this investment is to prioritize synthesis and reduce manual research overhead. By using predictions to optimize the synthesis of molecules with multi-parameter optimization, BMS aims to ensure that laboratory experiments are only performed on molecules with the highest probability of success. The company reports that this approach has already reduced some manual research work by 20% to 30%, with a target of reaching 50% in the coming years. This shift is best summarized by the transition from being able to process "10" items to processing "dozens." It represents a move away from small-scale feasibility testing toward a production-grade pipeline where compute power allows for a broader sweep of the chemical space while maintaining a high degree of selection accuracy.
The Reality of the Power Bill
The most telling detail in this announcement is the focus on performance per megawatt. BMS notes that the new cluster is expected to deliver up to 10 times the performance per megawatt of the infrastructure it replaces. This is a grounded acknowledgment of a harsh reality: "Electricity is not getting cheaper." For anyone who has actually deployed large-scale AI, you know that the biggest bottleneck is often not the raw chip speed, but the power availability and the cost of cooling. By highlighting a 10x efficiency gain, BMS is signaling that they are navigating the transition from "compute at any cost" to "sustainable compute." The real story here isn't just that they can predict more molecules; it's that they are building a system designed to survive the economic and physical constraints of modern data centers. It’s a move toward production-scale maturity where energy density is just as important as FLOPs.

Got a question about how this applies to you? →
Keep reading
Follow the thread
Middle-Mile Autonomy Gets Real: Inside Gatik's $200M Bet
$200 million is the headline; $600 million in contracted revenue against just $30 million recognized last year is the real story at Gatik. The company's bet on middle-mile autonomy only pays off if that pipeline converts into driverless trucks on schedule.
Read this noteSame lane, different angle
The Maintenance Debt of AI-Generated Code
We’re trading a minor speed boost for a massive, invisible tax on maintainers. Open-source projects are starting to ban AI contributions not because the code is "bad," but because the burden of auditing "code slop" is becoming unsustainable.
Jetson Orin Nano 2: Edge AI For Drones And Robots
NVIDIA's Jetson Orin Nano 2 doubles inference performance while using 40 percent less power in 15-watt mode - and it puts generative AI directly on drones and robots instead of in a data center. The specs look credible; what nobody has shown yet is independent benchmark data.