Tag: Dell

  • The New AI Math: Time-to-Token and Cost-per-Token Gets Highlighted at Dell Technologies World

    The New AI Math: Time-to-Token and Cost-per-Token Gets Highlighted at Dell Technologies World

    At Dell Technologies World this morning, Michael Dell introduced new metrics for measuring whether enterprise AI infrastructure is actually delivering. The AI infrastructure conversation has been dominated by GPU counts, cloud-versus-on-premises debates, and model benchmarks. Dell’s opening keynote added two measures that tie those inputs to outcomes. The first is  how quickly your infrastructure generates tokens, and the second is at what cost.

    Uptime, GPU capacity, and model advances are all foundational. GPUs are what get you to tokens in the first place. But they are not enough on their own to make infrastructure decisions. Organizations are now facing a more challenging optimization problem. Companies face a seemingly endless demand for AI compute and the energy required to support it. Businesses must also balance the performance of these evolving AI workloads within budgetary and location constraints.

     Time-to-token and cost per token are the metrics that tie these competing priorities together. They measure how quickly and how cheaply your infrastructure converts data and compute into intelligence that agents and models can act on. Jensen Huang reinforced this on stage alongside Dell, and OpenAI’s Greg Brockman made the same point independently on X today: “tokens are rapidly becoming the universal input for solving problems.” When the infrastructure providers and the model providers converge on the same metrics within weeks of each other, that is a directional signal worth acting on.

    What Michael Dell described was a two-year refinement of the Dell AI Factory with NVIDIA, informed by 5,000 enterprise customers now running production AI workloads on it. The announcements were substantial. What they mean for enterprise buyers is worth unpacking.

    The Data Bottleneck Is the Real Constraint

    Michael Dell said something on stage that every CIO needs to hear: “If your data is siloed, your agents are blind.” That is a concise description of why so many enterprise AI programs stall after the pilot. It also echoes what Irfan Khan and Muhammed Alam described as the need for business context at SAP’s Sapphire and in an online event about business data

    Most organizations are simultaneously trying to prepare data for AI and reengineer the data infrastructure needed to support it. Those are two hard problems happening at the same time, on top of each other. Dell’s announcements around the Dell AI Data Platform addressed both directly.

    One of the things that has changed is the data orchestration engine, the intelligent control center within the Dell AI Data Platform that turns raw, fragmented enterprise data into production-ready AI fuel. It indexes billions of unstructured files of all types, builds governed data pipelines, connects them to the models and agents that need them, and delivers structured outputs at speeds that make agentic workflows viable. Dell claims the platform now delivers 12 times faster vector indexing, 6 times faster data querying, and 19 times faster time to first token than prior generations. While the claims still need to be verified, the direction is exactly what enterprise buyers need.

    Why does this matter? AI agents need business context to be useful. An agent that can reason brilliantly but cannot reliably access your CRM, internal knowledge bases, operational systems, or proprietary data is not doing useful work. The data orchestration engine is what connects the model to the context. Without it, you have a powerful system with nothing meaningful to act on.

    Dell’s approach integrates orchestration, search, and governed pipelines natively into the platform. Getting that platform connected to your actual data sources still requires integration work. Budget for it before you buy the hardware.

    AI Infrastructure Is a Team Sport

    The broader lesson from the Dell Technologies keynote is not about any specific product. It is about what has changed in two years.

    When Dell announced the Dell AI Factory with NVIDIA in 2024, it was largely a hardware and partnership story. Today, with 5,000 enterprise customers running production workloads on it, the conversation has shifted to execution. How do you get from pilot to production? How do you keep agents from being blind to your actual business data? How do you manage cost curves as token consumption scales? How do you maintain security and governance when agents are operating autonomously at machine speed?

    Part of Dell’s answer is its ecosystem. The new Dell AI Ecosystem Program gives AI software providers a validated path to certify solutions on Dell infrastructure. For enterprise buyers, this speeds AI deployments by reducing integration some of the integration burden. Rather than assembling a custom stack from scratch, you get pre-validated blueprints that automate the deployment of software, services, and models together. That automation is a direct lever for reducing time-to-token at the program level. Dell claims it can deliver hundreds of AI racks a week to a given customer and have them generating outcomes within hours. Part of this is also achieved with ecosystem partner blueprints for automation. 

    The ecosystem also extends to the agent layer itself. Jensen Huang described on stage how agents do not run directly on the large language model. They run on a harness. The harness sits in a secure, governed container called a sandbox. It manages the agent’s reasoning loop, handles tool use, controls what data and systems the agent can access, and determines when to call the larger model and when to use a smaller local model instead. NVIDIA’s OpenShell is the open-source sandbox now supported across the entire Dell AI Factory. For enterprise buyers evaluating AI infrastructure, the harness is not a detail. It is a primary evaluation criterion. An infrastructure stack that does not clearly define how agent harnesses are deployed, secured, and governed is not production-ready.

    The ecosystem partners Dell named today include Google, Hugging Face, OpenAI, Palantir, ServiceNow, and SpaceXAI, among others. AI is not a solo deployment. The strength of the ecosystem around the infrastructure determines how fast you can actually move.

    Michael Dell put the security dimension plainly: “You can’t protect what you can’t see, and you can’t manage what you can’t see.” That applies to agents as much as it applies to data. Agents have credentials, memory, and access to systems. When they operate autonomously at machine speed, the blast radius of a security failure is no longer contained to one system. It can propagate across workflows and infrastructure. There are also tech tools such as X that help with confidential computing. 

    The Token Economics of Hybrid AI

    Sixty-seven percent of AI workloads already run outside the public cloud, on-premises, at the edge, or in co-location environments, according to Dell’s own survey data. Eighty-eight percent of organizations are running at least one AI workload on-premises. But, that does not mean cloud is going away. It means the real question for enterprise infrastructure leaders is not cloud versus on-premises. It is how to run both well.

    Hybrid AI is not a compromise. It is the architectural reality for most large enterprises. Some workloads need the cloud, which offers speed, training, high capacity, and flexibility. Others belong on-premises because it may access sensitive data that a company doesn’t want in the cloud or regulations require to be in a certain place. There may also be high-volume, continuous inference, where unpredictable cloud token costs create real budget exposure. The strategic challenge is matching the workload to the right environment and doing it consistently at scale.

    Jensen Huang described on stage why the compute requirements have shifted so dramatically. Agentic systems require 100x to 1,000x more computation than responding to a simple query because the agent has to reason, plan, use tools, evaluate results, and iterate. At that scale of consumption, every infrastructure decision has a direct cost consequence.

    Dell’s answer for high-volume on-premises workloads is what it calls “unmetered intelligence.” The idea is that owning infrastructure converts variable cloud API spend into a fixed infrastructure cost. Dell claims organizations can break even on API costs compared to the public cloud in as little as 3 months with desk-side agentic AI configurations.

    Balancing this correctly requires thinking about four variables simultaneously. Latency measures the time it takes for a system to process a request and return a response. Performance refers to the overall capability, accuracy, and capacity of the AI model to handle complex tasks. Cost is what you pay to get those outputs at the required latency and performance level. Energy is the fourth variable, and it is no longer theoretical. A single rack of NVIDIA Rubin GPUs can draw over 130 kilowatts.

    Granted, the average enterprise won’t be running a rack of Vera Rubin’s, but energy availability is becoming a real constraint regardless of sustainability goals. And if you’re using the cloud, you pay one way or the other for that energy. The right infrastructure solution varies by workload type. The time to token and the cost per token let you compare options on the same terms.

    Sovereign AI Is Becoming a Procurement Reality

    Two years ago, sovereign AI was a concept mostly discussed in European regulatory contexts and by a small number of governments building national AI infrastructure. Today, it shows up in enterprise procurement conversations across regulated and unregulated industries.

    Sovereign AI means the ability to independently develop, deploy, and govern AI systems entirely within an organization’s strategic, legal, and jurisdictional boundaries. For enterprises, this means your data does not leave your environment, your model choices are not constrained by a hyperscaler’s catalog, and your AI outputs are not subject to external policy changes.

    The ecosystem Dell announced is designed to both speed AI deployments and address sovereign AI requirements. Google’s Gemini 3 Flash models running on-premises via Google Distributed Cloud on Dell PowerEdge servers. OpenAI’s Codex is connected to the Dell AI Data Platform for agentic workflows on enterprise data. Palantir’s Foundry and AIP platform is deployed on-premises with Dell ObjectScale and PowerFlex as the data layer. SpaceXAI’s Grok is available in on-premises or hybrid enterprise deployments. Reflection’s open-source frontier models for regulated industries and sovereign entities.

    The pattern is consistent: bring the model to the data rather than the data to the model. For organizations in healthcare, financial services, defense, and government, this is not a preference. It is often a compliance requirement.

    Planning for Hybrid AI

    Today’s AI question is how to architect a hybrid AI solution that aligns with our organization’s specific workloads, data environment, cost constraints, and governance requirements. Some of that runs on-premises. Some runs in the cloud. The mix differs across organizations and will shift as workloads evolve and model costs change.

    The questions worth asking now: What is your time to first token across your most important workloads? What is your cost per token at scale? Does your data orchestration layer connect your proprietary data to the models that need it? How are your agent harnesses deployed and governed? And do you have the security architecture in place before your agents start making autonomous decisions?

    It’s not easy but nothing worthwhile ever is. 

     

  • Dell Shares AI Advances And New Metrics To Evaluate Infrastructure

    Dell Shares AI Advances And New Metrics To Evaluate Infrastructure

    At Dell Technologies World in Las Vegas, Dell Technologies chairman and CEO Michael Dell made a pointed argument to a room full of enterprise technology leaders: the metrics organizations use to evaluate infrastructure are evolving.

    GPU counts, cloud versus on-premises comparisons, and model benchmarks have dominated the conversation. Michael Dell’s day one keynote introduced two additional measures aimed at tying infrastructure decisions to actual outcomes: time to token, which measures how quickly a system processes a request and returns a usable AI output, and cost per token, which measures how cheaply that output is produced at scale.

    “Time to first token is incredibly important with investments of this scale,” Dell said on stage, noting that the company now has 5,000 enterprise customers running production AI workloads on its Dell AI Factory with NVIDIA platform. The figure represents a significant increase from the program’s launch two years ago.

    NVIDIA founder and CEO Jensen Huang, appearing alongside Dell, described why those metrics have taken on new urgency at both its NVIDIA GTC conference and at Dell Technologies World. Agentic AI systems, which reason, plan, and execute tasks autonomously over extended periods, require anywhere from 100 to 1,000 times more computation than a system simply responding to a query. “What took months now takes weeks, what took weeks now takes days, and what takes days now takes hours,” Huang said, describing the productivity transformation already underway at companies running agentic workflows. The demand implications for infrastructure are substantial.

    OpenAI president Greg Brockman echoed the framing independently on X.com the same day, writing that “tokens are rapidly becoming the universal input for solving problems.” The convergence of infrastructure vendors and model providers on the same metrics within weeks of each other signals a broader shift in how enterprise AI spending will be evaluated.

    The Data Problem Underneath the Infrastructure Problem

    One of Dell’s significant AI product announcements centered on a new data orchestration engine in the Dell AI Data Platform, which the company positioned as the missing layer between enterprise data and production-ready AI agents.

    The data orchestration engine is the platform’s intelligent control center. It indexes billions of unstructured files, builds governed data pipelines, and connects them to the models and agents that need them at speeds designed to make agentic workflows viable. Dell claims the updated platform delivers 12 times faster vector indexing, six times faster data querying, and 19 times faster time to first token compared to prior generations. While these claims still need to be validated, the proposed increase in performance is good news for enterprises looking to scale AI.

    The underlying problem the engine addresses is one most large organizations know well. Enterprises are simultaneously preparing existing data for AI use and reengineering the data infrastructure required to support AI workloads at scale. Those two efforts compete for the same resources and skills simultaneously.

    “If your data is siloed, your agents are blind,” Dell said. The statement is a precise description of why many enterprise AI pilots have not reached production. An agent operating without access to an organization’s proprietary data, internal knowledge bases, and operational systems cannot deliver the business context that makes agentic AI useful.

    Dell also announced GPU-accelerated SQL analytics through the Dell Data Analytics Engine, powered by Starburst, delivering up to six times faster query performance on NVIDIA Blackwell GPUs. Bank of America, which already has a partnership with Starburst, NVIDIA, and Dell, is among the institutions expected to use the capability.

    A Broad Ecosystem Built to Reduce Time to Production

    Dell announced a new Dell AI Ecosystem Program alongside a significant expansion of frontier model partnerships, positioning both as mechanisms for reducing the time between infrastructure procurement and production AI deployment.

    On the model side, Dell announced collaborations bringing several major AI providers on-premises to the Dell AI Factory. Google and Dell are collaborating to run Gemini 3 Flash models via Google Distributed Cloud on Dell PowerEdge XE9780 servers, enabling enterprises to run advanced generative AI workloads in a confidential computing environment that meets data residency and sovereignty requirements. OpenAI’s Codex will connect with the Dell AI Data Platform, giving enterprises a path to deploy agentic coding capabilities against their internal codebases, documentation, and business systems. SpaceXAI’s Grok is available in on-premises or hybrid enterprise deployments. Palantir’s Foundry and AIP platform is coming on-premises with its Ontology layer deployed on Dell ObjectScale and PowerFlex, allowing organizations to connect data sources and automate business workflows within their own environment.

    The Dell Enterprise Hub on Hugging Face gives enterprises on-premises access to a curated collection of open-weight models including MiniMax-M2.7, DeepSeek Pro, DeepSeek-V4, GLM 5.1, and Kimi K2.6, optimized for Dell AI Factory infrastructure.

    The Dell AI Ecosystem Program formalizes the partner relationship by providing software providers with a validated path to certify their solutions on Dell infrastructure. For enterprise buyers, the practical benefit is pre-validated deployment blueprints that automate the configuration of a specific software, service, or model, reducing integration work that has historically extended timelines from procurement to production.

    The Agent Harness: An Evaluation Criterion Enterprises Are Not Yet Asking About

    One of the more technically substantive moments in the keynote came from Huang’s description of how agents actually operate in production. Agents, he explained, do not run directly on the large language model. They run on a harness, a software layer that sits in a secure, governed sandbox. The harness manages the agent’s reasoning loop, controls tool access, determines when to call a large external model and when to use a smaller local model, and handles memory and context across multi-step tasks.

    NVIDIA’s OpenShell, the open-source sandbox, is now supported across the entire Dell AI Factory from deskside workstations through PowerEdge data center servers.  Dell also announced support for NVIDIA AIQ i.0 blueprints, which provide tested foundations for deploying multi-agent workflows.

    For CIOs evaluating AI infrastructure, the harness architecture is a meaningful addition to the evaluation checklist. Infrastructure that does not clearly define how agent harnesses are deployed, governed, and secured leaves a significant operational and security gap, particularly as agents acquire credentials, access enterprise systems, and take autonomous actions at machine speed.

    For example, “You can’t protect what you can’t see, and you can’t manage what you can’t see,” Dell said, framing the security challenge in terms that apply as directly to agents as to human users. An agent with compromised access or misconfigured permissions can propagate errors or security failures across workflows in ways that a single human user cannot.

    Hybrid AI Infrastructure and the Energy Constraint

    Dell’s survey data shows that 67% of AI workloads are already running outside the public cloud, and 88% of organizations are running at least one AI workload on-premises. The company positioned hybrid AI not as a transitional state but as the long-term architecture reality for most large enterprises.

    The new Dell PowerRack, announced Monday, is a fully integrated rack-scale system that combines compute, networking, and storage, engineered and validated as a single unit. It is designed to reduce the integration overhead of assembling AI infrastructure from components while supporting thermal management and power optimization at rack scale.

    Dell also introduced the Dell PowerCool CDU C7000, the first rack-mount cooling distribution unit designed to meet the cooling requirements of the NVIDIA Vera Rubin NVL72 platform, delivering more than 220 kilowatts of cooling capacity in a 4U form factor. A single rack of NVIDIA Rubin GPUs can draw over 130 kilowatts of power, and Dell noted that energy availability is an increasingly real constraint on AI deployment timelines, independent of sustainability considerations.

    For high-volume on-premises workloads, Dell introduced Dell Deskside Agentic AI, pairing high-performance Dell Pro Precision workstations with NVIDIA NemoClaw. The company claims the configuration enables enterprises to break even against public cloud API costs in as little as 3 months, converting variable token costs into a fixed infrastructure investment.

    What Changes for Enterprise Buyers

    The announcements from Dell Technologies World day one collectively continue to move the enterprise AI infrastructure conversation from capability to faster execution. The core questions are how quickly a given infrastructure configuration can reach first token on a production workload, at what cost per token, and with what governance architecture underpinning the agents running on it.

    The organizations best positioned to answer those questions are the ones that have already started rationalizing their data architecture, defined their hybrid workload placement strategy, and begun evaluating how agent harnesses will be secured and governed. The infrastructure improves almost daily, but the execution discipline required to use it remains the variable that separates AI programs that reach production from those that stay in pilot.

    Maribel Lopez is the founder and principal analyst at Lopez Research, a market research and strategy consulting firm specializing in enterprise AI, AI infrastructure, agentic systems, AI governance, and AI-driven customer experience. I version of this article of originally posted on Forbes.com.

  • Four Types of AI Agents With Dell’s John Roese. Most Enterprises Are Only Building One

    Four Types of AI Agents With Dell’s John Roese. Most Enterprises Are Only Building One

    Dell's CTO built a 4-category agent framework from real production deployments. Most enterprises are ignoring two of the categories that matter most.


    Full Show Notes

    Enterprise leaders are mapping AI agents to org charts — building digital employees, agentic teams, AI workers — and then wondering why the results fall short. Dell's Global CTO John Roese has been running agents in production long enough to know exactly why that framing fails, and what to do instead.

    In this episode, Roese shares a framework Dell developed from actual production deployments, not pilots. It identifies four categories of AI agents defined by two dimensions: how much autonomy you grant the agent, and how complex the underlying process is. Most enterprises are focused on one category. Two of the four are widely overlooked — and they may represent the fastest path to measurable ROI.

    This is a practical, grounded conversation about where agents are actually delivering value today, how to think about infrastructure cost in the context of agent economics, and why the sequence in which you deploy agents matters as much as which agents you build. If your organization is trying to move from AI experimentation to production, this episode is required listening.


    3. Chapter titles:

    • [00:00] — Introduction: Dell's dual role as tech vendor and enterprise AI user
    • [01:38] — Why the org chart model for agents fails
    • [03:12] — Decoupling human capacity from work capacity for the first time
    • [04:23] — The two-by-two framework: autonomy vs. process complexity
    • [06:14] — Productivity agents: what most enterprises already have
    • [07:00] — Hygiene agents: the overlooked category that fixes foundational data problems
    • [08:01] — The CRM data example: why every CRM is inaccurate and how agents fix it
    • [10:05] — Latent infrastructure capacity: running agents in GPU white space to cut costs to cents
    • [13:53] — Facilitation agents: removing entropy from complex cross-functional workflows
    • [17:30] — The sequencing insight: hygiene and facilitation as the path to expert agents
    • [19:24] — Why coordination agents aren't agentic bosses — and where human control actually lives
    • [22:21] — Roese's closing advice: become literate, pick a few, get them into production


    4. Guest Bio

    John Roese is the Global Chief Technology Officer and Chief AI Officer at Dell Technologies, where he is responsible for technology strategy, AI deployment, and research and development across the company. He has held senior technology leadership roles at Nortel, Enterasys Networks, Broadcom, and EMC. At Dell, he operates at a rare intersection: leading AI strategy for a major technology vendor while also deploying AI internally at enterprise scale — which means his frameworks are tested against real production constraints, not just market positioning.


    About This Podcast

    AI with Maribel Lopez is a podcast for enterprise technology leaders navigating AI adoption, agentic systems, AI infrastructure, and AI governance. Host Maribel Lopez covers enterprise technology and advises CIOs, CDOs, CMOs, and technology vendors on how to move from AI experimentation to measurable business outcomes. New episodes published bi-weekly.

    Subscribe on your platform of choice: buzzsprout.com/1947446

  • Dell AI Factory Expands with 40+ Enhancements for Enterprise AI Deployment

    Dell AI Factory Expands with 40+ Enhancements for Enterprise AI Deployment

    Dell Technologies unveiled a significant expansion of its AI Dell Factory platform at its annual Dell Technologies World conference today, announcing over 40 product enhancements designed to help enterprises deploy artificial intelligence workloads more efficiently across both on-premises environments and cloud systems.

    The Dell AI Factory is not a physical manufacturing facility but a comprehensive framework combining advanced infrastructure, validated solutions, services, and an open ecosystem to help businesses harness the full potential of artificial intelligence across diverse environments—from data centers and cloud to edge locations and AI PCs.

    The company has attracted over 3,000 AI Factory customers since launching the platform last year. In an earlier call with industry analysts, Dell shared research stating that 79% of production AI workloads are running outside of public cloud environments—a trend driven by cost, security, and data governance concerns. During the keynote, Michael Dell provided more color on the value of Dell’s AI factory concept. He said, “The Dell AI factory is up to 60% more cost effective than the public cloud, and recent studies indicate that about three-fourths of AI initiatives are meeting or exceeding expectations. That means organizations are driving ROI and productivity gains from 20% to 40% in some cases. 

    Making AI Easier to Deploy

    Organizations need the freedom to run AI workloads wherever makes the most sense for their business, without sacrificing performance or control. While IT leaders embraced the public cloud for their initial AI services, many organizations are now looking for a more nuanced approach where the company can control over their most critical AI assets while maintaining the flexibility to use cloud resources when appropriate. Over 80 percent of the companies Lopez Research interviewed said they struggled to find the budget and technical talent to deploy AI. These AI deployment challenges have only increased as more AI models and AI infrastructure services have been launched.

    Silicon Diversity and Customer Choice

    A central theme of Dell’s AI Factory message is how Dell makes AI easier to deploy while delivering choice. Dell is offering customers choice through silicon diversity in its designs, but also with ISV models. The company announced it has added Intel to its AI Factory portfolio with Intel Gaudi 3 AI accelerators and Intel Xeon processors, with a strong focus on inferencing workloads.

    Dell also announced its fourth update to the Dell AI Platform with AMD, rolling out two new PowerEdge servers—the XE9785 and the XE9785L—equipped with the latest AMD Instinct MI350 series GPUs. The Dell AI Factory with NVIDIA combines Dell’s infrastructure with NVIDIA’s AI software and GPU technologies to deliver end-to-end solutions that can reduce setup time by up to 86% compared to traditional approaches. The company also continues to strengthen its partnership with NVIDIA, announcing products leveraging NVIDIA’s Blackwell family and other updates launched at NVIDIA GTC. As of today, Dell supports choice by delivering AI solutions with all of the primary GPU and AI accelerator infrastructure providers.

    Client-Side AI Advancements

    At the edge of the AI Factory ecosystem, Dell announced enhancements to the Dell Pro Max in a mobile form factor, leveraging Qualcomm’s AI 100 discrete NPUs designed for AI engineers and data scientists who need fast inferencing capabilities. With up to 288 TOPs at 16-bit floating point precision, these devices can power up to a 70-billion parameter model, delivering 7x the inferencing speed and 4x the accuracy over a 40 TOPs NPU. Dell says the Pro Max Plus line can run a 109-billion-parameter AI model.

    The Pro Max and Plus launches follow Dell’s previous announcement of AI PCs featuring Dell Pro Max with GB 10 and GB 300 processors powered by NVIDIA’s Grace Blackwell architecture. Overall, Dell has simplified its PC portfolio but made it easier for customers to choose the right system for their workloads by providing the latest chips from AMD, Intel, Nvidia, and Qualcomm.

    On-Premise AI Deployment Gains Ecosystem Momentum

    Following the theme of choice, organizations need the flexibility to run AI workloads on-premises and in the cloud. Dell is making significant strides in enabling on-premise AI deployments with major software partners. The company announced it is the first provider to bring Cohere capabilities on-premises, combining Cohere’s generative AI models with Dell’s secure, scalable infrastructure for turnkey enterprise solutions.

    Similar partnerships with Mistral and Glean were also announced, with Dell facilitating their first on-premise deployments. Additionally, Dell is supporting Google’s Gemini on-premises with Google Distributed Cloud.

    To simplify model deployment, Dell now offers customers the ability to choose models on Hugging Face and deploy them in an automated fashion using containers and scripts. Enterprises increasingly recognize that while public cloud AI has its place, a hybrid AI infrastructure approach could deliver better economics and security for production workloads.

    The imperative for scalable yet efficient AI infrastructure at the edge is a growing need. As Michael Dell said during his Dell Technologies World keynote, “Over 75% of enterprise data will soon be created and processed at the edge, and AI will follow that data; it’s not the other way around. The future of AI will be decentralized, low latency, and hyper-efficient.”

    Dell’s ability to offer robust hybrid and fully on-premises solutions for AI is proving to be a significant advantage as companies increasingly seek on-premises support and even potentially air-gapped solutions for their most sensitive AI workloads. Key industries adopting the Dell AI Factory include finance, retail, energy, and healthcare providers.

    Scaling AI Requires a Focus on Energy Efficiency 

    Simplifying AI also requires product innovations that deliver cost-effective, energy-efficient technology. As AI workloads drive unprecedented power consumption, Dell has prioritized energy efficiency in its latest offerings. The company introduced the Dell PowerCool Enclosed Rear Door Heat Exchanger (eRDHx) with Dell Integrated Rack Controller (IRC). This new cooling solution captures nearly 100% of the heat coming from GPU-intensive workloads. This innovation lowers cooling energy requirements for a rack by 60%, allowing customers to deploy 16% more racks with the same power infrastructure.

    Dell’s new systems are rated to operate at 32 to 37 degrees Celsius, supporting significantly warmer temperatures than traditional air-cooled or water-chilled systems, further reducing power consumption for cooling. The PowerEdge XE9785L now offers Dell liquid cooling for flexible power management. Even if a company isn’t aiming for a specific sustainability goal, every organization wants to improve energy utilization. 

    Early Adopter Use Cases Highlight AI’s Opportunity

    With over 200 product enhancements to its AI Factory in just one year, Dell Technologies is positioning itself as a central player in the rapidly evolving enterprise AI infrastructure market. It offers the breadth of solutions and expertise organizations require to successfully implement production-grade AI systems in a secure and scalable fashion. However, none of this technology matters if enterprises can’t find a way to create business value by adopting it. Fortunately, examples from the first wave of enterprise early adopters highlight ways AI can deliver meaningful returns in productivity and customer experience. Let’s look at two use cases presented at Dell Tech World. 

    The Power of LLMs in Finance at JPMorgan Chase

    JPMorgan Chase took the stage to make AI real from a customer’s perspective. The financial firm uses Dell’s compute hardware, software-defined storage, client, and peripheral solutions. Larry Feinsmith, the Managing Director and Head of Global Tech Strategy, Innovation & Partnerships at JPMorgan Chase, said, “We have a hybrid, multi-cloud, multi-provider strategy. Our private cloud is an incredibly strategic asset for us. We still have many applications and data on-premises for resiliency, latency, and a variety of other benefits.” 

    Feinsmith also spoke of the company’s Large Language Model (LLM) strategy. He said, “Our strategy is to use a constellation of models, both foundational and open, which requires a tremendous amount of compute in our data centers, in the public cloud, and, of course, at the edge. The one constant thing, whether you’re training models, fine-tuning models, or finding a great use case that has large-scale inferencing, is that they all will drive compute. We think Dell is incredibly well positioned to help JPMorgan Chase and other companies in their AI journey.”

    Feinsmith noted that using AI isn’t new for JPMorgan Chase. For over a decade, JPMorgan Chase has leveraged various types of AI, such as machine learning models for fraud detection, personalization, and marketing operations. The company uses what Feinsmith called its LLM suite, which over 200,000 people at JPMorgan Chase use today. The generative AI application is used for QA summarization and content generation using JPMorgan Chase’s data in a highly secure way. Next, it has used the LLM suite architecture to build applications for its financial advisors, contact center agents, and any employee interacting with its clients. Its third use case highlighted changes in the software development area. JPMorgan Chase rolled out code generation AI capabilities to over 40,000 engineers. It has achieved as much as 20% productivity in the code creation and expects to leverage AI throughout the software development life cycle. Going forward, the financial firm expects to use AI agents and reasoning models to execute complex business processes end-to-end.

    Seemantini Godbole, EVP and Chief Digital and Information Officer at Lowe's, shared insights on designing the strategy for AI

    How AI Makes It Easier For Employees to Serve Customers at Lowe’s

    Lowe’s Home Improvement Stores provided another example of how companies are leveraging Dell Technology and AI to transform the customer and employee experience. Seemantini Godbole, EVP and Chief Digital and Information Officer at Lowe’s, shared insights on designing the strategy for AI when she said, “How should we deploy AI? We wanted to do impactful and meaningful things. We did not want to die a death of 1000 pilots, and we organized our efforts across how we sell, how we shop, and how we work. How we sell was for our associates. How we shop is for our customers, and how we work is for our headquarters employees. For whatever reason, most companies have begun with their workforce in the headquarters. We said, No, we are going to put AI in the hands of 300,000 associates.” For example, she described a generative AI companion app for store associates. “Every store associate now has on his or her zebra device a ChatGPT-like experience for home improvement.”, said Godbole. Lowe’s is also deploying computer vision algorithms at the edge to understand issues such as whether a customer in a particular aisle is waiting for help. The system will then send notifications to the associates in that department. Customers can also ask various home improvement questions, such as what paint finish to use in a bathroom, at Lowes.com/AI.

    Designing A World Where AI Delivers Human Opportunity

    Michael Dell said, “We are entering the age of ubiquitous intelligence, where AI becomes as essential as electricity, with AI, you can distill years of experience into instant insights, speeding up decisions and uncovering patterns in massive data. But it’s not here to replace humans. AI is a collaborator that frees your teams to do what they do best, to innovate, to imagine, and to solve the world’s toughest problems.” 

    While there are many AI deployment challenges ahead, the customer examples shared at Dell Technologies World provide a glimpse into a world where AI benefits both customers and employees. The challenge now is to do this sustainably and ethically at scale.