• Skip to main content
  • Skip to secondary menu
  • Skip to footer

Technologies.org

Technology Trends: Follow the Money

  • Technology Events 2026-2027
  • Sponsored Post
  • Technology Markets
  • About
    • GDPR
  • Contact

NVIDIA Unveils AI Platform to Minimize Downtime in Supercomputing Data Centers

June 22, 2020 By admin

NVIDIA today unveiled the NVIDIA® Mellanox® UFM® Cyber-AI platform, which minimizes downtime in InfiniBand data centers by harnessing AI-powered analytics to detect security threats and operational issues, as well as predict network failures.

This extension of the UFM platform product portfolio — which has managed InfiniBand systems for nearly a decade — applies AI to learn a data center’s operational cadence and network workload patterns, drawing on both real-time and historic telemetry and workload data. Against this baseline, it tracks the system’s health and network modifications, and detects performance degradations, usage and profile changes.

The new platform provides alerts of abnormal system and application behavior, and potential system failures and threats, as well as performs corrective actions. It is also targeted to deliver security alerts in cases of attempted system hacking to host undesired applications, such as cryptocurrency mining. The result is reduced data center downtime — which typically costs more than $300,000 an hour, according to research by ITIC.(1)

“The UFM Cyber-AI platform determines a data center’s unique vital signs and uses them to identify performance degradation, component failures and abnormal usage patterns,” said Gilad Shainer, senior vice president of marketing for Mellanox networking at NVIDIA. “It allows system administrators to quickly detect and respond to potential security threats and address upcoming failures, saving cost and ensuring consistent service to customers.”

Ecosystem Support
Organizations that have long been employing the UFM platform in their data centers have expressed strong interest in the latest offering.

Allan Williams, associate director of services and technology at the National Computational Infrastructure (NCI Australia), said: “NCI plays a pivotal role in the national research landscape. Our supercomputing infrastructure serves 5,000 researchers who use it for critical national and global activities. UFM enables us to effectively manage our supercomputers and to optimize performance. We look forward to utilizing the new capabilities of UFM Cyber-AI to enhance even further our supercomputing utilization and improve our return on investment.”

Douglas Johnson, associate director of the Ohio Supercomputer Center, said: “We have been using the UFM platform for years in our InfiniBand data centers. UFM and the expertise from the Mellanox networking team have been fundamental ingredients in the management of our network and the stability we’ve achieved. We see great advantages in the UFM Cyber-AI platform.”

Extending UFM Platform
The UFM Cyber-AI platform complements the UFM Enterprise platform, which provides network monitoring, management, performance optimization, configuration checks and secure cable management.

NVIDIA also added today a third member of the UFM family, the UFM Telemetry platform. This tool captures real-time network telemetry data, which is streamed to an on-premises or cloud-based database to monitor network performance and validate the network configuration.

Source: NVIDIA

Filed Under: Tech

Footer

Recent Posts

  • AltSql: An Embedded Database Engine That Syncs Devices and Gateways Without Conflicts
  • VPN Works: Uses Linux Network Namespaces to Isolate and Log Every Connection From an Agent
  • Precomputing: Materializes Dashboard Answers With SQLite Triggers as Data Arrives
  • Preconfiguration: Generates Reproducible Setup Across Multiple Cloud Platforms From One Spec
  • BareProxy: A Go Reverse Proxy That Makes Routing Decisions Explainable
  • AI Infrastructure Moves Beyond GPUs as Billions Flow Into Interconnects, Cloud, Robotics and Agent Systems
  • Supermicro Is Now Shipping NVIDIA Vera Rubin NVL72 Racks, With 1,152-GPU Scalable Units Ready to Order
  • Synopsys and TSMC Certify A14 Design Flows and Roll Out Agentic AI Chip Design Tools
  • Bird.com Secures $450M in Debt Financing and Opens Its Messaging Network to AI Agents
  • Meta AI Glasses Become the First Consumer Device to Record Dolby Atmos Audio

Media Partners

  • Market Analysis
  • BareProxy.com
  • App Coding
Screening Startup and Product Ideas After a Brainstorm: Four Questions, Money First
An AI Lab Is Paying Up Front for Atlas Energy’s (AESI) Generators as Agentic AI Multiplies Token Demand
Oracle’s Force Majeure Notice on Project Jupiter Shows Where AI Data Center Risk Is Landing
AI Infrastructure Credit Costs Rise as CoreWeave-Tied Bonds Price at 9.25% and China Chipmaker Profits Jump 620%
OpenAI and Anthropic Cut AI Model Prices as $1.75B in Funding Flows to Data, Security and Infrastructure
Semiconductor Revenue Hits Record $425B in Q2 2026, but Omdia’s $500B Q3 Forecast Implies Growth Halves
AI Extinction Warnings Went Global in Six Days. Nothing in the Technology Changed.
Anthropic Walks Away From $6 Billion Decart Acquisition: The Deal Was About Inference Cost, Not World Models
VR Status Report 2026: Quest Sales Keep Falling While Smart Glasses Take the Money
The Case That the US Can Grow Out of $40 Trillion in Debt: Three Conditions the Clinton Surpluses Actually Met
BareProxy 0.1 Alpha Adds Plan, Apply and Rollback
Introducing BareProxy, the Web Server and Reverse Proxy That Explains Itself
BareProxy Serves Hugo and Other Static Sites Straight From a Folder
How BareProxy Plan Shows What a Config Change Will Do Before It Goes Live
One Record per Request: How BareProxy Tracing Works
BareProxy Modules: Eight Add-Ons Around a Bare Core
Why BareProxy Routes on the Same Path It Forwards
About
Config Reference
Contact
A Background Job Queue in One SQLite File Covers What Most Apps Deploy Redis For
A Database Where Every Value Remembers Its Source, Confidence and Extraction Time
A Feature Store Small Enough to Embed: Point-in-Time Reads for Local and Edge ML
A Schema Diff That Warns Which Migration Will Lock a 40-Million-Row Table or Lose Data
A Single-File Log Database: Pipe Logs In, Query Them With SQL, Hand the File to Anyone
A Telemetry Database That Forgets on Purpose Keeps Rollups and Anomalies Instead of Raw Rows
A Tiny SQL Engine for JSON Streams Fits Between jq and a Stream Processing Cluster
An Embedded Sketch Database: Billions of Events, Megabytes of Storage, Error Bars on Every Answer
An Embedded Workflow Engine With Five Primitives: Trigger, Condition, Action, Wait and Retry
An Evidence Graph Built From Extracted Claims Keeps Contradictions Attached to Their Sources

Media Partners

  • AltSql.com
  • Technology Conferences
  • API Coding
AltSql Is a Native Hybrid of Key-Value and SQL, From the Sensor to the Gateway
AltSql Core and AltSql DB Are Now Open Source Under the Apache License 2.0
Introduction
AltSql DB 0.3 Adds Secondary Indexes and Statement Savepoints: A Lookup on a Million Rows Ran 198 Times Faster Than a Full Scan
How It Works
AltSql DB 0.2: SQL and Direct Calls Wrote Byte-Identical Files Over 100,000 Random Steps
AltSql DB Keeps a Whole Fleet in One File and Reads by Key 5 Times Faster Than SQLite Through SQL
AltSql Core and Six Prototypes Built Around It
AltSql Ask Puts One SQL Question to 500 Simulated Machines and Brings Back Only the Answers
Machine Monitoring Case Study: Overheat Alarm Handled on the Sensor While Wi-Fi Was Down
Cloudflare Connect 2026, October 19–21, Moscone West, San Francisco
JNUC 2026, September 23–25, Kansas City Convention Center, Kansas City
Startup World Cup Grand Finale 2026, November 4–6, Hilton San Francisco Union Square, San Francisco
FYUZ 2026, November 3–5, The Westin Seattle, Seattle
ONUG AI Networking Summit 2026, October 28–29, Penn District, New York
Networking Field Day 2026, October 6–9, San Jose
Nova Future Summit 2026, September 28–30, Napa
Breakbulk Americas 2026, September 22–23, George R. Brown Convention Center, Houston
Gartner CIO & IT Executive Conference 2026, September 21–23, Sheraton São Paulo WTC Hotel, São Paulo
ITC Vegas 2026, September 29–October 1, Mandalay Bay, Las Vegas
A Tiny ETL Binary Competes With curl, jq and SQLite in a Cron Job, So Build It That Small
An AI Agent Flight Recorder Belongs in One Portable File, the Way HAR Did It for HTTP
Benchmarking MCP Servers and Stateful API Workflows: Latency, Throughput and Tokens per Call
Diagnose a Slow API From One Request: DNS, Connect, TLS, Server Wait and Transfer
Finding API Endpoints, Tables and Flags Nobody Uses Takes Code and Traffic Together
Fuzzing MCP Servers: Generate Bad Arguments From the Tool Schema and Watch What Breaks
Give Any API a History: Poll It, Hash It and Query Old Versions With SQL
Infer Your API's Real Contract From Traffic, Then Diff It Against the Docs
Keeping Third-Party API Responses in SQLite Gets You a Cache, an Offline Mode and a History
LLM Response Caching Pays Off in CI, Evals and Agent Retries, Not in Chat

Copyright © 2026 Technologies.org

Media Partners: Market Analysis · Market Research · Referently · Photography