Back to Directory
NVIDIA A100 Tensor Core GPUs logo

NVIDIA A100 Tensor Core GPUs

2 use cases using this technology · by NVIDIA

2published use cases
2industries
1countries on record
AI Model Development & MLOps
top AI capability

Evidence mix: High 2 · Medium 0 · Low 0 — bands are computed from each record's evidence signals.

Industry
Country

2 use cases

Government & Public SectorAI Model Development & MLOps

Argonne National Laboratory accelerates cosmic discovery with AI-powered RADAR framework

Argonne National Laboratory and university collaborators developed RADAR, a federated, privacy-enhancing framework that lets observatories coordinate gravitational-wave and radio follow-up without moving or exposing proprietary data. Running site-local AI inference on NVIDIA GPUs across the Polaris, Delta and DeltaAI supercomputers, RADAR processed over an hour (4,096 seconds) of Advanced LIGO data in under 4.5 minutes, a 5-10x speedup versus CPU workloads. The RADAR paper has been accepted for publication in The Astrophysical Journal Supplement Series.

Argonne National Laboratory· United StatesNVIDIA A100 Tensor Core GPUs · NVIDIA A40 GPUs · NVIDIA GH200 GPUs
Technology & SoftwareGenerative AILarge Language Models

Accelerating Large Language Model Inference with NVIDIA in the Cloud

Perplexity built pplx-api, an API for developers to integrate open-source LLMs with fast inference, served on Amazon EC2 P4d instances powered by NVIDIA A100 Tensor Core GPUs and accelerated with NVIDIA TensorRT-LLM (with a planned move to Amazon P5 instances with NVIDIA H100 GPUs). pplx-api achieves up to 3.1X lower latency and up to 4.3X lower first-token latency versus other deployment platforms, and switching external inference-serving API references to pplx-api lowered costs 4X, saving $600,000 per year. Using NVIDIA H100 GPUs and FP8 precision on Amazon P5 instances cuts latency in half and boosts throughput by 200 percent versus A100 GPUs in the same configuration. Perplexity also uses AWS's Kubernetes integration to scale elastically beyond hundreds of GPUs.

PerplexityNVIDIA TensorRT-LLM · NVIDIA A100 Tensor Core GPUs · NVIDIA H100 Tensor Core GPUs +2