Industry background

Tractian builds hundreds of AI models faster, at lower cost using OCI

Company taps Oracle Cloud Infrastructure to ramp up AI model training and inference workloads as it fine-tunes its industrial monitoring systems.

United States | High Technology

Our AI agents run continuously across hundreds of thousands of machines, turning raw sensor data into real-time maintenance decisions. OCI and our NVIDIA GPU infrastructure is what makes that scale possible, and what lets us keep pushing our physical AI models further without disrupting the operations our customers rely on.
JP VoltaniCTO, Tractian

Headquartered in Atlanta, Tractian provides Physical AI solutions for asset-heavy industries across North America, Latin America, and Europe. As demand for industrial AI accelerates, the company is making one of the largest AI infrastructure commitments by a startup to advance multimodal asset health monitoring, large-scale model training, and real-time industrial intelligence. To support this next phase of growth, Tractian selected Oracle Cloud Infrastructure (OCI), leveraging NVIDIA GPU-powered infrastructure to train and deploy AI models at scale. With OCI, Tractian can process massive volumes of sensor, operational, and industrial data in real time, delivering actionable insights that help customers reduce unplanned downtime, improve asset reliability, and optimize maintenance operations across critical industrial environments. Today, Tractian is already delivering impact at scale, supporting more than 2,000 plants, monitoring over 200,000 assets, avoiding more than 172,000 equipment failures, and preventing over 97,000 hours of machine downtime in the last year alone.

Why Tractian chose Oracle

Tractian’s sophisticated AI models, which require terabytes of sensor data per machine, can take months to train. To accelerate that training, the company needed a high performance, bare metal cloud infrastructure with high-throughput storage. Tractian was also looking for low latency and predictable costs at scale to support its inferencing needs.

Tractian chose OCI because of Oracle’s dependable access to scarce NVIDIA GPUs and OCI’s mature Kubernetes engine, block storage, networking, and PostgreSQL integration. It also values Oracle’s responsive sales and engineering support as well as its continuous improvement of OCI. “On OCI, we saw that we could spin up and be successful fast,” says JP Voltani, Tractian’s VP of engineering. “Oracle is now our go-to for model training.”

Results

On OCI bare metal with NVIDIA H100 and GB300 GPUs, Tractian cut training costs by 20% and reduced training times by an average 35%, helping it deliver improved AI models to customers.

In production, inference costs were reduced by 15% and latency by 50%. This has helped Tractian’s manufacturer customers prioritize maintenance actions for vital assets.

As a result of these improvements, Tractian achieved a predictable scale for both training and inference. The company estimates that its systems are able to prevent a customer machine failure roughly every 15 minutes across its installed base, improving uptime and safety.

About the customer

Tractian builds hardware and software and develops its own AI models to help its manufacturer customers predict failures in their industrial equipment.

Learn more about Tractian