The case that AI will need far fewer data centre GPUs
As workloads shift from training to inference, smaller models and purpose-built silicon are loosening the hyperscale GPU cluster's grip on enterprise AI.
6 articles tagged Inference
As workloads shift from training to inference, smaller models and purpose-built silicon are loosening the hyperscale GPU cluster's grip on enterprise AI.
London-based Fractile has raised £165m ($220m) in Series B funding to build inference chips, hitting a $1bn post-money valuation as UK silicon start-ups draw fresh capital.
HPE survey reveals 22% of organizations have operationalized AI, up from 15% last year. IT leaders must focus on trusted inference at scale through data-centric execution and governance.