Small business
A few cards, one team's apps.
At 3 GPUs, pooling reclaims about 2 idle machines.
Your GPUs are booked around the clock and working a fraction of it. TorqHUB streams workloads onto the machines that can actually take them.
Reserved 24/7. Working, say, 18% of the time. Billed for 100% of it.
Bar chart: ten GPU racks. Two are actively computing, shown in blue. Eight sit idle, shown in grey. A coral line across all ten represents the full cost you pay. Illustrative figures — real utilisation varies.
A job holds the card for eight hours and computes for forty minutes. The card is busy. Nothing is happening.
How much of each reserved GPU is actually doing work. Drag to feel the gap.
Duty cycle slider. Current value 18 percent of reserved capacity. Range 5 to 60 percent. Use the left and right arrow keys to adjust by one percent.
Move the sliders. This is the money that turns into heat and nothing else.
Reserved GPUs burn money every hour they sit idle. Change the fleet — then drag the occupancy slider to watch the gap close.
Projected idle spend: $9,578 per month, $114,931 per year, at 18 percent duty cycle. less than one engineer / yr. Illustrative figures.
Different jobs, different shapes, different machines. Flip the switch and watch the gaps disappear.
Small shop or full datacenter, a typical fleet runs near 18% utilised — reserved 24/7, working a sliver. Here's the monthly burn without TorqHUB, and what pooling reclaims with it.
A few cards, one team's apps.
At 3 GPUs, pooling reclaims about 2 idle machines.
A room of GPUs, a dozen products.
At 8 GPUs, pooling reclaims about 6 idle machines.
A full rack at datacenter scale.
At 30 GPUs, pooling reclaims about 23 idle machines.
Illustrative — built from the same consolidation model as the live scheduler above, not a quote. Utilisation and per-GPU rates are tunable in config.
Monthly economics at three business scales, all at a typical 18% utilisation. Small business, 3 GPUs: about $3,290 per month wasted on idle metal without TorqHUB, about $2,680 per month saved by pooling — roughly 2 idle machines reclaimed. Mid business, 8 GPUs: about $9,280 per month wasted on idle metal without TorqHUB, about $8,490 per month saved by pooling — roughly 6 idle machines reclaimed. Big business, 30 GPUs: about $32,920 per month wasted on idle metal without TorqHUB, about $30,780 per month saved by pooling — roughly 23 idle machines reclaimed. Illustrative demo figures from the same model as the live scheduler, not a quote.
You write the logic. We route the compute.
One package, one token. No cluster to provision, no broker to configure.
Send a job with its type and payload. TorqHUB routes it to the GPU that fits.
Batch to hundreds, poll or await results. The router fills the gaps for you.
import { TorqClient } from '@torq/sdk'
const torq = new TorqClient({
token: process.env.TORQ_TOKEN,
})
// submit a job — the router picks the GPU that fits
const job = await torq.jobs.submit('llm-70b', {
model: 'meta-llama/Llama-3-70B',
prompt: 'Summarize this transcript…',
})
// wait for the result (or poll, or batch to hundreds)
const result = await torq.jobs.wait(job.id) Capacity where the work is. The router sees every node and every gap.
A single card — local, on your desk, fully private. TorqHUB schedules the queue and loads/unloads models automatically, so even one GPU stays busy instead of idling between tasks.
Even one card — local, on-prem, fully private. TorqHUB schedules the queue and loads/unloads models automatically, so a single GPU is reused evenly instead of idling between tasks.
Pool the GPUs you already own into a closed, private mesh. TorqHUB redistributes compute across the fleet inside your walls — your jobs, your machines, nothing leaves the network.
Or tap a crowdsourced pool: rent capacity on demand, or put your idle cards on the mesh and earn. The router streams every job to wherever the right GPU is free.
Network mesh: a central TorqHUB router connected to GPU worker nodes across multiple regions. Illustrative scale figures. TorqHUB runs two ways: as a public crowdsourced compute mesh, or deployed privately inside your own corporate systems — contact us for deployment details.
Tell us what you need. We'll come back with access, capacity or a deck.