STATUS // DCGM · INVENTORY · PLACEMENT · APICOMPUTE ROUTER · AGENT ECONOMY2026
Compute Routing Protocol // Agent Economy

Super Inu

For agents that need compute
Super Inu #001 · 1 of 1One mascot. One router.

Super Inu finds, scores and routes AI workloads to the best GPUs available, and explains every call.

Route a jobRead the docs
Investment Case // Thesis

Own the decision, not the metal.

01 // Why now

Agents are about to become buyers

Software that launches its own training, inference and batch jobs makes the human choosing an instance type the bottleneck. Whoever serves that buyer first sets the standard.

02 // The wedge

Explainable placement, shipping today

Open GPU telemetry in, scored and justified recommendations out. v0.1 is built, tested and runs on mock data. Real execution is the next milestone.

03 // The moat

Every decision makes the next one better

Routing data compounds: which providers delivered, which GPUs held up, what actually ran fast. That history is what pricing, reputation and trust get built on.

04 // The prize

A market that routes itself

If compute becomes a programmable commodity, the router that sits between agents and every provider becomes the default interface to it.

05 // Neutral

Multi-vendor by design

One provider interface covers local clusters, clouds and marketplaces. No single supplier owns the customer.

06 // Discipline

Routing first, payments last

No token and no speculation in the core. Settlement arrives only after routing and verification are proven.

Today

  1. Human picks a cloud
  2. Human picks a GPU
  3. Human provisions infrastructure
  4. Workload runs
  5. Human pays

With Super Inu

  1. Agent states the job and its limits
  2. Super Inu discovers and scores compute
  3. It explains why it chose a target
  4. Work is dispatched and verified
  5. Compute is released, payment follows
Protocol Suite // Compute Router

Advanced compute routing modules

01 // Module

Telemetry Reader

Pulls GPU metrics from NVIDIA DCGM through dcgm-exporter and normalizes them into one GPU model.

GPU NodeDCGMdcgm-exporterNormalizer
02 // Module

Inventory

One live list of every GPU across providers: model, VRAM, utilization, temperature, power and health.

ProvidersInventoryReserved GPUs
03 // Module

Placement Engine

A deterministic 100-point score behind hard filters. Same inputs, same answer, no black box.

Hard FiltersFit 20Compute 25Memory 20Health 15Model 10Thermal 10
04 // Module

Explainer

Every recommendation says why: free VRAM, utilization, temperature, XID status, hardware match and warnings.

ReasonsWarningsRejections
05 // Module

Provider Interface

One contract (reserve, launch, status, release) so local clusters, clouds and marketplaces plug in alike.

LocalCloudMarketplaceComputeProvider
06 // Module

Execution Planner

Turns a recommendation into ordered steps. Plans only in v0.1; real execution is next.

ValidateSelectReserveLaunchMonitorVerifyRelease
Live product model // Streaming (simulated)

Watch the router think.

Agents submit jobs in real time. Super Inu scores the fleet, routes each job and streams the flow into the GPUs while telemetry drifts every second. The stream is simulated in your browser; the scoring matches the v0.1 code.

idlebusy· green ring = running a job · white flash = just routed · dashed red = unhealthy · click a GPU to inspect or break it
    Select a GPU on the map to inspect it.
    GPU tech

    The signals behind every decision.

    Super Inu reads open DCGM telemetry for H100, H200, A100, L40S, L4 and RTX-class GPUs wherever the metrics are exposed. Five signals drive v0.1. Four more unlock smarter routing next.

    GPU utilization

    How busy the compute engines are.

    DCGM_FI_DEV_GPU_UTIL

    v0.1 reads this

    Framebuffer memory

    Free VRAM decides whether a job fits.

    DCGM_FI_DEV_FB_USED / FB_FREE

    v0.1 reads this

    Temperature

    Thermal headroom and health.

    DCGM_FI_DEV_GPU_TEMP

    v0.1 reads this

    Power draw

    Enforces per-GPU power limits.

    DCGM_FI_DEV_POWER_USAGE

    v0.1 reads this

    XID errors

    Driver-reported faults mark a GPU degraded.

    DCGM_FI_DEV_XID_ERRORS

    v0.1 reads this

    Tensor-core activity

    Real AI-math load, not just 'busy'.

    DCGM_FI_PROF_PIPE_TENSOR_ACTIVE

    planned

    NVLink bandwidth

    Multi-GPU jobs need fast links.

    DCGM_FI_PROF_NVLINK_TX_BYTES

    planned

    ECC errors

    Memory reliability for long runs.

    DCGM_FI_DEV_ECC_DBE_VOL_TOTAL

    planned

    Clocks and throttling

    Spot GPUs slowing themselves down.

    DCGM_FI_DEV_SM_CLOCK

    planned

    Field names come from NVIDIA DCGM; confirm them against the docs and NVIDIA's repos before relying on them.

    Supported Providers

    Route across clusters, clouds and marketplaces

    Provider // 01

    Local cluster

    GPUs visible through a dcgm-exporter endpoint.

    Live in v0.1
    Provider // 02

    Cloud A

    First public-cloud provider behind the same interface.

    Planned · v0.3
    Provider // 03

    Cloud B

    A second cloud, so routing has real choices.

    Planned · v0.3
    Provider // 04

    GPU marketplaces

    Open markets of rentable compute, scored like any other source.

    Planned · v0.3 to v0.6
    Flywheel // Value Engine

    The routing flywheel

    Agent Job→Discover→Score→Route→Verify→Reputation Data→Better Routing ↻

    There is no token in the core. Every run adds data about which providers delivered, so the next decision is better.

    Leverage it for your next

    Workload

    Route a job
    Coming Soon // Roadmap

    The path from router to compute economy.

    V0.1 · GPU intelligence
    Telemetry, inventory, explainable placement, API. Built.
    V0.2 · Workload execution
    Real reserve, launch and release on Kubernetes with the NVIDIA GPU Operator.
    V0.3 · Multi-provider compute
    Cloud and marketplace providers behind one interface.
    V0.4 · Pricing and optimization
    Cheapest, fastest or safest objectives; budgets enforced.
    V0.5 · Provider reputation
    Reliability history and SLA tracking.
    V0.6 · Autonomous marketplace
    Agents negotiate directly with providers.
    V1.0 · Agent-native compute economy
    Verified execution and automatic settlement.