Declare the service, not a benchmark.
Model, request mix, SLO, hardware inventory, power envelope and failure policy become one versioned study brief.
- Input
- Workload + SLO
- Output
- Study manifest
- Guardrail
- Verify before action
FROM SIMULATION TO AUTONOMY
The future OpenFabric agent treats SimLLM and hardware as one evolving system. It predicts before deployment, learns from measured residuals, then verifies every change before acting.
Vision page / current boundaries remain visible belowINTERACTIVE WORKFLOW
Model, request mix, SLO, hardware inventory, power envelope and failure policy become one versioned study brief.
THE ARCHITECTURAL MENTAL MODEL
OpenFabric does not learn one opaque correction factor. Each boundary has an owner, an interface, a measurement and a residual. The agent can update one layer without hiding a missing mechanism in another.
Arrivals / prefix reuse / SLO / framework scheduler
currentKernels / GPU / HBM / DMA / NCCL / WQE / NIC
nextCollectives / flows / queues / packets / recovery
currentTTFT / TPOT / FCT / counters / traces
currentSearch / calibrate / deploy / observe / retune
futureWorkload queueing, real scheduler records, compute estimates, collective traffic, packet backends and request metrics.
GPU and NIC resource queues, PD KV transfer, capture-driven kernel tables and hardware residual accounting.
An agent proposes, simulates, deploys within policy, observes and rolls back when evidence leaves its support envelope.
THE PRODUCT PROMISE