@mlbottleneck/engine
MITPhysics-based LLM inference planner: decode/prefill tokens per second, memory fit, multi-GPU strategy, and speculative-decoding gains for any model on any hardware, calibrated on community benchmarks.
100
Security score
0 known advisories in v0.6.0
Weekly downloads
—
Unpacked size
1.0 MB
Dependencies
0
Last publish
27 days ago
Security advisories
No known vulnerabilities affect v0.6.0.