WORKFLOW
The Model Optimization Platform,
Run by Agents
Tell the agent your target model, deployment environment, and goal metrics. It handles the entire optimization pipeline.
Or skip the agent and drive it yourself with CLI. Same pipeline, every knob exposed.
1
$ np run |
CORE CAPABILITIES
Everything You Need to Optimize and Deploy AI
Real Results on Real Hardware
On device benchmarks:
Before vs. After optimization
Inside the Demo
Experience the full NetsPresso optimization pipeline in a real environment and validate performance improvements firsthand.
✓
No setup required
✓
Hands-on CLI-based optimization workflow
✓
Real benchmarks measured on target hardware
NetsPresso
ok, run the experiment
Optimization complete: smollm2-rpi5
Advanced Quantize — AWQ / W4-A8 dynamic · 4m 6s
Graph Optimize — 8 patterns applied automatically, including QKV fusion · 14m 37s
Graph Quantize — graph-level requantization · 5m 36s
I profiled and evaluated all 4 intermediate models. The best configuration lands at 1,186.7ms / PPL 63.13.
Your files are ready for deployment.
export files

Visualize and Compare Model Changes Across Every Iteration
A high-performance graph visualizer for instant topology comparison. No installation, 100% free.

Model Diff
View topology changes between two models side-by-side, including new or removed nodes
Synchronized Graph Navigation
pan and zoom both graph views simultaneously

Custom Node Coloring
highlight the nodes that matter most to your team









