A Model Zoo branded as yours. Optimized for your silicon.
Live partner portals
Edge AI Foundation — community model portal
Ceva — NeuPro-Nano model platform
Infineon — DeepCraft AI Hub
What's in a partner portal
Vendor-branded UI
Matches your design system. Hosted on a modelnova.ai subdomain (yourname.modelnova.ai) or on your own domain.
Developer registration & access controls
Account creation, role-based access, and entitlement tiers if needed.
AI playground workflows
In-browser model exploration before download — try-before-deploy.
Developer onboarding flow
Guided first-run from sign-up to first inference on the target.
Pre-optimized models for your silicon
Curated from the ModelNova catalog, ported and validated on your NPU.
Quantized inference pipelines
INT8/INT4 pipelines tuned to the compute and memory profile of the target.
Hardware-specific benchmarks
Per-model latency, throughput, and memory footprint on the target. Published with each model.
SDK integration
Models packaged for your toolchain so developers stay in your workflow.
Reference applications
End-to-end working examples per model — clone, build, run.
Deployment documentation
Integration guides, troubleshooting, and silicon-specific notes per model.
The engineering underneath
How models actually get to your silicon. The optimization stack that runs under every portal we operate.
TFLite Micro
- Operator selection and graph rewrites for resource-constrained targets running LiteRT.
ExecuTorch
- Native PyTorch export and runtime integration for production embedded deployment.
NPU-targeted compilation
- Graph compilation, operator mapping, and runtime integration tuned to the target NPU's architecture and toolchain.
NPU-aware quantization
- INT8 and INT4 quantization tuned to the target NPU's compute and memory profile.
Memory-aware inference
- Layer fusion and operator scheduling to fit within SRAM and TCM budgets.
Hardware benchmarking
- On-device latency, throughput, and power measurement under realistic workloads.
Embedded firmware integration
- Inference pipelines integrated with the target's RTOS, drivers, and runtime.
Recognition
Partnership agreement
Define silicon targets, model scope, branding, and developer access model. Optional MOU available with ModelNova
Silicon targets Model scope Branding Access model
Model porting & validation
Catalog, compile, optimize, and benchmark models on your hardware. Validate accuracy, latency, and power metrics against agreed targets.
Compiled models Optimization Benchmarking Validation
SDK integration
Package models for your toolchain and integrate with your developer workflow. Publish sample apps, API wrappers, and build-system hooks.
SDK packaging API wrappers Sample apps Build hooks
Portal build
Implement branding, access controls, registration flow, and developer onboarding. Build documentation, quickstarts, and support channels.
Branding Optimization Access controls Documentation Quickstarts
Launch
Deploy to your developer community. Execute the joint announcement, co-marketing campaign, and partner-facing launch event.
Deployment Announcement Co-marketing Launch event
Ongoing support
Quarterly model additions, performance updates, and roadmap reviews. Continuous feedback loop with your developer community to prioritize improvements.
Quarterly updates Model additions Perf updates Roadmap reviews