Your chosen model
Optimise your chosen open-weight or internal model, subject to evaluation and access to the necessary artefacts.
Enterprise deployments
Minima optimises the models your enterprise already uses and delivers the optimised weights and mnma runtime for deployment in your environment.
Core offer
Optimise your chosen open-weight or internal model, subject to evaluation and access to the necessary artefacts.
Deliver the optimised model and compatible runtime for the agreed hardware.
Deploy within your infrastructure, including the existing supported on-premises and VPC paths.
Benchmark quality, throughput, memory and GPU requirements against your baseline and acceptance criteria.
What you receive
Featured deployment result
Four NVIDIA B200 nodes → one NVIDIA B200 node
Four NVIDIA B200 nodes for the baseline become one NVIDIA B200 node with Minima.
Minima internal benchmark and featured deployment result. Achievable savings for a customer's model, workload and hardware are established through evaluation.
Workflow
Commercial model
We scope the model, hardware and workload with your team, validate the improvement, and agree the production licence.
Enterprise FAQ
Deployment runs within the customer's infrastructure through the agreed supported on-premises or VPC path.
Minima supplies the agreed optimised model artefacts, the mnma serving runtime for the validated configuration, and deployment instructions and integration requirements.
Quality is benchmarked against the customer's baseline and agreed acceptance criteria alongside throughput, memory and GPU requirements.
Minima scopes the model, hardware and workload with the customer, validates the improvement, and prices the annual software licence on agreed GPU-hour usage.
Enterprise benchmark
Minima will scope your model, hardware and workload with your team and validate the improvement against the agreed baseline and acceptance criteria.
Request received
Your Enterprise benchmark request has been accepted and sent to the Minima team.