Our fleet, at the core.
Murakumo runs primarily on our distributed fleet. Available models, capacity, and response times vary by endpoint and workload. Choose the service conditions that fit your application.
AWAI NETWORK / DISTRIBUTED AI
Distributed AI inference, designed around affordability rather than instant responses. Murakumo puts our own fleet first, for work that can tolerate waiting.
Batch processing, experimentation, and background tasks are the intended fit. Latency varies with model size and fleet load. If you need a firm response-time or capacity commitment, agree it with us before integration.
Murakumo runs primarily on our distributed fleet. Available models, capacity, and response times vary by endpoint and workload. Choose the service conditions that fit your application.
Check the selected endpoint or checkout for its current price and limits. Partner billing, capacity, and service commitments are agreed separately. “Cheap” is our design direction, not a promise to beat every provider for every workload.
Requests pass through our gateway and the selected inference backend. Operational and billing records are distinct from prompt and output content. We do not claim service-wide zero data retention. Contact us for retention and deletion requirements before sending data that needs specific contractual controls.
This is a registered-office and mailing address, not a data-centre location.
For partnerships, billing, support, or data requests, contact our team. Include the service and request ID where available; do not email passwords, payment-card details, or private prompts.