Sail Research screenshot
Sail Research logo

Sail Research

The most efficient inference for long-horizon agents

Inference built for agents that run for hours, pricing latency tolerance as a dial and pairing it with persistent VM sandboxes for long-horizon workloads.

Pricing

Usage based, no flat monthly plan

No free tier

Usage-based pricing per million tokens with no seat or subscription fee, and $5 of free credits monthly - rates vary by model and completion window, for example GLM-5.2 runs $0.80 input and $3.00 output per million tokens on asap, dropping to $0.40 and $1.80 on flex - cached input is billed separately - enterprise volume pricing is custom

From the vendor pricing page, 2026-08-28

About Sail Research

Tags

autonomous-agentagent-frameworkinfrastructureinferencellm-api