Sells inference on open-weight models as a hosted, OpenAI-compatible API at three latency tiers — realtime, async, and 24-hour batch — priced per token, and…
Sells inference on open-weight models as a hosted, OpenAI-compatible API at three latency tiers — realtime, async, and 24-hour batch — priced per token, and separately ships an open-source gateway (Control Layer) that teams run themselves.
| Industry | AI & Machine Learning |
| Website | Visit Doubleword |
Where Doubleword sits against the other names we cover on this beat. Each line is that company’s verdict, not a summary of it.