Company intelligence
FriendliAI
The Frontier AI Inference Cloud. Deploy frontier open models with unmatched efficiency—maximizing tokens and margins.
About FriendliAI
FriendliAI is the Frontier Inference Cloud for Agents, delivering high throughput, low latency, and reliability at scale for agentic workloads. Through vertically optimized inference infrastructure, it achieves 2–5× faster output token speed and 50–90% lower inference cost, backed by a 99.99% uptime SLA built for high-volume agent traffic. The platform powers low-latency streaming for real-time agents, reliable long-context inference, and robust tool calling.
Verified activity
Signals from FriendliAI
6 published signals
Presence & Recognition
FriendliAI is attending The AI Conference in San Francisco from September 29 to October 1.
Reported by FriendliAI
People
FriendliAI is growing its Discord community for AI agent builders, developers, engineers, and researchers.
Reported by FriendliAI
Products & Services
FriendliAI delivers 301 tokens per second throughput for the Z.ai GLM-5.3 model, nearly 45% higher than the runner-up.
Reported by FriendliAI
Research & Knowledge
FriendliAI published an August 28 snapshot for GLM-5.3 featuring a 1M-token context window with 1M-token max output and the highest P50 output throughput among OpenRouter providers.
Reported by FriendliAI
Products & Services
FriendliAI achieved #1 on OpenRouter with GLM-5.3 runs fastest.
Reported by FriendliAI
Products & Services
FriendliAI shipped new models on Model APIs, FDE, Reserve GPUs, FDE Quickstart tutorial video, and GLM-5.3 vs. Kimi K3 blog in August.
Reported by FriendliAI