Connecting
Overview
…·deepseek-v4.1-flash·OpenAI-compatibleRequests (24h)
0 completed
Success rate
0 failed · 0 cancelled
Throughput
avg output tokens per second
P50 generation
0 cooldowns
Live queue
Strict FIFO — exactly one upstream request is in flight at a time.
0 active · 0 queued
Queue is empty
New requests appear here the moment they arrive.
Model
Served through the Anthropic Messages transport, which preserves image input and cache accounting.
Active
Context
0
payload dependent
Max output
0
tokens per request
ReasoningTool callingVisionStreaming
Cached prompt tokens are reported per request. OpenAI requests and responses are translated to the provider dialect at the gateway.
Requests over time
Completed, failed and cancelled requests per bucket.
Input 0Output 0Cached —
Activity
Gateway lifecycle events.
No events yet
Traffic events stream in here.