DownForAI

Vellum AI status: API, auth, latency & outage reports

Vellum AI is operational right now. DownForAI checks Vellum AI every ~75 minutes across 1 monitored surface, with 0 community reports in the last 24 hours.

Operational
Last probe 24 min ago·1 surface
Probe-monitoredLow confidence
📚 Docs →
100.0%
24h Uptime
183ms
p50 Network Latency
358ms
p95 Network Latency
0
Incidents (30d)

DownForAI monitors Vellum AI via a single endpoint — Vellum AI Web. Current status: operational — all surfaces are responding normally. Measured uptime over the last 24 hours: 100.0%. Median response time (24 h): 183 ms; p95: 358 ms.

No incidents have been logged for Vellum AI, and no community reports were received in the last 24 hours. Probes run continuously across the configured surfaces.

Data confidence: LowLast checked 24 min ago · 1 monitored surface · 1 with 24 h latency data

Current Status

Operational
Everything works normally
Checked by our probes · 401ms
Verified 24m ago
No major incident reported
No reports in the last 24h

Having issues?

Help other users by reporting if Vellum AI is not working for you.

0reports in the last 24 hours

Surface Health

Vellum AI WebOperational
HTTP 200p50 183ms24m ago

Uptime — last 24h

100.0%

Network latency — last 20 probes

Loading…

What DownForAI can verify about Vellum AI

No confirmed provider-side issue detected
DownForAI currently sees Vellum AI as operational based on available monitoring signals.
Signal source: public surface check · Monitoring confidence: Low
Still having trouble?

The issue may be account-specific, regional, network-related, or related to authentication, billing, quota, or rate limits.

Incident history (30d)

No incidents recorded in the past 30 days.

Reported symptoms

No user reports for Vellum AI in the last 24 hours.

Known error signatures

Common failure patterns and how to diagnose them

Provider details

Vellum is a platform for building, evaluating and deploying LLM workflows and prompts, with deployments served from Vellum's API. Applications that call Vellum-deployed prompts at runtime depend on its API availability.

What we monitor
Vellum APIDeployed prompts and workflows
Vellum appBuilder and evals
Model providersBehind deployments
Ecosystem dependencies
Model providers

Fallback alternatives

What to use if this service is down

Vellum is down
Langfuse or PromptLayer (monitored on DownForAI) manage prompts; call providers directly with cached prompts
Moderate effort

How we monitor

DownForAI checks monitored AI service surfaces on a rotating schedule, roughly every 75 minutes per surface, from a single centralized probe infrastructure. For services with an official machine-readable status page, we use that official status signal when available. For other services, we perform a basic public-surface check. Some services block datacenter probes; in those cases we mark them as “Monitoring limited” instead of treating a failed probe as a confirmed outage. Status is classified as Operational, Degraded, or Outage. Network latency (check response time) measures how long the monitored endpoint took to respond to our probe — it is not model inference speed, time-to-first-token, or tokens-per-second performance. We are independent of all providers listed and receive no compensation to report any particular status.

Vellum AI statusEmbed this status badgeREADME · docs ▾
Markdown
[![Vellum AI status](https://downforai.com/api/badge/vellum-ai.svg)](https://downforai.com/vellum-ai)
HTML
<a href="https://downforai.com/vellum-ai"><img src="https://downforai.com/api/badge/vellum-ai.svg" alt="Vellum AI status" /></a>

Community Discussion

0/1000
No comments yet. Be the first to share your experience!

Still having issues with Vellum AI?

Let the community know and help others experiencing the same problem.