Operations
Observability
See the health of every Project in your workspace on one screen, and spot slowdowns and errors before your users do.
Metrics & health
Open Metrics & health from the console menu. It covers every Project and hosted application in the current workspace.
| Card | Shows |
|---|---|
Projects | How many are registered in the workspace |
Running | How many are up, and how many answer health checks |
Needs attention | Projects that failed or can't be reached |
Compute age | Total running time of the workspace's Projects |
Fleet health refreshes every 30 seconds, listing each Project's stack, region, state, health, and last update.
What to watch
- Error rate: a jump right after a change usually points to that change.
- Response time: slow list pages usually need an index or a smaller perPage.
- Function failures: open the function's logs for the exact error.
- Traffic: a sudden spike can be a launch, or abuse. Rate limits protect the Project either way.
Handle an incident
- Find when things changed in the Activity chart on the Logs page.
- Filter logs to Error for that time.
- Undo or fix the cause.
- Confirm the Project shows healthy again before closing the incident.
- Write down what happened and what will stop it happening again.
Found something wrong on this page?
Fix it yourself. The link below opens this file in GitHub's editor and forks the repository for you if you need one, and your change becomes a pull request without leaving the browser.