Getting Started
All CloudWatcher app pages
A route-by-route guide to what each CloudWatcher page does and how it fits into the product.
This page is useful when you are new to the app or auditing coverage. It explains the purpose of every major page in plain language.
Pages and responsibilities
- Landing page: explains CloudWatcher, cloud cost optimization, multicloud monitoring, AI observability, infrastructure simulations, integrations, savings, and product entry points.
- Signup and verification: account creation, signup verification, login handoff, and protected access after verification.
- Login and 2FA: email/password login, two-factor verification, resend flow, recovery, forgot password, reset password, and first-login password reset.
- Dashboard: consolidated provider overview for spend, health, resources, alerts, and operational status.
- Service dashboards: provider-aware detail views for compute, storage, database, network, security, cost, AI insights, health, and alerts.
- Watchdog: fleet-level operational status, active resources, security findings, optimization signals, and investigation starting points.
- Recommendations: cost, reliability, security, performance, cloud, and AI recommendations with action or simulation follow-up where supported.
- Cost Savings: optimization opportunities, estimated savings, validation, completed savings records, and realized savings tracking.
- Actions: action planning, previews, approvals, execution status, rollback attempts, audit trail, savings verification, and registry information.
- AI Observability Overview: request volume, tokens, cost, latency, errors, providers, models, environments, and recent activity.
- AI Traces: trace list, trace grouping, request detail, spans, metadata, tokens, latency, cost, and failure inspection.
- AI Cost: daily cost trends, provider split, model split, token cost, and cost-drivers.
- AI Models: model usage, token volume, latency, errors, cost, and provider grouping.
- AI Errors: grouped rate limits, timeouts, client errors, server errors, failed requests, and root-cause signals.
- AI Alerts: cost, token, latency, error, anomaly, and budget alert review.
- AI Recommendations: routing recommendations, prompt optimization, model choice, cost reduction, and reliability suggestions.
- AI Prompts: repeated prompt patterns, oversized prompts, prompt efficiency, and prompt-level optimization signals.
- AI Playground: prompt testing with provider, model, temperature, output, latency, token, cost, and error comparison.
- AI Evaluations: model evaluation, custom judge settings, score review, pass/fail analysis, and failed-example inspection.
- Settings > AI Keys: provider key status, OpenAI usage, OpenAI logs, Anthropic usage, Gemini usage, pricing lookup, key save, and key delete.
- Settings > AI Observability: ingest keys, trace/event scopes, SDK setup, telemetry endpoint, and privacy defaults.
- Settings > AWS: IAM role setup, external ID, credential status, disconnect, and AWS validation.
- Settings > Azure: tenant, subscription, service-principal connection, onboarding template, validation, and disconnect.
- Settings > GCP: project, service-account connection, billing export setup, validation, callback behavior, and disconnect.
- Settings > GitHub: OAuth status, connect, disconnect, repository list, branch list, and deployment source selection.
- Settings > Reports: report preferences, cadence, recipients, sections, and test delivery.
- Simulation Builder: visual canvas, service nodes, config panel, cost box, Terraform preview, deployment status, autosave, manual save, and provider selection.
- Simulation History: saved simulation list, detail, update, delete, destroy, PEM download, deployment records, and cleanup status.
- Live Infrastructure: provider inventory sync, service-group canvases, resource detail, status, metrics, code view for Lambda where supported, and live action safety checks.
- Resize Migration: scope, source discovery, target sizes, job creation, task transition, classification confirmation, access configuration, resume, explanation, report, and delete.
- VPS Logs: agent registration, agent config, secure ingest, recent log review, summary, clear recent logs, alarm rules, alert policy, and mail test.
- Profile: user profile, preferences, integration awareness, security controls, and account settings.
- SaaS Admin: system stats, tenant list, compliance alerts, support visibility, and administrator-only management.
- Global AI Agent: product help, docs help, and grounded operational assistant entry point.
Suggested daily workflow
- Start on Dashboard to check the selected provider, spend, resource health, and active alerts.
- Open Watchdog to see broad operational risk and fleet status.
- Open service dashboards for the resource area that needs investigation.
- Open Billing Metrics, Recommendations, and Cost Savings for cost changes.
- Open AI Observability pages if AI spend, model latency, or errors changed.
- Use Actions, Simulation, Live Infrastructure, or Resize Migration only after reading the relevant detail page.
Need more help?
Keep moving with the right next guide
If your team still runs into setup issues, empty dashboards, billing delays, callback failures, or alerting problems, continue with troubleshooting before re-running the entire onboarding flow.