Quick guides

Quick guide: Start an operational investigation

Use a short checklist to follow a problem from the visible symptom to operations, logs, and audit records.

Before you use a printed copy. Check the documentation version and verification date at the top of the page. Use the online guide if this copy is older.

When to use this guide

Use it when work is waiting, failed, missing, or giving an unexpected result.

Checklist

  1. Record what the person sees and the time it happened.
  2. Record the tenant, product area, and safe item ID.
  3. Keep the request ID when one is shown.
  4. Check System operations or Tenant operations.
  5. Find the related service, queue, or job.
  6. Open logs for the same time and service.
  7. Search by request ID when one is available.
  8. Open the matching audit record when the action changes important data.
  9. Write one evidence based explanation.
  10. Apply only the safe action that matches the cause.
  11. Repeat the original check once.

Stop and get help when

  • Data versions do not match.
  • A database or file store is unavailable.
  • Several tenants or service groups are affected.
  • The next action could lose or expose data.

Record

Record times, service group, state, queue or job ID, request ID, error code, action taken, and result. Remove personal information and secrets.

Full guide

Read Investigate a problem and Use logs and audit records.

Pūnaha Docs

Search the guides

Enter at least two characters.

    Product screen

    View the full screenshot