azure-diagnostics

microsoft/azure-skills · updated Apr 8, 2026

$npx skills add https://github.com/microsoft/azure-skills --skill azure-diagnostics
0 commentsdiscussion
summary

Systematic diagnosis and remediation for Azure production issues using AppLens, Monitor, and resource health.

  • Covers Container Apps, Function Apps, and AKS troubleshooting with service-specific guides for image pulls, cold starts, health probes, invocation failures, and node/pod issues
  • Integrates AppLens (MCP) for AI-powered root cause analysis and Azure Monitor (MCP) for KQL-based log and metric queries
  • Provides a five-step diagnostic flow: identify symptoms, check resource health,
skill.md

Azure Diagnostics

AUTHORITATIVE GUIDANCE — MANDATORY COMPLIANCE

This document is the official source for debugging and troubleshooting Azure production issues. Follow these instructions to diagnose and resolve common Azure service problems systematically.

Triggers

Activate this skill when user wants to:

  • Debug or troubleshoot production issues
  • Diagnose errors in Azure services
  • Analyze application logs or metrics
  • Fix image pull, cold start, or health probe issues
  • Investigate why Azure resources are failing
  • Find root cause of application errors
  • Troubleshoot Azure Function Apps (invocation failures, timeouts, binding errors)
  • Find the App Insights or Log Analytics workspace linked to a Function App
  • Troubleshoot AKS clusters, nodes, pods, ingress, or Kubernetes networking issues

Rules

  1. Start with systematic diagnosis flow
  2. Use AppLens (MCP) for AI-powered diagnostics when available
  3. Check resource health before deep-diving into logs
  4. Select appropriate troubleshooting guide based on service type
  5. Document findings and attempted remediation steps
  6. Route AKS incidents to the dedicated AKS troubleshooting document

Quick Diagnosis Flow

  1. Identify symptoms - What's failing?
  2. Check resource health - Is Azure healthy?
  3. Review logs - What do logs show?
  4. Analyze metrics - Performance patterns?
  5. Investigate recent changes - What changed?

Troubleshooting Guides by Service

Service Common Issues Reference
Container Apps Image pull failures, cold starts, health probes, port mismatches container-apps/
Function Apps App details, invocation failures, timeouts, binding errors, cold starts, missing app settings functions/
AKS Cluster access, nodes, kube-system, scheduling, crash loops, ingress, DNS, upgrades AKS Troubleshooting

Routing

  • Keep Container Apps and Function Apps diagnostics in this parent skill.
  • Route active AKS incidents, AKS-specific intake, evidence gathering, and remediation guidance to AKS Troubleshooting.

Quick Reference

Common Diagnostic Commands

# Check resource health
az resource show --ids RESOURCE_ID

# View activity log
az monitor activity-log list -g RG --max-events 20

# Container Apps logs
az containerapp logs show --name APP -g RG --follow

# Function App logs (query App Insights traces)
az monitor app-insights query --apps APP-INSIGHTS -g RG \
  --analytics-query "traces | where timestamp > ago(1h) | order by timestamp desc | take 50"

AppLens (MCP Tools)

For AI-powered diagnostics, use:

mcp_azure_mcp_applens
  intent: "diagnose issues with <resource-name>"
  command: "diagnose"
  parameters:
    resourceId: "<resource-id>"

Provides:
- Automated issue detection
- Root cause analysis
- Remediation recommendations

Azure Monitor (MCP Tools)

For querying logs and metrics:

mcp_azure_mcp_monitor
  intent: "query logs for <resource-name>"
  command: "logs_query"
  parameters:
    workspaceId: "<workspace-id>"
    query: "<KQL-query>"

See kql-queries.md for common diagnostic queries.


Check Azure Resource Health

Using MCP

mcp_azure_mcp_resourcehealth
  intent: "check health status of <resource-name>"
  command: "get"
  parameters:
    resourceId: "<resource-id>"

Using CLI

# Check specific resource health
az resource show --ids RESOURCE_ID

# Check recent activity
az monitor activity-log list -g RG --max-events 20

References

Discussion

Product Hunt–style comments (not star reviews)
  • No comments yet — start the thread.
general reviews

Ratings

4.849 reviews
  • Chaitanya Patil· Dec 24, 2024

    Keeps context tight: azure-diagnostics is the kind of skill you can hand to a new teammate without a long onboarding doc.

  • Kaira Zhang· Dec 20, 2024

    Keeps context tight: azure-diagnostics is the kind of skill you can hand to a new teammate without a long onboarding doc.

  • Sofia Khanna· Dec 16, 2024

    We added azure-diagnostics from the explainx registry; install was straightforward and the SKILL.md answered most questions upfront.

  • Liam Menon· Dec 8, 2024

    I recommend azure-diagnostics for anyone iterating fast on agent tooling; clear intent and a small, reviewable surface area.

  • Anika Agarwal· Nov 27, 2024

    Solid pick for teams standardizing on skills: azure-diagnostics is focused, and the summary matches what you get after install.

  • Piyush G· Nov 15, 2024

    Registry listing for azure-diagnostics matched our evaluation — installs cleanly and behaves as described in the markdown.

  • Naina Choi· Nov 11, 2024

    Registry listing for azure-diagnostics matched our evaluation — installs cleanly and behaves as described in the markdown.

  • Anika Patel· Oct 18, 2024

    azure-diagnostics has been reliable in day-to-day use. Documentation quality is above average for community skills.

  • Shikha Mishra· Oct 6, 2024

    azure-diagnostics reduced setup friction for our internal harness; good balance of opinion and flexibility.

  • Anika Ndlovu· Oct 2, 2024

    azure-diagnostics reduced setup friction for our internal harness; good balance of opinion and flexibility.

showing 1-10 of 49

1 / 5