An Azure service that provides an event-driven serverless compute platform.
@Daniel We noticed that this problem was observed during the scale-out process of your Function App. When a new instance was added, it failed to start as expected, and shortly after, the other instances also began to fail during their operation. Following several attempts to start and allocate jobs, these instances gradually shut down.
Unfortunately, there are no logs that definitively confirm the root cause. What we can conclude is that the problem stems from the platform in some form. As a temporary work around until the team resolves this issue the suggestion is to implement the Auto-Heal option as below
To configure Auto-Heal for detecting and recovering from such random Function App offline failures, you can set up rules based on specific conditions like high memory usage, frequent crashes, or lack of activity. Here’s a basic example of how you can configure Auto-Heal:
- Navigate to your App Service in the Azure portal.
- Select “Diagnose and solve problems” from the left-hand menu.
- Under “Diagnostic Tools,” select “Auto-Heal.”
- Add a rule to monitor specific conditions (500 HTTP status codes in this case).
- Set the action to “Recycle” the worker process or “Restart” the app service.
For more detailed steps, you can refer to the Auto-Heal documentation on Microsoft Learn.
Announcing the New Auto Healing Experience in App Service Diagnostics - Azure App Service
I will update this thread once the underlying issue has been fixed. I apologize for the inconvenience this issue has caused you and appreciate your patience in trying to resolve the issue.