> For the complete documentation index, see [llms.txt](https://docs.distribute.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.distribute.ai/distribute-for-enterprise/enterprise-inference-api/async-api/chat.md).

# Chat

The Asynchronous Chat Completion API allows you to submit chat-based LLM tasks that are processed in the background and retrieved later. Ideal for high-latency or large-scale workloads, this custom implementation lets you queue chat requests via a `POST` endpoint and retrieve results using a unique task ID. It supports robust job tracking, delayed responses, and retry-safe workflows—making it well-suited for batch processing, long-running prompts, and serverless environments where real-time responses aren't required.

{% content-ref url="/pages/FqXZBVvFlsSdrEgmZy9Y" %}
[Chat Create](/distribute-for-enterprise/enterprise-inference-api/async-api/chat/chat-create.md)
{% endcontent-ref %}

{% content-ref url="/pages/fLEF8XhUOyUVSt1K4Bnw" %}
[Chat Result](/distribute-for-enterprise/enterprise-inference-api/async-api/chat/chat-result.md)
{% endcontent-ref %}
