Currently, all AI upstream services are simulated using this fake server method.
Author: shreemaan-abhishekCreated May 8, 2026Updated Aug 11, 2026
Labelstech debt
Currently, all AI upstream services are simulated using this fake server method. I'm worried that the difference between the fake server and the real LLM request here is too big.
Should we introduce a container specifically for LLM Fake Server?
Originally posted by @membphis in https://github.com/apache/apisix/pull/13307#discussion_r3201089261
Source: apache/apisix