Price: 0
Number of applications: 10
05.08.26 (inclusive)
contractual
MVP
ICT tasks
Robotics
Neurotechnology and artificial Intelligence
Software/ IS,
Configuration and settings of the language model
When using LLM language models in customer service dialog scenarios in a telephone channel, there is a high delay in generating the first response in each specific user conversation session. Given the criticality to delays in real-time voice interaction, this has a negative impact on the overall perception of the person (the interlocutor) of the dialogue with the AI assistant. The delay is related to the formation of the context and caching of the request in the language model. It is possible to insert "stubs" in the time intervals while the LLM response is being generated, but this is not a technological solution to the problem.
Configuration and settings of the language model to reduce the delay in generating the first response.
Torchik V.V.
Purpose and description of task (project)
Search for ways to reduce delays at the first step of the dialogue when using language models. Language models: ChatGPT cloud services, QWEN