A computer running AI can work on several people’s requests together. That can let it produce more answers with the same equipment.
But waiting to form a group, or working through a bigger one, can delay one person’s answer.
Engineers have to balance how many answers the computer produces with how soon each person receives one. That choice can make services using the same AI feel different.