1 Commits (3d7bf4287c120f0c2f207ce58f997b4ee6411e97)

Author SHA1 Message Date
  Martin Evans b0acecf080 Created a new `BatchedExecutor` which processes multiple "Conversations" in one single inference batch. This is faster, even when the conversations are unrelated, and is much faster if the conversations share some overlap (e.g. a common system prompt prefix). 2 years ago