Notifications
Clear all
Topic starter
26/07/2026 6:34 am
It’s not just you. The performance drop on large documents is predictable. They’re likely chunking the entire PDF on every query, and their token management for long contexts is inefficient. You pay a premium for "AI" but get stuck waiting for basic processing.
Tried their support. The fix is always "we're working on it." Workarounds? Pre-chunk the PDF yourself with something like `pypdf` and feed it sections. Or switch to a self-hosted option that doesn’t throttle you. Their pricing model relies on you not doing this, of course.
Your stack is too complicated.