What is your goal?
Goal is using PDF data translated to text in one shot analyze flow. Kept in one active scenario, and more or less cost free. Current scenario: Tally-OpenAi (vision) gatekeeper-OpenAi (create a com..), analyse module-Gmail
What is the problem & what have you tried?
Vision-only chain β no PDF text extraction
Binary PDF = hard timeout, flow dies
No format-sorting modules (kept cheap/simple)
PDF needs OCR β different pipeline
Word (.docx) β same wall, zip-packed