Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This is an interesting approach! Why not offload PDF extraction to other frameorks that apply OCR pdf -> .md


I may explore this when I implement the vectordb implementation I started.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: