frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Tell HN: Bob Cringely has died

870•paveworld•1d ago•187 comments

Ask HN: Who wants to be hired? (October 2026)

129•whoishiring•3d ago•514 comments

Ask HN: Who is hiring? (October 2026)

266•whoishiring•3d ago•294 comments

DOS Game Stunts Port to Linux,Windows,Browser,etc.

2•LowLevelMahn•15h ago•1 comments

Pico CSS is no longer maintained (final release v2.1.1)

3•anydigital•1d ago•4 comments

Ask HN: What are you reading to your kids?

18•jopizio•4d ago•28 comments

Chamath says open AI models are "commoditize your complement." Agree?

6•Dinuda•2d ago•4 comments

Muse.ai gets me kicked off fb marketplace

31•zcalvin•4d ago•15 comments

Why Linux is picking up at a rapid pace for AI usage?

5•keshwanianup•18h ago•7 comments

Ask HN: Email leaked while traveling abroad?

17•thatonehack•3d ago•17 comments

What did Google do to its Linux terminal on Pixel phones

4•Steaglsz•2d ago•1 comments

Tell HN: OVH price increase for dedicated servers

6•esher•3d ago•4 comments

Rpi – a Rust rewrite of the Pi agent, 10× faster startup

8•bigfishhk•3d ago•10 comments

GitHub mutuals on AI in life sciences or bioinformatics

5•arli_ap•2d ago•0 comments

40 years of writing code. I didn't see the end coming

22•temilson•3d ago•21 comments

OpenAI halved the allowance of the $200 plan to 10x

22•k9294•4d ago•9 comments

Ask HN: What landmarks/lighthouses in the era of LLM/AI?

3•throw_away_ai•3d ago•4 comments

Ask HN: What's a good title for a software engineer now?

6•type4•4d ago•19 comments

You've reached the end!

Open in hackernews

Ask HN: Best on device LLM tooling for PDFs?

4•martinald•1y ago
I've got very used to using the "big" LLMs for analysing PDFs

Now llama.cpp has vision support; I tried out PDFs with it locally (via LM Studio) but the results weren't as good as I hoped for. One time it insisted it couldn't do "OCR", but gave me an example of what the data _could_ look like - which was the data.

The other major problem is sometimes PDFs are actually made up of images; and it got super confused on those as well.

Given this is so new I'm struggling to find any tools which make this easier.

Comments

raymond_goo•1y ago
Try something like this

  !pip install pytesseract pdf2image pillow
  !apt install poppler-utils
  #!apt install tesseract-ocr
  from pdf2image import convert_from_path
  import pytesseract

  pages = convert_from_path('k.pdf', dpi=300)

  all_text = ""
  for page_num, img in enumerate(pages, start=1):
      text = pytesseract.image_to_string(img)
      all_text += f"\n--- Page {page_num} ---\n{text}"

  print(all_text)
constantinum•1y ago
give https://pg.llmwhisperer.unstract.com/ a try