frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

An Enterprise-Level Retrieval-Augmented Generation System

https://comfyai.app/article/llm-applications/enterprise-level-rag-hands-on-practice-II
6•zljdanceholic•1y ago

Comments

zljdanceholic•1y ago
How can we search the wanted key information from 10,000+ pages of PDFs within 2.5 hours? For fact check, how do we implement it so that answers are backed by page-level references, minimizing hallucinations?

RAG-Challenge-2 is a great open-source project by Ilya Rice that ranked 1st at the Enterprise RAG Challenge, which has 4500+ lines of code for implementing a high-performing RAG system. It might seem overwhelming to newcomers who are just beginning to learn this technology. Therefore, to help you get started quickly—and to motivate myself to learn its ins and outs—I’ve created a complete tutorial on this.

We have a complete graph to explain its workflow where multiple tools are used: Docling for parsing PDFs, LangChain for chunking text, faiss for vectorization and similarity searching, and chatgpt for LLMs.

Besides, I also outline the codeflow, demonstrating the running logic involving multiple python files where starters can easily get lost. Different files are colored differently. The purpose of showing this is not letting you memorize all of these file relationships. It works better for you to check the source code yourself and use this as a reference if you find yourself lost in the code.

The original project of Ilya Rice design its RAG system for answering questions of annual reports from companies, so he only designed three types of question response format for that challenge: a name, a number, or a boolean. But to ask questions about technical stuff, we absolutely ask general questions like How does RoPE work? to know about some concepts and the like Therefore, I further modify the system logic to fit this need by customizing an AnswerWithRAGContextExplanationPrompt class and automatically matching the most related chapter and corresponding pages via searching through all faiss databases (only retrieve the top-1) The final performance is demonstrated below (not cherry-picked, only tested once).

- What's the steps to train a nanoGPT from scratch? The answer is:

Training nanoGPT from scratch involves several clearly defined steps. First, set up the environment by installing necessary libraries, using either Anaconda or Google Colab, and then download the dataset (e.g., tinyShakespeare). Next, tokenize the text into numerical representations and split the data into training and validation sets. Define the model architecture including token/positional embeddings, transformer blocks with multi-head self-attention and feed-forward networks, and layer normalization. Configure training hyperparameters and set up an optimizer (such as AdamW). Proceed with a training loop that performs forward passes, computes loss, backpropagates, and updates parameters, while periodically evaluating performance on both training and validation data. Finally, use the trained model to generate new text from a given context.

All code are provided on Colab and the tutorial is referenced here. Hope this helps!

What can go wrong with GenAI? 21 risk categories from 53 studies

https://onlinelibrary.wiley.com/doi/pdf/10.1155/hbe2/2320511
1•yousufsaif•37s ago•0 comments

I tried connecting a 3.5mm jack directly to a speaker: My first audio amplifier

https://ajseven.me/blog/first-electronic-project
1•its_ajseven•1m ago•0 comments

Research;Open Sourcing Health Research with AI – Accelerating Evidence Synthesis

https://zenodo.org/records/23199360
1•xxiioiixx•8m ago•1 comments

AWS Large Payment Verification Amount

https://www.reddit.com/r/aws/comments/1wzh3yp/large_payment_verification_amount/
3•ericdiao•13m ago•0 comments

Calling It Quits on ServerFault

https://sysadmin1138.net/mt/blog/2026/10/calling-it-quits-on-serverfault.shtml
4•zdw•13m ago•0 comments

Guide to Using Generative AI in Programming – Textbook by Antti Laaksonen

https://link.springer.com/book/10.1007/978-3-032-07453-9
1•vismit2000•21m ago•1 comments

Ask HN: Does AI do better with CSS or Tailwind?

1•pupppet•26m ago•0 comments

Backdooring Sparse Autoencoders

https://arxiv.org/abs/2610.06049
2•sbulaev•28m ago•0 comments

Claude subscription plan provides more value than OpenAI's, study says

https://www.theregister.com/ai-and-ml/2026/10/06/anthropic-claude-subscription-plan-provides-more...
1•galaxyLogic•28m ago•0 comments

The Eternal Complement

https://openai.com/index/the-eternal-complement/
1•gmays•31m ago•0 comments

Local AI is real now. And it's blowing my mind

https://hive.technology/lab-notes/local-ai-is-real/
1•tylermhall•33m ago•1 comments

Show HN: Git extension for controlling Git worktree mess

https://getzit.org/
2•ironfootnz•33m ago•0 comments

Claude Isn't Allowed to Write Me Prose

https://blog.kvit.app/posts/agent-not-allowed-to-write-prose/
4•skolos•50m ago•0 comments

The art of defusing a second world war bomb

https://www.theguardian.com/news/ng-interactive/2026/oct/06/it-could-knock-a-whole-street-down-th...
6•sandebert•52m ago•1 comments

Ichneumon eumerus

https://en.wikipedia.org/wiki/Ichneumon_eumerus
2•GaryBluto•55m ago•0 comments

Show HN: A browser-only checker for text hiding under PDF black boxes

https://camlabs.ai/redaction-checker/
1•camlabsai•56m ago•0 comments

Revisiting Soundness for Occurrence Typing, Semantically

https://arxiv.org/abs/2609.16299
2•azhenley•56m ago•0 comments

Show HN: EasyToSTL – Convert 2D images and text to watertight, print-ready STLs

https://easytostl.com/
1•dawn3727•59m ago•0 comments

Show HN: ImgKit – 62 image tools that run in the browser

https://imgkit.xyz
3•jaydippatel83•1h ago•0 comments

Jev isn't a better judge than Claude

https://deeplearningguy.github.io/blog/jev-vs-claude-judge/
2•deepbuilder•1h ago•0 comments

Show HN: NanoMuse – An open-source AI agent for your phone and computer

https://github.com/nano-muse/nanoMuse
2•ilreb•1h ago•0 comments

Show HN: Vim-like, hyper-efficient, LLM power tool – 20-80k tokens not 350k+

https://easiest.ai/
1•skhameneh•1h ago•0 comments

Multiplayer AI coding workspace for teams and agents

https://www.tryfridaywork.com
1•arunjdass•1h ago•0 comments

Biology, Buddhism and AI: the Platonic space model and the rogue agent swarm

https://okossa.com/biology-buddhism-and-ai-ff213a852eb4
1•okwe•1h ago•0 comments

An Open Challenge to GitHub Users: Test Das Architecture with Meta Muse Sentinel

https://zenodo.org/records/23201404
1•sangamdas•1h ago•0 comments

Can VLMs recognize famous videos from just their colors?

https://loganbolton.github.io/blog/videocolorbench/
1•septisum•1h ago•0 comments

OpenWAM: An Open Framework for Composable World-Action Models

https://openwam.stanford.edu/
2•ilreb•1h ago•0 comments

La Cueva BBS in Mexico in 1993 (session replay)

https://nanochess.org/la_cueva_bbs.html
11•nanochess•1h ago•1 comments

How do you handle client feedback and revisions on website projects?

1•tweakpage•1h ago•0 comments

Agentty is still the best way to manage agentic sessions

https://github.com/agentty-xyz/agentty
2•minev-dev•1h ago•0 comments