frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

An Enterprise-Level Retrieval-Augmented Generation System

https://comfyai.app/article/llm-applications/enterprise-level-rag-hands-on-practice-II
6•zljdanceholic•1y ago

Comments

zljdanceholic•1y ago
How can we search the wanted key information from 10,000+ pages of PDFs within 2.5 hours? For fact check, how do we implement it so that answers are backed by page-level references, minimizing hallucinations?

RAG-Challenge-2 is a great open-source project by Ilya Rice that ranked 1st at the Enterprise RAG Challenge, which has 4500+ lines of code for implementing a high-performing RAG system. It might seem overwhelming to newcomers who are just beginning to learn this technology. Therefore, to help you get started quickly—and to motivate myself to learn its ins and outs—I’ve created a complete tutorial on this.

We have a complete graph to explain its workflow where multiple tools are used: Docling for parsing PDFs, LangChain for chunking text, faiss for vectorization and similarity searching, and chatgpt for LLMs.

Besides, I also outline the codeflow, demonstrating the running logic involving multiple python files where starters can easily get lost. Different files are colored differently. The purpose of showing this is not letting you memorize all of these file relationships. It works better for you to check the source code yourself and use this as a reference if you find yourself lost in the code.

The original project of Ilya Rice design its RAG system for answering questions of annual reports from companies, so he only designed three types of question response format for that challenge: a name, a number, or a boolean. But to ask questions about technical stuff, we absolutely ask general questions like How does RoPE work? to know about some concepts and the like Therefore, I further modify the system logic to fit this need by customizing an AnswerWithRAGContextExplanationPrompt class and automatically matching the most related chapter and corresponding pages via searching through all faiss databases (only retrieve the top-1) The final performance is demonstrated below (not cherry-picked, only tested once).

- What's the steps to train a nanoGPT from scratch? The answer is:

Training nanoGPT from scratch involves several clearly defined steps. First, set up the environment by installing necessary libraries, using either Anaconda or Google Colab, and then download the dataset (e.g., tinyShakespeare). Next, tokenize the text into numerical representations and split the data into training and validation sets. Define the model architecture including token/positional embeddings, transformer blocks with multi-head self-attention and feed-forward networks, and layer normalization. Configure training hyperparameters and set up an optimizer (such as AdamW). Proceed with a training loop that performs forward passes, computes loss, backpropagates, and updates parameters, while periodically evaluating performance on both training and validation data. Finally, use the trained model to generate new text from a given context.

All code are provided on Colab and the tutorial is referenced here. Hope this helps!

Show HN

https://topfloor.company/
1•sankalpdomore•1m ago•0 comments

Schlep Blindness (2012)

https://www.paulgraham.com/schlep.html
1•poly2it•1m ago•0 comments

My retrieval can't tell a question it can answer from one it can't

https://github.com/asanabrial/leteo/blob/main/docs/nothing-worth-tuning.md
1•asanabrial•3m ago•0 comments

God's Eye View

https://github.com/bilawalsidhu/gods-eye-view
1•pelagicAustral•3m ago•0 comments

FDA Failed to Protect Consumers Again. Here's the Latest Proof

https://efoodalert.com/2026/08/27/fda-failed-to-protect-consumers-again-heres-the-latest-proof/
1•speckx•5m ago•0 comments

The Linux Kernel Is Approaching 2k CVEs per Release

https://www.phoronix.com/news/Linux-Kernel-CVEs-Nearly-2000
2•aniviacat•5m ago•0 comments

In Africa, a Growing Underground Solar Boom

https://e360.yale.edu/digest/africa-solar-2026
2•Brajeshwar•5m ago•0 comments

Show HN: I built a Mac app that replaces identifiers with stable tokens

https://clipscrub.com/
1•mhay•6m ago•0 comments

You Can Now Carry a Teaching Mainframe in Your Bag – Planet Mainframe

https://planetmainframe.com/2026/08/gibson-zos-security-simulator/
1•rbanffy•6m ago•0 comments

Show HN: AgentConnect, shared agents with separate permissions

https://github.com/agentconnect-md/agentconnect
2•Nicole9•6m ago•0 comments

Show HN: Grith – syscall-level supervision for AI coding agents on Linux

https://github.com/grith-ai/grith
2•edf13•7m ago•0 comments

Ask HN: Launch board built only for iOS apps?

1•appshunter•7m ago•0 comments

What is RSS Chat – and what not?

https://twowavesandadot.blog/what-is-rss-chat-and-what-not.html
1•tjansen•8m ago•0 comments

Wikivibes – Write the encyclopedia you wished existed

https://wikivibes.org/
1•dhtgbdgthb•11m ago•0 comments

AI-proof? Younger workers desert the digital world for traditional crafts

https://www.theguardian.com/money/2026/aug/27/ai-proof-jobs-traditional-crafts
2•pseudolus•12m ago•0 comments

Amazon Aurora DSQL now supports foreign key constraints

https://aws.amazon.com/about-aws/whats-new/2026/08/aurora-dsql-foreign-key-constraints/
1•alexbilbie•13m ago•0 comments

Show HN: Coordination Layer for Coding Agents

https://twing.dev/
1•imayank•13m ago•1 comments

Spring Boot on a 256 MB VPS: JDK 25 Compact Headers Put to the Test

https://pvrlabs.xyz/articles/spring-boot-256mb-jdk25.html
2•theanonymousone•13m ago•0 comments

Gong Zizheng, who led China's planetary defence project, dies at 61

https://www.scmp.com/news/china/science/article/3365533/gong-zizheng-who-led-chinas-planetary-def...
3•bookofjoe•14m ago•0 comments

Cloudflare frees 100TB of RAM by shrinking DNS cache entries

https://www.tomshardware.com/tech-industry/big-tech/cloudflare-frees-100tb-of-ram-by-shrinking-dn...
1•sbulaev•15m ago•0 comments

Vibe Coding and Desktop Publishing

https://blog.numericcitizen.me/2026/08/28/vibe-coding-and-desktop-publishing.html
1•speckx•15m ago•0 comments

Packman at 25 – where do we go from here?

https://lists.links2linux.de/pipermail/packman/2026-August/018314.html
1•xiaoyu2006•18m ago•0 comments

iPhone Photography Awards 2026

https://ippawards.com/winners/showcase
2•my2centsWorth•20m ago•0 comments

China's Oil Demand 'Likely' Peaked Last Year, Sinopec Says

https://www.bloomberg.com/news/articles/2026-08-24/china-s-oil-demand-very-likely-peaked-last-yea...
3•toomuchtodo•20m ago•1 comments

Waymo imported over 3,200 Zeekr robotaxis in the US despite 102.5% tariffs

https://carnewschina.com/2026/08/12/why-alphabets-waymo-is-importing-3200-chinese-zeekr-robotaxis...
2•mnming•21m ago•1 comments

How weaving helped invent modern computing – and is shaping its future

https://theconversation.com/how-weaving-helped-invent-modern-computing-and-is-shaping-its-future-...
2•surprisetalk•21m ago•0 comments

Show HN: dmx – MCP server for adding configurable, gated loops to coding agents

https://dmx.deepmodel.ai/
2•hpieris•21m ago•1 comments

Shared Kitchen Summit

https://www.sharedkitchensummit.com
2•mooreds•21m ago•0 comments

Apache Iggy Graduates to a Top-Level Project

https://iggy.apache.org/blogs/2026/08/24/apache-iggy-top-level-project-tlp-graduation/
2•aray07•22m ago•0 comments

I accidentally turned LLM memory into program analysis

https://pwning.systems/posts/llm-memory-program-analysis/
3•jordyzomer•22m ago•0 comments