frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Outdoor displays fail in places you don't expect

https://duobond-display.com/news/display-interfaces-system-integration/532.html
1•Todorov•51s ago•0 comments

Hackers Don't Break in Anymore, They Log In

https://redmondmag.com/articles/2026/03/10/hackers-dont-break-in-anymore.aspx
1•janandonly•1m ago•0 comments

Show HN: Shri-io – a privacy-first, self-hosted URL shortener

https://github.com/barats/shrl-io
1•baratsemet•1m ago•0 comments

AI and the collapse of the intelligence-based hierarchy of merit

https://mattbruenig.com/2026/08/31/more-thoughts-on-ai/
1•rzk•7m ago•0 comments

Do you code in silence or do you listen to music?

https://textlog.cc/post/1078
1•stagas•7m ago•0 comments

Reticulum Is Not a Democracy

1•supernihil•10m ago•0 comments

Code Review Is Dead?

https://gioorgi.com/2026/code-review-dead/
1•daitangio•10m ago•1 comments

False Completion Is the Real Failure Mode of Coding Agents

https://medium.com/@giladha/false-completion-is-the-real-failure-mode-of-coding-agents-744be5a8c9c4
1•giladha•12m ago•0 comments

Apple Speech vs. Whisper: How Good Is Apple's New SpeechAnalyzer?

https://whispernotes.app/blog/apple-speech-vs-whisper
1•mazzystar•13m ago•0 comments

AI agents carried out every step of this ransomware attack – then left

https://www.theregister.com/security/2026/09/02/ai-agents-carried-out-every-step-of-this-ransomwa...
1•sbulaev•14m ago•0 comments

Why Ilya Sutskever's $32B SSI Valuation Highlights a Fatal Epistemic Vacuum

https://zenodo.org/records/22162749
1•AaRubinstein_CA•15m ago•0 comments

Three schoolgirls in Kinsale pulled up a pea plant covered in warts

https://scienceblog.com/b-three-schoolgirls-in-kinsale-pulled-up-a-pea-plant-covered-in-warts-and...
2•DamonHD•16m ago•1 comments

The Download: AI puzzles and a path to our nearest star system

https://www.technologyreview.com/2026/09/02/1143283/the-download-ai-puzzles-alpha-centauri-mission/
1•joozio•18m ago•0 comments

Sandbox

1•cyteeditor•18m ago•0 comments

Chat Bots Compared: An interactive quiz to find AI companions

https://chatbotscompared.com/
1•Sharanxxxx•19m ago•0 comments

Pre-Release of Polars 2.0

https://pola.rs/posts/announcing-polars-2/
2•komape•22m ago•0 comments

The birthday paradox and why hashes are 256 bits, not 128

https://sslog.dpdns.org/birthday-attack.html
1•Shaurya_Sharma•25m ago•0 comments

An open-source design plugin that helps your AI build less generic websites

https://github.com/MickeyAlton33/web-designer-plugin
1•Nina_antalpha•28m ago•0 comments

Sex After AGI

https://twitter.com/Joshbocanegra/status/2093122061116035105
1•jbai•30m ago•0 comments

Terence Tao: Protecting math problems from automated solvers

https://mathstodon.xyz/@tao/117204929023813310
3•bertman•30m ago•0 comments

No Vibe Coding: Discussion Coding on Termux – Phone Only, Zero Trust

https://m.blog.naver.com/amadale/224399795090
1•garlicfarmer•33m ago•0 comments

Show HN: Self-hosted CRM for manufacturing (Django and Htmx)

https://github.com/OlegUshakov-pl/CRM
1•OlegUshakov•34m ago•0 comments

Markdown Editors Weren't Built for Coding Agents, So I Made One

https://marginal.md/blog.html
2•kishiguro•35m ago•0 comments

Show HN: Tiny Markdown CLI renderer – Streaming, 10x faster, 2MB flat RAM

https://github.com/cjccjj/mdflow
2•cjcj•35m ago•1 comments

Lex Speedman: We love you, Lex. but You talk too slow

https://lexspeedman.com/
1•mellosouls•40m ago•0 comments

Show HN: 2nd Release Canddiate for Back in Time 2.0.0

https://github.com/bit-team/backintime/releases/tag/v2.0.0-rc2
1•buhtz•51m ago•0 comments

How Delhi Built a World-Class Metro for a Bargain

https://www.nytimes.com/2026/09/02/headway/india-delhi-metro-public-transit.html
2•thunderbong•52m ago•0 comments

Show HN: A register of agent-payment code that pays once when the reply is lost

https://aurumflux.co/retry-safety/
1•aurumflux20•53m ago•0 comments

Gemini 3.8 Flash in Google Antigravity – Google Antigravity Blog

https://antigravity.google/blog/gemini-3-8-flash-in-google-antigravity
1•humphrey_liu•55m ago•0 comments

Claude Code skill: no emdashes

https://github.com/woutersf/claude-skill-no-emdashes
1•woutersf•58m ago•0 comments
Open in hackernews

Open-Source Refact.ai Agent is #1 on SWE-bench Lite With a 59.7% Score

https://refact.ai/blog/2025/sota-on-swe-bench-lite-open-source-refact-ai/
3•kate_at_refact•1y ago

Comments

kate_at_refact•1y ago
Open-source Refact.ai achieves #1 on SWE-bench Lite with a 59.7% score. Our approach: fully autonomous Agent, no manual intervention needed.

How we did this:

• Prompt strategy: https://github.com/smallcloudai/refact/blob/swe-boosted-prom... • Claude 3.7 Sonnet as orchestrator • deep_analysis() tool (powered by o4-mini) for reasoning • Tool suite for repository exploration, code modification, and testing. Used dynamically based on task needs • One correct solution through iteration!

Autonomy = our core strength.

Refact.ai Agent completes the entire dev workflow independently: plans, executes, tests, self-corrects, and delivers a production-ready result. For each task, it made one multi-step run to generate a single correct solution, creating custom strategies rather than following rigid scripts.

You can read tech details on our SWE-bench approach: https://refact.ai/blog/2025/sota-on-swe-bench-lite-open-sour...

Your questions are welcome! Also, welcome to try Refact.ai Agent in VS Code and Jet Brains: https://linktr.ee/refactai