frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Ask HN: What's the best hands-on path to learn ML inference infrastructure?

3•censor5•42m ago•1 comments

Ask HN: What are the most promising RL fields for a new master student?

31•zecice•7h ago•12 comments

BSD seq(1) on Linux

3•1vuio0pswjnm7•2h ago•0 comments

Robotic Adaptation Suffers from Data Collection Bottleneck

2•aunmesh•2h ago•0 comments

Ask HN: Anyone suffering from the tyranny of the Waiting Equation?

2•akersten•2h ago•1 comments

Ask HN: Is an in-page speed reader useful?

2•jvbotelho•5h ago•4 comments

Ask HN: I spent personal money on game dev. How to turn stalled project around?

2•pixelbyindex•6h ago•3 comments

Ask HN: Are you using Rust on embedded devices yet? If not, why?

9•mempirate•8h ago•2 comments

Ask HN: Graal VM Relevance

5•sid_op•7h ago•0 comments

Ask HN: Something fishy is going on with Google feed on my Android

4•motbus3•8h ago•0 comments

Ask HN: Protecting your Sites/Services from Unwanted Traffic?

10•prologic•16h ago•7 comments

Ask HN: What livestream do you keep open in a tab?

48•missmoss•1d ago•16 comments

Searchable, leveled learning map of AI/ML tools

2•maneeshthakur•9h ago•1 comments

Ask HN: Anyone else experiencing extreme Codex limits token burn rate?

2•throwaway2027•3h ago•1 comments

Connection-Agnostic Presence Tracking for Stateless Distributed Back Ends

5•duckydude20•14h ago•0 comments

ASK HN: Why has technology become so unreliable?

6•prmph•9h ago•10 comments

I built a jav data API, need honest feedback

2•javinfo•11h ago•0 comments

Ask HN: Which Jabber clients support SCRAM+ and XEP-0474

12•Bender•1d ago•0 comments

YouTube Comments in .csv

2•cristyg0101•13h ago•1 comments

Ask HN: Is Your Work Relaxing?

3•julienreszka•14h ago•1 comments

Timed YouTube video transcripts in .txt and .pdf

7•cristyg0101•1d ago•4 comments

What happens behind the scenes when we change effort for same LLM models?

13•tbharath•1d ago•8 comments

Ask HN: HotPin – lossless 120B MoE inference on 24GB RAM (CPU, 50 loc)

10•LozzKappa•1d ago•0 comments

Tell HN: Namecheap gave my account to an unverified third party

491•Thrashed•3d ago•175 comments

Who's making money from using/exploiting AI?

10•jonthepirate•1d ago•4 comments

Telemetry.sh – Telemetry Built for Agents

2•flurly•6h ago•1 comments

Ask HN: What do you do when your AI agents are working?

8•mavsman•1d ago•7 comments

Ask HN: How to get your open-source project gain popularity?

4•Hussain04•1d ago•12 comments

Ask HN: What happens when we do compress the context in Claude Code?

6•tbharath•1d ago•4 comments

Ask HN: Is neuromorphic computing going to replace traditional AI?

5•lennart-rth•1d ago•2 comments
Open in hackernews

Ask HN: How to use agents via API cheaply?

3•thih9•1d ago
I’d like to code with an AI agent via a BYOK open source client. I want to be able to inspect everything, check the system prompt, etc. Also, I’d like to avoid vendor lock in.

Is anyone here using a setup like this?

Does this get close in pricing and efficiency to the subsidized plans offered by model providers (like Claude Pro, etc)?

In particular Deepseek and Kimi models are being praised as both efficient and cheap; I’m curious if anyone built a coding setup around them.

Comments

ryanxsim•1d ago
Outside claude and codex, I use kilo code and its good so far
thih9•1d ago
What model do you use with kilo code? What are the typical usage costs?

How would you compare it to the provider plans (in costs and efficiency)?

ryanxsim•1d ago
Any local llm or third party outside gpt and claude. It depends on the trends sometimes deepseek, kimi or glm or sometimes gemini 3.1 pro was decent when it was decent