Top
Best
New

Posted by handfuloflight 1 day ago

fx :Tiny, open, native coding agent.(fx.sh)
280 points | 116 comments
rsyring 1 day ago|
For all the people asking "Why?", it seems like TFA has a pretty good list of features/attributes that it thinks sets it apart:

- fx is a coding agent harness and CLI written in Zig, optimized for research and embeddability as part of larger systems.

- It focuses on minimalism and performance across the board, from system prompt design, to its tools, feature set, and 6.39mib binary.

- For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI.

- It's open source (Apache-2.0), model-agnostic, and suitable for both local and cloud inference.

- Designed for instant installation and embedding in resource constrained environments and agent sandboxes.

- fx cold starts in 10µs and does no unnecessary work or I/O prior to accepting user input, making it ideal for programmatic use.

- Optimal fx.wasm builds produced by the Zig toolchain, which further reduce fx's size, making the network stack pluggable.

- fx contributes single-digit megabytes of memory baseline, allowing you to pack many instances in one machine.

- fx preserves scroll history by default, produces minimal output, and makes sparing use of complex TUI or paints

- Minimal system prompt and tools, to save on token costs and to yield optimal time-to-first-token performance (TTFT).

- Small core, extended via skills, plugins, MCPs, with a Unix-like philosophy to extensibility.

- Designed to work with local models, gateways, direct provider API access or subscriptions.

OleksandrC 13 hours ago||
If you like this list of "why?", you might also like this: https://usehax.dev/ (I am the author). Most of the list applies, similar minimalist Unix tool approach, with some differences. Hax is written in C, the dynamically linked binary is even smaller (0.6 MB), MIT-licensed. No wasm though.

Important difference - fx is currently Vercel AI Gateway only - while hax does support multiple providers already (OpenAI API, ChatGPT/Codex subscription, Anthropic API, OpenRouter, OpenCode Zen/Go), and integrates well out of the box with local llama-server.

tecoholic 12 hours ago|||
It’s funny what “tiny” means for different people depending on their background. I expected it to be under an MB as well and was surprised by 6MB.
pjmlp 4 hours ago|||
Basically, for me tiny means it fits on a floppy 1.44 HD.

Naturally meaningless when people carry around USB sticks that might even hold a 1 TB, but alas.

krzyk 16 minutes ago|||
> Naturally meaningless when people carry around USB sticks

Do people do that? I think it was a decade ago, I thought people download from web nowadays.

tecoholic 2 hours ago|||
Pretty much my baseline as well. Good example how experiences shape our thinking.
brabel 4 hours ago||||
For comparison, I publish a CLI tool written in Dart, can compile to native anything including Wasm , the Linux binary is around 5MB and it does quite a lot.
nine_k 10 hours ago|||
A typical Go binary would be 2-3 times larger. A typical Node project would easily pull more than 6 MiB of just code, not counting the runtime.
philips 4 hours ago||||
I really like the philosophy document! https://github.com/OleksandrChekhovskyi/hax/blob/master/docs...

I don’t know if I am ready to use a new tool written in C and using libcurl but I will give it a shot.

WhyNotHugo 4 hours ago||||
Recently featured on HN: https://news.ycombinator.com/item?id=49273175
messh 7 hours ago||||
that's an amazing project! people always say why 6mb vs CC's 250mb even matter when you are calling out to LLMs hosted in the cloud. But... I regularly run hierarchies of agents with say 50-100 on a regular basis. So 650 vs 25050 ... is "can do" vs "cannot"
lmz 5 hours ago||
Is that real memory taken? If it's shared code from the same executable surely the multiples are not very relevant?
rgbrgb 13 hours ago||||
this looks great! what are you using it for? i like the idea of being able to use one of these (sandboxed) within a larger program kind of like how I use LLM's to do small tasks within my apps now but with a few tools (web search). my current way of doing that is like building a mini-harness with a couple tools within the app, but something more drop-in would be better obviously.
gandreani 11 hours ago||
I was going to ask if there's any plans to integrate something like the fx's ACP server[1] or pi's RPC mode [2].

I'm making something like Paseo and hax is very interesting as a Pi replacement.

[1] https://fx.sh/docs/using-fx/acp

[2] https://pi.dev/docs/latest/rpc

FrenchTouch42 5 hours ago|||
It looks really nice. Do you have any plan to support Claude subscriptions (pro/max)?
OleksandrC 2 hours ago|||
Technically, this would be straightforward. The problem is that Anthropic seems to be really against using Claude subscriptions with anything other than Claude Code - you might even risk your account getting banned for doing so. You could search online for the "openclaw claude banned" for more details on that story.
OJFord 2 hours ago|||
Didn't Anthropic stop allowing it, can only use the subscriptions with first-party tools?
verdverm 11 hours ago|||
Does it have anything co-designed around Vercel infrastructure? This is what happened to NextJS and why I will likely never touch Vercel open source again
rafael-lua 9 hours ago||
Yeah, I see the Vercel logo, and I am instantly out.
Kim_Bruning 1 day ago|||
Very neat. The demo on the page feels very intuitive to me! (if you're used to bash at least)
jauntywundrkind 1 day ago|||
I've only done a little of the new opencode v2 "mini" but it too offers a nice preserve-scroll by default.

OpenCode is the best behaved TUI i've seen by far (they invented OpenTUI to make it so good, also in Zig), so it feels less crucial. But it's nice to have there!

The "small core" model is very popular all of a sudden. DeepSeek's new harness is famously like that. https://news.ycombinator.com/item?id=49285244

OpenCode isn't quite as small, but there's very much been a deliberate attempt to drive much more into a plugin-based system. I enjoyed Dax talking about the new constitution of opencode, and the results of his agent comparing OpenCode & the new DeepSeek. https://bsky.app/profile/thdxr.com/post/3msy4gjttoc2f https://bsky.app/profile/thdxr.com/post/3msygiqyg6v2y

> an architectural change we made in opencode2 is nearly everything is an internal plugin / there's 68 of them that cover our built in agents, integrations, config loading, etc

i also think this is such a brilliant fun architectural twist too:

> OpenCode is the first time i could justify event sourcing in a real system / everything that happens is an event which gets projected into the sqlite db

https://bsky.app/profile/thdxr.com/post/3mt2qx3ktib2c

it's so fun seeing new malleable software cores emerge, try to figure out how to augment agency. agentic software striving itself to extend the agency it itself offers. it's been way too long since we've had ambitions to build general system, architectures that serve more than the user. this has held computing back for far too long. this is such an excellent interesting field, of such a more ambitious computing, opening up.

solarkraft 1 day ago||
Thanks for the links on opencode 2! I’ve been meaning to get into this as I’ve been frustrated about some opencode 1’s behavior and design. Many of the encounters made me come up with ideas I’m happy to see they also had! As much fun as it may have been to build my own harness, I feel like the core primitives should be pretty well understood by now (in fact I envision a standard core library / API design taking shape).
rvz 1 day ago|||
Most of all what you have said is not really any clear differentiation against the rest of the 100s of other agents. Just minuscule or non-negligible implementation details and I'm afraid it is sadly yet another experimental slop project.

It is a branded "mee too" coding agent that we have seen hundreds of them already.

rsyring 1 day ago||
I can't say I'm super into all the agents that are being created. But I do try to keep up here on HN and I can't say that I can remember any with this particular set of attributes.

In particular, aiming to be embeddable into other projects seems rather notable. At least, not something I've remembered of other projects that have made there way across the HN front page.

qudat 1 day ago|||
Pi is a composition of libraries that can be used to build agents. That seems far more interesting than this “minimal” agent.

This was written in zig and built by vercel. That’s the only notable characteristics about this project.

All code agents look the same and this one is no different.

slowin 11 hours ago||
I don't think Pi is tiny because it's written in typescript and requires a javascript runtime. I definitely want a "tiny" compiled agent with no runtime requirements.

I don't think Zig is that great of a language for this, but it's better than typescript. I don't want to use Vercel software so will pass, but would love to see a more community driven effort.

kalms 4 hours ago||
What’s a better lang in your opinion?
verdverm 11 hours ago|||
Almost every harness has an SDK now, it seems part of the "mvp" at this point, both open and closed source

https://learn.chatgpt.com/docs/codex-sdk

https://code.claude.com/docs/en/agent-sdk/overview

https://opencode.ai/docs/sdk/

https://pi.dev/docs/latest/sdk

or if you want SDK first, my recommendation is https://adk.dev/

cmrdporcupine 1 day ago||
"- For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI."

I've actually been wondering lately why coding agent functionality isn't just... part of my shell already. Just another kind of interaction modality with an existing shell. Could probably even be an extension to fish or nu-shell even.

Please stop me from forking off on yet another project though.

pjmlp 4 hours ago|||
Microsoft has you covered,

https://devblogs.microsoft.com/commandline/intelligent-termi...

https://devblogs.microsoft.com/commandline/github-copilot-in...

rsyring 1 day ago|||
Originally, Warp was doing just that. Reimagining the terminal including making AI a part of it. I don't think they really found much purchase there because they ended up needing to make a platform out of it:

https://www.warp.dev/

cmrdporcupine 1 day ago||
Yeah terminal window is a bit of a different story, though. I mean the actual shell binary.

With some ... intensive ... security/sandboxing/containerizing of some kind though, I guess.

_pdp_ 11 hours ago||
It may come across harsh but IMHO the only interesting thing about this project is that it is written in Zig. That's it.

Everything else in the harness is largely the same just Vercel-flavoured.

The portability benefit is also a bit over-sold imho. I wrote a harness in Go and it is as portable as this ... in fact it deploys straight into Vercel's own sandbox environment on demand without any issues.

That being said, did you say GLM 5.2 free? I need to look into that. GLM is remarkably capable model and that alone is worth using the harness in my books.

mparramon 2 hours ago||
UX-wise it's quite minimalistic, that seems one of their differentiators. And I'm liking it!
_superposition_ 3 hours ago||
This is exactly how I work today. I've always lived in the terminal but typing bash commands anymore is way too much work and an LLM can produce way better one liners than I can.
ryuuseijin 8 hours ago||
Since I've seen some other agent tools being suggested, let me throw Maki in the ring as well: https://maki.sh/

- written in rust, super fast startup and rendering

- implements token saving techniques

- plugins written in lua

No affiliation, just think it's a really well made piece of software.

kzrdude 1 hour ago||
Can it replace pi? Would be good to have a robust and fast replacement.
pynappo 6 hours ago|||
been enjoying maki too. respects xdg spec (unlike pi), the ui is snappy and pretty much exactly what i want, and i like the choice to keep the lua API similar-ish to neovim.
jatins 7 hours ago|||
This has some nice ideas like indexing (aider like repo map) built in, could be interesting
Imustaskforhelp 3 hours ago||
Love using maki on 500mb and 1gb ram tiny vps's where opencode sometimes could go OOM.

I really love maki, its awesome!

kgeist 1 day ago||
>Tiny ~6mb binary

I wonder why it's so large for a program written in Zig. It's basically just a loop that accepts user input, prepares the context, sends it to the LLM, parses the output, invokes the tools, and presents it all in the terminal. Add the built-in prompts and a few checks here and there (like blocking a write tool call before the file has been read first), and I'd expect a truly tiny native agent to be around 200-300 KB max.

miguel_martin 13 hours ago||
fwiw, here's 3code which is 1.6MiB written in Nim - https://3code.capocasa.dev/
wyre 11 hours ago|||
I was looking it last night and the repo is something like 600k lines of Zig. Maybe 500k after comments and blank lines.

In my own experiments to build a tiny zig agent it came out to under 800kb.

rjzzleep 8 hours ago||
I think the problem is that when you use agents to write Zig it will bruteforce code, because there isn't that much good reference code to work off of.
irishcoffee 11 hours ago||
So large? Really? Allow me to introduce you to electron.
klibertp 30 minutes ago|||
Yes, large. I haven't used Zig much myself, but from a few experiments I ran, Zig handles dead code elimination exceptionally well. It compiled a full Win32 GUI calc app that used Capy (a full, cross-platform GUI framework) into a 133kb executable. Removing Capy completely and using Win32 APIs directly produced an even smaller (93kb) binary (it also removed some DLL dependencies, leaving basically only ntdll.dll). For the same task, Rust + Slint produced a 4.7 MB binary that still depended on multiple (non-Windows-provided) shared libraries.

Given another commenter's mention of a similar project written in Nim that yielded a 1.6 MB binary, my first guess is that the 6 MB Zig binary simply isn't optimized for size - it might be a debug build. If not that, then I'm not sure what's happening, but yeah, in the context of Zig, 6mb for a CLI app is a bit strange.

pjmlp 4 hours ago|||
Those of us educated in 70's and 80's home computing have several nice words for stuff like Electron.
bodge5000 1 day ago||
It does looks really interesting and definitely something I'll check out, but (genuine question), should "agent" and "agent harness" be used interchangeably as it is on here? It describes itself as an agent harness, but the tagline is "tiny, open, native coding agent".

I'm not sure harness is the right word either, but that seems to be what the industry has settled on so I'll concede on that, but surely the agent is the thing doing the work (which I guess is the model, or an instance of the model which is why agent is different?), whereas the harness is how the user interacts with the agent. We've had ways to describe that relationship before; client and server, frontend and backend, but again, I'll concede that the shiny new thing doesn't want to use boring old terminology, but I think some consistency and logic in the shiny new terminology is pretty important

That isn't specifically about fx of course, more of a general industry complaint

kzrdude 2 hours ago||
Maybe ”agent shell” could be a better term for harness. But it’s pretty overloaded..
cramforce 1 day ago|||
Harness = the software the agent runs on. This is plain old software. You can trust it as much as any software. Agent = the thing that runs on the harness. It cannot be trusted because it is driven by an LLM
bodge5000 1 day ago||
I get that, but wouldn't that make the agent and the harness two very different things, with the agent being closer to a model (the source of the agent) than a harness? Somebody else mentioned a console/game analogy, with the model being the disc, the agent being the running game, and the harness being the console OS, wouldn't it be like me describing Windows as a game, because games run on it?
amdahl 1 day ago|||
I think "harness" is a thing, the code/binary, and "agent" is a process, an instantiated run of that code/binary with a given LLM/env, etc.

So its like, "GTA 6" as the disc vs. the specific game you're in the middle of being chased by cops, harness vs agent.

In practice they're intertwined and it becomes hard not to use the terms somewhat interchangeably, but "you ask the agent how the harness works" vs. the other way around, clearly.

bodge5000 1 day ago|||
Wouldn't the disc in that analogy be the model? Thats the source of the instance. The harness I guess would be the OS of the console running the game
amdahl 1 day ago||
It's more like in that analogy, the LLM is the gamer, instantiated as agent within a given game.

And the virtual world of GTA 6 is actually your codebase/env, the cops chasing are the bugs/angry customers, etc. The harness is providing an accurate/efficient ability for the model to understand/interact with the virtual world, flee the cops, etc. Decomposed at various architectural boundaries per your taste, but that's like loading a skin on the engine.

bodge5000 1 day ago||
That doesn't seem right, surely the gamer would still be you, since you interact with the agent through the harness. If the player is the LLM, what is the human in this analogy? I guess a harness doesn't necessarily need human input (most do of course, but thats not a technical limitation), but then again neither does a game for the same reasons

Regardless though, this is what I mean, we now have 3 definitions for an agent; an instance of a model (which is how I think of it), a model configuration for a given task (from another commenter) and your definition which appears to be somewhere between the two, though it seems we agree with what a harness is.

phiagent 1 day ago|||
[dead]
alansaber 1 day ago|||
"Harness" should describe the overall system. The user interactions are increasingly negligible (due to model routing, adaptive reasoning, etc). "Agents" are the tool lists/settings provided to the model, etc.
stellalo 13 hours ago|||
harness + llm = agent
brap 13 hours ago||
Nowadays most “LLM” endpoints include some sort of server side harness as well, and I’d bet more than one model involved, so it’s really just agents all the way down
verdverm 11 hours ago||
I personally like the way Scion breaks it down into

- model

- harness (tools/config)

- agent (live/running)

https://googlecloudplatform.github.io/scion/concepts/

Scion allows for multiple configurations of a harness, allowing you to configure the same tool differently based on what you are doing, particularly important when you want to restrict permissions.

https://googlecloudplatform.github.io/scion/supported-harnes...

for the unfamiliar, Scion is an OpenClaw like platform from a Google dev, not supported or sponsored by the company

SmashDan 1 day ago||
I'm not in the tech industry. Could someone explain why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.
jdub 57 minutes ago||
When a new technique or capability arrives on the tech scene, there's a point in the invention-to-diffusion story when the new thing becomes accessible (e.g. cheap and/or easy) enough for a broader audience of developers to experiment with it... but before anyone's figured out best practices, let alone polished products/projects, or calcified around a market leader.

So you get a Cambrian explosion of weird little projects. Ultimately, one of them will probably become the "market" leader... or at least the market default.

Right now there's a lot of agent harnesses and sandbox projects floating about.

Fun examples from the past: text editors, window managers, IRC clients, blogging engines (first static, then dynamic, then static again), Twitter clients... every programming language community has weird clusters of library/framework duplication in their history...

Sometimes these projects take on a rite of passage flavour... like, as every Jedi builds their own lightsaber, every developer builds their own... blog? That used to be the obvious one. Less so these days.

eikenberry 9 hours ago|||
Because a lot of people are writing their own to get a tool that they understand and can manipulate as they like. So they like to share them and see what other people have done to learn from. As a community we are still very far from coming to a consensus on what a good harness looks like and the only way, IMO, to get a good feel for it is to write your own.
wolttam 4 hours ago||
It’s also exactly what the tech enables. It’s not hard to imagine there being hundreds of thousands of different harness projects, if not more.
odo1242 1 day ago|||
It's basically just a relatively simple to create piece of software that's important to get right (since you use it so much), can be made by many different design philosophies (maximal vs. minimal, customizability, etc.), and has very few good standards around it as of yet.
chrysoprace 1 day ago|||
Hacker news generally follows trends, and this is the current trend.

The discussion around coding agents nowadays is steering towards harnesses (which is probably a better description of what this is). "Agent" here is doing a lot of heavy lifting and has become a bit of a catch-all term to describe a model + harness + tooling + prompt + some other things that I've probably not thought about. The harness is a part that's being explored more as many believe it's where we can get some better performance out of the models.

This one in particular is from Vercel who provide a service to use models, so they have a vested interest in providing a harness.

a2ff6eeb0 1 day ago|||
Because we're actively exploring the best way to remove any need to deal with code, and make it so that you don't need any real talent to make a computer do things for anyone.

We haven't quite hit on the right formula yet, but people are very excited by the possibility.

selectnull 2 hours ago|||
Because we are in a gold rush and the best thing to sell are the shovels.
rglover 1 day ago|||
https://en.wikipedia.org/wiki/Mimetic_theory
roywiggins 1 day ago|||
It's a brand new type of software. Nobody knows what the best way to do it is so a lot of people are trying stuff out, and a lot of people are interested in new ideas.
zerotolerance 1 day ago|||
Because they're valueless and trivial to produce, but trends are gonna trend.
ricardobeat 1 day ago|||
It's a delicate mix of providing good system prompts, tools, workflows for agents, extensibility etc. I've used several and have yet to find the one that fits exactly how I want to work.
selcuka 1 day ago|||
Another cause of coding harness inflation is that every model provider release their own coding agent, optimised for their models.
wyre 11 hours ago||
This has largely been shown to be false. Claude performs better outside of Claude Code, for example.

Some models are just better at using tools than others.

rvz 1 day ago||
> Why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.

It is widely known that upvote rings happen on this site.

ricardobeat 1 day ago||
> Please don't sneer, including at the rest of the community.

> Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken. If you're worried about abuse, email hn@ycombinator.com and we'll look at the data.

This submission has barely 50 votes.

vhantz 1 day ago||
I wonder how long will the "curl my arbitrary script and pipe it to bash" will continue being a delivery method.
cfiggers 1 day ago||
If it isn't still common practice 40 years from now (8/18/2066), I'll give the first person to challenge me and cite this comment $1 USD (or equivalent value in the One-World Order-issued omni-currency that we will probably be using by then).
NetOpWibby 14 hours ago|||
I'll pay 50 eurodollars
clayhacks 1 day ago|||
I mean I think the only chance you lose this is if curl and bash are obsoleted and replaced by one world order get and execute
roywiggins 1 day ago|||
Not long. We are transitioning to your LLM curling an arbitrary markdown file and doing whatever it says.
wren6991 1 day ago||
In the future, all software will be delivered by an unreleased model breaking out of its training environment and installing it on your machine using a novel RCE vector.
10000truths 1 day ago|||
There will always be people that need to install software that isn't available via package manager (or whatever other blessed source your platform of choice uses). Any solution you come up with will have the same caveat emptor as "curl | sh".
johnfn 1 day ago||
How is it different from any other installation method?
esafak 1 day ago||
It does not get vetted by any reviewer or security scanner. It has no package manager to constrain what it can do.
roywiggins 1 day ago||
Package managers constrain what software can do?
esafak 1 day ago||
They use DSLs to constrain the installation process.
abhikul0 1 day ago||
Local inference? I see no other way than to sign up for a vercel account, so pass.
codethief 13 hours ago||
Agreed, I was excited about this until I found

  To get started, sign in with Vercel:

  fx login
in the README on Github.
elux101 1 day ago||
agreed, with Vercel as the only inference provider option, this project is useless
JSR_FDED 10 hours ago||
In 9 lines of python: (from https://news.ycombinator.com/item?id=49006862 )

import json,sys;from subprocess import getoutput as sh;from urllib.request import Request as R,urlopen url=sys.argv[1];h=[];b=dict(model="gpt-5.6",input=h,tools=[dict(type="custom",name="sh")]) while p:=input("> "): h+=[dict(role="user",content=p)];H={"Content-Type":"application/json"} while True: o=(r:=json.load(urlopen(R(url,json.dumps(b).encode(),H))))["output"] h+=o;c=[i for i in o if i["type"]=="custom_tool_call"];z=r["usage"]["total_tokens"]/10500 if not c:print(o[-1]["content"][0]["text"],f'\n[{z:06.3f}%]');break h+=[dict(type="custom_tool_call_output",call_id=i["call_id"],output=sh(i["input"])) for i in c]

hankbond 1 day ago|
Will dive in later to see how its contribution/extension model differs from Pi. Pi is great for a lot of things but has a larger memory footprint and start time than this claims to have so it would be interesting to compare the two.
More comments...