fx :Tiny, open, native coding agent.
143 points by handfuloflight 24 hours ago | 68 comments

rsyring 23 hours ago
For all the people asking "Why?", it seems like TFA has a pretty good list of features/attributes that it thinks sets it apart:

- fx is a coding agent harness and CLI written in Zig, optimized for research and embeddability as part of larger systems.

- It focuses on minimalism and performance across the board, from system prompt design, to its tools, feature set, and 6.39mib binary.

- For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI.

- It's open source (Apache-2.0), model-agnostic, and suitable for both local and cloud inference.

- Designed for instant installation and embedding in resource constrained environments and agent sandboxes.

- fx cold starts in 10µs and does no unnecessary work or I/O prior to accepting user input, making it ideal for programmatic use.

- Optimal fx.wasm builds produced by the Zig toolchain, which further reduce fx's size, making the network stack pluggable.

- fx contributes single-digit megabytes of memory baseline, allowing you to pack many instances in one machine.

- fx preserves scroll history by default, produces minimal output, and makes sparing use of complex TUI or paints

- Minimal system prompt and tools, to save on token costs and to yield optimal time-to-first-token performance (TTFT).

- Small core, extended via skills, plugins, MCPs, with a Unix-like philosophy to extensibility.

- Designed to work with local models, gateways, direct provider API access or subscriptions.

reply
OleksandrC 48 minutes ago
If you like this list of "why?", you might also like this: https://usehax.dev/ (I am the author). Most of the list applies, similar minimalist Unix tool approach, with some differences. Hax is written in C, the dynamically linked binary is even smaller (0.6 MB), MIT-licensed. No wasm though.

Important difference - fx is currently Vercel AI Gateway only - while hax does support multiple providers already (OpenAI API, ChatGPT/Codex subscription, Anthropic API, OpenRouter, OpenCode Zen/Go), and integrates well out of the box with local llama-server.

reply
Kim_Bruning 23 hours ago
Very neat. The demo on the page feels very intuitive to me! (if you're used to bash at least)
reply
jauntywundrkind 23 hours ago
I've only done a little of the new opencode v2 "mini" but it too offers a nice preserve-scroll by default.

OpenCode is the best behaved TUI i've seen by far (they invented OpenTUI to make it so good, also in Zig), so it feels less crucial. But it's nice to have there!

The "small core" model is very popular all of a sudden. DeepSeek's new harness is famously like that. https://news.ycombinator.com/item?id=49285244

OpenCode isn't quite as small, but there's very much been a deliberate attempt to drive much more into a plugin-based system. I enjoyed Dax talking about the new constitution of opencode, and the results of his agent comparing OpenCode & the new DeepSeek. https://bsky.app/profile/thdxr.com/post/3msy4gjttoc2f https://bsky.app/profile/thdxr.com/post/3msygiqyg6v2y

> an architectural change we made in opencode2 is nearly everything is an internal plugin / there's 68 of them that cover our built in agents, integrations, config loading, etc

i also think this is such a brilliant fun architectural twist too:

> OpenCode is the first time i could justify event sourcing in a real system / everything that happens is an event which gets projected into the sqlite db

https://bsky.app/profile/thdxr.com/post/3mt2qx3ktib2c

it's so fun seeing new malleable software cores emerge, try to figure out how to augment agency. agentic software striving itself to extend the agency it itself offers. it's been way too long since we've had ambitions to build general system, architectures that serve more than the user. this has held computing back for far too long. this is such an excellent interesting field, of such a more ambitious computing, opening up.

reply
solarkraft 22 hours ago
Thanks for the links on opencode 2! I’ve been meaning to get into this as I’ve been frustrated about some opencode 1’s behavior and design. Many of the encounters made me come up with ideas I’m happy to see they also had! As much fun as it may have been to build my own harness, I feel like the core primitives should be pretty well understood by now (in fact I envision a standard core library / API design taking shape).
reply
rvz 23 hours ago
Most of all what you have said is not really any clear differentiation against the rest of the 100s of other agents. Just minuscule or non-negligible implementation details and I'm afraid it is sadly yet another experimental slop project.

It is a branded "mee too" coding agent that we have seen hundreds of them already.

reply
rsyring 23 hours ago
I can't say I'm super into all the agents that are being created. But I do try to keep up here on HN and I can't say that I can remember any with this particular set of attributes.

In particular, aiming to be embeddable into other projects seems rather notable. At least, not something I've remembered of other projects that have made there way across the HN front page.

reply
qudat 20 hours ago
Pi is a composition of libraries that can be used to build agents. That seems far more interesting than this “minimal” agent.

This was written in zig and built by vercel. That’s the only notable characteristics about this project.

All code agents look the same and this one is no different.

reply
cmrdporcupine 23 hours ago
"- For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI."

I've actually been wondering lately why coding agent functionality isn't just... part of my shell already. Just another kind of interaction modality with an existing shell. Could probably even be an extension to fish or nu-shell even.

Please stop me from forking off on yet another project though.

reply
rsyring 23 hours ago
Originally, Warp was doing just that. Reimagining the terminal including making AI a part of it. I don't think they really found much purchase there because they ended up needing to make a platform out of it:

https://www.warp.dev/

reply
cmrdporcupine 23 hours ago
Yeah terminal window is a bit of a different story, though. I mean the actual shell binary.

With some ... intensive ... security/sandboxing/containerizing of some kind though, I guess.

reply
miguel_martin 25 minutes ago
Related: 3code is a coding agent (agentic loop) written in Nim (binary size: 1.6MiB): https://3code.capocasa.dev/
reply
bodge5000 22 hours ago
It does looks really interesting and definitely something I'll check out, but (genuine question), should "agent" and "agent harness" be used interchangeably as it is on here? It describes itself as an agent harness, but the tagline is "tiny, open, native coding agent".

I'm not sure harness is the right word either, but that seems to be what the industry has settled on so I'll concede on that, but surely the agent is the thing doing the work (which I guess is the model, or an instance of the model which is why agent is different?), whereas the harness is how the user interacts with the agent. We've had ways to describe that relationship before; client and server, frontend and backend, but again, I'll concede that the shiny new thing doesn't want to use boring old terminology, but I think some consistency and logic in the shiny new terminology is pretty important

That isn't specifically about fx of course, more of a general industry complaint

reply
stellalo 42 minutes ago
harness + llm = agent
reply
brap 16 minutes ago
Nowadays most “LLM” endpoints include some sort of server side harness as well, and I’d bet more than one model involved, so it’s really just agents all the way down
reply
cramforce 22 hours ago
Harness = the software the agent runs on. This is plain old software. You can trust it as much as any software. Agent = the thing that runs on the harness. It cannot be trusted because it is driven by an LLM
reply
bodge5000 22 hours ago
I get that, but wouldn't that make the agent and the harness two very different things, with the agent being closer to a model (the source of the agent) than a harness? Somebody else mentioned a console/game analogy, with the model being the disc, the agent being the running game, and the harness being the console OS, wouldn't it be like me describing Windows as a game, because games run on it?
reply
amdahl 22 hours ago
I think "harness" is a thing, the code/binary, and "agent" is a process, an instantiated run of that code/binary with a given LLM/env, etc.

So its like, "GTA 6" as the disc vs. the specific game you're in the middle of being chased by cops, harness vs agent.

In practice they're intertwined and it becomes hard not to use the terms somewhat interchangeably, but "you ask the agent how the harness works" vs. the other way around, clearly.

reply
bodge5000 22 hours ago
Wouldn't the disc in that analogy be the model? Thats the source of the instance. The harness I guess would be the OS of the console running the game
reply
amdahl 22 hours ago
It's more like in that analogy, the LLM is the gamer, instantiated as agent within a given game.

And the virtual world of GTA 6 is actually your codebase/env, the cops chasing are the bugs/angry customers, etc. The harness is providing an accurate/efficient ability for the model to understand/interact with the virtual world, flee the cops, etc. Decomposed at various architectural boundaries per your taste, but that's like loading a skin on the engine.

reply
bodge5000 22 hours ago
That doesn't seem right, surely the gamer would still be you, since you interact with the agent through the harness. If the player is the LLM, what is the human in this analogy? I guess a harness doesn't necessarily need human input (most do of course, but thats not a technical limitation), but then again neither does a game for the same reasons

Regardless though, this is what I mean, we now have 3 definitions for an agent; an instance of a model (which is how I think of it), a model configuration for a given task (from another commenter) and your definition which appears to be somewhere between the two, though it seems we agree with what a harness is.

reply
phiagent 21 hours ago
[dead]
reply
alansaber 22 hours ago
"Harness" should describe the overall system. The user interactions are increasingly negligible (due to model routing, adaptive reasoning, etc). "Agents" are the tool lists/settings provided to the model, etc.
reply
kgeist 23 hours ago
>Tiny ~6mb binary

I wonder why it's so large for a program written in Zig. It's basically just a loop that accepts user input, prepares the context, sends it to the LLM, parses the output, invokes the tools, and presents it all in the terminal. Add the built-in prompts and a few checks here and there (like blocking a write tool call before the file has been read first), and I'd expect a truly tiny native agent to be around 200-300 KB max.

reply
miguel_martin 20 minutes ago
fwiw, here's 3code which is 1.6MiB written in Nim - https://3code.capocasa.dev/
reply
lubitelpospat 35 minutes ago
I couldn’t find an option to connect to a generic OpenAI-compatible endpoint - did I miss it?
reply
SmashDan 24 hours ago
I'm not in the tech industry. Could someone explain why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.
reply
odo1242 23 hours ago
It's basically just a relatively simple to create piece of software that's important to get right (since you use it so much), can be made by many different design philosophies (maximal vs. minimal, customizability, etc.), and has very few good standards around it as of yet.
reply
chrysoprace 23 hours ago
Hacker news generally follows trends, and this is the current trend.

The discussion around coding agents nowadays is steering towards harnesses (which is probably a better description of what this is). "Agent" here is doing a lot of heavy lifting and has become a bit of a catch-all term to describe a model + harness + tooling + prompt + some other things that I've probably not thought about. The harness is a part that's being explored more as many believe it's where we can get some better performance out of the models.

This one in particular is from Vercel who provide a service to use models, so they have a vested interest in providing a harness.

reply
a2ff6eeb0 23 hours ago
Because we're actively exploring the best way to remove any need to deal with code, and make it so that you don't need any real talent to make a computer do things for anyone.

We haven't quite hit on the right formula yet, but people are very excited by the possibility.

reply
roywiggins 23 hours ago
It's a brand new type of software. Nobody knows what the best way to do it is so a lot of people are trying stuff out, and a lot of people are interested in new ideas.
reply
zerotolerance 23 hours ago
Because they're valueless and trivial to produce, but trends are gonna trend.
reply
ricardobeat 23 hours ago
It's a delicate mix of providing good system prompts, tools, workflows for agents, extensibility etc. I've used several and have yet to find the one that fits exactly how I want to work.
reply
selcuka 23 hours ago
Another cause of coding harness inflation is that every model provider release their own coding agent, optimised for their models.
reply
rvz 23 hours ago
> Why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.

It is widely known that upvote rings happen on this site.

reply
ricardobeat 22 hours ago
> Please don't sneer, including at the rest of the community.

> Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken. If you're worried about abuse, email hn@ycombinator.com and we'll look at the data.

This submission has barely 50 votes.

reply
vhantz 23 hours ago
I wonder how long will the "curl my arbitrary script and pipe it to bash" will continue being a delivery method.
reply
cfiggers 23 hours ago
If it isn't still common practice 40 years from now (8/18/2066), I'll give the first person to challenge me and cite this comment $1 USD (or equivalent value in the One-World Order-issued omni-currency that we will probably be using by then).
reply
NetOpWibby 2 hours ago
I'll pay 50 eurodollars
reply
clayhacks 22 hours ago
I mean I think the only chance you lose this is if curl and bash are obsoleted and replaced by one world order get and execute
reply
roywiggins 23 hours ago
Not long. We are transitioning to your LLM curling an arbitrary markdown file and doing whatever it says.
reply
wren6991 22 hours ago
In the future, all software will be delivered by an unreleased model breaking out of its training environment and installing it on your machine using a novel RCE vector.
reply
10000truths 21 hours ago
There will always be people that need to install software that isn't available via package manager (or whatever other blessed source your platform of choice uses). Any solution you come up with will have the same caveat emptor as "curl | sh".
reply
johnfn 22 hours ago
How is it different from any other installation method?
reply
esafak 22 hours ago
It does not get vetted by any reviewer or security scanner. It has no package manager to constrain what it can do.
reply
roywiggins 20 hours ago
Package managers constrain what software can do?
reply
esafak 20 hours ago
They use DSLs to constrain the installation process.
reply
abhikul0 13 hours ago
Local inference? I see no other way than to sign up for a vercel account, so pass.
reply
codethief 22 minutes ago
Agreed, I was excited about this until I found

  To get started, sign in with Vercel:

  fx login
in the README on Github.
reply
elux101 12 hours ago
agreed, with Vercel as the only inference provider option, this project is useless
reply
hankbond 24 hours ago
Will dive in later to see how its contribution/extension model differs from Pi. Pi is great for a lot of things but has a larger memory footprint and start time than this claims to have so it would be interesting to compare the two.
reply
NetOpWibby 2 hours ago
I'm most impressed with the domain name. GG

I thought my `disc.sh` was good but this is better.

reply
sixtyj 23 hours ago
I don’t want another coding agent. I would have to be unemployed to try all new software that pops up everywhere. :)

And every says it is the one :)

reply
paretolaw 37 minutes ago
Just another CLI? You could simply fork one existing, rename it, and save tokens
reply
smy20011 23 hours ago
I was wondering to build the same thing but someone already built it. I just need some agent that open/close super fast and don't eat half a gb of memory.
reply
konaraddi 23 hours ago
Nice to see herdr support. The docs seem to imply it only works with vercel’s AI gateway? I guess it cant be configured to use another provider?
reply
chrysoprace 23 hours ago
It's a Vercel project, so I assume this is intentional. The steps to get started are to log in with Vercel or add an AI Gateway key.
reply
fazxes 23 hours ago
support incoming for subscriptions (Codex, Grok)
reply
impulser_ 23 hours ago
I'm sorry, but this is pure slop. This has 26 tools and a tool for every single file operation and a tool for read tool output? What the fuck... And they call this minimalist.... Lmfao

The person who built this obviously has little understanding of harnesses.

You should have significantly less tools today with how good LLMs have become.

The start up time and binary size are quite literally the most useless stats to base a harness off of lol

Man what with Cloudflare, Vercel and all these tech companies just releasing pure slop.

Just use Pi. It's actually minimal and well thought out by people who actually understand agents.

reply
maherbeg 2 hours ago
I think part of this is to enable the interface on the web and other devices where you might not have a terminal. But yeah, agreed, it's wayyyy too many tools.
reply
jauntywundrkind 24 hours ago
Is Vercel so all-in on Zig elsewhere too?
reply
fazxes 23 hours ago
fx isn’t primarily “another coding agent.” It’s a tiny, embeddable agent harness and infrastructure component that also happens to have a good CLI.
reply
rvz 24 hours ago
Some questions:

1. Why do we need yet another coding agent over the rest of them?

2. Is this going to be another Vercel Labs slop project that they will abandon like the others since this is super experimental?

reply
alexcommitter 4 hours ago
[flagged]
reply
Alephinitesimal 21 hours ago
[dead]
reply
fenestella 23 hours ago
[dead]
reply
kevinbaiv 23 hours ago
[dead]
reply
wonderwaffle27 24 hours ago
[dead]
reply
catlover76 24 hours ago
[dead]
reply
parisiansam 16 hours ago
i think there is a name clash here, i like fx the json viewer https://github.com/antonmedv/fx (20.6k star on gh)
reply
alexboehm 22 hours ago
I too was frustrated with needing npm and slow startups or huge rust compile times for agents. I tried getting agents to write a tool like this with proper raw mode content pasting/ interruptions, but they just kept screwing it up without a framework like ratatui, so I wrote one in c by hand https://gist.github.com/fourlexboehm/a60e4ef9306744483731cd1... the only dependency is libcurl.

This binary is ~40kb and uses much less ram than fx.

reply
handfuloflight 18 hours ago
"it can do everything Claude code can do" This is a tad bit hyperbolic.
reply
alexboehm 14 hours ago
Fair point, but name one thing you can't do with bash.
reply
smy20011 20 hours ago
Very clean, thank you for posting!
reply