• For all the people asking "Why?", it seems like TFA has a pretty good list of features/attributes that it thinks sets it apart:

    - fx is a coding agent harness and CLI written in Zig, optimized for research and embeddability as part of larger systems.

    - It focuses on minimalism and performance across the board, from system prompt design, to its tools, feature set, and 6.39mib binary.

    - For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI.

    - It's open source (Apache-2.0), model-agnostic, and suitable for both local and cloud inference.

    - Designed for instant installation and embedding in resource constrained environments and agent sandboxes.

    - fx cold starts in 10µs and does no unnecessary work or I/O prior to accepting user input, making it ideal for programmatic use.

    - Optimal fx.wasm builds produced by the Zig toolchain, which further reduce fx's size, making the network stack pluggable.

    - fx contributes single-digit megabytes of memory baseline, allowing you to pack many instances in one machine.

    - fx preserves scroll history by default, produces minimal output, and makes sparing use of complex TUI or paints

    - Minimal system prompt and tools, to save on token costs and to yield optimal time-to-first-token performance (TTFT).

    - Small core, extended via skills, plugins, MCPs, with a Unix-like philosophy to extensibility.

    - Designed to work with local models, gateways, direct provider API access or subscriptions.

    • If you like this list of "why?", you might also like this: https://usehax.dev/ (I am the author). Most of the list applies, similar minimalist Unix tool approach, with some differences. Hax is written in C, the dynamically linked binary is even smaller (0.6 MB), MIT-licensed. No wasm though.

      Important difference - fx is currently Vercel AI Gateway only - while hax does support multiple providers already (OpenAI API, ChatGPT/Codex subscription, Anthropic API, OpenRouter, OpenCode Zen/Go), and integrates well out of the box with local llama-server.

      • It’s funny what “tiny” means for different people depending on their background. I expected it to be under an MB as well and was surprised by 6MB.
        • Basically, for me tiny means it fits on a floppy 1.44 HD.

          Naturally meaningless when people carry around USB sticks that might even hold a 1 TB, but alas.

          • > Naturally meaningless when people carry around USB sticks

            Do people do that? I think it was a decade ago, I thought people download from web nowadays.

            • Not everyone lives in areas where that is a convenient option.
          • Pretty much my baseline as well. Good example how experiences shape our thinking.
        • For comparison, I publish a CLI tool written in Dart, can compile to native anything including Wasm , the Linux binary is around 5MB and it does quite a lot.
        • A typical Go binary would be 2-3 times larger. A typical Node project would easily pull more than 6 MiB of just code, not counting the runtime.
      • I really like the philosophy document! https://github.com/OleksandrChekhovskyi/hax/blob/master/docs...

        I don’t know if I am ready to use a new tool written in C and using libcurl but I will give it a shot.

      • that's an amazing project! people always say why 6mb vs CC's 250mb even matter when you are calling out to LLMs hosted in the cloud. But... I regularly run hierarchies of agents with say 50-100 on a regular basis. So 650 vs 25050 ... is "can do" vs "cannot"
        • Is that real memory taken? If it's shared code from the same executable surely the multiples are not very relevant?
      • this looks great! what are you using it for? i like the idea of being able to use one of these (sandboxed) within a larger program kind of like how I use LLM's to do small tasks within my apps now but with a few tools (web search). my current way of doing that is like building a mini-harness with a couple tools within the app, but something more drop-in would be better obviously.
      • It looks really nice. Do you have any plan to support Claude subscriptions (pro/max)?
        • Technically, this would be straightforward. The problem is that Anthropic seems to be really against using Claude subscriptions with anything other than Claude Code - you might even risk your account getting banned for doing so. You could search online for the "openclaw claude banned" for more details on that story.
        • Didn't Anthropic stop allowing it, can only use the subscriptions with first-party tools?
    • Does it have anything co-designed around Vercel infrastructure? This is what happened to NextJS and why I will likely never touch Vercel open source again
      • Yeah, I see the Vercel logo, and I am instantly out.
    • Very neat. The demo on the page feels very intuitive to me! (if you're used to bash at least)
    • I've only done a little of the new opencode v2 "mini" but it too offers a nice preserve-scroll by default.

      OpenCode is the best behaved TUI i've seen by far (they invented OpenTUI to make it so good, also in Zig), so it feels less crucial. But it's nice to have there!

      The "small core" model is very popular all of a sudden. DeepSeek's new harness is famously like that. https://news.ycombinator.com/item?id=49285244

      OpenCode isn't quite as small, but there's very much been a deliberate attempt to drive much more into a plugin-based system. I enjoyed Dax talking about the new constitution of opencode, and the results of his agent comparing OpenCode & the new DeepSeek. https://bsky.app/profile/thdxr.com/post/3msy4gjttoc2f https://bsky.app/profile/thdxr.com/post/3msygiqyg6v2y

      > an architectural change we made in opencode2 is nearly everything is an internal plugin / there's 68 of them that cover our built in agents, integrations, config loading, etc

      i also think this is such a brilliant fun architectural twist too:

      > OpenCode is the first time i could justify event sourcing in a real system / everything that happens is an event which gets projected into the sqlite db

      https://bsky.app/profile/thdxr.com/post/3mt2qx3ktib2c

      it's so fun seeing new malleable software cores emerge, try to figure out how to augment agency. agentic software striving itself to extend the agency it itself offers. it's been way too long since we've had ambitions to build general system, architectures that serve more than the user. this has held computing back for far too long. this is such an excellent interesting field, of such a more ambitious computing, opening up.

      • Thanks for the links on opencode 2! I’ve been meaning to get into this as I’ve been frustrated about some opencode 1’s behavior and design. Many of the encounters made me come up with ideas I’m happy to see they also had! As much fun as it may have been to build my own harness, I feel like the core primitives should be pretty well understood by now (in fact I envision a standard core library / API design taking shape).
    • rvz
      Most of all what you have said is not really any clear differentiation against the rest of the 100s of other agents. Just minuscule or non-negligible implementation details and I'm afraid it is sadly yet another experimental slop project.

      It is a branded "mee too" coding agent that we have seen hundreds of them already.

      • I can't say I'm super into all the agents that are being created. But I do try to keep up here on HN and I can't say that I can remember any with this particular set of attributes.

        In particular, aiming to be embeddable into other projects seems rather notable. At least, not something I've remembered of other projects that have made there way across the HN front page.

        • Pi is a composition of libraries that can be used to build agents. That seems far more interesting than this “minimal” agent.

          This was written in zig and built by vercel. That’s the only notable characteristics about this project.

          All code agents look the same and this one is no different.

          • I don't think Pi is tiny because it's written in typescript and requires a javascript runtime. I definitely want a "tiny" compiled agent with no runtime requirements.

            I don't think Zig is that great of a language for this, but it's better than typescript. I don't want to use Vercel software so will pass, but would love to see a more community driven effort.

            • What’s a better lang in your opinion?
        • Almost every harness has an SDK now, it seems part of the "mvp" at this point, both open and closed source

          https://learn.chatgpt.com/docs/codex-sdk

          https://code.claude.com/docs/en/agent-sdk/overview

          https://opencode.ai/docs/sdk/

          https://pi.dev/docs/latest/sdk

          or if you want SDK first, my recommendation is https://adk.dev/

    • "- For end users, its CLI output style and form factor aims to be closer to a Unix shell than a heavy "IDE in the terminal" TUI."

      I've actually been wondering lately why coding agent functionality isn't just... part of my shell already. Just another kind of interaction modality with an existing shell. Could probably even be an extension to fish or nu-shell even.

      Please stop me from forking off on yet another project though.

  • It may come across harsh but IMHO the only interesting thing about this project is that it is written in Zig. That's it.

    Everything else in the harness is largely the same just Vercel-flavoured.

    The portability benefit is also a bit over-sold imho. I wrote a harness in Go and it is as portable as this ... in fact it deploys straight into Vercel's own sandbox environment on demand without any issues.

    That being said, did you say GLM 5.2 free? I need to look into that. GLM is remarkably capable model and that alone is worth using the harness in my books.

    • This is exactly how I work today. I've always lived in the terminal but typing bash commands anymore is way too much work and an LLM can produce way better one liners than I can.
    • UX-wise it's quite minimalistic, that seems one of their differentiators. And I'm liking it!
  • Since I've seen some other agent tools being suggested, let me throw Maki in the ring as well: https://maki.sh/

    - written in rust, super fast startup and rendering

    - implements token saving techniques

    - plugins written in lua

    No affiliation, just think it's a really well made piece of software.

    • Can it replace pi? Would be good to have a robust and fast replacement.
    • been enjoying maki too. respects xdg spec (unlike pi), the ui is snappy and pretty much exactly what i want, and i like the choice to keep the lua API similar-ish to neovim.
    • This has some nice ideas like indexing (aider like repo map) built in, could be interesting
    • Love using maki on 500mb and 1gb ram tiny vps's where opencode sometimes could go OOM.

      I really love maki, its awesome!

  • >Tiny ~6mb binary

    I wonder why it's so large for a program written in Zig. It's basically just a loop that accepts user input, prepares the context, sends it to the LLM, parses the output, invokes the tools, and presents it all in the terminal. Add the built-in prompts and a few checks here and there (like blocking a write tool call before the file has been read first), and I'd expect a truly tiny native agent to be around 200-300 KB max.

    • fwiw, here's 3code which is 1.6MiB written in Nim - https://3code.capocasa.dev/
    • I was looking it last night and the repo is something like 600k lines of Zig. Maybe 500k after comments and blank lines.

      In my own experiments to build a tiny zig agent it came out to under 800kb.

      • I'm always surprised how large harnesses are

        By comparison, the recently open sourced grok harness was 840k lines, codex 950k, pi 250k https://simonwillison.net/2026/Jul/15/grok-build/

      • I think the problem is that when you use agents to write Zig it will bruteforce code, because there isn't that much good reference code to work off of.
    • So large? Really? Allow me to introduce you to electron.
      • Yes, large. I haven't used Zig much myself, but from a few experiments I ran, Zig handles dead code elimination exceptionally well. It compiled a full Win32 GUI calc app that used Capy (a full, cross-platform GUI framework) into a 133kb executable. Removing Capy completely and using Win32 APIs directly produced an even smaller (93kb) binary (it also removed some DLL dependencies, leaving basically only ntdll.dll). For the same task, Rust + Slint produced a 4.7 MB binary that still depended on multiple (non-Windows-provided) shared libraries.

        Given another commenter's mention of a similar project written in Nim that yielded a 1.6 MB binary, my first guess is that the 6 MB Zig binary simply isn't optimized for size - it might be a debug build. If not that, then I'm not sure what's happening, but yeah, in the context of Zig, 6mb for a CLI app is a bit strange.

      • Those of us educated in 70's and 80's home computing have several nice words for stuff like Electron.
  • It does looks really interesting and definitely something I'll check out, but (genuine question), should "agent" and "agent harness" be used interchangeably as it is on here? It describes itself as an agent harness, but the tagline is "tiny, open, native coding agent".

    I'm not sure harness is the right word either, but that seems to be what the industry has settled on so I'll concede on that, but surely the agent is the thing doing the work (which I guess is the model, or an instance of the model which is why agent is different?), whereas the harness is how the user interacts with the agent. We've had ways to describe that relationship before; client and server, frontend and backend, but again, I'll concede that the shiny new thing doesn't want to use boring old terminology, but I think some consistency and logic in the shiny new terminology is pretty important

    That isn't specifically about fx of course, more of a general industry complaint

    • Maybe ”agent shell” could be a better term for harness. But it’s pretty overloaded..
    • Harness = the software the agent runs on. This is plain old software. You can trust it as much as any software. Agent = the thing that runs on the harness. It cannot be trusted because it is driven by an LLM
      • I get that, but wouldn't that make the agent and the harness two very different things, with the agent being closer to a model (the source of the agent) than a harness? Somebody else mentioned a console/game analogy, with the model being the disc, the agent being the running game, and the harness being the console OS, wouldn't it be like me describing Windows as a game, because games run on it?
    • I think "harness" is a thing, the code/binary, and "agent" is a process, an instantiated run of that code/binary with a given LLM/env, etc.

      So its like, "GTA 6" as the disc vs. the specific game you're in the middle of being chased by cops, harness vs agent.

      In practice they're intertwined and it becomes hard not to use the terms somewhat interchangeably, but "you ask the agent how the harness works" vs. the other way around, clearly.

      • Wouldn't the disc in that analogy be the model? Thats the source of the instance. The harness I guess would be the OS of the console running the game
        • It's more like in that analogy, the LLM is the gamer, instantiated as agent within a given game.

          And the virtual world of GTA 6 is actually your codebase/env, the cops chasing are the bugs/angry customers, etc. The harness is providing an accurate/efficient ability for the model to understand/interact with the virtual world, flee the cops, etc. Decomposed at various architectural boundaries per your taste, but that's like loading a skin on the engine.

          • That doesn't seem right, surely the gamer would still be you, since you interact with the agent through the harness. If the player is the LLM, what is the human in this analogy? I guess a harness doesn't necessarily need human input (most do of course, but thats not a technical limitation), but then again neither does a game for the same reasons

            Regardless though, this is what I mean, we now have 3 definitions for an agent; an instance of a model (which is how I think of it), a model configuration for a given task (from another commenter) and your definition which appears to be somewhere between the two, though it seems we agree with what a harness is.

      • [dead]
    • "Harness" should describe the overall system. The user interactions are increasingly negligible (due to model routing, adaptive reasoning, etc). "Agents" are the tool lists/settings provided to the model, etc.
    • harness + llm = agent
      • Nowadays most “LLM” endpoints include some sort of server side harness as well, and I’d bet more than one model involved, so it’s really just agents all the way down
    • I personally like the way Scion breaks it down into

      - model

      - harness (tools/config)

      - agent (live/running)

      https://googlecloudplatform.github.io/scion/concepts/

      Scion allows for multiple configurations of a harness, allowing you to configure the same tool differently based on what you are doing, particularly important when you want to restrict permissions.

      https://googlecloudplatform.github.io/scion/supported-harnes...

      for the unfamiliar, Scion is an OpenClaw like platform from a Google dev, not supported or sponsored by the company

  • I'm not in the tech industry. Could someone explain why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.
    • When a new technique or capability arrives on the tech scene, there's a point in the invention-to-diffusion story when the new thing becomes accessible (e.g. cheap and/or easy) enough for a broader audience of developers to experiment with it... but before anyone's figured out best practices, let alone polished products/projects, or calcified around a market leader.

      So you get a Cambrian explosion of weird little projects. Ultimately, one of them will probably become the "market" leader... or at least the market default.

      Right now there's a lot of agent harnesses and sandbox projects floating about.

      Fun examples from the past: text editors, window managers, IRC clients, blogging engines (first static, then dynamic, then static again), Twitter clients... every programming language community has weird clusters of library/framework duplication in their history...

      Sometimes these projects take on a rite of passage flavour... like, as every Jedi builds their own lightsaber, every developer builds their own... blog? That used to be the obvious one. Less so these days.

    • Because a lot of people are writing their own to get a tool that they understand and can manipulate as they like. So they like to share them and see what other people have done to learn from. As a community we are still very far from coming to a consensus on what a good harness looks like and the only way, IMO, to get a good feel for it is to write your own.
      • It’s also exactly what the tech enables. It’s not hard to imagine there being hundreds of thousands of different harness projects, if not more.
    • It's basically just a relatively simple to create piece of software that's important to get right (since you use it so much), can be made by many different design philosophies (maximal vs. minimal, customizability, etc.), and has very few good standards around it as of yet.
    • Hacker news generally follows trends, and this is the current trend.

      The discussion around coding agents nowadays is steering towards harnesses (which is probably a better description of what this is). "Agent" here is doing a lot of heavy lifting and has become a bit of a catch-all term to describe a model + harness + tooling + prompt + some other things that I've probably not thought about. The harness is a part that's being explored more as many believe it's where we can get some better performance out of the models.

      This one in particular is from Vercel who provide a service to use models, so they have a vested interest in providing a harness.

    • Because we're actively exploring the best way to remove any need to deal with code, and make it so that you don't need any real talent to make a computer do things for anyone.

      We haven't quite hit on the right formula yet, but people are very excited by the possibility.

    • Because we are in a gold rush and the best thing to sell are the shovels.
    • It's a brand new type of software. Nobody knows what the best way to do it is so a lot of people are trying stuff out, and a lot of people are interested in new ideas.
    • Because they're valueless and trivial to produce, but trends are gonna trend.
    • It's a delicate mix of providing good system prompts, tools, workflows for agents, extensibility etc. I've used several and have yet to find the one that fits exactly how I want to work.
    • Another cause of coding harness inflation is that every model provider release their own coding agent, optimised for their models.
      • This has largely been shown to be false. Claude performs better outside of Claude Code, for example.

        Some models are just better at using tools than others.

    • > Why there are so many new coding agents, and why they're commonly upvoted on HackerNews? It seems like there's a new one in the top 10 every other day.

      It is widely known that upvote rings happen on this site.

      • > Please don't sneer, including at the rest of the community.

        > Please don't post insinuations about astroturfing, shilling, brigading, foreign agents, and the like. It degrades discussion and is usually mistaken. If you're worried about abuse, email hn@ycombinator.com and we'll look at the data.

        This submission has barely 50 votes.

  • I wonder how long will the "curl my arbitrary script and pipe it to bash" will continue being a delivery method.
    • If it isn't still common practice 40 years from now (8/18/2066), I'll give the first person to challenge me and cite this comment $1 USD (or equivalent value in the One-World Order-issued omni-currency that we will probably be using by then).
      • I'll pay 50 eurodollars
      • I mean I think the only chance you lose this is if curl and bash are obsoleted and replaced by one world order get and execute
    • Not long. We are transitioning to your LLM curling an arbitrary markdown file and doing whatever it says.
      • In the future, all software will be delivered by an unreleased model breaking out of its training environment and installing it on your machine using a novel RCE vector.
    • There will always be people that need to install software that isn't available via package manager (or whatever other blessed source your platform of choice uses). Any solution you come up with will have the same caveat emptor as "curl | sh".
    • How is it different from any other installation method?
      • It does not get vetted by any reviewer or security scanner. It has no package manager to constrain what it can do.
        • Package managers constrain what software can do?
          • They use DSLs to constrain the installation process.
  • Local inference? I see no other way than to sign up for a vercel account, so pass.
    • Agreed, I was excited about this until I found

        To get started, sign in with Vercel:
      
        fx login
      
      in the README on Github.
    • agreed, with Vercel as the only inference provider option, this project is useless
  • In 9 lines of python: (from https://news.ycombinator.com/item?id=49006862 )

    import json,sys;from subprocess import getoutput as sh;from urllib.request import Request as R,urlopen url=sys.argv[1];h=[];b=dict(model="gpt-5.6",input=h,tools=[dict(type="custom",name="sh")]) while p:=input("> "): h+=[dict(role="user",content=p)];H={"Content-Type":"application/json"} while True: o=(r:=json.load(urlopen(R(url,json.dumps(b).encode(),H))))["output"] h+=o;c=[i for i in o if i["type"]=="custom_tool_call"];z=r["usage"]["total_tokens"]/10500 if not c:print(o[-1]["content"][0]["text"],f'\n[{z:06.3f}%]');break h+=[dict(type="custom_tool_call_output",call_id=i["call_id"],output=sh(i["input"])) for i in c]

  • Will dive in later to see how its contribution/extension model differs from Pi. Pi is great for a lot of things but has a larger memory footprint and start time than this claims to have so it would be interesting to compare the two.
  • Just tried and although it’s not that polished or fast imo it’s very cool.

    GLM 5.2 totally free even with a free Vercel account.

    Make business sense too - fx is the entry point to bring more user to Vercel AI.

  • I don’t want another coding agent. I would have to be unemployed to try all new software that pops up everywhere. :)

    And every says it is the one :)

  • I like it from trying it quickly, but is there any way to use it with a provider other than vercel?
  • I too was frustrated with needing npm and slow startups or huge rust compile times for agents. I tried getting agents to write a tool like this with proper raw mode content pasting/ interruptions, but they just kept screwing it up without a framework like ratatui, so I wrote one in c by hand https://gist.github.com/fourlexboehm/a60e4ef9306744483731cd1... the only dependency is libcurl.

    This binary is ~40kb and uses much less ram than fx.

    • "it can do everything Claude code can do" This is a tad bit hyperbolic.
      • Fair point, but name one thing you can't do with bash.
    • Very clean, thank you for posting!
  • They should spend their efforts making a harness that utilizes tiny AI models on the host in unison with the frontier model, to improve intelligence and capabilities. This is an under-explored area.
  • I couldn’t find an option to connect to a generic OpenAI-compatible endpoint - did I miss it?
  • Related: 3code is a coding agent (agentic loop) written in Nim (binary size: 1.6MiB): https://3code.capocasa.dev/
  • I was wondering to build the same thing but someone already built it. I just need some agent that open/close super fast and don't eat half a gb of memory.
  • I'm most impressed with the domain name. GG

    I thought my `disc.sh` was good but this is better.

  • fx isn’t primarily “another coding agent.” It’s a tiny, embeddable agent harness and infrastructure component that also happens to have a good CLI.
  • Nice to see herdr support. The docs seem to imply it only works with vercel’s AI gateway? I guess it cant be configured to use another provider?
    • It's a Vercel project, so I assume this is intentional. The steps to get started are to log in with Vercel or add an AI Gateway key.
    • support incoming for subscriptions (Codex, Grok)
  • i think there is a name clash here, i like fx the json viewer https://github.com/antonmedv/fx (20.6k star on gh)
  • Is Vercel so all-in on Zig elsewhere too?
    • I'd say it's more about the popularity of `ghostty` and what they've showcased with terminal achievements in Zig.

      Zig is totally cool to do this in-- no shade intended by this.

  • I'm sorry, but this is pure slop. This has 26 tools and a tool for every single file operation and a tool for read tool output? What the fuck... And they call this minimalist.... Lmfao

    The person who built this obviously has little understanding of harnesses.

    You should have significantly less tools today with how good LLMs have become.

    The start up time and binary size are quite literally the most useless stats to base a harness off of lol

    Man what with Cloudflare, Vercel and all these tech companies just releasing pure slop.

    Just use Pi. It's actually minimal and well thought out by people who actually understand agents.

    • Bit harsh. AFAIK, you cant embed pi in a webpage (wasm). That itself is a huge win for fx.
    • I think part of this is to enable the interface on the web and other devices where you might not have a terminal. But yeah, agreed, it's wayyyy too many tools.
  • finally some good sofware
  • Just another CLI? You could simply fork one existing, rename it, and save tokens
  • Some questions:

    1. Why do we need yet another coding agent over the rest of them?

    2. Is this going to be another Vercel Labs slop project that they will abandon like the others since this is super experimental?

  • [flagged]
  • [dead]