\

Show HN: Recurse – Develop and deploy specialist agents faster

5 points - yesterday at 10:06 PM


Hi HN! We are looking to gather some feedback on our serverless agent harness. The admittedly not-so-specific use case is to accelerate agent development and deployment. After building several custom/special-purpose agents for a few customers, we built this to accelerate our workflow at first, and now we are trying to understand whether it could be useful to others.

Our driver use case was development of specialist agents with a request/response lifecycle. Think of agents that have a well established input/output contract where they are expected to produce high-quality output (artifacts, responses etc.). Especially when the problem is in some verifiable domain and the LLM can iteratively refine a result to a final value that satisfies constraints or optimizes some goal.

The product is a coding agent skill + a serverless execution runtime with a harness that takes in a system prompt + Python functions as tools. The coding agent takes in the requirements from the user, and tries agent variants by executing prompt/tool variants it creates.

It works best for cases where you can think of how you can evaluate a candidate agent - when you describe this information to your coding agent, it often does a decent job at building candidate prompts, tools and even benchmarks.

Prompts and Python tools that the coding agent creates integrate with a harness that implements an FSM that is tuned to drive an iterative refinement process for verifiable domains. This tuning enables one to use small models like Luna to produce high quality results while keeping costs at a manageable level.

Our website is not 100% complete yet (some examples are missing write-ups, not all use cases we tried are there etc.), but the system is operational and docs are there.

Thanks!

Source
  • adityamishra241

    today at 11:38 AM

    Interesting approach. The iterative refinement + verifiable constraints combination seems especially useful. Curious what kinds of tasks have worked best in practice.

    • beecasthurlbow

      today at 2:29 AM

      It seems like you're saying "This is a platform to iterate on your agents and make them better over time", but the headline of your site does not communicate that at all.

      Your website is full of AI-isms, which makes it hard to parse what you're doing. (3 punchy sentence headline, "X is not Y." , "never able to claim").

      Sounds like you have a cool product, try writing at least the front page of your site yourself and it'll be much much more engaging.

      What does "changing the specialist" mean? changing the model? The prompt?

      • williamse

        today at 8:56 AM

        [flagged]

        • yesterday at 10:06 PM

          • ailephant

            today at 3:59 AM

            [dead]

            • logan-hogg

              today at 1:42 AM

              [flagged]