11 comments

  • YuechenLi 9 minutes ago ago

    Have you compared this to using GPT-6.1 Sol instead of GPT 6 Astra + Deepseek? From my test, 6.1 Sol is a lot more token efficient than 6 Sol while being similar to Astra in performance, and I don't really find 6 Astra to be significantly better than 6/6.1 Sol for general coding as I feel 6 Astra is only noticeably better at spatial reasoning/vision compared to 6 Sol, and 6.1 Sol really closed the gap on that front.

  • ajspig1 32 minutes ago ago

    How do you handle provider variance on OpenRouter for the opensource models? Or do you use your own hosted version to mitigate this?

    And for both opensource and closed source, does the router account for provider quality, or catch it when a provider degrades?

  • rirze an hour ago ago

    How does this choose which models to use with any arbitrary set of model providers to work from? And why is an openrouter necessary for self-hosting?

  • jamesforestwest 2 hours ago ago

    How do you define the model buckets, and what happens when a session genuinely needs a model that isn't in the bucket the HMM picked?

  • thefourthchime 3 hours ago ago

    Interesting work, and thanks for describing how your router works internally. It's definitely a fascinating subject. How would you say this compares to Cursor's auto mode?

    • adchurch 2 hours ago ago

      Absolutely!

      Conceptually very similar to Cursor's auto mode. The key distinctions are:

      - We plug into any harness (e.g. Claude Code, Codex, OpenCode, Pi)

      - We aren't incentivized to route to our own model, we're incentivized to route to the best model whatever it may be

    • aschla 3 hours ago ago

      And similarly, Copilot’s Auto mode?

  • 1minusp an hour ago ago

    Does this allow for a predefined budget?

  • redrove 3 hours ago ago

    Is the model you trained available as open weights?

    • adchurch 2 hours ago ago

      It is not sorry!

  • aminsamir45 2 hours ago ago

    AGI is here!