Charles Muehlberger
Founder of Conifer, Least cost routing system to reduce 70%+ token spend
Company
Charles Muehlberger is listed as a founder of Conifer. There are thousands of models, providers, and harnesses, each with its own pricing and strengths, leaving an overwhelming number of choices when it comes to managing token spend and intelligence. Conifer uses intelligent routing and orchestration logic to handle the full path of each query: model selection, provider choice, and cache management. It all runs inside your existing harness, so tools like Claude Code and Codex work without changing your setup. By centralizing where inference occurs, Conifer lets companies save on their inference bill while leaving teams free to focus on what they are building.
Founder traction evidence
Public posts and activity attributed directly to this founder.
- X
Little peak into what we've been working on: - Run any model from any harness from one API key (claude code, codex, open code, pi, etc) - Access to new fusion models pushing the cost and performance frontier - Local-cloud routing with configurable model pools! Very very soon
Little peak into what we've been working on: - Run any model from any harness from one API key (claude code, codex, open code, pi, etc) - Access to new fusion models pushing the cost and performance frontier - Local-cloud routing with configurable model pools! Very very soon
- X
The team at @TryOpenTag is truly N of 1 It's great to see more apps building around model agnostic endpoints and we're happy we get to be a part of their journey.
The team at @TryOpenTag is truly N of 1 It's great to see more apps building around model agnostic endpoints and we're happy we get to be a part of their journey. Check them out at https://t.co/TrE6LDoNwx and if you're looking to hedge against singleton model buildout check out
- X
Today we’re bringing a better user experience to model selection and agent runtime.
Today we’re bringing a better user experience to model selection and agent runtime. Everyone agrees you should use the best model. But the current infrastructure makes this impossible. Look at what actually exists: the leading harnesses offer a handful of models. Open-weight and
- X
Reading over an old report from @EpochAIResearch.
Reading over an old report from @EpochAIResearch. The trend lines suggest we'll achieve a model of comparable intelligence to Fable/Sol that you'll be able to host on a single consumer GPU.
- X
Its been super fun playing with these models - post training has been so well documented and easy to use: - 220 tok/s on-device - open weights - continuous local embedding at near zero marginal cost You can really make these customizable for any specialized task
Its been super fun playing with these models - post training has been so well documented and easy to use: - 220 tok/s on-device - open weights - continuous local embedding at near zero marginal cost You can really make these customizable for any specialized task
- X
Fresh essay on scaling laws. Sara Hooker spent years inside DeepMind and Cohere watching scaling laws get treated as gospel. Then she wr...
Fresh essay on scaling laws. Sara Hooker spent years inside DeepMind and Cohere watching scaling laws get treated as gospel. Then she wrote down the heresy: Falcon 180B was state of the art in 2023. One year later it lost to a model 22x smaller. The $700B question isn't whether...
- X
I'm seeing more and more detailed conversations about a new model that neither person has used.
I'm seeing more and more detailed conversations about a new model that neither person has used. That didn't used to be possible. Benches made it possible.
- X
Super excited to open this up to the public.
Super excited to open this up to the public. Any model in any harness (Pi, Codex, Claude Code, Hermes...) Tools, subagents, plugins, everything works out of the box. Super excited to hear feedback and let us know if you break it or we missed your workspace. macOS: curl -fsSL