HomeCompaniesConifer
Conifer

Least cost routing system to reduce 70%+ token spend

There are thousands of models, providers, and harnesses, each with its own pricing and strengths, leaving an overwhelming number of choices when it comes to managing token spend and intelligence. Conifer uses intelligent routing and orchestration logic to handle the full path of each query: model selection, provider choice, and cache management. It all runs inside your existing harness, so tools like Claude Code and Codex work without changing your setup. By centralizing where inference occurs, Conifer lets companies save on their inference bill while leaving teams free to focus on what they are building.
Active Founders
Charles Muehlberger
Charles Muehlberger
Founder/CTO
Founder at Conifer. Engineering researcher at Princeton, accelerating multi-modal inference on edge devices and designing physics simulation software for coupled laser reservoirs and machine learning with physical neural networks. Previously built custom edge AI devices for RF-based brain injury modeling at DoW, and helped engineer the world record holding fastest electric speedboat at PES.
Michael Jeffords
Michael Jeffords
Founder/CGO
Co-founder at Conifer (YC S26). Previously co-founded The HeartCheck Foundation, scaling free hypertension screenings to 20+ barbershops and over 10,000 individuals. Engineered an ML-driven computer vision pipeline using pose estimation for early-onset ALS and Parkinson's detection. Former clinical researcher at the Dell Med Institute of Vascular Surgery and UTHSC, specializing in hematopoetic stem cell transplantation, achieving infiltration levels >90% in mouse models.
Company Launches
Conifer: Local-first least cost routing system
See original launch post

Hi all 👋, We’re Michael and Charles, founders of Conifer.

TLDR: Conifer is one interface for both local and cloud models, available today as Juniper, our consumer-facing app. It routes each request to the most cost-effective option, beginning with your own hardware. Most requests are processed locally without API fees, reducing paid token volume by up to 80%.

📺 Launch Video: https://www.youtube.com/watch?v=QLVNISmtet0

❓ The problem: Model labs direct every request to the cloud, with no incentive to route your queries to a cheaper competitor or a free, on-device model. That may benefit them, but it’s unnecessarily costly for you. Whether it’s a simple typo fix or a complex architecture problem, the request is sent to a data center, exposing your data to third parties and increasing your monthly spend. This quickly adds up for teams running coding agents, customer support, and other high-volume workloads. When handling financials, patient records, or customer data, teams must also trust that a cloud provider’s security is as robust as promised.

uploaded image

🔨 What Conifer does: Conifer routes every request from the bottom up:

Tier 0: Your own hardware, with $0 in API fees

Tier 1: An efficient cloud model

Tier 2: A frontier model for the most demanding queries

The majority of requests are processed locally, ensuring that cloud pricing applies only when actually necessary, drastically reducing paid token volume. Additionally, Conifer consolidates disparate subscriptions, dashboards, and API keys within a single unified interface.

For sensitive workloads 🕵️, local-only mode disables cloud routing entirely, so all conversations and data remain completely on-device.

🏃‍♂️ Ensuring the local tier is fast enough: This routing only works if the local tier is reliable, so we built Conifer's inference engine from the ground up in Rust. On Apple Silicon, decode speeds reach up to 60% faster than llama.cpp.

uploaded image

On-device operation with Conifer is now the default starting point for inference, rather than a feature limited to hobbyist use.

Asks:

  1. Try Juniper with a one-click download @ conifer.build, or reach out to us at contact@conifer.build to have a model running in under two minutes.
  2. Introduce us to teams with high monthly token spend. If teams are incurring significant API costs and want to optimize unit economics without rebuilding their stack, we would like to connect.
Conifer
Founded:2026
Batch:Summer 2026
Team Size:4
Status:
Active
Location:San Francisco
Primary Partner:Gustaf Alstromer